Member of Technical Staff — Developer Technology
RadixArkPalo Alto, California, United States · Posted 1 month agoDescription
About the Role
Key Responsibilities
-
Accelerate AI workloads. Profile and optimize GPU performance for real production workloads on current and next-generation hardware, root-causing bottlenecks from kernels to distributed multi-node systems.
-
Go deep in one or two focus areas. The team collectively covers the full stack; each engineer specializes in one or two tracks:
-
Inference performance: engine tuning, benchmarking, long-context and multi-turn optimization, parallelism strategy, production debugging
-
Kernels and model/hardware enablement: custom CUDA/ROCm/Triton kernels, low-precision quantization, day-0 support for new models on new silicon
-
Speculative decoding: draft-model training, acceptance-rate tuning, cross-platform kernel adaptation
-
Training systems: RL post-training with Miles, FP8 training, elasticity, long-rollout and long-context efficiency
-
-
Partner directly with the ecosystem. Turn ambiguous, high-stakes problems from expert engineers at our key partners into concrete wins, clear technical guidance, and reproducible cookbooks.
-
Enhance SGLang and Miles. Feed user-driven improvements back into our open-source systems and roadmap, so every win compounds across the ecosystem.
Qualifications
-
4+ years of experience in GPU systems, LLM infrastructure, or performance engineering.
-
Strong profiling and debugging skills: able to root-cause performance and correctness issues across the stack.
-
Hands-on GPU programming experience in at least one of CUDA, ROCm, or Triton, and willingness to work across platforms.
-
Strong programming skills in Python plus C++ or CUDA.
-
Comfortable making progress on hard, ambiguous problems with little context to start from, and fast to ramp into unfamiliar systems, codebases, and domains.
-
Ability to translate ambiguous asks into clear technical plans, verified cookbooks, and actionable recommendations, and to communicate credibly with expert engineering audiences.
-
Deep familiarity with LLM inference internals: distributed serving, parallelism, routing, KV-cache management, scheduling.
-
Experience with low-precision quantization and inference/training (FP8, INT8/INT4; NVFP4 or MXFP4 a strong plus).
-
Experience writing and optimizing custom GPU kernels.
-
Practical familiarity with speculative decoding methods such as Eagle, DFlash, or DSpark.
-
Working knowledge of large-scale distributed training: pre-training, SFT, RL post-training, elasticity, long-context workloads.
-
Experience optimizing across both NVIDIA and AMD platforms.
-
Hands-on experience with SGLang, Miles, vLLM, TensorRT-LLM, Megatron, or comparable frameworks; contributions to open-source AI/ML projects.
About RadixArk
RadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Founded by AI infrastructure veterans from xAI and NVIDIA, we're on a mission to democratize frontier-level AI infrastructure by building world-class open systems for inference and training. Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.
Compensation
Depending on background, skills, and experience, the expected annual salary range for this position is $200,000 - $400,000 USD + equity.
Equal Opportunity
RadixArk is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.
Similar jobs
Est. 300,000 USD
About the Role RadixArk is seeking a Developer Advocate to build and engage our technical community around SGLang, Miles, and our open source infrastructure. SGLang already has 30K+ GitHub stars and serves billions of to…
Est. 300,000 USD
About the Role RadixArk is looking for a Member of Technical Staff — Backend/API Platform Engineer to build the API layer, control plane, and platform services that power SGLang and Miles in production. You'll design and…
Est. 300,000 USD
About the Role RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference. You will work on the core systems that serve frontier models at scale, optimizing performance, la…
Est. 300,000 USD
About the Role RadixArk is hiring a Member of Technical Staff — Performance in Palo Alto, CA — someone who can push LLM inference and training systems to the limit across real production workloads. You’ll work on the per…
Est. 300,000 USD
About the Role RadixArk is seeking experienced product-focused engineers to join our team in building the developer-facing surfaces of our inference and training infrastructure. As a Member of Technical Staff — Product,…
Est. 300,000 USD
About the Role RadixArk is seeking a Member of Technical Staff — Kernel / Compiler / Communication to push the limits of performance for frontier AI systems. You will work at the lowest layers of the stack — kernels, run…
Est. 300,000 USD
About the Role RadixArk is seeking a Member of Technical Staff - Inference-Multi-Hardware to push the limits of performance for frontier AI systems. Most performance engineering assumes a single vendor's stack. This role…
Est. 165,000 USD
About the Role As a Technical Program Manager at RadixArk, you'll drive the execution of complex, cross-functional programs across our inference and training infrastructure. You'll partner closely with Product Management…
Est. 300,000 USD
About the Role RadixArk is looking for a Member of Technical Staff Cluster Infrastructure to architect and scale the core compute platform that powers frontier-level AI training and inference. You will design and operate…
Est. 300,000 USD
About the Role RadixArk is seeking a Member of Technical Staff — Inference-Multimodal & Diffusion to advance the frontier of generative modeling. You will work on cutting-edge diffusion and flow-based models for imag…
Est. 140,000 USD
About The Role RadixArk is launching a full-time, paid, 1-year residency program for aspiring AI infrastructure engineers. You'll rotate across inference, training, kernels, compilers, and cluster infrastructure, working…
Est. 300,000 USD
About the Role As a Member of Technical Staff, Training, you will design, build, and operate the distributed systems behind large-scale model post-training — spanning training, inference, and orchestration, with a focus…
Est. 300,000 USD
About the Role RadixArk is hiring a Member of Technical Staff — CI / Infrastructure to own the infrastructure that keeps SGLang moving. Our CI system runs 300+ GPU tests across NVIDIA, AMD, Intel, and Ascend hardware poo…
Est. 155,000 USD
About the Role RadixArk is seeking a Product Marketing Manager to own how SGLang, Miles, and our open source infrastructure are positioned and perceived across the market. SGLang already has 20K+ GitHub stars and serves…
Est. 144,000 USD
Key Responsibilities Product Strategy & Roadmap Define, prioritize, and drive the product roadmap for inference and training infrastructure. Stay ahead of AI trends, including new model architectures, hardware optimi…
Est. 300,000 USD
About the Role RadixArk is looking for a Member of Technical Staff — TPU Systems to build high-performance inference and training systems using JAX, XLA, and Pallas. You'll push model workloads to their limits on TPU har…
Est. 150,000 USD
About the Role We're looking for a Head of Business Development to build the BD function at RadixArk from the ground up. The BD team is the institutional memory of this company — maintaining active relationships across e…
Est. 159,000 USD
About the Role We're looking for a hands-on Talent Operations Specialist to build and run the machinery behind talent and people ops as we scale. This isn't a traditional HR generalist role - it's for someone who treats…
Est. 115,000 USD
About the Role RadixArk builds the open-source AI infrastructure behind SGLang and Miles, used by developers and enterprises around the world. We're looking for a visual designer to join our design team and to give our b…
Est. 289,800 USD
XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and…
Est. 235,200 USD
XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and…
Est. 310,000 USD
SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organi…
Est. 328,650 USD
XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and…
Est. 328,650 USD
XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and…
Est. 239,000 USD
Your Impact at LILA The AI Research team is tackling one of the most exciting, open problems in AI: training LLMs to run long-horizon scientific discovery tasks. Our approach spans the full post-training stack - from SFT…
Est. 232,000 USD
Your Impact at LILA The Staff/Principal DevOps Engineer - AI Inference will drive the design, implementation, and optimization of infrastructure purpose-built for serving machine learning models at scale. This role bridg…
Est. 205,000 USD
Gradial is the marketing operations system of work that helps marketers and creatives move from idea to execution faster. Our platform orchestrates across martech stacks, workflows, and people to automate marketing execu…
Est. 310,000 USD
At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructu…
Est. 237,500 USD
Who We Are Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying AI systems—designed to take ideas from research to production with…
Est. 229,500 USD
ABOUT US E-commerce got real-time data infrastructure decades ago. Physical stores still have not. RADAR is changing that.RADAR is building the data infrastructure layer for the physical world, starting with retail. Our…