Skip to main content
← Back to search
R

AI Infra Resident (1-Year Program)

RadixArk

CA$200,000 / year

Palo Alto, CAFull-timeOn-siteIntermediate

Need a reasonable accommodation to apply or interview? Contact us.

JobMinglr uses automated technology to recommend jobs based on profile information and job preferences. Match Score does not determine eligibility for a position, prevent a user from viewing or applying to a job, or make hiring decisions on behalf of an employer.

Description

 

About The Role

RadixArk is launching a full-time, paid, 1-year residency program for aspiring AI infrastructure engineers. You'll rotate across inference, training, kernels, compilers, and cluster infrastructure, working on real production systems alongside senior engineers. This is an opportunity to learn from the team behind SGLang and Miles while owning meaningful projects that impact thousands of developers.

Requirements

  • Bachelor's or Master's degree in Computer Science, Engineering, or related field (recent graduates welcome)

  • Strong foundations in systems programming, ML systems, or high-performance computing

  • Experience with Python and/or C++; exposure to CUDA, JAX, or distributed systems is a plus

  • Demonstrated ability to learn quickly and debug complex technical problems

  • Curiosity, grit, and willingness to dive deep into unfamiliar layers of the stack

  • Desire to grow into a world-class AI infrastructure engineer

  • Strong communication skills and ability to work collaboratively

Responsibilities

  • Join a full-time, paid, 1-year residency focused on AI infrastructure

  • Rotate across inference, training, kernels, compilers, and cluster infrastructure

  • Work on real production systems: LLM serving (SGLang), diffusion inference (Flux, Wan), RL training (Miles), schedulers, and performance tooling

  • Learn to debug failures across the full stack: model → runtime → kernel → hardware

  • Own meaningful projects with close mentorship from senior engineers

  • Contribute to SGLang, Miles, and other open-source projects

  • Document learnings and create guides for the community

  • Present technical deep-dives and share knowledge with the team

  • Participate in code reviews and learn engineering best practices

 

About RadixArk

RadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Founded by AI infrastructure veterans from xAI and NVIDIA, we're on a mission to democratize frontier-level AI infrastructure by building world-class open systems for inference and training. Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.

Compensation

We offer competitive compensation and comprehensive health benefits. The expected annual salary for this position is $200,000 USD.

Equal Opportunity

RadixArk is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.

 

About this role

RadixArk's AI Infrastructure Residency is a structured one-year program designed to take engineers with foundational systems knowledge and rotate them through the full stack of AI infrastructure—from inference and training to kernels, compilers, and cluster management. You'll work on production systems including SGLang (their high-performance LLM serving engine), Flux and Wan (diffusion inference), and Miles (their RL training framework), with close mentorship from senior engineers who've previously optimized systems serving billions of tokens daily and coordinated training across thousands of GPUs.

The role suits recent graduates or early-career engineers with a CS or engineering degree who have solid grounding in systems programming, ML systems, or high-performance computing, plus hands-on experience with Python and/or C++. You should be comfortable debugging across multiple layers of the stack and genuinely curious about learning unfamiliar technical territory. Beyond individual contributions, you'll document your learnings, present technical findings to the team, and participate in code reviews as part of the broader engineering culture.

This is a full-time, in-office position in Palo Alto with a fixed salary of $200,000 and health benefits. It's explicitly framed as a residency and learning opportunity rather than a traditional hire, making it well-suited for someone committed to becoming a world-class AI infrastructure engineer over the next year.

How this employer is doing

average

  • Sponsors Visas

This role's local market on JobMinglr

Pay for this role

$200,000 per year

That is the employer's posted range, and where you land in it is usually decided in one conversation. Before that conversation, run the numbers: analyze an offer for this role.

How JobMinglr reads this job

Every listing here is scored against your profile before you apply: skills overlap, experience level, location and work arrangement, each weighted and explained. You see the score and the reasons, not just a list. How the matching works.