Mistral AISingaporejust now
SambaNova SystemsPosted 3mo ago
Senior ML Infrastructure Engineer at SambaNova Systems scores 94 out of 100 on AI centrality, which makes it AI Level 4 of 4 (Builds AI) on this board. The level measures how much of the work is AI, not seniority.
AI in this role
SambaNova is a leader in next-generation AI infrastructure, delivering a full-stack inference platform for customers worldwide. At the core of SambaNova's technology is the RDU (Reconfigurable Dataflow Unit) — a chip built on a dataflow architecture rather than the traditional GPU model. Its decode performance is especially strong for agentic workloads like multi-turn agents, code generation, and long-running applications. RDUs are packaged into SambaRack, rack-scale hardware that lets customers deploy state-of-the-art models with better performance, greater energy efficiency, and faster time to value.
About the team
The ML Infrastructure team builds and operates the inference stack that serves SambaNova's models on RDU accelerators, from request scheduling and caching through the public APIs in SambaStack and SambaCloud. We take inference techniques like speculative decoding, constrained decoding, and long-context serving from prototype to production, and own the accuracy infrastructure that gates every feature we ship. We work alongside the ML, compiler, runtime, and product teams, since most of what we build touches all four.
About the role
The Senior Software Engineer, ML Infrastructure will be responsible for designing, building, and operating the production-grade inference infrastructure that powers SambaNova's serving stack on our Reconfigurable Dataflow Unit (RDU) architecture. SambaNova is an inference-first company, and this role sits at the heart of that mission: turning state-of-the-art inference techniques into reliable, high-throughput, low-latency services exposed to customers through SambaStack and SambaCloud. The engineer will own end-to-end systems spanning request scheduling, advanced decoding algorithms, caching layers, API surfaces, and the accuracy infrastructure that keeps the stack trustworthy. This role partners closely with ML, compiler, runtime, and product teams to ship inference features from prototype to production.
Responsibilities
- Design and productionize advanced inference techniques on RDU to optimize for performance and cost. Key areas include speculative decoding, constrained decoding, function/tool calling, prompt caching, and long-context inference.
- Own SambaNova's integration with vLLM and adjacent serving frameworks, adapting them to RDU's architecture.
- Own the public inference API surface exposed through SambaStack and SambaCloud.
- Build and maintain the accuracy verification and regression infrastructure that gates every inference feature shipped to customers.
- Partner with ML, compiler, runtime, and product teams to take inference features from prototype to production.
- Contribute to technical design discussions, code reviews, and architectural decisions as a senior individual contributor.
Required Qualifications
- B.S. in Computer Science, Electrical Engineering, or related field
- 5+ years of industry experience building and operating large-scale distributed systems, ideally in ML serving
- Strong software engineering fundamentals: algorithms, data structures, concurrency, and systems design
- Experience designing and maintaining production services with strict latency, throughput, and availability requirements
- Working knowledge of modern LLM inference techniques and familiarity with open-source serving stacks such as vLLM, TensorRT-LLM, or SGLang
- Proficiency in Python
- Experience collaborating across teams to deliver complex, system-level engineering solutions
Base Salary Range:
Base Pay Range$200,000—$275,000 USDSubmission Guidelines
Please note that in order to be considered an applicant for any position at SambaNova Systems, you must submit an application form for each position for which you believe you are qualified.
EEO Policy
SambaNova Systems is an Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard basis of age (40 and over), color, disability, gender identity, genetic information, marital status, military or veteran status, national origin/ancestry, race, religion, creed, sex (including pregnancy, childbirth, breastfeeding), sexual orientation, and any other applicable status protected by federal, state, or local laws.
Benefits Summary for US-Based, Full-Time Employment Positions
SambaNova offers a competitive total rewards package, including the base salary, plus equity and benefits. We cover 95% premium coverage for employee medical insurance, and 77% premium coverage for dependents and offer a Health Savings Account (HSA) with employer contribution. We also offer Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life, and AD&D insurance plans in addition to Flexible Spending Account (FSA) options like Health Care, Limited Purpose, and Dependent Care. Our library of well-being benefits available to you and your dependents includes a full subscription to Headspace, Gympass+ membership with access to physical gyms, One Medical membership, counseling services with an Employee Assistance Program, and much more.
Prepare for this job
A free preview built only from this posting: what it asks for, what you could be asked in an interview, and how to adjust your resume.
Skills and AI tools this role asks for
Questions you could be asked
- What's a project where you used vLLM hands-on?
- How would you decide a model or AI system is ready to ship?
- Tell me about a time a model underperformed in production. How did you find out, and what did you change?
Adapt your resume
- List these exact terms on your resume: vLLM. An applicant tracking system matches the wording, not the idea.
- Attach one line of real, concrete experience to at least one of them — a tool named with nothing behind it rarely survives a human read.
- Lead with what you built, trained or shipped — this role is judged on the AI system itself, not the tools around it.
Want your resume actually rewritten for this job?
The free preview above is everything we have today. A full resume rewrite is not live yet and has no price set. Join the waitlist and we will email you if we open it.
Get new AI jobs at AI Level 4+ by email
One email a week with the new AI jobs at AI Level 4+, each rated AI Level 1 to 4 for how much AI is in the work. No recruiter spam, unsubscribe in one click.
Free. One email a week. Unsubscribe in one click.
Similar roles
Software Engineering roles rated AI Level 4 at other companies.
SmartsheetRemote · Bangalore, INDIA9h ago
OpenAIRemote · San Francisco$266k-$445k13h ago
Anduril IndustriesWaltham, Massachusetts, United States$191k-$253k13h ago
AmazonTW, TPE, Taipei14h ago
What kind of AI work fits you?
Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.
Find my next stepMore jobs at SambaNova Systems
SambaNova SystemsLondon, England, United Kingdom£120k-£146k17h ago
SambaNova SystemsSan Jose, California, United States$196k-$240k17h ago
SambaNova SystemsSan Jose, California, United States$167k-$204k17h ago







