SambaNova Systems
Senior Principal Runtime Engineer
San Jose, California, United States
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →hirly's read of this role
- Seniority
- Lead / management
- Country
- US
- Work mode
- Remote-friendly
- First seen by hirly
- 2 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
SambaNova is a leader in next-generation AI infrastructure, delivering a full-stack inference platform for customers worldwide. At the core of SambaNova's technology is the RDU (Reconfigurable Dataflow Unit) — a chip built on a dataflow architecture rather than the traditional GPU model. Its decode performance is especially strong for agentic workloads like multi-turn agents, code generation, and long-running applications. RDUs are packaged into SambaRack, rack-scale hardware that lets customers deploy state-of-the-art models with better performance, greater energy efficiency, and faster time to value.
About the role
The Runtime Team builds a high-performance, distributed and scalable software execution environment for SambaNova SambaStack and Cloud platforms to support data-flow applications such as ML training and inference and HPC applications. We are searching for a software engineer who will work on all parts of the runtime stacks, supporting AI, ML, and scientific applications in high-performance distributed systems. You will participate in building, testing and deploying next-generation high-performance compute systems for AI applications at scale.
Responsibilities
- div]:bg-bg-000/50 [&_pre>div]:border-0.5 [&_pre>div]:border-border-400 [&_.ignore-pre-bg>div]:bg-transparent [&_.standard-markdown_:is(p,blockquote,h1,h2,h3,h4,h5,h6)]:pl-2 [&_.standard-markdown_:is(p,blockquote,ul,ol,h1,h2,h3,h4,h5,h6)]:pr-8 [&_.progressive-markdown_:is(p,blockquote,h1,h2,h3,h4,h5,h6)]:pl-2 [&_.progressive-markdown_:is(p,blockquote,ul,ol,h1,h2,h3,h4,h5,h6)]:pr-8">
- _*]:min-w-0 gap-3 [&_>_*:last-child]:mb-0 print:block print:[&_>_*_+_*]:mt-3 standard-markdown">
New and enhanced features to support high-performance, scalable ML inference and training applications
Drivers and kernels for next generation silicon
Eliminate networking bottlenecks to enable high performance distributed systems
User-space libraries for high performance and high utilization of HW resources
User-facing tools (analysis, job and HW management, profiling, debugging, etc) for Datascale systems
Cross-functional collaboration including Hardware, ML Application, Compiler, and DevOps
Required Qualifications
B.S. in Computer Science, Engineering, or related field
5+ years of software engineering experience, often with emphasis on distributed systems, networking, or cloud infrastructure
Proven experience building, testing, and tuning software for distributed, high-performance systems. In-depth knowledge of user libraries, and runtime stacks
Solid understanding of Switching and Routing and ability to configure and debug switches and routers for maximum application performance
Significant experience with RDMA and RoCE networking stacks, such as RDMA based verbs and congestion management
Hands-on experience with kernel drivers and system-level software that directly interfaces with hardware
Expertise in designing and optimizing systems that handle massive parallel workloads, including machine learning training and inference tasks that involve billions of operations per second
Deep understanding of hardware-software interaction, including registers, device memory management, and the intricacies of accelerator design. Experience working with ASIC accelerators is highly desirable
Familiarity with distributed systems architecture, including networking, communication protocols, and the challenges of scaling compute resources efficiently
Hands-on experience with software development tools such as Git, Jenkins, and Jira, with an ability to drive automation and continuous integration efforts
Ability to work at the intersection of hardware and software, designing systems that optimize both performance and reliability
Preferred Qualifications
Experience designing or working closely with custom hardware accelerators (ASICs, FPGAs, etc.) and understanding low-level interactions
Knowledge of SDN, DPDK, SONiC, or modern networking frameworks and background with distributed communication libraries (e.g., UCX, MPI, NCCL).
Familiarity with deploying high-performance systems in distributed, cloud, or data center environment
- Submission Guidelines
- Please note that in order to be considered an applicant for any position at SambaNova Systems, you must submit an application form for each position for which you believe you are qualified.
- EEO Policy
- SambaNova Systems is an Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard basis of age (40 and over), color, disability, gender identity, genetic information, marital status, military or veteran status, national origin/ancestry, race, religion, creed, sex (including pregnancy, childbirth, breastfeeding), sexual orientation, and any other applicable status protected by federal, state, or local laws.
- Benefits Summary for US-Based, Full-Time Employment Positions
- SambaNova offers a competitive total rewards package, including the base salary, plus equity and benefits. We cover 95% premium coverage for employee medical insurance, and 77% premium coverage for dependents and offer a Health Savings Account (HSA) with employer contribution. We also offer Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life, and AD&D insurance plans in addition to Flexible Spending Account (FSA) options like Health Care, Limited Purpose, and Dependent Care. Our library of well-being benefits available to you and your dependents includes a full subscription to Headspace, Gympass+ membership with access to physical gyms, One Medical membership, counseling services with an Employee Assistance Program, and much more.
Similar jobs
- Senior Manager, Kubernetes Runtime EngineeringNvidia · 2 LocationsFirst seen today
- Runtime EngineerMatX · Mountain View (HQ) or RemoteFirst seen 2d ago
- Runtime Engineer - KernelizeKernelize · Remote early stage startup that prefers employees located in the USA or contractors in other locations.First seen 22d agoremote
- Runtime EngineerLemurianlabs · Santa Clara, CA - Toronto, CanadaFirst seen 22d agoremote
- AI Runtime EngineerEnchargeai36 · U.S., Canada, Germany, NorwayFirst seen 26d ago
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job