DDN
Staff Engineer
Santa Clara Office
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Engineering
- Seniority
- Lead / management
- Stated salary
- $185,000 – $250,000 per year
- Country
- US
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
DDN is seeking a Staff Engineer to join our Infinia Core team. This is a hands-on technical role combining deep distributed-systems engineering with direct engagement with customers running Infinia in production.
You'll own complex technical escalations end-to-end, from root-cause analysis and incident response through to mitigation, customer communication and product improvements. You'll also help shape engineering best practice, mentor other engineers and drive the use of AI and automation to improve reliability and diagnostics.
If you love the technical depth but want to stay behind the curtain, this probably isn't the right fit but if you want to combine serious engineering with real customer impact, read on.
About Infinia
Infinia is DDN's next-generation, software-defined storage platform, built from the ground up for AI and accelerated computing. It combines separate control and data planes, all-flash performance, sub-millisecond latency and multi-tenancy for demanding enterprise and hyperscale AI and GPU workloads.
What You'll Do
Communicate technical issues clearly to customers, engineers and senior stakeholders, including executive audiences.
Own complex customer escalations from diagnosis through to resolution, mitigation and RCA.
Lead live incident response, war rooms and cross-functional investigations with Engineering, QA and Field teams.
Debug complex distributed-systems, storage and performance issues across the system, protocol and application layers.
Reproduce customer issues and feed findings into product and reliability improvements.
Develop runbooks, troubleshooting guidance and performance-tuning practices.
Act as a technical authority on Infinia internals, mentoring engineers and influencing architectural best practice.
Partner with Field CTOs, Solutions Architects and Sales Engineers on strategic customer issues.
Use AI, automation and observability to improve diagnostics, reliability and MTTR.
Communicate technical issues clearly to customers, engineers and senior stakeholders, including executive audiences.
This position requires participation in an on-call rotation to provide after-hours support as needed.
What You'll Bring
Must-Haves
Significant experience in enterprise storage, distributed systems or cloud infrastructure, with technical leadership at Senior or Staff level.
Deep understanding of file systems and storage technologies, including S3, POSIX, NFS and storage performance.
Strong Linux systems knowledge, including kernel-level troubleshooting and debugging.
Strong coding ability in Python or C++.
Proven ability to diagnose complex issues using tools such as strace, tcpdump and perf.
Genuine interest in working directly with customers and taking ownership of complex problems through to resolution.
Nice-to-Haves
Experience with DDN, VAST, Weka or similar scale-out storage/file systems.
Familiarity with observability platforms such as Prometheus, Grafana, ELK or OpenTelemetry.
Knowledge of replication, consistency models and data integrity mechanisms.
Experience supporting AI/ML, LLM training or other high-performance computing environments.
Experience using AI tools for log analysis, troubleshooting, automated RCA or reducing MTTR.
Similar jobs
- Information Security Engineer, PrincipalBSC · El Dorado Hills, CA, United States; WA, United States; RI, United States; OH, United States; MO, United States; CA, United States; Long Beach, CA, United States; AZ, United States; CO, United States; FL, United States; GA, United States; MD, United States; MN, United States; NV, United States; OR, United States; Lodi, CA, United States; Rancho Cordova, CA, United States; San Diego, CA, United States; AL, United States; IL, United States; VA, United States; WI, United States; TX, United States; NY, United StatesFirst seen today
- Information Security Engineer, PrincipalBSC · El Dorado Hills, CA, United States; CA, United States; Long Beach, CA, United States; Lodi, CA, United States; Oakland, CA, United States; Rancho Cordova, CA, United States; San Diego, CA, United StatesFirst seen today
- Lead QA Engineer - Tax Product DevelopmentBDO USA Experienced · New York, NY, United States; Atlanta, GA, United States; Austin, TX, United States; Baltimore, MD, United States; Boston, MA, United States; Charlotte, NC, United States; Cherry Hill, NJ, United States; Chicago, IL, United States; Cincinnati, OH, United States; Cleveland, OH, United States; Columbus, OH, United States; Dallas, TX, United States; Detroit, MI, United States; Fort Lauderdale, FL, United States; Fort Worth, TX, United States; Grand Rapids, MI, United States; Greenville, SC, United States; Houston, TX, United States; Indianapolis, IN, United States; Jacksonville, FL, United States; Kalamazoo, MI, United States; Melville, NY, United States; Madison, WI, United States; McLean, VA, United States; Memphis, TN, United States; Miami, FL, United States; Milwaukee, WI, United States; Minneapolis, MN, United States; Nashville, TN, United States; Norfolk, VA, United States; Oak Brook, IL, United States; Omaha, NE, United States; Orlando, FL, United States; Philadelphia, PA, United States; Pittsburgh, PA, United States; Potomac, MD, United States; Raleigh, NC, United States; Rosemont, IL, United States; Richmond, VA, United States; St Louis, MO, United States; Stamford, CT, United States; Tampa, FL, United States; Tulsa, OK, United States; Washington, DC, United States; West Palm Beach, FL, United States; Wilmington, DE, United States; Woodbridge, NJ, United StatesFirst seen today
- Senior Staff Embedded Software EngineerFord Global · Long Beach, CA, United StatesFirst seen today
- Principal Software Engineer, Core InfrastructureOracle · Nashville, TN, United StatesFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job