Base Compute
Founding ML Engineer
Melbourne
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.6M live jobs from 190,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Data & ML
- Seniority
- Mid level
- Country
- AU
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
About Us
Base Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device.
We’re building the infrastructure for the next generation of on-device AI, from silicon-level optimizations to distributed inference systems.
We’re working on hard problems at the intersection of inference efficiency, model intelligence and autonomous research.
The Role
We’re looking for a Founding ML Engineer to work at the frontier of on-device AI. This role is for someone who lives at the intersection of systems engineering and machine learning, turning state-of-the-art research into hyper-optimized, production-ready infrastructure.
You’ll have significant ownership over our entire inference stack and direct influence on the technical bets the company makes.
What You’ll Work On
Inference engine development: Building and scaling our custom inference engine, handling everything from weight loading and KV-cache management to efficient request scheduling
Cross-platform silicon optimization : Writing and tuning custom kernels and leveraging hardware-specific instructions to squeeze maximum performance out of diverse architectures, including Apple Silicon, NVIDIA, AMD, Snapdragon, and other edge platforms
Systems architecture: Developing robust, low-latency serving runtimes in C++ to manage model routing, continuous batching, and novel decoding strategies under strict thermal and memory constraints
Performance profiling: Identifying and eliminating bottlenecks across the entire stack, from memory bandwidth ceilings to kernel interleaving
What We’re Looking For
3+ years of experience in ML engineering or systems programming (Rust, C/C++), with a strong track record of building performance-critical software
Expertise in GPU programming and hardware optimization across various platforms (CUDA, ROCm, Metal, Triton, or similar)
Solid understanding of modern LLM architectures, including parsing formats and implementing optimization techniques (quantization, speculative decoding, etc.)
A strong sense of ownership and autonomy: the ability to take ambiguous architectural challenges and drive them from research translation directly into production-ready infrastructure
Good communication: the ability to explain complex architectural decisions simply, give honest feedback and document systems cleanly
Nice-to-haves:
Familiarity with ML compilers (torch.compile, custom operators)
Experience with low-precision inference (INT8/FP8/FP4)
Knowledge of Edge LLMOps
What We Offer
Founding team equity and strong base salary
Direct influence on technical direction: your ideas will shape the roadmap
Work on genuinely hard problems that haven't been solved yet
Small team, fast iteration, low bureaucracy
Location
The team is based in Melbourne and Berlin and works in-person from the office most days. We require strong written and spoken English, since the team collaborates across time zones.
Listed on hirly, a job board. hirly is not the employer: Base Compute is hiring for this role.
Similar jobs
- AI/ML EngineerDowner · Melbourne, VIC, AustraliaFirst seen 7d ago
- ML EngineerKogan · South Melbourne, VictoriaFirst seen 12d ago
- Generative AI Data Scientist & ML EngineerRbc · TORONTO, Ontario, CanadaFirst seen today
- AI/ML Engineer - TS/SCI w/polyGdit · USA VA HerndonFirst seen today
- AI / ML Engineer (Agent Developer)Dxctechnology · TUN - ARIANAFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job