Constellation
Research Engineer
San Francisco
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Mid level
- Stated salary
- $180,000 – $250,000 per year
- Country
- US
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
About Us
Constellation is creating the AI-human translation layer that ensures humanity evolves alongside our technology. Our mission is to leverage AI towards addressing deep and meaningful problems at the core of the human experience: empowering people towards their goals, augmenting our cognition and emotional wellness, and understanding ourselves and each other. Our path forward is to move away from AI that captures human knowledge towards AI that truly understands what it is to be human. We are generating the richest multimodal dataset ever collected to build a new class of foundation models and we're seeking the team of researchers that will build them.
The Role
We're looking for a Research Engineer to sit between our data and our models and make the whole loop faster. You will orchestrate and optimize training runs on long-horizon multimodal sequences, build the pipelines that turn a messy, daily-growing corpus into something our models can learn from, and write the research code (libraries, dataloaders, evaluation harnesses) that lets the rest of the team try ideas quickly and trust what they see. You'll work closely with both our engineering and research teams.
Responsibilities
Orchestrate and optimize training: distributed configuration (DDP/FSDP), mixed precision, checkpointing and recovery on multi-day runs, and profiling to find which of kernel, dataloader, communication or pipeline is actually the bottleneck before touching anything.
Build high-throughput data loading over TB-scale multimodal data: sharding, prefetching, caching and format choices that keep GPUs saturated rather than waiting on I/O.
Wrangle complex data: align and synchronize multi-stream time series, handle changing collection protocols, and own dataset versioning and lineage as the corpus grows.
Build pipelines that derive rich features from raw recordings so every derived dataset is versioned, reproducible from its inputs, and cheap to recompute when upstream data changes.
Write research code that makes research better: clean, well-tested, well-documented libraries for datasets, models, transforms and metrics that the team builds on rather than around.
Make every run traceable: code version, config, dataset version and environment recoverable from any result.
Build evaluation harnesses that run automatically on new checkpoints and surface regressions before anyone goes looking.
Turn research prototypes into repeatable pipelines without flattening the flexibility researchers need to keep exploring.
Qualifications
3+ years building ML systems or research infrastructure, including distributed training runs on 100+ GPUs.
Expert-level Python and deep knowledge of PyTorch internals: DDP and FSDP, mixed precision, gradient accumulation, and the profiling tools to tell a slow model from a starved one.
A track record of building or maintaining research libraries others depend on. Contributions to packages like torch_geometric, torchaudio, torcheeg, torch_brain, neuralsets, or comparable internal tooling are exactly what we're looking for.
Experience with data pipelines over large unstructured and multimodal datasets, and familiarity with columnar and streaming formats (Zarr, Parquet, Arrow, Lance, WebDataset, Vortex) and the tradeoffs between random access and sequential throughput.
Hands-on experience with experiment tracking and dataset versioning tooling (ClearML, Weights & Biases, MLflow, or similar).
Enough research fluency to read a paper, reimplement a component, and tell whether a loss curve is broken.
Debugging range that spans everything from a corrupted shard to a silently wrong collate function to a training run that quietly diverged on day two.
You care about the well-being of the people this technology touches, and you thrive in a high-bandwidth, collaborative environment.
Nice to Have
Custom kernel work in CUDA or Triton, or compiler-level optimization (torch.compile, TensorRT, ONNX).
Experience with time-series or multimodal data where synchronization across streams is a prerequisite for analysis and learning.
Low-latency or edge inference, quantization, or distillation.
Comfort dropping into Rust or C++ when Python is the wrong tool.
Interest in neurotech, mental health, or human-AI interaction.
Benefits
Comprehensive, high-quality health, dental, and vision insurance with premiums fully covered.
A renovated, light-filled office with a quirky layout in the heart of San Francisco's Mission District, surrounded by world-class food and coffee.
Relocation assistance for those joining us in SF and workspace setup stipend.
Monthly team dinners and outings, plus twice-yearly off-sites and retreats.
Fully stocked kitchen (snacks!)
Sponsorship for travel to relevant conferences and regular professional development activities to help you level up.
A minimum of 12 weeks of fully paid parental leave with a “soft-landing” transition back.
Paid time off (PTO)
Listed on hirly, a job board. hirly is not the employer: Constellation is hiring for this role.
Similar jobs
- Research Engineering/Scientist Associate IIUtaustin · UT MAIN CAMPUSFirst seen today
- Research Engineering/ Scientist Associate IUtaustin · PORT ARANSAS, TXFirst seen today
- Research EngineerThomsonreutersFirst seen today
- R&D Research EngineerPpg · USA - Huntsville PlantFirst seen today
- Robotics Research Engineer - Robot Simulation and EvaluationNvidia · US, CA, Santa ClaraFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job