This role has closed. Npv has taken the posting down.
hirly last saw it live on 8 September 2026. See similar open roles below, or browse all jobs in Paris.
Npv
Multimodal ML Engineer
Paris
Similar open jobs
- ML Engineer - F/HOnepoint · Paris, FranceFirst seen 2d ago
- Stage 2027 - AI/ML Engineer - 50% Client 50% R&DEkimetrics · ParisFirst seen 5d ago
- ML Engineer - PowerKpler · ParisFirst seen 6d ago
- Multimodal ML EngineerWhitecircle · ParisFirst seen 18d ago
- AI/ML Engineer, ParisAirapps · ParisFirst seen 27d agoremote
- Product ML EngineerPhotoroom · ParisFirst seen 28d ago
- ML engineer confirmé (H/F)IpponFirst seen 8d ago
- ML Engineer H/FIpponFirst seen 8d ago
- Senior AI & ML engineerEkimetrics · ParisFirst seen 19d ago
- Staff ML EngineerYubo · Paris - Full RemoteFirst seen 27d agoremote
- AI/ML EngineerBah · Chantilly, VAFirst seen today
- AI and ML Engineer and Data ScientistBah · Fort Belvoir, VAFirst seen today
- AI / ML EngineerAccenture · IndiaFirst seen today
- AI / ML EngineerAccenture · IndiaFirst seen today
- AI/ML EngineerCorning · Pune, Maharashtra, IndiaFirst seen today
hirly's read of this role
- Role family
- Data & ML
- Seniority
- Mid level
- Stated salary
- $100,000 – $250,000 per year
- Country
- FR
- Work mode
- Remote-friendly
- First seen by hirly
- 6 Sept 2026
Derived automatically from the posting.
the posting
We're looking for a Multimodal ML Engineer to join White Circle , an AI Safety company building the policy enforcement and optimization layer for AI systems. Backed by $11M from senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, and DeepMind, White Circle processes 100M+ API calls monthly and runs its own LLMs in production.
You will
Train and fine-tune large-scale multimodal models (vision-language, audio, speech, video) from scratch and from pretrained checkpoints.
Design experiments, build multimodal data pipelines, and train MoE architectures.
Build alignment pipelines (SFT, DPO, GRPO), optimize for production (quantization, distillation, streaming), and deploy end-to-end.
Define evaluation metrics that actually matter for the product.
Requirements
3+ years training large-scale multimodal models.
Strong PyTorch and distributed training experience (DeepSpeed, FSDP).
Deep familiarity with multimodal architectures – LLaVA, Qwen-VL, InternVL, Audio Flamingo, Whisper, HuBERT, Conformer or similar.
Hands-on RLHF/alignment across modalities (GRPO, DPO, reward modeling).
Both audio and video experience required – sequence modeling for each, plus large-scale dataset curation and production inference optimization.
Relocation to Paris or London (hybrid) required.
Bonus
Audio signal processing fundamentals – spectrograms, mel features, noise reduction.
MoE architecture experience.
We offer
$100k–$250k/year salary + equity; higher figures can be negotiated.
Official employment, visa and relocation help.
Compensation: $100K – $250K • Higher figures and equity are negotiable
• $100K – $250K • Higher figures and equity are negotiable
Find Jobs in France on Arbeitnow