Multimodal ML Engineer
Npv · 🌍 Paris
Sponsorship verdict
No sponsorship evidence yet
No government record and no wording either way. Not a refusal — ask the recruiter.
- No government sponsor record hereThis employer posted directly and does not match a government sponsor register.
- The posting doesn’t mention sponsorshipSilence isn’t a refusal — ask the recruiter before investing much time.
- Can’t check pay against the visa rulesWe don’t have visa salary rules for this country yet.
- Confirmed live todayWhen a source last listed this job as open.
A verdict summarises public evidence; it is not legal advice and never a guarantee — the employer and the immigration authority decide. Sign in to factor in where you can already work.
Or apply yourself on the official page →
Why not apply?
No register record and no sponsorship wording in the posting. Worth asking the employer before investing significant time.
SponsorApply flags time-wasters so your applications go where they can land. These come from the posting's own wording — read the original listing to confirm. See better-fit alternatives →
Sponsor Radar — Npv
This employer posted directly and does not match a government sponsor register.
Past sponsorship or register membership never guarantees sponsorship for this vacancy or for you. Full Sponsor Radar for Npv →
About the role
We're looking for a Multimodal ML Engineer to join White Circle , an AI Safety company building the safety, reliability, and optimization layer for AI systems through natural-language policies it automatically tests, enforces, and improves at scale. Backed by $70M (Series A) from top funds and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, and others, White Circle processes 100M+ API calls monthly and fine-tunes and trains its own LLMs to run faster and cheaper than open or proprietary models. You will • Train and fine-tune large-scale multimodal models (vision-language, audio, speech, video) from scratch and from pretrained checkpoints. • Design experiments, build multimodal data pipelines, and train MoE architectures. • Build alignment pipelines (SFT, DPO, GRPO), optimize for production (quantization, distillation, streaming), and deploy end-to-end. • Define evaluation metrics that actually matter for the product. Requirements • 3+ years training large-scale multimodal models. • Strong PyTorch and distributed training experience (DeepSpeed, FSDP). • Deep familiarity with multimodal architectures – LLaVA, Qwen-VL, InternVL, Audio Flamingo, Whisper, HuBERT, Conformer or similar. • Hands-on RLHF/alignment across modalities (GRPO, DPO, reward modeling). • Both audio and video experience required – sequence modeling for each, plus large-scale dataset curation and production inference optimization. • Relocation to Paris or London (hybrid) required. Bonus • Audio signal processing fundamentals – spectrograms, mel features, noise reduction. • MoE architecture experience. We offer • Competitive salary + equity. • Official employment, visa and relocation help. Find Jobs in France on Arbeitnow