Source-listed Job

Multimodal ML Engineer

We're looking for a Multimodal ML Engineer to join White Circle , an AI Safety company building the safety, reliability, and optimization layer for AI systems through natural-language policies it automatically tests, enforces,...

Job Source description available
Npv Source published Oct 9, 2026 Source retrieved Oct 9, 2026
Source: arbeitnow · A retrieval date records when our system last obtained the source record. It does not guarantee the vacancy is still open or that every detail has been independently checked.
Description from the source The source description is formatted below for discovery. The provider owns the original wording and may change its requirements or close applications.

Overview

We're looking for a Multimodal ML Engineer to join White Circle , an AI Safety company building the safety, reliability, and optimization layer for AI systems through natural-language policies it automatically tests, enforces, and improves at scale. Backed by $70M (Series A) from top funds and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, and others, White Circle processe

Full job description

We're looking for a Multimodal ML Engineer to join White Circle , an AI Safety company building the safety, reliability, and optimization layer for AI systems through natural-language policies it automatically tests, enforces, and improves at scale. Backed by $70M (Series A) from top funds and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, and others, White Circle processes 100M+ API calls monthly and fine-tunes and trains its own LLMs to run faster and cheaper than open or proprietary models. You will Train and fine-tune large-scale multimodal models (vision-language, audio, speech, video) from scratch and from pretrained checkpoints. Design experiments, build multimodal data pipelines, and train MoE architectures. Build alignment pipelines (SFT, DPO, GRPO), optimize for production (quantization, distillation, streaming), and deploy end-to-end. Define evaluation metrics that actually matter for the product. Requirements 3+ years training large-scale multimodal models. Strong PyTorch and distributed training experience (DeepSpeed, FSDP). Deep familiarity with multimodal architectures – LLaVA, Qwen-VL, InternVL, Audio Flamingo, Whisper, HuBERT, Conformer or similar. Hands-on RLHF/alignment across modalities (GRPO, DPO, reward modeling). Both audio and video experience required – sequence modeling for each, plus large-scale dataset curation and production inference optimization. Relocation to Paris or London (hybrid) required. Bonus Audio signal processing fundamentals – spectrograms, mel features, noise reduction. MoE architecture experience. We offer Competitive salary + equity. Official employment, visa and relocation help. Find more English Speaking Jobs in France on Arbeitnow

Tips for this job

Practical JobOpportunity guidance. These tips do not replace official rules or create new eligibility requirements.

  1. Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
  2. Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
  3. Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
  4. Apply through the original employer or official recruitment destination shown on this page.
Original authoritative source

JobOpportunity.info helps you discover and organize source listings. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.

Apply through JobOpportunity →

Browse current JobOpportunity listings from arbeitnow →

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books