Verified current Job

Machine Learning Researcher

Job Description: We are seeking a Machine Learning Researcher to join our team and help advance the state of the art in human-centric generative video models. Your work will focus on improving expression control, lip synchronisati...

Job Remote Full source details
BRAHMA United Kingdom Source published Sep 8, 2025 Verified 2 weeks ago
✓ 90% verification score · Source: Brahma Careers · Always confirm final requirements on the original source.
Complete source information imported The available role or programme description, requirements, benefits and source facts were imported from the public official endpoint and formatted for reading.
Machine Learning Researcher opportunity at BRAHMA
EmploymentHiring Now
Work modeRemote / location-flexible
CountryUnited Kingdom

Overview

Job Description: We are seeking a Machine Learning Researcher to join our team and help advance the state of the art in human-centric generative video models. Your work will focus on improving expression control, lip synchronisation, and overall realism in models such as WAN and Hunyuan. You’ll collaborate with a world-class team of researchers and engineers to build systems that can generate lifelike talking-head videos from text, audio, or motion signals—pushing the boundaries of neural rendering and avatar animation. We are hiring remotely across the EMEA region. Key Responsibilities Research and develop cutting-edge generative video models, with a focus on controllable facial expression, head motion, and audio-driven lip synchronisation. Fine-tune and extend video diffusion models such as WAN and Hunyuan for better visual realism and audio-visual alignment. Design robust training pip

Full job description

Full Job Description

Job Description:

We are seeking a Machine Learning Researcher to join our team and help advance the state of the art in human-centric generative video models. Your work will focus on improving expression control, lip synchronisation, and overall realism in models such as WAN and Hunyuan. You’ll collaborate with a world-class team of researchers and engineers to build systems that can generate lifelike talking-head videos from text, audio, or motion signals—pushing the boundaries of neural rendering and avatar animation. We are hiring remotely across the EMEA region.

Key Responsibilities

  • Research and develop cutting-edge generative video models, with a focus on controllable facial expression, head motion, and audio-driven lip synchronisation.
  • Fine-tune and extend video diffusion models such as WAN and Hunyuan for better visual realism and audio-visual alignment.
  • Design robust training pipelines and large-scale video/audio datasets tailored for talking-head synthesis.
  • Explore techniques for controllable expression editing, multi-view consistency, and high-fidelity lip sync from speech or text prompts.
  • Work closely with product and creative teams to ensure models meet quality and production constraints.
  • Stay current with the latest research in video generation, speech-driven animation, and 3D-aware neural rendering.

Must Haves

  • Strong background in machine learning and deep learning, especially in generative models for video, vision, or speech.
  • Hands-on experience with video synthesis tasks such as face reenactment, lip sync, audio-to-video generation, or avatar animation.
  • Proficient in Python and PyTorch; familiar with libraries like MMPose, MediaPipe, DLIB, or image/video generation frameworks.
  • Experience training large models and working with high-resolution audio/video datasets.
  • Deep understanding of architectures such as transformers, diffusion models, GANs and motion representation techniques.
  • Proven ability to work independently and drive research from idea to implementation.
  • Strong problem-solving skills, ability to work autonomously in a remote-first environment.

Nice to Have

  • PhD in Computer Vision, Machine Learning, or a related field, with publications in top-tier conferences (CVPR, ICCV, ICLR, NeurIPS, etc.).
  • Familiarity with or contributions to open-source projects in lip sync, video generation, or 3D face modelling.
  • Experience with real-time inference, model optimisation, or deployment for production applications.
  • Knowledge of adjacent areas like emotion modelling, multimodal learning, or audio-driven animation.
  • Experience working with or adapting models like WAN, Hunyuan or similar.

Tips for this job

Practical Job and Scholarship guidance. These tips do not replace official rules or create new eligibility requirements.

  1. Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
  2. Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
  3. Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
  4. Apply through the original employer or official recruitment destination shown on this page.

Verification notes

Verified from public schema.org JobPosting structured data on the official source page. The complete published description, responsibilities, requirements and benefits were normalized when present; unstated facts were not inferred.

Original authoritative source

Job and Scholarship is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.

Brahma Careers ↗

Browse current Job and Scholarship listings from Brahma Careers →

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books