Verified current Job

AI Researcher - Robot Learning

About 1X We're building humanoid robots that work in home - doing the chores, handling the tasks, and giving people their time back. Simple, but it's not. To do this right, we have to solve robotics, AI, manufacturing - at the sam...

Job Full source details
1X Careers San Carlos, California Source published Jun 17, 2026 Verified 7 hours ago
✓ 92% verification score · Source: 1X Careers · Always confirm final requirements on the original source.
Complete source information imported The available role or programme description, requirements, benefits and source facts were imported from the public official endpoint and formatted for reading.
EmploymentFull Time
Published compensationUSD 200000.00 – 300000.001 YEAR
DepartmentArtificial Intelligence

Overview

About 1X We're building humanoid robots that work in home - doing the chores, handling the tasks, and giving people their time back. Simple, but it's not. To do this right, we have to solve robotics, AI, manufacturing - at the same time, at scale, in a form factor that has to be safe enough to live with your family. If you're inspired by this, you'll thrive here. We've been at this since 2014 and we're at the point where the hard problems are behind us and the hard work is in front of us. NEO is our flagship - a home robot designed to move, learn, and operate in the real world alongside real people. We're not demoing it - we're shipping it. We're excited to meet you, if this excites you. If you've spent your career working on problems that matter and want to see them actually reach the world - this is that moment. We're scaling, we're hiring with intention, and we need people who want to

Full job description

Description

Full Job Description

About 1X

We're building humanoid robots that work in home - doing the chores, handling the tasks, and giving people their time back. Simple, but it's not.

To do this right, we have to solve robotics, AI, manufacturing - at the same time, at scale, in a form factor that has to be safe enough to live with your family. If you're inspired by this, you'll thrive here. We've been at this since 2014 and we're at the point where the hard problems are behind us and the hard work is in front of us.

NEO is our flagship - a home robot designed to move, learn, and operate in the real world alongside real people. We're not demoing it - we're shipping it. We're excited to meet you, if this excites you.

If you've spent your career working on problems that matter and want to see them actually reach the world - this is that moment. We're scaling, we're hiring with intention, and we need people who want to build something that will genuinely change how humans spend their time - safely creating abundance for all.

About the Team

The Motion team enables NEO to move through and interact with the world. We build the perception NEO needs to understand its surroundings and locomote through any environment, and the control that lets it use its whole body to accomplish real tasks - crawling, bracing, climbing, lifting with more than just its arms. Because NEO operates around people, safe and compliant motion is a design constraint on everything we ship, not a feature layered on top.

Your Charter

Own the full pipeline from RL algorithm development through production deployment: training NEO on manipulation and locomotion tasks in simulation, closing the sim-to-real gap, and shipping policies that work reliably in real-world environments. This is critical-path work: the range of tasks NEO can perform safely and reliably is a direct function of the quality of RL policies your team ships. You will collaborate closely with the hardware and world model teams, and measure your impact by what NEO can do in the field.

Key Outcomes

  • Train and deploy RL policies for manipulation and locomotion tasks that perform reliably in real-world home environments measured by field task success rates, not just simulation benchmarks

  • Advance sim-to-real transfer techniques that measurably narrow the gap between simulation training performance and real-world policy behavior, enabling faster iteration cycles

  • Build training and evaluation infrastructure that lets the team iterate on policies faster with standardized benchmarks, automated regression detection, and clear connections between training metrics and field performance

  • Partner with hardware, controls, data, and QA teams to ship RL-trained skills to production customer sites, owning the handoff from research to deployment

Key Competencies

  • Sim-to-real practitioner closing the sim-to-real gap on physical systems; understands domain randomization, reward shaping, and the engineering required to make simulated policies transfer reliably to real hardware

  • RL algorithms depth with strong foundation in RL algorithms (PPO, SAC, TD-MPC, or similar); can choose the right approach for the task and modify or extend it when standard methods fall short

  • Full-stack ownership owning data engineering, model architecture, and deployment; treats a promising training curve as the beginning of the job, not the end

  • Effective cross-functional partner working closely with hardware, controls, QA, and data teams to translate RL research into deployed robot skills, and communicates technical constraints clearly across disciplines

Minimum Requirements

  • Strong Python and/or C++ with experience in large codebases and build tools (Bazel or equivalent)

  • Proficiency with PyTorch for RL policy training and experimentation

  • Hands-on experience with simulation platforms (Isaac Sim, MuJoCo, or equivalent) for policy training at scale

  • Demonstrated experience training RL policies for manipulation or locomotion tasks, including addressing the sim-to-real gap on physical hardware

Preferred Skills

  • Experience with model-based RL or world-model-guided policy learning that leverages predictive models to improve sample efficiency

  • Familiarity with imitation learning or learning from demonstration (behavior cloning, GAIL, IQL) as a complement or bootstrap to RL

  • Experience deploying RL-trained policies to physical robots in production environments, including monitoring, failure analysis, and iterative improvement

  • Background in legged locomotion, dexterous manipulation, or contact-rich control for physical systems

Compensation

  • Salary Range: $200,000 - $300,000 + Equity

Benefits

  • Comprehensive medical, dental, and vision coverage

  • Generous paid time off, company holidays, and parental leave

  • 401(k) plan with company match (100% on the first 3% of contributions, 50% on the next 2%)

  • Flexible Spending Accounts (FSA) and Health Savings Accounts (HSA) options

  • Commuter benefits (transit and parking)

  • Short-term and long-term disability, and life insurance

  • Employee Assistance Program (EAP) for mental health, financial, and personal support

  • Onsite snacks and catered lunches

Equal Opportunity Employer

1X is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, gender, gender identity or expression, sexual orientation, national origin, ancestry, citizenship, age, marital status, medical condition, genetic information, disability, military or veteran status, justice system impact, or any other characteristic protected under applicable federal, state, or local law.

Tips for this job

Practical Job and Scholarship guidance. These tips do not replace official rules or create new eligibility requirements.

  1. Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
  2. Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
  3. Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
  4. Apply through the original employer or official recruitment destination shown on this page.

Verification notes

Discovered directly from the employer’s public Ashby Job Postings API. The complete public role content and compensation metadata were normalized into safe candidate-facing sections. Complete structured details were extracted from the public authoritative source while preserving the original application link.

Original authoritative source

Job and Scholarship is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.

1X Careers ↗

Browse current Job and Scholarship listings from 1X Careers →

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books