Source-listed Remote Opportunities

Search worldwide by keyword, country, category, job type and location. Select any result to review its complete source-backed details.

Clear filters
77,932 source-listed opportunitiesSelect a card to preview the complete details
Applied Research Scientist, LLM Evaluation & Post-Training Source-linked
Innodata Inc.
United States
Job Remote
2w ago
Applied Research Scientist, LLM Evaluation & Post-Training Source-linked
Innodata Inc.
United States
Job Remote
2w ago
Applied Data Scientist, Health AI Evaluation & Datasets Source-linked
Innodata Inc.
United States
Job Remote
2w ago
Applied Data Scientist, Health AI Evaluation & Datasets Source-linked
Innodata Inc.
United States
Job Remote
2w ago
Applied Data Scientist, Finance AI Evaluation & Datasets Source-linked
Innodata Inc.
United States
Job Remote
2w ago
Applied Data Scientist, Finance AI Evaluation & Datasets Source-linked
Innodata Inc.
United States
Job Remote
2w ago
Applied AI Associate - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
Applied AI Analyst - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Workflow Specialist - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Workflow Associate - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Training Specialist - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Solutions Engineer Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Quality Analyst - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI/ML Research Engineer, LLM Post-Training & Evaluation Source-linked
Innodata Inc.
United States
Job Remote
2w ago
AI/ML Research Engineer, LLM Post-Training & Evaluation Source-linked
Innodata Inc.
United States
Job Remote
2w ago
AI Evaluator - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Data Specialist - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Content Specialist - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
AI Content - Flexible Hours Source-linked
Innodata Inc.
Global / location varies
Job Remote
1d ago
Account Executive, Enterprise Sales, Federal Practice Source-linked
Innodata Inc.
Global / location varies
Job Remote
4d ago
Regional Account Executive Source-linked
AMAROK
Global / location varies
Job Remote
3d ago
Regional Account Executive Source-linked
AMAROK
Global / location varies
Job Remote
3d ago
National Account Development Manager Source-linked
AMAROK
Global / location varies
Job Remote
1w ago
Field Service Technician Source-linked
AMAROK
Global / location varies
Job Remote
1w ago
Loading opportunity details…
Job Source-linked

Applied Research Scientist, LLM Evaluation & Post-Training

Innodata Inc.
⌖ INN - Remote - US, United States Remote Posted 2 weeks ago
CountryUnited States
Job typeJob
Work / event modeRemote
DeadlineNot specified

About this job

Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the

Responsibilities & complete job details

Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers. Scope of the Role:  Innodata is expanding its GenAI research capability to advance state-of-the-art evaluation and post-training methods for LLM and multimodal systems. As an Applied Research Scientist, LLM Evaluation & Post-Training, you will lead research and experimentation on how evaluation design, measurement strategies, and feedback signals influence model improvement. This role is ideal for a technically rigorous researcher who is deeply fluent in modern LLM evaluation and post-training, and who can turn research insight into practical methods for customer solutions and internal platform innovation. You will work across human-in-the-loop and AI-augmented workflows, partnering with Language Data Scientists and AI/ML Research Engineers to design and validate evaluation frameworks that drive measurable model gains. The ideal candidate combines strong experimental and statistical judgment with hands-on technical ability and can engage as a peer with research and engineering stakeholders at leading AI companies. What You’ll Own: As an Applied Research Scientist, LLM Evaluation & Post-Training, you will help define the next generation of evaluation-driven model improvement workflows. You will study how different evaluation approaches (human, automated, hybrid) shape model selection and post-training outcomes, and you will design experiments that produce credible, actionable conclusions. Your work may include designing benchmark datasets, developing evaluation taxonomies and protocols, defining metrics and scoring methodologies, analyzing failure modes, and testing how changes in evaluation setup affect downstream fine-tuning results. You will also support customer engagements by bringing scientific rigor to evaluation strategy, methodology review, and technical recommendations. This is a highly collaborative role that sits at the intersection of research, engineering, and language/data operations. Additional responsibilities include (but are not limited to): Define and execute a research agenda focused on LLM evaluation and post-training, especially evaluation-driven model improvement Design rigorous experiments to study how evaluation methodologies impact fine-tuning and post-training outcomes Develop and validate evaluation frameworks for LLM and multimodal systems, including: benchmark/task design scoring methods judge/model-assisted evaluation human evaluation protocols robustness/stress testing Lead research on advanced evaluation domains, including long-context, cross-modal, and dynamic multi-turn evaluations Study the effectiveness and limitations of existing evaluation techniques, and propose improved methodologies with clear validity and scalability tradeoffs Analyze model behavior and failure patterns; generate actionable recommendations for model improvement and evaluation redesign Collaborate with AI/ML Research Engineers to translate research methods into scalable evaluation and post-training pipelines Collaborate with Language Data Scientists to integrate human-in-the-loop and synthetic data/evaluation strategies into research programs Engage with customer technical stakeholders to understand evaluation goals, review methodologies, and provide expert recommendations Contribute to internal benchmark datasets, evaluation frameworks, and reusable research assets Produce high-quality technical documentation, internal research reports, and client-facing materials explaining methods, results, assumptions, and limitations Contribute to thought leadership and best practices in LLM evaluation, post-training, and GenAI quality measurement You’ll Thrive in This Role If You Have: MS/PhD in Computer Science, Machine Learning, Statistics, Applied Mathematics, AI, or a related quantitative scientific field (PhD strongly preferred) 5+ years of relevant experience in applied research / research science in ML/AI, with substantial work in LLMs or foundation models Demonstrated experience with LLM evaluation, benchmarking, alignment, post-training, or model quality research Strong foundation in experimental design, statistical analysis, and scientific reasoning for ML systems Strong coding skills in Python for research experimentation and analysis (e.g., data processing, evaluation pipelines, statistical analysis, visualization) Experience working with modern ML tooling/frameworks (e.g., PyTorch, Hugging Face, JAX/TensorFlow as applicable) sufficient to design and execute model/evaluation experiments Ability to evaluate and compare human and automated evaluation methods, including tradeoffs in cost, reliability, validity, and scalability Experience designing evaluation studies and protocols that are reproducible across datasets, model versions, and evaluation runs Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data scientists, and customer technical counterparts Strong communication skills and ability to present nuanced technical conclusions, assumptions, and limitations clearly The expected salary range for this position is $175,000 – $225,000 USD per year, based on experience, skills, and qualifications.   Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at  https://consumer.ftc.gov/articles/job-scams.   If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at  verifyjoboffer@innodata.com  and consider reporting it to the FTC at  ReportFraud.ftc.gov .

About Innodata Inc.

Innodata Inc. is the organization associated with this source-listed listing. JobOpportunity keeps the original authoritative source attached to every record so applicants can verify final requirements directly.

Source-first verification. JobOpportunity helps you discover and organize opportunities. Always confirm final eligibility, compensation/funding, dates and application instructions on the official source before submitting.

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books