Talent.com
Mercor
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)Mercor • Yonkers, New York, US
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Mercor • Yonkers, New York, US
7 days ago
Job type
  • Full-time
  • Remote
Job description
We're looking for experienced machine learning researchers with hands-on experience training and improving language models end-to-end. You'll work on well-scoped empirical open-ended LLM research problems. * * * ### **Responsibilities** - Train transformer-based language models from scratch and fine-tune open-weight models. - Get the most out of limited data and compute. - Construct training corpora from raw web-scale sources. - Build post-training pipelines. - Diagnose and resolve training issues. * * * ### **Requirements** We are looking for candidates with strong expertise in one or more of the following areas: **Foundation Model Pre-training** Experience with: - Training transformer-based language models from scratch, end-to-end. - Data- and compute-constrained regimes: allocating a fixed budget across model size, tokens, and epochs. - Diagnosing optimisation failures, convergence issues, and training instabilities. **Pre-training Data** Experience with: - Corpus construction from raw web crawls and other large unfiltered sources. - Data filtering, deduplication, quality classification, and mixture/ordering optimisation. - Measuring data interventions rigorously. **LLM Post-Training** Hands-on experience with one or more of: - Supervised fine-tuning, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. - Preference optimisation (DPO, RLHF, RLAIF) and reward modelling / human-preference prediction. - Alignment fine-tuning: shaping refusal behaviour, truthfulness, and unbiased reasoning while preserving general capability. - Fine-tuning for narrow, verifiable domains (math, code, games, structured prediction) where outputs can be checked programmatically. **Additional Areas of Interest** Experience in any of the following is a plus: - Scaling laws and training-efficiency research. - Curriculum learning and data ordering. - LLM evaluation: benchmark construction, contamination control, statistically sound comparisons. - Reinforcement learning for language models. - Model alignment and AI safety. **General Qualifications** - 3+ years of machine learning research experience (PhD research counts toward this requirement). - Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. - Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. * * * ### **Why Join** - Work on cutting-edge foundation model research. - Collaborate with leading AI researchers on challenging, high-impact projects. - Flexible, project-based work with competitive compensation.
Create a job alert for this search

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour) • Yonkers, New York, US

Similar jobs

AI Workflow Developer (Remote) - Media & Entertainment

Tandym GroupEnglewood Cliffs, NJ, United States
Remote
Full-time

AI Workflow Developer (Remote) - Media & EntertainmentGet AI-powered advice on this job and more exclusive features.A recognized media organization offers a remote opportunity for an innovative... Show more

 • Promoted

Market Research Associate - No Experience

Clubshop | USStony Point, New York
Full-time

Discover The System Surveoo That’s Helping Newbies Earn Daily.Turn your free time into daily earnings by answering simple surveys.No product or tech skills needed.Join For Free And Start Earn... Show more

Remote | Machine Learning & NLP Research Specialist — $75–$105/hour

24-MAGNew York, New York, United States
$75.00 hourly
Remote
Full-time +1
Quick Apply

We are sharing a specialised part-time consulting opportunity for US-based machine learning and natural language processing professionals with hands-on experience in Python, model training and eval... Show more

Flexible remote AI work. Your schedule. Paid weekly, straight to your bank account.

Meridian.aiPeekskill, NY, US
Remote
Full-time

Review and label digital content including text, images, and documents.Every task you complete helps improve how technology interprets information and performs in practical settings.Detail-oriented... Show more

 • Promoted

Earn Side Money at Home Testing Products

Product Review JobsHAVERSTRAW, NY, United States
Full-time

Compensation: Varies per assignment.Location: Remote (USA) Company: ProductReviewJobs Thank you for your interest in becoming a Paid Product Tester.This opportunity is for completing market res... Show more

 • Promoted

Senior AI Scientist: LLM Eval & Prompt Design (Remote)

LexisNexisNew York City, NY, United States
Remote
Full-time

A leader in legal technology is seeking a Data Scientist specializing in AI Evaluation and Prompt Engineering.This role involves applying data science techniques to enhance AI-driven legal research... Show more

 • Promoted

Remote | ML Model Development & MLOps Expert — $95–$135/hour

24-MAGNew York, New York, United States
$95.00 hourly
Remote
Full-time +1
Quick Apply

We are sharing a specialised part-time consulting opportunity for professionals experienced in machine learning engineering, model development, Python, ML frameworks, model deployment, MLOps, and s... Show more

Remote Investment Analyst - AI Trainer ($50-$60 / hour)

Data AnnotationClifton, NJ, United States
Remote
Full-time +1

We are looking for a finance professional to join our team to train AI models.You will measure the progress of these AI chatbots, evaluate their logic, and solve problems to improve the quality of ... Show more

 • Promoted

Remote Product Tester - Flexible Studies

BuzzTestersHaverstraw, NY
Remote
Full-time

Earn up to $400/week, depending on the number and type of studies you complete.Share honest feedback on products and services from independent brands.Assignments include short online surveys (10–20... Show more

 • Promoted

AI internship - Remote

The Upskill AcademyNew York City, NY, United States
Remote
Internship

We are a dynamic and innovative company in the software development industry, as an AI Intern.We're a small but passionate team of 1-10 professionals dedicated to creating impactful digital solutio... Show more

 • Promoted

Sr Machine Learning / AI Engineer - Remote

Nava Software Solutions LLCJersey City, NJ, United States
Remote
Full-time

NAVA Software solutions is looking for a Senior Machine Learning / AI EngineerDetails :Senior Machine Learning / AI EngineerLocation :RemoteDuration :1yearAbout the ProjectThe team has developed an... Show more

 • Promoted

Retail Sr Data Science Analytics Remote

IT Search CorpNorwood, NJ, United States
Remote
Full-time

Benefits :401(k) matchingBonus based on performanceCompetitive salaryDental insuranceHealth insuranceStock options planVision insuranceAssociate Director, Data Science & AnalyticsRetail backgro... Show more