Talent.com
Mercor
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)Mercor • Georgetown, Texas, US
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Mercor • Georgetown, Texas, US
9 days ago
Job type
  • Full-time
  • Remote
Job description
We're looking for experienced machine learning researchers with hands-on experience training and improving language models end-to-end. You'll work on well-scoped empirical open-ended LLM research problems. * * * ### **Responsibilities** - Train transformer-based language models from scratch and fine-tune open-weight models. - Get the most out of limited data and compute. - Construct training corpora from raw web-scale sources. - Build post-training pipelines. - Diagnose and resolve training issues. * * * ### **Requirements** We are looking for candidates with strong expertise in one or more of the following areas: **Foundation Model Pre-training** Experience with: - Training transformer-based language models from scratch, end-to-end. - Data- and compute-constrained regimes: allocating a fixed budget across model size, tokens, and epochs. - Diagnosing optimisation failures, convergence issues, and training instabilities. **Pre-training Data** Experience with: - Corpus construction from raw web crawls and other large unfiltered sources. - Data filtering, deduplication, quality classification, and mixture/ordering optimisation. - Measuring data interventions rigorously. **LLM Post-Training** Hands-on experience with one or more of: - Supervised fine-tuning, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. - Preference optimisation (DPO, RLHF, RLAIF) and reward modelling / human-preference prediction. - Alignment fine-tuning: shaping refusal behaviour, truthfulness, and unbiased reasoning while preserving general capability. - Fine-tuning for narrow, verifiable domains (math, code, games, structured prediction) where outputs can be checked programmatically. **Additional Areas of Interest** Experience in any of the following is a plus: - Scaling laws and training-efficiency research. - Curriculum learning and data ordering. - LLM evaluation: benchmark construction, contamination control, statistically sound comparisons. - Reinforcement learning for language models. - Model alignment and AI safety. **General Qualifications** - 3+ years of machine learning research experience (PhD research counts toward this requirement). - Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. - Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. * * * ### **Why Join** - Work on cutting-edge foundation model research. - Collaborate with leading AI researchers on challenging, high-impact projects. - Flexible, project-based work with competitive compensation.
Create a job alert for this search

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour) • Georgetown, Texas, US

Similar jobs

Mira Safety is hiring: Remote Creative Internship - Video Production & Motion Gr

MIRA SafetyCedar Park, TX, United States
Remote
Internship

Job DescriptionJob DescriptionRemote Creative Internship - Video Production & Motion GraphicsLocation :Remote (must have stable internet and personal equipment)Commitment :20 hours / week (flex... Show more

 • Promoted

Flexible remote AI work. Your schedule. Paid weekly, straight to your bank account.

Meridian.aiGeorgetown, TX, US
Remote
Full-time

Review and label digital content including text, images, and documents.Every task you complete helps improve how technology interprets information and performs in practical settings.Detail-oriented... Show more

 • Promoted

Work from home - Market Research Study

Earn HausRound Rock, Texas, United States
Remote
Full-time +1

We are urgently seeking people interested in taking market research studies for well known brands.If you are a self-starter, looking for flexible hours throughout the week, this may be for you! Ear... Show more

 • Promoted

Remote Mathematics Expert (PhD) - $100+ per hour

Turing Cedar Park, Texas
Remote
Full-time
Quick Apply

Remote contract for PhDs in Mathematics, Statistics, or related fields.Work on cutting-edge projects with top AI labs while earning $100+/hour, fully remote, with flexible weekly hours.Help fine-tu... Show more

 • Promoted

Sales and Onboarding Specialist - AI Tech - USA remote

Affinda GroupRound Rock, TX, United States
Remote
Full-time

The Opportunity Affinda Group is transforming how enterprises work with their data and documents through cutting-edge AI technology.We are trusted by 1,750customers across 75countries, with 100%ann... Show more

 • Promoted

Remote Physics Expert (PhD) - $100+ per hour

Turing Brushy Creek, Texas
Remote
Full-time
Quick Apply

Remote contract for PhDs in Physics, Applied Physics, or related fields.Work on cutting-edge projects with top AI labs while earning $100+/hour, fully remote, with flexible weekly hours.Help fine-t... Show more