Talent.com
Mercor
Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AIMercor • Hurst, Texas, US
Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AI

Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AI

Mercor • Hurst, Texas, US
26 days ago
Job type
  • Full-time
  • Remote
Job description

We're looking for experienced machine learning researchers with hands-on experience training and improving deep learning models end-to-end, across vision and language. You'll work on well-scoped empirical open-ended ML research problems. ## Responsibilities - Train image classifiers and generative image models from scratch, and fine-tune open-weight language models. - Get the most out of limited data, compute, and model-size budgets. - Make models robust — to adversarial inputs and to adversarial conversations. - Compress models to meet hard size and latency constraints without sacrificing accuracy. - Diagnose and resolve training issues. ## Requirements We are looking for candidates with strong expertise in one or more of the following areas: **Adversarial Robustness** Experience with: - Adversarial training of image classifiers (e.g. PGD-based training, TRADES). - Evaluating robust accuracy under standard threat models (e.g. L∞ attacks, AutoAttack) and avoiding gradient-masking pitfalls. - Managing the robustness–accuracy trade-off and robust overfitting. **Efficient Computer Vision** Experience with: - Training image classifiers end-to-end, especially for fine-grained recognition (many visually similar classes, few examples per class). - Model compression: quantization, pruning, and knowledge distillation from large teachers into small students. - Deploying models under hard size or latency budgets (on-device, edge, or embedded settings). **Generative Image Modeling** Experience with: - Training image generative models from scratch: diffusion models, GANs, VAEs, or flow-based models. - Iterating against sample-quality metrics such as FID. - Training-efficiency tricks that produce good generators quickly and at small parameter counts. **LLM Post-Training & Behavioral Robustness** Hands-on experience with one or more of: - Supervised fine-tuning and preference optimisation (DPO, RLHF, RLAIF) of open-weight language models, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. - Shaping conversational behaviour over multiple turns: resistance to persuasion and sycophancy, calibrated confidence, and knowing when to accept corrections. - Alignment-style fine-tuning that changes a specific behaviour while preserving general capability. **Multilingual Pre-training** Experience with: - Training multilingual or low-resource-language models from scratch. - Tokenizer design across scripts and typologically diverse languages. - Balancing highly unequal per-language data (sampling temperatures, cross-lingual transfer) in data-constrained regimes. **Additional Areas of Interest** Experience in any of the following is a plus: - Scaling laws and training-efficiency research. - Curriculum learning and data ordering. - Model evaluation: benchmark construction, contamination control, statistically sound comparisons. - Uncertainty estimation and model calibration. - Data augmentation and synthetic data for robustness. ## General Qualifications - 3+ years of machine learning research experience (PhD research counts toward this requirement). - Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. - Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. ## Why Join - Work on cutting-edge machine learning research. - Collaborate with leading AI researchers on challenging, high-impact projects. - Flexible, project-based work with competitive compensation.

Create a job alert for this search

Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AI Trainer ($100-$120 per hour) • Hurst, Texas, US

Similar jobs

Remote Consumer Research Participant

GL IncCleburne, Texas
$15.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

WFH - Product Assessments - $25-$45 per hour (No Experience)

Online Consumer Panels AmericaCleburne, Texas, US
$25.00–$45.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Data Scientist

TradeJobsWorkforce76244 Fort Worth, TX, US
Full-time

Data Scientist Job Duties: Formulates and leads guided, multifaceted analytic studies again... Show more

 • Promoted

Remote Work – Product Assessments - $25-$45 per hour (No Experience)

Online Consumer Panels AmericaJoshua, Texas, US
$25.00–$45.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Lead Scientist, Data Science (AI Driven Pricing Optimization) - Remote

XPOFort Worth, TX, United States
Remote
Full-time

What you'll need to succeed as Lead Scientist, Data Science at XPOMinimum Qualifications:Bachelor's degree or equivalent related work or military experience 4 years of experience in data science, o... Show more

 • Promoted

Remote Research Panel Participant

GL IncCleburne, Texas
$15.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Online Consumer Research Participant

GL IncCleburne, Texas
$15.00 hourly
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Remote Product Tester - Flexible Studies

BuzzTestersCovington, TX
Remote
Full-time

Earn up to $400/week, depending on the number and type of studies you complete.Share honest feedback on products and services from independent brands.Assignments include short online surveys (10–20... Show more

 • Promoted

Remote Data Entry - Product Support - $45 per hour

GL Inc.Cleburne, Texas
$45.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Online Work-From-Home - $45 per hour - No Experience

Online Consumer Panels AmericaJoshua, Texas, US
$45.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Research Assistant

TradeJobsWorkForce76182 North Richland Hills, TX, US
Full-time

Research Assistant Job Duties: Generates hypotheses and designs and performs experiments to test th... Show more

 • Promoted

Medical Billing and Coding - Entry Level Training Program

Dreambound Inc.Cleburne, Texas, US
Full-time

Note : This is an educational program, not a job.Successful completion of the program does not guarantee employment but will equip you with valuable skills for the healthcare job market.Looking to ... Show more

 • Promoted

Consumer Research Associate (Remote)

GL IncCleburne, Texas
$15.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Remote Physics Expert

Micro1Cleburne, Texas, US
Remote
Full-time

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Show more

 • Promoted

Online Market Research Participant

GL IncCleburne, Texas
$15.00 hourly
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Remote Product Tester - $25-45 per hour

Online Consumer Panels AmericaJoshua, Texas, US
$25.00–$45.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Product Tester – Remote

OCPAJoshua, Texas, US
$18.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home in the UK to fulfill upcoming contracts with local and international companies.We guarantee 15-25 hours per week with an hourly pay of between £18/hr.Th... Show more

 • Promoted

Online Remote Work

Online Consumer Panels AmericaJoshua, Texas, US
$25.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Remote Online Product Support - No Experience

GLOCPACleburne, Texas
$15.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home nationwide in the US to fulfill upcoming contracts with national and international companies.We guarantee 15-25 hours per week with an hourly pay of bet... Show more

 • Promoted

Remote Product Tester - No Experience

OCPAJoshua, Texas, US
$18.00 hourly
Remote
Part-time +1

Product Testers are wanted to work from home in the UK to fulfill upcoming contracts with local and international companies.We guarantee 15-25 hours per week with an hourly pay of between £18/hr.Th... Show more