Talent.com
Mercor
Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AIMercor • Midland, Texas, US
Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AI

Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AI

Mercor • Midland, Texas, US
12 days ago
Job type
  • Full-time
  • Remote
Job description
We're looking for experienced machine learning researchers with hands-on experience training and improving deep learning models end-to-end, across vision and language. You'll work on well-scoped empirical open-ended ML research problems. ## Responsibilities - Train image classifiers and generative image models from scratch, and fine-tune open-weight language models. - Get the most out of limited data, compute, and model-size budgets. - Make models robust — to adversarial inputs and to adversarial conversations. - Compress models to meet hard size and latency constraints without sacrificing accuracy. - Diagnose and resolve training issues. ## Requirements We are looking for candidates with strong expertise in one or more of the following areas: **Adversarial Robustness** Experience with: - Adversarial training of image classifiers (e.g. PGD-based training, TRADES). - Evaluating robust accuracy under standard threat models (e.g. L∞ attacks, AutoAttack) and avoiding gradient-masking pitfalls. - Managing the robustness–accuracy trade-off and robust overfitting. **Efficient Computer Vision** Experience with: - Training image classifiers end-to-end, especially for fine-grained recognition (many visually similar classes, few examples per class). - Model compression: quantization, pruning, and knowledge distillation from large teachers into small students. - Deploying models under hard size or latency budgets (on-device, edge, or embedded settings). **Generative Image Modeling** Experience with: - Training image generative models from scratch: diffusion models, GANs, VAEs, or flow-based models. - Iterating against sample-quality metrics such as FID. - Training-efficiency tricks that produce good generators quickly and at small parameter counts. **LLM Post-Training & Behavioral Robustness** Hands-on experience with one or more of: - Supervised fine-tuning and preference optimisation (DPO, RLHF, RLAIF) of open-weight language models, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. - Shaping conversational behaviour over multiple turns: resistance to persuasion and sycophancy, calibrated confidence, and knowing when to accept corrections. - Alignment-style fine-tuning that changes a specific behaviour while preserving general capability. **Multilingual Pre-training** Experience with: - Training multilingual or low-resource-language models from scratch. - Tokenizer design across scripts and typologically diverse languages. - Balancing highly unequal per-language data (sampling temperatures, cross-lingual transfer) in data-constrained regimes. **Additional Areas of Interest** Experience in any of the following is a plus: - Scaling laws and training-efficiency research. - Curriculum learning and data ordering. - Model evaluation: benchmark construction, contamination control, statistically sound comparisons. - Uncertainty estimation and model calibration. - Data augmentation and synthetic data for robustness. ## General Qualifications - 3+ years of machine learning research experience (PhD research counts toward this requirement). - Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. - Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. ## Why Join - Work on cutting-edge machine learning research. - Collaborate with leading AI researchers on challenging, high-impact projects. - Flexible, project-based work with competitive compensation.
Create a job alert for this search

Remote LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness) - AI Trainer ($100-$120 per hour) • Midland, Texas, US

Similar jobs

Trainer, Intermediate

Kodiak Gas Services, LLCMidland, TX, United States
Full-time

Kodiak understands that our most valuable resource is our employees, and in order to provide industry-leading service and runtime, you must attract and retain premier talent.To accomplish this, Kod... Show more

 • Promoted

Remote Work From Home Online - Paid Research Panel - Data Entry Clerk Welcome

ApexFocusGroupMidland, TX, United States
Remote
Full-time +1

Remote Work From Home Online - Paid Research Panel - Data Entry Clerk Welcome.Our company is looking for qualified candidates to take part in paid national and local focus groups, clinical trials, ... Show more

 • Promoted

Sales Manager in Training (100% Remote)

Global EliteMidland, TX, United States
Remote
Full-time

Sales Manager In Training (100% Remote).We're looking for enthusiastic, hard-working, friendly individuals to come work at AO and support a huge network of clients.This position relies on outstandi... Show more

 • Promoted

Remote Resident Medical Specialist (MD/DO)

Turing Midland, Texas
Remote
Full-time
Quick Apply

Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems.Turing s... Show more

 • Promoted

Consumer Insights Analyst

Earn HausMidland, Texas, United States
Part-time

We're currently seeking motivated individuals to participate in online surveys and paid gaming opportunities for well-known national brands.If you're looking for a flexible way to earn extra cash f... Show more

 • Promoted

CS/Sales Agent - Entry Level & REMOTE, work by Appointments

Global EliteMidland, TX, United States
Remote
Full-time

CS/Sales Agent - Entry Level & Remote, Work by Appointments.With consistent growth year over year, we're looking to add more talented individuals to our rapidly growing company.This career allows y... Show more

 • Promoted

Remote Product Tester - Flexible Studies

BuzzTestersMidland, TX
Remote
Full-time

Earn up to $400/week, depending on the number and type of studies you complete.Share honest feedback on products and services from independent brands.Assignments include short online surveys (10–20... Show more

 • Promoted

Remote Chemistry Expert (PhD) - $100+ per hour

Turing Midland, Texas
Remote
Full-time
Quick Apply

Remote contract for PhDs in Chemistry, Chemical Engineering, or related fields.Work on cutting-edge projects with top AI labs while earning up to $100+/hour, fully remote, with flexible weekly hour... Show more

 • Promoted

Managers in Training (Virtual/ Work from home)

Global EliteMidland, TX, United States
Remote
Full-time

Managers In Training (Virtual/ Work From Home).With our company growing faster than ever, we are looking for individuals with amazing people skills to join our 100% remote team.We want everyone who... Show more

 • Promoted

Work from home - Market Research Study

Earn HausMidland, Texas, United States
Remote
Full-time +1

We are urgently seeking people interested in taking market research studies for well known brands.If you are a self-starter, looking for flexible hours throughout the week, this may be for you! Ear... Show more

 • Promoted

Trainer, Technology and Systems

Kodiak Gas Services, LLCMidland, TX, United States
Full-time

Kodiak understands that our most valuable resource is our employees, and in order to provide industry-leading service and runtime, you must attract and retain premier talent.To accomplish this, Kod... Show more

 • Promoted

Market Research Associate - No Experience

Clubshop | USMidland, Texas
Full-time

Discover The System Surveoo That’s Helping Newbies Earn Daily.Turn your free time into daily earnings by answering simple surveys.No product or tech skills needed.Join For Free And Start Earn... Show more