Talent.com
Mercor
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)Mercor • Laredo, Texas, US
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Mercor • Laredo, Texas, US
5 days ago
Job type
  • Full-time
  • Remote
Job description
We're looking for experienced machine learning researchers with hands-on experience training and improving language models end-to-end. You'll work on well-scoped empirical open-ended LLM research problems. * * * ### **Responsibilities** - Train transformer-based language models from scratch and fine-tune open-weight models. - Get the most out of limited data and compute. - Construct training corpora from raw web-scale sources. - Build post-training pipelines. - Diagnose and resolve training issues. * * * ### **Requirements** We are looking for candidates with strong expertise in one or more of the following areas: **Foundation Model Pre-training** Experience with: - Training transformer-based language models from scratch, end-to-end. - Data- and compute-constrained regimes: allocating a fixed budget across model size, tokens, and epochs. - Diagnosing optimisation failures, convergence issues, and training instabilities. **Pre-training Data** Experience with: - Corpus construction from raw web crawls and other large unfiltered sources. - Data filtering, deduplication, quality classification, and mixture/ordering optimisation. - Measuring data interventions rigorously. **LLM Post-Training** Hands-on experience with one or more of: - Supervised fine-tuning, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. - Preference optimisation (DPO, RLHF, RLAIF) and reward modelling / human-preference prediction. - Alignment fine-tuning: shaping refusal behaviour, truthfulness, and unbiased reasoning while preserving general capability. - Fine-tuning for narrow, verifiable domains (math, code, games, structured prediction) where outputs can be checked programmatically. **Additional Areas of Interest** Experience in any of the following is a plus: - Scaling laws and training-efficiency research. - Curriculum learning and data ordering. - LLM evaluation: benchmark construction, contamination control, statistically sound comparisons. - Reinforcement learning for language models. - Model alignment and AI safety. **General Qualifications** - 3+ years of machine learning research experience (PhD research counts toward this requirement). - Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. - Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. * * * ### **Why Join** - Work on cutting-edge foundation model research. - Collaborate with leading AI researchers on challenging, high-impact projects. - Flexible, project-based work with competitive compensation.
Create a job alert for this search

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour) • Laredo, Texas, US

Similar jobs

Remote Computer Science Expert (PhD)

Micro1Laredo, Texas, US
$80.00 hourly
Remote
Full-time

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Show more

 • Promoted

Remote Biology Expert

Micro1Laredo, Texas, US
Remote
Full-time

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Show more

 • Promoted

Remote Physics Expert

Micro1Laredo, Texas, US
Remote
Full-time

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Show more

 • Promoted

Remote Data Scientist

Micro1Laredo, Texas, US
$65.00 hourly
Remote
Full-time

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Show more

 • Promoted

Earn up to $3,000/study Work at Home (Remote) Data Entry Position

MaxionLaredo, TX, US
Remote
Full-time +2

Want to make extra money on YOUR schedule? Join our exclusive list of research study participants and .Perfect for anyone seeking remote, part-time, or temporary work, these opportunities require .... Show more

 • Promoted

Remote Market Research Assistant – Earn Up to $50/Task

Occupons QuebecLaredo, TX, United States
Remote
Full-time

Compensation: $45–$90 per completed task.The Consumer Insights Division of our organization is expanding its nationwide research team.We are seeking dependable individuals to support ongoing studie... Show more

 • Promoted

Flexible remote AI work. Your schedule. Paid weekly, straight to your bank account.

Meridian.aiLaredo, TX, US
Remote
Full-time

Review and label digital content including text, images, and documents.Every task you complete helps improve how technology interprets information and performs in practical settings.Detail-oriented... Show more

 • Promoted

Remote Product Tester - Flexible Studies

BuzzTestersLaredo, TX
Remote
Full-time

Earn up to $400/week, depending on the number and type of studies you complete.Share honest feedback on products and services from independent brands.Assignments include short online surveys (10–20... Show more