Talent.com
Mercor
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)Mercor • Laredo, Texas, US
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Mercor • Laredo, Texas, US
5 days ago
Job type
  • Full-time
  • Remote
Job description
We're looking for experienced machine learning researchers with hands-on experience training and improving language models end-to-end. You'll work on well-scoped empirical open-ended LLM research problems. * * * ### **Responsibilities** - Train transformer-based language models from scratch and fine-tune open-weight models. - Get the most out of limited data and compute. - Construct training corpora from raw web-scale sources. - Build post-training pipelines. - Diagnose and resolve training issues. * * * ### **Requirements** We are looking for candidates with strong expertise in one or more of the following areas: **Foundation Model Pre-training** Experience with: - Training transformer-based language models from scratch, end-to-end. - Data- and compute-constrained regimes: allocating a fixed budget across model size, tokens, and epochs. - Diagnosing optimisation failures, convergence issues, and training instabilities. **Pre-training Data** Experience with: - Corpus construction from raw web crawls and other large unfiltered sources. - Data filtering, deduplication, quality classification, and mixture/ordering optimisation. - Measuring data interventions rigorously. **LLM Post-Training** Hands-on experience with one or more of: - Supervised fine-tuning, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. - Preference optimisation (DPO, RLHF, RLAIF) and reward modelling / human-preference prediction. - Alignment fine-tuning: shaping refusal behaviour, truthfulness, and unbiased reasoning while preserving general capability. - Fine-tuning for narrow, verifiable domains (math, code, games, structured prediction) where outputs can be checked programmatically. **Additional Areas of Interest** Experience in any of the following is a plus: - Scaling laws and training-efficiency research. - Curriculum learning and data ordering. - LLM evaluation: benchmark construction, contamination control, statistically sound comparisons. - Reinforcement learning for language models. - Model alignment and AI safety. **General Qualifications** - 3+ years of machine learning research experience (PhD research counts toward this requirement). - Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. - Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. * * * ### **Why Join** - Work on cutting-edge foundation model research. - Collaborate with leading AI researchers on challenging, high-impact projects. - Flexible, project-based work with competitive compensation.
Create a job alert for this search

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour) • Laredo, Texas, US

Similar jobs

Online Survey Participant: Work Remote and Earn Up To $25 Per Survey

Earn HausLaredo, TX, US
Remote
Full-time +1

Looking for people to participate in taking online surveys for Fortune 500 brands.All you need to do is complete online surveys by sharing your opinion.You will help influence brand decisions on se... Show more

 • Promoted

Survey Taker: Earn up to $25 per survey (Remote)

Earn HausLaredo, TX, US
Remote
Full-time +1

Looking for people to participate in taking online surveys for Fortune 500 brands.All you need to do is complete online surveys by sharing your opinion.You will help influence brand decisions on se... Show more

 • Promoted

Trainee Entry Level Home Daily, Weekly, and Every 2 weeks

SkillConnect LLCLAREDO, TX, USA
Full-time

Recent CDL-A Grads – Start Your Trucking Career with Paid Training! * *Just got your CDL-A, We’ve got the perfect opportunity to launch your career!* We’re looking for motivated, safety-mind... Show more

Trainee Entry Level CDL-A Truck Driver Training

Class A SolutionsLAREDO, TX, USA
Full-time

Entry Level CDL-A Truck Driver Training  - Dedicated Accounts - We do not want to offer you just another truck driving job, but a long-term, prosperous career in the transportation indus... Show more

CDL Trainee Entry Level Home Daily, Weekly, or Every 2 weeks

SkillConnect LLCLAREDO, TX, USA
Full-time

Recent CDL-A Grads – Start Your Trucking Career with Paid Training!  CDL Trainee Entry Level Home Daily, Weekly, or Every 2 weeks * *Just got your CDL-A, We’ve got the perfect opportunity to... Show more

Earn up to $3,000/study Work at Home (Remote) Data Entry Position

MaxionLaredo, TX, US
Remote
Full-time +2

Want to make extra money on YOUR schedule? Join our exclusive list of research study participants and .Perfect for anyone seeking remote, part-time, or temporary work, these opportunities require .... Show more

 • Promoted

Consumer Insights Analyst

Earn HausLaredo, Texas, United States
Full-time +1

We are urgently seeking people interested in taking market research studies for well known brands.If you are a self-starter, looking for flexible hours throughout the week, this may be for you! Ear... Show more

 • Promoted

Remote Market Research Assistant – Earn Up to $50/Task

Occupons QuebecLaredo, TX, United States
Remote
Full-time

Compensation: $45–$90 per completed task.The Consumer Insights Division of our organization is expanding its nationwide research team.We are seeking dependable individuals to support ongoing studie... Show more