Talent.com
Mercor
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)Mercor • Laredo, Texas, US
Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour)

Mercor • Laredo, Texas, US
Hace 5 días
Tipo de contrato
  • A tiempo completo
  • Teletrabajo
Descripción del trabajo
We're looking for experienced machine learning researchers with hands-on experience training and improving language models end-to-end. You'll work on well-scoped empirical open-ended LLM research problems. * * * ### **Responsibilities** - Train transformer-based language models from scratch and fine-tune open-weight models. - Get the most out of limited data and compute. - Construct training corpora from raw web-scale sources. - Build post-training pipelines. - Diagnose and resolve training issues. * * * ### **Requirements** We are looking for candidates with strong expertise in one or more of the following areas: **Foundation Model Pre-training** Experience with: - Training transformer-based language models from scratch, end-to-end. - Data- and compute-constrained regimes: allocating a fixed budget across model size, tokens, and epochs. - Diagnosing optimisation failures, convergence issues, and training instabilities. **Pre-training Data** Experience with: - Corpus construction from raw web crawls and other large unfiltered sources. - Data filtering, deduplication, quality classification, and mixture/ordering optimisation. - Measuring data interventions rigorously. **LLM Post-Training** Hands-on experience with one or more of: - Supervised fine-tuning, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. - Preference optimisation (DPO, RLHF, RLAIF) and reward modelling / human-preference prediction. - Alignment fine-tuning: shaping refusal behaviour, truthfulness, and unbiased reasoning while preserving general capability. - Fine-tuning for narrow, verifiable domains (math, code, games, structured prediction) where outputs can be checked programmatically. **Additional Areas of Interest** Experience in any of the following is a plus: - Scaling laws and training-efficiency research. - Curriculum learning and data ordering. - LLM evaluation: benchmark construction, contamination control, statistically sound comparisons. - Reinforcement learning for language models. - Model alignment and AI safety. **General Qualifications** - 3+ years of machine learning research experience (PhD research counts toward this requirement). - Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. - Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. * * * ### **Why Join** - Work on cutting-edge foundation model research. - Collaborate with leading AI researchers on challenging, high-impact projects. - Flexible, project-based work with competitive compensation.
Crear una alerta de empleo para esta búsqueda

Remote LLM Research Scientist (Pre-training & Post-Training) - AI Trainer ($100-$120 per hour) • Laredo, Texas, US

Ofertas similares

Remote Computer Science Expert (PhD)

Micro1Laredo, Texas, US
80,00 US$ por hora
Teletrabajo
A tiempo completo

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Mostrar más

 • Oferta promocionada

Remote Biology Expert

Micro1Laredo, Texas, US
Teletrabajo
A tiempo completo

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Mostrar más

 • Oferta promocionada

Remote Physics Expert

Micro1Laredo, Texas, US
Teletrabajo
A tiempo completo

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Mostrar más

 • Oferta promocionada

Remote Data Scientist

Micro1Laredo, Texas, US
65,00 US$ por hora
Teletrabajo
A tiempo completo

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Mostrar más

 • Oferta promocionada

Earn up to $3,000/study Work at Home (Remote) Data Entry Position

MaxionLaredo, TX, US
Teletrabajo
A tiempo completo +2

Want to make extra money on YOUR schedule? Join our exclusive list of research study participants and .Perfect for anyone seeking remote, part-time, or temporary work, these opportunities require .... Mostrar más

 • Oferta promocionada

Remote Market Research Assistant – Earn Up to $50/Task

Occupons QuebecLaredo, TX, United States
Teletrabajo
A tiempo completo

Compensation: $45–$90 per completed task.The Consumer Insights Division of our organization is expanding its nationwide research team.We are seeking dependable individuals to support ongoing studie... Mostrar más

 • Oferta promocionada

Flexible remote AI work. Your schedule. Paid weekly, straight to your bank account.

Meridian.aiLaredo, TX, US
Teletrabajo
A tiempo completo

Review and label digital content including text, images, and documents.Every task you complete helps improve how technology interprets information and performs in practical settings.Detail-oriented... Mostrar más

 • Oferta promocionada

Remote Product Tester - Flexible Studies

BuzzTestersLaredo, TX
Teletrabajo
A tiempo completo

Earn up to $400/week, depending on the number and type of studies you complete.Share honest feedback on products and services from independent brands.Assignments include short online surveys (10–20... Mostrar más