Talent.com

Program evaluation Empleos en Sunnyvale ca

Crear una alerta de empleo para esta búsqueda

Program evaluation • sunnyvale ca

Última actualización: hace 9 horas

AIML - Sr Machine Learning Engineering Manager, Evaluation

AppleCupertino, CA, United States
A tiempo completo

AIML - Sr Machine Learning Engineering Manager, Evaluation.Apple's AIML Evaluation team builds the systems and methodologies that measure and improve the quality of foundation models and agentic ex... Mostrar más

Program Supervisor - Adult Program

Gardner Health ServicesSan Jose, CA, United States
98.700,00 US$–110.000,00 US$ anual
A tiempo completo

Gardner Health Services is currently recruiting to fill a Program Supervisor position.This is a full-time position in the Adult/Older Adult Division of the Behavioral Health Department based out of... Mostrar más

Program Manager

Glint Tech Solutions LLCSanta Clara, CA, USA
A tiempo completo
Quick Apply

We are seeking an experienced Program Manager to support customer programs and drive cross-functional coordination within the Flex PCB / electronics manufacturing industry.This role serves as the p... Mostrar más

Program Director

E-SolutionsSan Jose, CA, United States
A tiempo completo

Develop and oversee strategic program roadmaps, ensuring alignment with Blue Planets business objectives and customer expectations.Lead the creation of comprehensive, cross-functional execution pla... Mostrar más

Program Aide

LifeMovesMountain View, CA, United States
A tiempo completo +2

HomeKey Mountain View - Mountain View, CA 94043.Schedule: Sunday-Thursday; 8:00AM-4:30PM.LifeMoves is the largest and most effective provider of housing and services for neighbors experiencing home... Mostrar más

Program Manager, Learning Evaluation

TeslaPalo Alto, CA, United States
A tiempo completo

We're looking for a Program Manager with proven experience designing and delivering impactful learning interventions, establishing robust learning infrastructure, and using data-driven insights to ... Mostrar más

Program Manager Evaluation & Learning

Catholic Charities of Santa Clara CountySan Jose, CA, United States
A tiempo completo

Program Manager Of Evaluation & Learning.The Program Manager of Evaluation & Learning is a core member of CCSCC's Organizational Effectiveness team, responsible for leading CCSCC's outcomes measure... Mostrar más

Manager, AV Evaluation

General MotorsSunnyvale, CA, United States
A tiempo completo

Work Arrangement: This role is categorized as Remote/hybrid.Remote: This role is based remotely but if you live within a 50-mile radius of [Austin, Detroit, Warren, Milford, Sunnyvale, CA], you are... Mostrar más

Program Manager

Micross ComponentsMilpitas, CA, United States
A tiempo completo +1

We are seeking a Program Manager that will be responsible for the following:.The STS Program Manager will handle a wide variety of low volume, hi-reliability products to schedule, to specification.... Mostrar más

Program Manager

Amphenol TCSSanta Clara, CA, US
100.000,00 US$–189.260,00 US$ anual
A tiempo completo

Amphenol Communications Solutions (ACS), a division of Amphenol Corporation, is a world leader in interconnect solutions for Communications, Mobile, RF, Optics, and Commercial electronics markets.A... Mostrar más

Technical Program Manager, Digital Test $ Evaluation

Advanced Automation CorporationMountain View, CA, US
180.000,00 US$–215.000,00 US$ anual
Teletrabajo
A tiempo completo
Quick Apply

The incumbent will serve as a Technical Program Manager and Subject Matter Expert (SME) in Digital Test & Evaluation (T&E), with responsibility for helping translate traditional defense T&a... Mostrar más

 • Nueva oferta

Program Manager | (Program management) | Hybrid |

SamprasoftSunnyvale, CA, United States
A tiempo completo

Coordinates projects and ensures company resources are utilized appropriately.Compiles project status reports, coordinates project schedules, manages project meetings, and identifies and resolves t... Mostrar más

Head of AI Quality, Assurance and Evaluation

AgilentSanta Clara, CA, United States
A tiempo completo

Agilent Enterprise AI Solutions Developer.Agilent inspires and supports discoveries that advance the quality of life by providing life science, diagnostic, and applied market laboratories worldwide... Mostrar más

Staff Tech Lead Manager, Machine Learning, Simulator Evaluation

WaymoMountain View, CA, United States
A tiempo completo

Staff Tech Lead Manager, Machine Learning, Simulator Evaluation.Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver.Since its start as the Googl... Mostrar más

Humanities Evaluation Specialist - Remote

YO AI LabsPalo Alto, CA, USA
Teletrabajo
A tiempo completo
Quick Apply

Humanities Evaluation Specialist.Humanities Evaluation Specialists.AI training project involving research, critical analysis, question development, and evaluation.No prior AI experience is required... Mostrar más

Senior Project Manager, Clinical Risk Evaluation

AbbottSanta Clara, CA, United States
A tiempo completo

Clinical Risk Evaluation (CRE) Senior Scientist/Program Manager.Abbott is a global healthcare leader that helps people live more fully at all stages of life.Our portfolio of life-changing technolog... Mostrar más

Director, Simulation and Evaluation - Autonomous Driving

Bosch GroupSunnyvale, California, United States
A tiempo completo

As the Director for Simulation and Evaluation, you will sit at the center of the Global AI Backbone, architecting the multi-level simulation ecosystems required to train, evaluate and validate next... Mostrar más

Director, Evaluation - Graduate School of Business

StanfordStanford, CA, United States
162.809,00 US$–197.109,00 US$ anual
A tiempo completo

Residing in Silicon Valley, the heart of innovation, Stanford Graduate School of Business has built a global reputation based on its immersive and innovative management programs.We provide students... Mostrar más

Senior Research Manager, World Model Evaluation

NVIDIASanta Clara, CA, United States
A tiempo completo

At NVIDIA, we're not just building the future, we're generating it! Our world model team is pushing the boundaries of multimodal AI, robotics, and world foundation models for Physical AI.We are loo... Mostrar más

Program Management - Program Manager I

Apex SystemsSunnyvale, CA, United States
49,00 US$–59,00 US$ por hora
A tiempo completo

Program Management - Program Manager I.We are seeking a Program Manager to support sustaining product operations for consumer hardware products post-launch.This role is responsible for managing rec... Mostrar más

AIML - Sr Machine Learning Engineering Manager, Evaluation

AIML - Sr Machine Learning Engineering Manager, Evaluation

AppleCupertino, CA, United States
Hace más de 30 días
Tipo de contrato
  • A tiempo completo
Descripción del trabajo

AIML - Sr Machine Learning Engineering Manager, Evaluation

Apple's AIML Evaluation team builds the systems and methodologies that measure and improve the quality of foundation models and agentic experiences. We are looking for a senior, hands-on Machine Learning Engineering Manager to lead a small team working at the intersection of model evaluation, agent optimization, and data generation. In this role, you will help define how evaluation closes the loop with model and product development, turning observed quality gaps into targeted improvements to prompts, agent harnesses, datasets, and models. You will combine technical depth with people leadership. You should be comfortable moving from research papers and experimental results to production-quality ML pipelines, while mentoring engineers and aligning teams around a clear technical direction. Your work will span Apple Foundation Models and product teams, with the goal of creating repeatable evaluation and refinement loops that improve the quality of Apple intelligence experiences.

As a Senior Machine Learning Engineering Manager in AIML Evaluation, you will lead the technical strategy and execution for agent evaluation and automatic optimization. You will own systems that evaluate foundation models and agents, diagnose failure modes, and use those signals to drive automated prompt, context, tool, rubric, and agent-harness improvements. You will also help establish the interfaces between evaluation and post-training so that high-value failures can be converted into targeted data, environments, reward signals, and measurable model improvements. This is a hands-on leadership role. You will prototype new approaches, participate in architecture and code reviews, design experiments, and help your team translate emerging research into scalable evaluation and optimization pipelines. You will partner closely with Apple Foundation Models, product engineering teams, and other AIML groups to build an evaluation flywheel that connects real product behavior with model and agent refinement. You will also work across the organization to advance synthetic data generation for both evaluation and post-training, with strong attention to data quality, representativeness, privacy, and reproducibility.

Responsibilities

  • Architects and builds scalable evaluation systems for foundation models and agents, including benchmarks, LLM-based evaluators, simulation environments, trajectory analysis, and regression testing.
  • Establishes an end-to-end evaluation flywheel with Apple Foundation Models and product teams that connects observed failures to diagnosis, targeted refinement, post-training, and measurable quality improvement.
  • Leads, mentors, and grows a small team of machine learning engineers while remaining deeply involved in technical design, experimentation, implementation, and review.
  • Defines the technical strategy and roadmap for automatic prompt, context, tool, rubric, and agent-harness optimization for agentic development and model evaluation.
  • Develops methods that convert evaluation findings into actionable model-improvement signals, including targeted datasets, synthetic trajectories, reward or preference signals, and optimization objectives.
  • Partners across AIML to design and scale synthetic data generation pipelines for evaluation and post-training.
  • Applies and adapts recent research in LLM and agent evaluation, automatic optimization, LLM-as-judge, reward modeling, test-time search, and post-training to production-quality workflows.

Minimum Qualifications

  • 8+ years of professional experience in machine learning, applied research, or software engineering, including experience building production ML systems or large-scale experimentation platforms.
  • 3+ years of technical leadership experience, including direct people management of machine learning or software engineers and a demonstrated ability to mentor and grow strong technical talent.
  • Master's or PhD in Computer Science, Machine Learning, Artificial Intelligence, or a related technical field.
  • Strong hands-on programming and software engineering skills, particularly in Python, with experience building reliable ML pipelines using modern machine learning or deep learning frameworks.
  • Deep experience with large language models or agentic systems, including evaluation of multi-turn behavior, tool use, planning, reasoning, or other action-taking workflows.
  • Experience building automated evaluation methods such as LLM-based judges, rubrics, reward models, simulation-based evaluation, or scalable benchmark infrastructure.
  • Experience with at least one model or agent refinement area such as automatic prompt or context optimization, post-training, preference optimization, reinforcement learning, or agent-harness optimization.
  • Excellent communication and collaboration skills, with demonstrated ability to align research, engineering, and product teams around ambiguous technical problems.

Preferred Qualifications

  • Track record of applying recent machine learning research to production systems or high-impact product development.
  • Experience with automatic prompt or context optimization, agent-search methods, evaluator optimization, or multi-objective optimization for agentic systems.
  • Experience generating and evaluating synthetic datasets, tool-use trajectories, or multi-turn agent interactions, including methods for filtering, deduplication, diversity, and quality control.
  • Experience designing evaluation systems that combine offline benchmarks, simulation, human evaluation, and product- or usage-derived signals.
  • Experience with privacy-preserving or on-device machine learning and evaluation.
  • Demonstrated ability to influence technical strategy across organizational boundaries and communicate complex model-quality tradeoffs to senior technical leaders.

Pay & Benefits

At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $237,600 and $401,700, and your base pay will depend on your skills, qualifications, experience, and location. Apple employees also have the opportunity to become an Apple shareholder through participation in Apple's discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple's Employee Stock Purchase Plan. You'll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program. Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant At Apple, we believe accessibility is a fundamental human right. You'll find that idea reflected in everything here in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong. Learn about accessibility in Apple's workplace Learn about reasonable accommodations for job applicants Apple accepts applications to this posting on an ongoing basis.