Talent.com
Jobleads-US
Staff GenAI Inference Engineer: Optimize LLM Serving LatencyJobleads-US • San Francisco, CA, United States
No se aceptan más aplicaciones
Staff GenAI Inference Engineer: Optimize LLM Serving Latency

Staff GenAI Inference Engineer: Optimize LLM Serving Latency

Jobleads-US • San Francisco, CA, United States
Hace 28 días
Salario
190.900,00 US$ anual
Tipo de contrato
  • A tiempo completo
Descripción del trabajo

A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong software engineering background and a proven ability to collaborate with researchers and drive architectural decisions. Competitive compensation is offered, with a salary range of $190,900 to $232,800 USD. #J-18808-Ljbffr

Crear una alerta de empleo para esta búsqueda

Staff GenAI Inference Engineer: Optimize LLM Serving Latency • San Francisco, CA, United States

Ofertas similares

Applied AI / ML Engineer - Ship GenAI Features (Remote)

IdeogramSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A leading design technology company is seeking a Machine Learning Engineer to lead applied AI initiatives to transform advanced generative models into production features.You will collaborate acros... Mostrar más

 • Oferta promocionada

RL Environments Engineer (Remote)

Verita HRSan Francisco, CA, United States
Teletrabajo
A tiempo completo

Czym bedziesz sie zajmowac?About the company:US-based AI startup focused on building the next generation of training data for LLMs.The team partners with top AI labs to create realistic RL environm... Mostrar más

 • Oferta promocionada

Senior MLOps Engineer - Remote-First, High-Impact ML

CompScienceSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A technology-driven startup in California is seeking an experienced Sr MLOps Engineer to build and maintain the infrastructure that powers their machine learning products.In this high-impact role, ... Mostrar más

 • Oferta promocionada

NLP Engineer - Production ML for PII Redaction (Remote)

TonicAISan Francisco, CA, United States
Teletrabajo
A tiempo completo

A leading data privacy firm in San Francisco is seeking a hands-on Machine Learning Engineer to develop production-grade NLP systems.The ideal candidate will have over 3 years of experience in appl... Mostrar más

 • Oferta promocionada

Senior AI / ML Data Engineer - Remote Healthcare Platform

Crossing HurdlesSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A forward-thinking healthcare AI company is seeking a Senior Data / AI / ML Software Engineer to build the intelligence behind clinical workflows.This role requires 7years of experience in AI / ML ... Mostrar más

 • Oferta promocionada

Lead AI / ML Staff Engineer

Dream Technologies IncSan Francisco, CA, United States
A tiempo completo

Lead AI & Algorithms EngineerWe are looking for a Lead AI & Algorithms Engineer to own the full lifecycle of novel spatial model development; defining the right framing of complex inference... Mostrar más

 • Oferta promocionada

Staff ML Engineer, AI Platform

Ambience HealthcareSan Francisco, CA, United States
A tiempo completo

Ambience Clinical AI EngineerHere at Ambience, we never set out to be just another scribe.We're building the AI intelligence platform that restores humanity to healthcare and drives meaningful ROI ... Mostrar más

 • Oferta promocionada

Staff ML Infrastructure Engineer - Remote

Darwin RecruitmentSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A fast-growing AI company is seeking a Staff / Principal ML Infrastructure Engineer to lead the design and deployment of large language model infrastructure.Responsibilities include optimizing mode... Mostrar más

 • Oferta promocionada

Senior/Staff Machine Learning Engineer

DexterityRedwood City, CA, United States
A tiempo completo

Redwood CityEngineering – Platform /FT /On-siteAbout DexterityAt Dexterity, we believe robots can positively transform the world.Our breakthrough technology frees people to do the creative, inspiri... Mostrar más

 • Oferta promocionada

Senior Staff ML Engineer - Recommender Systems (Remote)

PinterestSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A leading social media platform is seeking a Tech Lead Architect in San Francisco to drive engineering efforts on ML-driven product experiences.This role demands expertise in Python, Java, and larg... Mostrar más

 • Oferta promocionada

Senior AI Engineer -- Agentic LLMs Lead & Mentor (Remote)

LiveRampSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A leading data collaboration platform in San Francisco is seeking a Senior Staff AI Engineer to lead the development of advanced AI agents.This role entails mentoring a talented team and collaborat... Mostrar más

 • Oferta promocionada

Remote Senior MLOps Engineer: Scalable ML Pipelines

Clariti Cloud Inc.San Francisco, CA, United States
Teletrabajo
A tiempo completo

A cloud technology company is looking for a Senior MLOps Engineer to join their team in San Francisco, CA.This role involves designing, building, and scaling ML systems, ensuring models transition ... Mostrar más

 • Oferta promocionada

Senior AI/ML Engineer

Sigma ComputingSan Francisco, CA, United States
A tiempo completo

About the RoleAt Sigma, we’re not just adding AI—we’re building the future of how people work with data.Our platform already lets users explore billions of rows of data in seconds with a spreadshee... Mostrar más

 • Oferta promocionada

Senior ML Engineer - Generative Vision, Remote

The Rundown AI, Inc.San Francisco, CA, United States
Teletrabajo
A tiempo completo

A tech company specializing in AI systems is seeking a Machine Learning Engineer to develop innovative research projects in computer vision.The role involves collaboration with a top engineering te... Mostrar más

 • Oferta promocionada

Staff ML Engineer, Personalization AI - Remote

BlockSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A global technology company is seeking a Staff Machine Learning Engineer for the Tidal Personalization AI team.This role involves leading initiatives in Personalization and AI Engineering, developi... Mostrar más

 • Oferta promocionada

Senior ML Engineer - RAG & LLM Platform (Remote)

Zendesk, Inc.San Francisco, CA, United States
Teletrabajo
A tiempo completo

A leader in customer experience solutions is looking for a Senior ML Engineer to enhance their retrieval-augmented generation (RAG) platform.In this role, you will leverage the latest in LLM techno... Mostrar más

 • Oferta promocionada

Staff AI / ML Engineer, Payments & Risk -- Remote

GustoSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A tech company specializing in payments and risk is seeking a Staff Applied AI and Machine Learning Engineer.In this role, you will build and deploy models to identify and mitigate risks while coll... Mostrar más

 • Oferta promocionada

Staff Machine Learning Engineer, Infrastructure

Rad AISan Francisco, CA, United States
A tiempo completo

About Rad AIAt Rad AI, we're on a mission to transform healthcare with artificial intelligence.Founded by a radiologist, our AI-driven solutions are revolutionizing radiologysaving time, reducing b... Mostrar más

 • Oferta promocionada

Lead ML Platform Engineer -- Remote

Autodesk, Inc.San Francisco, CA, United States
Teletrabajo
A tiempo completo

A leading software company is seeking a Principal Software Engineer for their Machine Learning Platform in San Francisco, CA.The role involves designing and developing software for an ML lifecycle ... Mostrar más

 • Oferta promocionada

Senior ML Platform Engineer -- Remote, Unlimited PTO

ParafinSan Francisco, CA, United States
Teletrabajo
A tiempo completo

A financial tech company in San Francisco is seeking a Software Engineer for its Infrastructure team.The role involves developing the ML Platform, handling model experimentation and deployment.Cand... Mostrar más