Talent.com
Jobleads-US
Staff GenAI Inference Engineer: Optimize LLM Serving LatencyJobleads-US • San Francisco, CA, United States
Staff GenAI Inference Engineer: Optimize LLM Serving Latency

Staff GenAI Inference Engineer: Optimize LLM Serving Latency

Jobleads-US • San Francisco, CA, United States
7 days ago
Salary
$190,900.00 yearly
Job type
  • Full-time
Job description

A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong software engineering background and a proven ability to collaborate with researchers and drive architectural decisions. Competitive compensation is offered, with a salary range of $190,900 to $232,800 USD. #J-18808-Ljbffr

Create a job alert for this search

Staff GenAI Inference Engineer: Optimize LLM Serving Latency • San Francisco, CA, United States

Similar jobs

Applied AI / ML Engineer - Ship GenAI Features (Remote)

IdeogramSan Francisco, CA, United States
Remote
Full-time

A leading design technology company is seeking a Machine Learning Engineer to lead applied AI initiatives to transform advanced generative models into production features.You will collaborate acros... Show more

 • Promoted

Senior Model Inference Engineer for Production-Scale AI

OpenAISan Francisco, CA, United States
Full-time

A leading AI research company in San Francisco seeks an engineer to optimize their powerful AI models for high-volume production environments.The ideal candidate has over 5 years of software engine... Show more

 • Promoted

Technology & AI Strategy Leader

SB EnergyRedwood City, CA, United States
Full-time

A technology solutions company in Redwood City seeks a Technology & AI Solutions Partner.This critical role will lead the organization's technology infrastructure, governance, and security while de... Show more

 • Promoted

Founding ML Engineer for AI-Driven Hedge Fund

PoesisSan Francisco, CA, United States
Full-time

A pioneering hedge fund in San Francisco is seeking a Founding ML Engineer to architect and build machine learning systems for investment decisions.This hands-on role requires 5–10+ years of experi... Show more

 • Promoted

Director, AI Platforms & Infra — Accelerate ML

MetaMenlo Park, CA, United States
Full-time

Meta invites a seasoned Product Management leader to own the end-to-end strategy for AI Platforms and Infrastructure in Menlo Park, CA.You will guide how engineers train, serve, and iterate on reco... Show more

 • Promoted

Aerospace Engineer

TradeJobsWorkforce94707 Berkeley, CA, US
Full-time

Aerospace Engineer Job Duties: Contributes to the design, manufacturing, and testing of aircraft and a... Show more

 • Promoted

Staff GenAI Inference Engineer: Optimize LLM Serving Latency

Menlo VenturesSan Francisco, CA, United States
Full-time

A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine.The role requires expertise in CUDA, GPU pro... Show more

 • Promoted

Senior ML Engineer - Real-Time Engine Optimization

RobloxSan Mateo, CA, United States
Full-time

A leading gaming platform is seeking experts in machine learning to enhance real-time optimization and user experience.Candidates will define ML strategies for resource management and solve complex... Show more

 • Promoted

Senior ML Engineer: Search & Recommendations in Production

TaskRabbitSan Francisco, CA, United States
Full-time

A leading technology firm is seeking a Senior Machine Learning Engineer to join their team in San Francisco.This hybrid role requires expertise in the entire machine learning lifecycle, including m... Show more

 • Promoted

Director of Software Engineering - AI & Robotics Growth Leader

Radical AIEmeryville, CA, United States
Full-time

Radical AI in Emeryville is seeking a Director of Software Engineering to build and lead a talented software team.This role combines hands-on coding with leadership, focusing on project stability a... Show more

 • Promoted

Senior Staff ML Engineer - Recommender Systems (Remote)

PinterestSan Francisco, CA, United States
Remote
Full-time

A leading social media platform is seeking a Tech Lead Architect in San Francisco to drive engineering efforts on ML-driven product experiences.This role demands expertise in Python, Java, and larg... Show more

 • Promoted

Senior AI Engineer -- Agentic LLMs Lead & Mentor (Remote)

LiveRampSan Francisco, CA, United States
Remote
Full-time

A leading data collaboration platform in San Francisco is seeking a Senior Staff AI Engineer to lead the development of advanced AI agents.This role entails mentoring a talented team and collaborat... Show more

 • Promoted

AI Safety & Safeguards ML Engineer

anthropicSan Francisco, CA, United States
Full-time

An innovative AI company based in New York is looking for ML Engineers and Research Engineers to enhance AI safety through misuse detection systems.Candidates should have at least 4 years of experi... Show more

 • Promoted

Co-Founder & COO — Deep-Tech Robotics & Ops Leader

TerranovaBerkeley, CA, United States
Full-time

Terranova, located in Berkeley, California, is seeking a COO or VP Operations to lead execution across multiple functions in a high-growth environment.This is a critical role focused on operations,... Show more

 • Promoted

GenAI Forward-Deployed Engineer: Data Infra Impact

Scale AISan Francisco, CA, United States
Full-time

A leading AI data company is seeking a Forward Deployed Engineer to drive impactful solutions in the advancement of AI.You will collaborate with technical customers to deliver high-quality data sol... Show more

 • Promoted

Senior ML Engineer — Applied AI Research & Model Development

LightfieldSan Francisco, CA, United States
Full-time

A leading AI technology firm in San Francisco seeks a Machine Learning Engineer to join their AI/ML team.The role involves creating innovative AI experiences, training models, and developing unique... Show more

 • Promoted

Senior ML Engineer - RAG & LLM Platform (Remote)

Zendesk, Inc.San Francisco, CA, United States
Remote
Full-time

A leader in customer experience solutions is looking for a Senior ML Engineer to enhance their retrieval-augmented generation (RAG) platform.In this role, you will leverage the latest in LLM techno... Show more

 • Promoted

Sofware Engineer

TradeJobsWorkForce94706 Albany, CA, US
Full-time

Analyze, design and develop tests and test-automation suites.Design, create and develop a processing platform using various configuration management technologies.Test software development methodolo... Show more

 • Promoted

Director of Software Engineering & AI-Powered Delivery

Braviant HoldingsEmeryville, CA, United States
Full-time

Braviant Holdings is seeking a Head of Software Engineering based in Addison, TX.In this role, you will lead a team of engineers, oversee full-stack development, and implement AI-assisted developme... Show more

 • Promoted

Sr. Facilities Planner (Regulated Industry)

Mentor Technical GroupRedwood City, CA, United States
Full-time

Mentor Technical Group (MTG) provides a comprehensive portfolio of technical support and solutions for the FDA-regulated industry.As a world leader in life science engineering and technical solutio... Show more