Talent.com
GPU Performance Engineer
GPU Performance EngineerGenmo • San Francisco, CA, US
No longer accepting applications
GPU Performance Engineer

GPU Performance Engineer

Genmo • San Francisco, CA, US
25 days ago
Job type
  • Full-time
Job description

Job Description

Job Description

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.

We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and optimize our model serving stack to its absolute limits.

The Role

You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure our infrastructure delivers world-class performance. This role is perfect for someone who gets excited about microsecond optimizations and pushing hardware to its theoretical limits.

Key Responsibilities

Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation

Write high-performance CUDA and Triton kernels for critical model operations

Optimize cold start latency from seconds to milliseconds for our serving infrastructure

Tune memory access patterns, kernel fusion, and GPU utilization

Collaborate with ML engineers to optimize model implementations

Debug performance issues across the full stack from application to hardware

Implement custom memory pooling and allocation strategies

Share optimization techniques and build performance culture across teams

Qualifications

Bachelor's or Master's degree in Computer Science, Electrical Engineering, or related field

5+ years systems programming experience with 3+ years focused on GPU optimization

Expert proficiency with GPU profiling tools (Nsight Systems, nvprof)

Strong CUDA programming skills with production kernel development

Deep understanding of GPU architecture (memory hierarchy, SMs, warps)

Track record of achieving significant performance improvements (5-10x)

Experience with Python and C++ in production environments

We Value

Experience with Triton kernel development

Knowledge of CUTLASS or similar high-performance libraries

Background in ML-specific optimizations (attention, transformers)

RDMA / InfiniBand optimization experience

Contributions to GPU libraries or frameworks

Low-level debugging skills (PTX / SASS reading)

Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.

Create a job alert for this search

GPU Performance Engineer • San Francisco, CA, US

Similar jobs
Performance ML Engineer : CUDA, GPU Systems

Performance ML Engineer : CUDA, GPU Systems

Relace • San Francisco, CA, United States
Full-time
A tech company specializing in ML infrastructure is seeking a Machine Learning Engineer who excels at making models faster and more efficient through performance tuning and optimization.The ideal c...Show more
Last updated: 30+ days ago • Promoted
Senior HPC Engineer — GPU Clusters & Infra Automation

Senior HPC Engineer — GPU Clusters & Infra Automation

San Francisco Compute Co. • San Francisco, CA, United States
Full-time
A leading compute solutions provider located in San Francisco is looking for a dedicated professional to manage GPU training clusters. You will ensure the smooth operation of high-performance comput...Show more
Last updated: 30+ days ago • Promoted
Performance Engineer LMTS

Performance Engineer LMTS

Salesforce, Inc. • San Francisco, CA, United States
Full-time
To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. Job CategorySoftware EngineeringJob Details • • • •Abo...Show more
Last updated: 15 days ago • Promoted
GPU Systems Engineer : High-Performance C++

GPU Systems Engineer : High-Performance C++

10X Recruiting Partners • San Francisco, CA, United States
Full-time
A technology consulting firm is seeking a highly skilled Software Engineer (C++ Systems) to join their client's team in San Francisco. This role focuses on optimizing GPU virtualization performance ...Show more
Last updated: 30+ days ago • Promoted
Distinguished Engineer

Distinguished Engineer

Scale AI, Inc. • San Francisco, CA, United States
Full-time
Our mission is to develop reliable AI systems for the world's most important decisions.The Enterprise AI business delivers performant, reliable, and production-grade AI applications to enterprises ...Show more
Last updated: 30+ days ago • Promoted
GPU Performance Engineer

GPU Performance Engineer

Genmo Inc. • San Francisco, CA, United States
Full-time
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the bo...Show more
Last updated: 30+ days ago • Promoted
Assembly Technician

Assembly Technician

Akash Systems • Emeryville, California, United States
Full-time
Job Description We are seeking hands-on Assembly Technicians to build, inspect, and integrate state-of-the-art electronic and electromechanical products using drawings, schematics, and work instruc...Show more
Last updated: 5 hours ago • Promoted • New!
Rides Maintenance Supervisor $80,000-$95,000

Rides Maintenance Supervisor $80,000-$95,000

Six Flags Discovery Kingdom Careers • VALLEJO, CA, United States
Full-time
Responsible for the operation of the Ride Maintenance areas including : repair and maintenance, preventative maintenance, service calls, training, scheduling of work and adherence to all safety, par...Show more
Last updated: 4 hours ago • Promoted • New!
GPU / CPU Accelerated Bioinformatics Engineer

GPU / CPU Accelerated Bioinformatics Engineer

Antler • San Francisco, CA, United States
Full-time
Prima Mente is a frontier biology AI lab.We generate our own data, build general purpose biological foundation models, and translate discoveries into research and clinical outcomes.Our first goal i...Show more
Last updated: 7 days ago • Promoted
Performance Engineer, GPU

Performance Engineer, GPU

Anthropic • San Francisco, CA, United States
Full-time
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group ...Show more
Last updated: 30+ days ago • Promoted
Radiology / Cardiology-X-Ray Tech

Radiology / Cardiology-X-Ray Tech

Zenex Partners • Berkeley, CA, United States
Full-time
Job Opportunity : Radiology / Cardiology - X-Ray Tech.Facility : Sutter Health Alta Bates Summit Medical Center Ashby.Employment Type : Travel / Contract. Shift : Night (5x8 Hours) 23 : 00 7 : 00 / Evening (5...Show more
Last updated: 30+ days ago • Promoted
Senior HPC & GPU Infrastructure Engineer

Senior HPC & GPU Infrastructure Engineer

Sciforium • San Francisco, CA, United States
Full-time
Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary, high-efficiency serving platform. Backed by multi-million-dollar funding and direct spons...Show more
Last updated: 22 days ago • Promoted
Wireless Power FEA & Thermal Simulation Engineer

Wireless Power FEA & Thermal Simulation Engineer

Apple Inc. • San Francisco, CA, United States
Full-time
A leading technology company in California is seeking an experienced engineer to work on innovative wireless power products. The role involves design, simulation, and analysis using advanced techniq...Show more
Last updated: 30+ days ago • Promoted
IC CAD Engineer - Analog Mixed-Signal Flow Automation

IC CAD Engineer - Analog Mixed-Signal Flow Automation

Capgemini • San Francisco, CA, United States
Full-time
At Capgemini Engineering, the world leader in engineering services, we bring together a global team of engineers, scientists, and architects to help the world’s mostinnovative companies unleash the...Show more
Last updated: 12 days ago • Promoted
Civil Engineer (0480U), Facilities Services - 81815

Civil Engineer (0480U), Facilities Services - 81815

InsideHigherEd • Berkeley, California, United States
Full-time
Civil Engineer (0480U), Facilities Services - 81815.At the University of California, Berkeley, we are dedicated to fostering a community where everyone feels welcome and can thrive.Our culture of o...Show more
Last updated: 30+ days ago • Promoted
GPU Fleet Engineer – Hyperscale Infra, Kubernetes & AI

GPU Fleet Engineer – Hyperscale Infra, Kubernetes & AI

OpenAI • San Francisco, CA, United States
Full-time
Join a forward-thinking company as an engineer in the fleet infrastructure team, where you'll design and operate systems for one of the largest GPU fleets globally. This role offers the chance to wo...Show more
Last updated: 30+ days ago • Promoted
Junior Attraction Maintenance Engineer (VALLEJO)

Junior Attraction Maintenance Engineer (VALLEJO)

Six Flags Discovery Kingdom • VALLEJO, California, US
Part-time
Hourly overtime eligible position and you get paid weekly!.Guaranteed hours, benefits eligible, and paid vacation days!.Learn valuable skills about rides and attractions. Promotional and growth oppo...Show more
Last updated: 16 hours ago • Promoted • New!
Senior Avionics Hardware Engineer - Fighters Mission Systems (Level 5)

Senior Avionics Hardware Engineer - Fighters Mission Systems (Level 5)

Boeing • Berkeley, CA, US
Permanent +1
At Boeing, we innovate and collaborate to make the world a better place.We’re committed to fostering an environment for every teammate that’s welcoming, respectful and inclusive, with great opportu...Show more
Last updated: 6 days ago • Promoted