Talent.com
GPU Performance Engineer
GPU Performance EngineerGenmo • San Francisco, CA, US
No longer accepting applications
GPU Performance Engineer

GPU Performance Engineer

Genmo • San Francisco, CA, US
30+ days ago
Job type
  • Full-time
Job description

Job Description

Job Description

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.

We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and optimize our model serving stack to its absolute limits.

The Role

You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure our infrastructure delivers world-class performance. This role is perfect for someone who gets excited about microsecond optimizations and pushing hardware to its theoretical limits.

Key Responsibilities

Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation

Write high-performance CUDA and Triton kernels for critical model operations

Optimize cold start latency from seconds to milliseconds for our serving infrastructure

Tune memory access patterns, kernel fusion, and GPU utilization

Collaborate with ML engineers to optimize model implementations

Debug performance issues across the full stack from application to hardware

Implement custom memory pooling and allocation strategies

Share optimization techniques and build performance culture across teams

Qualifications

Bachelor's or Master's degree in Computer Science, Electrical Engineering, or related field

5+ years systems programming experience with 3+ years focused on GPU optimization

Expert proficiency with GPU profiling tools (Nsight Systems, nvprof)

Strong CUDA programming skills with production kernel development

Deep understanding of GPU architecture (memory hierarchy, SMs, warps)

Track record of achieving significant performance improvements (5-10x)

Experience with Python and C++ in production environments

We Value

Experience with Triton kernel development

Knowledge of CUTLASS or similar high-performance libraries

Background in ML-specific optimizations (attention, transformers)

RDMA / InfiniBand optimization experience

Contributions to GPU libraries or frameworks

Low-level debugging skills (PTX / SASS reading)

Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.

Create a job alert for this search

Performance Engineer • San Francisco, CA, US

Related jobs
High Performance AI Engineer

High Performance AI Engineer

VirtualVocations • Santa Clara, California, United States
Full-time
A company is looking for a High Performance AI Engineer to build groundbreaking multi-agent systems for the CUDA ecosystem. Key Responsibilities Design, build, and optimize agentic AI systems for ...Show more
Last updated: 4 days ago • Promoted
PLS-CADD Engineer

PLS-CADD Engineer

VirtualVocations • Oakland, California, United States
Full-time
A company is looking for a PLS-CADD Engineer (LiDAR-Based Power Line Modeling).Key Responsibilities Process LiDAR data to create accurate 3D models of power line networks Build and analyze overh...Show more
Last updated: 22 days ago • Promoted
Sr. System Engineer - GPU Servers (27156)

Sr. System Engineer - GPU Servers (27156)

Supermicro • San Jose, CA, United States
Full-time
Supermicro is a top-tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT, Hadoop / Big Data, Hyperscale, HPC, and IoT / Embedded customers...Show more
Last updated: 1 day ago • Promoted
Combat Engineer

Combat Engineer

United States Army • San Francisco, CA, US
Temporary
Combat Engineer Job Overview : Jump start your career in engineering with our world class training program earning up to 45 advanced certifications. As a Combat Engineer, you will gain construction a...Show more
Last updated: 2 days ago • Promoted
Performance Engineer

Performance Engineer

VirtualVocations • San Francisco, California, United States
Full-time
A company is looking for a Performance Engineer / Tester.Key Responsibilities Develop, execute, and maintain performance test plans, scripts, and scenarios for drilling applications Analyze system...Show more
Last updated: 30+ days ago • Promoted
Performance Engineer - GPU

Performance Engineer - GPU

Menlo Ventures • San Francisco, CA, United States
Full-time
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group ...Show more
Last updated: 20 days ago • Promoted
Floating Point Verification Engineer

Floating Point Verification Engineer

VirtualVocations • Oakland, California, United States
Full-time
A company is looking for a Floating Point Formal Verification Engineer.Key Responsibilities Perform formal verification of high-speed floating-point designs, including data-path, assertion, and p...Show more
Last updated: 4 days ago • Promoted
Senior Plastics Design EngineerMechanical Design Engineering • Berkeley, CA • Full time • On-site

Senior Plastics Design EngineerMechanical Design Engineering • Berkeley, CA • Full time • On-site

Form Energy • Berkeley, CA, United States
Full-time
Are you ready to build America's energy future? Form Energy is an American manufacturing and energy technology company.We're revolutionizing energy storage with cost-effective, multi-day technology...Show more
Last updated: 5 days ago • Promoted
GPU Software Engineer

GPU Software Engineer

Disability Solutions • Santa Clara, CA, US
Full-time
At Roche you can show up as yourself, embraced for the unique qualities you bring.Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted ...Show more
Last updated: 23 days ago • Promoted
Principal GPU Software Engineer II

Principal GPU Software Engineer II

F. Hoffmann-La Roche Gruppe • Santa Clara, CA, United States
Full-time
At Roche you can show up as yourself, embraced for the unique qualities you bring.Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted ...Show more
Last updated: 30+ days ago • Promoted
GPU RTL / FW Engineer

GPU RTL / FW Engineer

Mastech Digital • San Jose, CA, US
Temporary
Digital Transformation Services for all American Corporations.We value our professionals, providing comprehensive benefits and the opportunity for growth. San Jose, CA; San Diego, CA; Austin, TX - H...Show more
Last updated: 30+ days ago • Promoted
Performance Engineer

Performance Engineer

Menlo Ventures • San Francisco, CA, United States
Full-time
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group ...Show more
Last updated: 20 days ago • Promoted
Software Performance Engineer

Software Performance Engineer

VirtualVocations • Fremont, California, United States
Full-time
A company is looking for a Software Performance Engineer.Key Responsibilities Develop and maintain custom benchmark tools and automation frameworks for bare-metal and virtualized environments Ex...Show more
Last updated: 3 days ago • Promoted
Power Systems Simulation Engineer

Power Systems Simulation Engineer

VirtualVocations • Fremont, California, United States
Full-time
A company is looking for an Innovation Engineer - Power Systems Simulation.Key Responsibilities Model, simulate, and analyze MV and LV power distribution systems using advanced tools Conduct pha...Show more
Last updated: 4 days ago • Promoted
Principal GPU Software Engineer II

Principal GPU Software Engineer II

Disability Solutions • Santa Clara, CA, US
Full-time
At Roche you can show up as yourself, embraced for the unique qualities you bring.Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted ...Show more
Last updated: 23 days ago • Promoted
Border Patrol Agent - Experienced (GL9 / GS11)

Border Patrol Agent - Experienced (GL9 / GS11)

U.S. Customs and Border Protection • Montara, CA, US
Full-time
Check out these higher-salaried federal law enforcement opportunities with the U.Your current or prior law enforcement experience may qualify you for this career opportunity with the nation's premi...Show more
Last updated: 1 day ago • Promoted
Senior Performance Engineer

Senior Performance Engineer

VirtualVocations • Oakland, California, United States
Full-time
A company is looking for a Senior Performance and Development Engineer.Key Responsibilities Build AI models, tools, and frameworks for real-time application performance metrics Develop automatio...Show more
Last updated: 30+ days ago • Promoted
Sr. Hardware Design Engineer - x86 / GPU / HPC Servers (27733)

Sr. Hardware Design Engineer - x86 / GPU / HPC Servers (27733)

Supermicro • San Jose, CA, United States
Full-time
Supermicro is a Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT, Hadoop / Big Data, Hyperscale, HPC and IoT / Embedded customers...Show more
Last updated: 1 day ago • Promoted