Talent.com

Performance engineer Jobs in Berkeley, CA

Create a job alert for this search

Performance engineer • berkeley ca

Last updated: 10 hours ago
  • Promoted
GPU Performance Engineer

GPU Performance Engineer

GenmoSan Francisco, CA, US
Full-time
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the bo...Show moreLast updated: 9 days ago
  • Promoted
visionOS Performance Engineer

visionOS Performance Engineer

AppleSan Francisco, CA, United States
Full-time
Apple is where individual imaginations gather together, committing to the values that lead to great work.Every new product we build, service we create, or Apple Store experience we deliver is the r...Show moreLast updated: 12 days ago
  • Promoted
Software Engineer (AI Performance)

Software Engineer (AI Performance)

Gimlet Labs, IncSan Francisco, CA, United States
Full-time
Gimlet Labs is building the foundation for the next generation of AI applications.As generative AI workloads rapidly scale, inference efficiency is becoming the critical bottleneck.Gimlet is redefi...Show moreLast updated: 30+ days ago
  • Promoted
Staff Infrastructure and Performance Engineer

Staff Infrastructure and Performance Engineer

NashSan Francisco, CA, United States
Full-time
Staff Infrastructure & Performance Engineer.Staff Infrastructure Performance & Engineer.You’ll work directly with the Engineering Leadership team, platform, and product engineering teams to design ...Show moreLast updated: 7 days ago
Building Performance Engineer

Building Performance Engineer

Harrison Consulting SolutionsSan Francisco, California, USA
Full-time
Leading architectural and engineering firm is adding a.Senior Building Performance Engineer.Lead commissioning and optimization of building systems (HVAC systems air / water distribution systems buil...Show moreLast updated: 21 days ago
  • Promoted
GPU Performance Engineer

GPU Performance Engineer

Genmo Inc.San Francisco, CA, United States
Full-time
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the bo...Show moreLast updated: 30+ days ago
  • Promoted
AI Performance Engineer

AI Performance Engineer

Cornelis NetworksSan Francisco, CA, United States
Full-time
Cornelis Networks delivers the world's highest performance scale-out networking solutions for AI and HPC datacenters.Our differentiated architecture seamlessly integrates hardware, software and sys...Show moreLast updated: 12 days ago
  • Promoted
Sr Building Performance Engineer

Sr Building Performance Engineer

HGASan Francisco, CA, United States
Full-time
Sr Building Performance Engineer.US-CA-San Francisco | US-CA-San Diego | US-CA-San Jose | US-CA-Sacramento.Engineering - Building Performance. Help Shape the Future of Sustainable Building Performan...Show moreLast updated: 12 days ago
  • Promoted
Performance Modelling Engineer

Performance Modelling Engineer

Flux ComputingSan Francisco, CA, United States
Permanent
We're searching for a Staff Performance Modelling Engineer to create and own the analytical and simulation models that steer OTPU architecture and software evolution. You will build functional simul...Show moreLast updated: 12 days ago
  • Promoted
Performance Engineer

Performance Engineer

Menlo VenturesSan Francisco, CA, United States
Full-time
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group ...Show moreLast updated: 30+ days ago
  • Promoted
  • New!
Performance Engineer LMTS

Performance Engineer LMTS

SalesforceSan Francisco, CA, United States
Full-time
To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. Salesforce is the #1 AI CRM, where humans with age...Show moreLast updated: 10 hours ago
  • Promoted
Performance Modelling Engineer

Performance Modelling Engineer

PageBolt WordPressSan Francisco, CA, United States
Permanent
We’re searching for a Staff Performance Modelling Engineer to create and own the analytical and simulation models that steer OTPU architecture and software evolution. You will build functional simul...Show moreLast updated: 30+ days ago
HPC / AI Data Performance Engineer

HPC / AI Data Performance Engineer

Lawrence Berkeley National LaboratoryBerkeley, CA, United States
Full-time +1
In this exciting role, you will serve as a Data Performance Engineer in NERSC's Application Performance Group, architecting HPC and AI data services that advance fundamental science.You'll optimize...Show moreLast updated: 30+ days ago
  • Promoted
Performance Engineer - Michigan

Performance Engineer - Michigan

VirtualVocationsSan Francisco, California, United States
Full-time
A company is looking for a Performance Engineer & Optimization Specialist.Key Responsibilities Collaborate with teams to define performance testing objectives and establish key performance metric...Show moreLast updated: 2 days ago
  • Promoted
Performance Engineer SMTS

Performance Engineer SMTS

Salesforce, Inc.San Francisco, CA, United States
Full-time
To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. Job CategorySoftware EngineeringJob Details • • • •Abo...Show moreLast updated: 30+ days ago
  • Promoted
Performance Engineer MTS / SMTS

Performance Engineer MTS / SMTS

Salesforce.Com IncSan Francisco, CA, United States
Full-time
To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. Salesforce is the #1 AI CRM, where humans with age...Show moreLast updated: 12 days ago
Product Performance Engineer

Product Performance Engineer

OpenAISan Francisco
Full-time
We bring OpenAI's technology to the world through products like ChatGPT and the OpenAI API.We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool...Show moreLast updated: 30+ days ago
  • Promoted
HPC / AI Data Performance Engineer

HPC / AI Data Performance Engineer

Lawrence Berkeley LabBerkeley, CA, United States
Full-time +1
In this exciting role, you will serve as a Data Performance Engineer in NERSC's Application Performance Group, architecting HPC and AI data services that advance fundamental science.You'll optimize...Show moreLast updated: 12 days ago
People also ask
GPU Performance Engineer

GPU Performance Engineer

GenmoSan Francisco, CA, US
9 days ago
Job type
  • Full-time
Job description

Job Description

Job Description

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.

We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and optimize our model serving stack to its absolute limits.

The Role

You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure our infrastructure delivers world-class performance. This role is perfect for someone who gets excited about microsecond optimizations and pushing hardware to its theoretical limits.

Key Responsibilities

Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation

Write high-performance CUDA and Triton kernels for critical model operations

Optimize cold start latency from seconds to milliseconds for our serving infrastructure

Tune memory access patterns, kernel fusion, and GPU utilization

Collaborate with ML engineers to optimize model implementations

Debug performance issues across the full stack from application to hardware

Implement custom memory pooling and allocation strategies

Share optimization techniques and build performance culture across teams

Qualifications

Bachelor's or Master's degree in Computer Science, Electrical Engineering, or related field

5+ years systems programming experience with 3+ years focused on GPU optimization

Expert proficiency with GPU profiling tools (Nsight Systems, nvprof)

Strong CUDA programming skills with production kernel development

Deep understanding of GPU architecture (memory hierarchy, SMs, warps)

Track record of achieving significant performance improvements (5-10x)

Experience with Python and C++ in production environments

We Value

Experience with Triton kernel development

Knowledge of CUTLASS or similar high-performance libraries

Background in ML-specific optimizations (attention, transformers)

RDMA / InfiniBand optimization experience

Contributions to GPU libraries or frameworks

Low-level debugging skills (PTX / SASS reading)

Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.