Parasail is redefining AI infrastructure by enabling seamless deployment across a distributed network of GPUs, optimizing for cost, performance, and flexibility. Our mission is to empower AI developers with a fast, cost-efficient, and scalable cloud experience-free from vendor lock-in and designed for the next generation of AI workloads.
Job Description :
The AI Performance Engineer plays a crucial role in delivering a competitive platform by focusing on efficiently scheduling, executing, and managing AI workloads on distributed compute systems. This role is deeply technical, spanning from low-level GPU kernels to distributed AI orchestration and Kubernetes (K8s) deployments. It is about more than optimization; it's about pioneering efficient infrastructure that supports AI's transformative role in reshaping productivity, revolutionizing industries, and addressing some of the world's most challenging problems. You'll ensure that generative AI - including large language models (LLMs), multi-modal models, and diffusion models - operates efficiently at enterprise scale while driving continuous improvements in cost, performance, and sustainability.
Responsibilities :
Qualifications :
What You Bring to the Table : We are looking for people who are eager to learn and master the lower-level compute concepts that are critical for the AI revolution. With us, your skills will not only contribute to coding but will also have a significant impact on the scalability and efficiency of AI applications at large. If you're geared up for the challenge of optimizing AI performance and eager to push our technological prowess to new heights, we're excited to welcome you aboard.
Performance Engineer • San Francisco, CA, United States