About the role
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.
We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and optimize our model serving stack to its absolute limits.
The Role
You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure our infrastructure delivers world-class performance. This role is perfect for someone who gets excited about microsecond optimizations and pushing hardware to its theoretical limits.
Key Responsibilities
Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation
Write high-performance CUDA and Triton kernels for critical model operations
Optimize cold start latency from seconds to milliseconds for our serving infrastructure
Tune memory access patterns, kernel fusion, and GPU utilization
Collaborate with ML engineers to optimize model implementations
Debug performance issues across the full stack from application to hardware
Implement custom memory pooling and allocation strategies
Share optimization techniques and build performance culture across teams
Qualifications
Bachelor's or Master's degree in Computer Science, Electrical Engineering, or related field
5+ years systems programming experience with 3+ years focused on GPU optimization
Expert proficiency with GPU profiling tools (Nsight Systems, nvprof)
Strong CUDA programming skills with production kernel development
Deep understanding of GPU architecture (memory hierarchy, SMs, warps)
Track record of achieving significant performance improvements (5-10x)
Experience with Python and C++ in production environments
We Value
Experience with Triton kernel development
Knowledge of CUTLASS or similar high-performance libraries
Background in ML-specific optimizations (attention, transformers)
RDMA/InfiniBand optimization experience
Contributions to GPU libraries or frameworks
Low-level debugging skills (PTX/SASS reading)
Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.
Aplyr's read
Genmo is at the forefront of AI-enhanced video technology, attracting tech-savvy individuals passionate about machine learning and creative content innovation.
What's promising
- •Genmo offers opportunities to work with cutting-edge AI in video technology.
- •The company is actively hiring diverse roles, indicating growth and innovation.
- •Employees can contribute to pioneering advancements in AI-driven video content.
What to watch
- •The niche focus on AI video tech may limit broader tech exposure.
- •Startups like Genmo can face funding and market competition challenges.
- •Rapid tech changes may require constant skill updates and adaptability.
Why Genmo
- •Genmo specializes in integrating AI with video content, a niche field.
- •The company emphasizes research, attracting those interested in AI innovation.
- •Genmo's roles span from engineering to design, offering diverse career paths.
Aplyr’s read is generated by AI from public sources. Was it useful?
About Genmo
Genmo is an AI-driven platform that focuses on creating and enhancing video content using advanced machine learning techniques.
Similar roles
Senior System Software Engineer - GPU Performance
NVIDIA
Senior System Software Engineer - GPU Performance Profiling Tools
NVIDIA
Senior Systems Software Engineer - GPU Performance at Scale
NVIDIA
GPU Software Performance Engineer
Intel
GPU Performance Engineer - Neural Reconstruction
NVIDIA
Software Engineer, GPU Performance Tools
NVIDIA