What are the responsibilities and job description for the Software Engineer (serving API) position at Genmo Inc.?
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.
Ensure all your application information is up to date and in order before applying for this opportunity.
Role Overview
We are looking for a senior / staff software engineer to join our inference team. In this role, you will be responsible for designing and scaling our inference systems as they grow to support over millions of users across more than 20 different data centers.
Key Responsibilities
- Develop high-performance, high-throughput, efficient, and low-latency inference pipelines.
- Design, develop, and maintain scalable backend services that support our AI-powered content creation platform.
- Implement and optimize model serving infrastructure using Kubernetes and other cloud-native technologies.
- Collaborate with ML engineers to transition models from research to production.
- Design APIs for integrating our AI capabilities into our partner ecosystem.
- Implement monitoring, logging, and alerting systems for backend services and model inference.
- Develop monitoring infrastructure for our ML serving pipeline and apply advanced model compression and optimization techniques (quantization, pruning, distillation) to improve inference performance.
Qualifications
Strong past experience with Ray or Kubernetes.
Experience with GPU programming is a plus.
Additional Information
The role is based in the Bay Area (San Francisco). Candidates are expected to be located near the Bay Area or open to relocation.
Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company.
J-18808-Ljbffr