What are the responsibilities and job description for the Member of Technical Staff - ML Inference Engineer position at Genmo?

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.

Role overview:

We are looking for a senior software engineer to join our inference team. In this role, you will be responsible for designing and scaling our inference systems as they grow to support over 1 million users across more than 20 different data centers.

Key responsibilities:

Develop high-performance, high-throughput, efficient, and low-latency inference pipelines.
Design, develop, and maintain scalable backend services that support our AI-powered content creation platform.
Implement and optimize model serving infrastructure using Kubernetes and other cloud-native technologies.
Collaborate with ML engineers to transition models from research to production.
Design APIs for integrating our AI capabilities into our partner ecosystem.
Implement monitoring, logging, and alerting systems for backend services and model inference.
Develop monitoring infrastructure for our ML serving pipeline and apply advanced model compression and optimization techniques (quantization, pruning, distillation) to improve inference performance.

Qualifications:

Bachelor's or Master's degree in Computer Science, Software Engineering, or a related field
5 years of experience in software engineering, with at least 3 years focusing on backend systems and ML infrastructure
Strong past experience with Ray or Kubernetes
Strong proficiency in Python and Go
Experience with GPU programming is a plus
Solid understanding of model serving frameworks (e.g., TensorFlow Serving, NVIDIA Triton)
Experience with a ML framework such as TensorFlow, PyTorch, or JAX
Experience with model compression and optimization techniques
Strong knowledge of cloud platforms (AWS, GCP, or Azure) and their ML-specific services
Familiarity with distributed systems and microservices architectures
Experience with high-performance, low-latency systems

Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.

Apply for this job

Receive alerts for other Member of Technical Staff - ML Inference Engineer job openings

What is the career path for a Member of Technical Staff - ML Inference Engineer?

Sign up to receive alerts about other jobs on the Member of Technical Staff - ML Inference Engineer career path by checking the boxes next to the positions that interest you.

Software Engineer II

Income Estimation:

$97,257 - $120,701

Software Engineer III

Income Estimation:

$123,167 - $152,295

Software Engineer III

Income Estimation:

$123,167 - $152,295

Software Engineer IV

Income Estimation:

$146,673 - $180,130

Software Engineer IV

Income Estimation:

$146,673 - $180,130

Software Engineer V

Income Estimation:

$176,149 - $220,529

Software Engineer I

Income Estimation:

$77,657 - $95,021

Software Engineer II

Income Estimation:

$97,257 - $120,701

Job openings at Genmo

Research Scientist (post-training)

Genmo

San Francisco, CA Full Time

Job Details Job Description Job Description We are Genmo, a research lab dedicated to building open, state-of-the-art mo...

Research Scientist (diffusion)

Genmo

San Francisco, CA Full Time

Job Details Job Description Job Description We are Genmo, a research lab dedicated to building open, state-of-the-art mo...

Founding Product Designer

Genmo

San Francisco, CA Full Time

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking ...

Software Engineer (serving API)

Genmo

San Francisco, CA Full Time

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking ...

Not the job you're looking for? Here are some other Member of Technical Staff - ML Inference Engineer jobs in the San Francisco, CA area that may be a better fit.

Member of Technical Staff - ML Inference Engineer

What are the responsibilities and job description for the Member of Technical Staff - ML Inference Engineer position at Genmo?

What is the career path for a Member of Technical Staff - ML Inference Engineer?

Job openings at Genmo

Not the job you're looking for? Here are some other Member of Technical Staff - ML Inference Engineer jobs in the San Francisco, CA area that may be a better fit.

We don't have any other Member of Technical Staff - ML Inference Engineer jobs in the San Francisco, CA area right now.

AI Assistant is available now!