Demo

Software Engineer (serving API)

Genmo Inc.
San Francisco, CA Full Time
POSTED ON 1/15/2025
AVAILABLE BEFORE 3/28/2025

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.

Ensure all your application information is up to date and in order before applying for this opportunity.

Role Overview

We are looking for a senior / staff software engineer to join our inference team. In this role, you will be responsible for designing and scaling our inference systems as they grow to support over millions of users across more than 20 different data centers.

Key Responsibilities

  • Develop high-performance, high-throughput, efficient, and low-latency inference pipelines.
  • Design, develop, and maintain scalable backend services that support our AI-powered content creation platform.
  • Implement and optimize model serving infrastructure using Kubernetes and other cloud-native technologies.
  • Collaborate with ML engineers to transition models from research to production.
  • Design APIs for integrating our AI capabilities into our partner ecosystem.
  • Implement monitoring, logging, and alerting systems for backend services and model inference.
  • Develop monitoring infrastructure for our ML serving pipeline and apply advanced model compression and optimization techniques (quantization, pruning, distillation) to improve inference performance.

Qualifications

  • Bachelor's or Master's degree in Computer Science, Software Engineering, or a related field.
  • 5 years of experience in software engineering, with at least 3 years focusing on backend systems and ML infrastructure.
  • Must Have :
  • Strong past experience with Ray or Kubernetes.

  • Strong proficiency in Python and at least one systems programming language (Rust, C or Go).
  • Solid understanding of model serving frameworks (e.g., TensorFlow Serving, NVIDIA Triton).
  • Experience with a ML framework such as TensorFlow, PyTorch, or JAX.
  • Experience with model compression and optimization techniques.
  • Strong knowledge of cloud platforms (AWS, GCP, or Azure) and their ML-specific services.
  • Familiarity with distributed systems and microservices architectures.
  • Experience with high-performance, low-latency systems.
  • Ideal candidate will have :
  • Experience with GPU programming is a plus.

    Additional Information

    The role is based in the Bay Area (San Francisco). Candidates are expected to be located near the Bay Area or open to relocation.

    Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company.

    J-18808-Ljbffr

    If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
    Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

    What is the career path for a Software Engineer (serving API)?

    Sign up to receive alerts about other jobs on the Software Engineer (serving API) career path by checking the boxes next to the positions that interest you.
    Income Estimation: 
    $97,257 - $120,701
    Income Estimation: 
    $123,167 - $152,295
    Income Estimation: 
    $97,257 - $120,701
    Income Estimation: 
    $123,167 - $152,295
    Income Estimation: 
    $146,673 - $180,130
    Income Estimation: 
    $176,149 - $220,529
    Income Estimation: 
    $77,657 - $95,021
    Income Estimation: 
    $97,257 - $120,701
    Income Estimation: 
    $123,167 - $152,295
    Income Estimation: 
    $146,673 - $180,130
    View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

    Not the job you're looking for? Here are some other Software Engineer (serving API) jobs in the San Francisco, CA area that may be a better fit.

    Software Engineer

    Software Resources, San Francisco, CA

    Software Engineer

    Software Aspekte, San Francisco, CA

    AI Assistant is available now!

    Feel free to start your new journey!