Backend Engineer - Studio Media Platform
Sarvam AI
Sarvam AI
Sarvam is pioneering Sovereign AI for India by developing a comprehensive full-stack AI platform. Our mission is to make AI truly impactful for India, focusing on research, models, infrastructure, and applications. We collaborate with leading enterprises and public institutions, supported by prominent investors. Our work involves partnering with major Indian brands to drive innovation.
We are seeking a Backend Engineer to join our Studio Media Platform team. This role will involve working on cutting-edge AI dubbing, live translation, and the underlying shared services powering these products. You will be instrumental in developing and maintaining production services, ML pipeline libraries, and platform SDKs to enable scalable, multilingual media processing for our enterprise clients and Studio users.
Design and optimize high-performance FastAPI services for dubbing and live translation, incorporating multi-stage task orchestration and robust backpressure controls.
Develop and manage distributed worker architectures that allow for independent scaling and automatic recovery of tasks.
Take ownership of the data layer, including async ORM models, schema migrations, and query optimization using PostgreSQL.
Implement real-time features such as WebSocket-based job tracking and streaming audio pipelines for live translation.
Manage Kubernetes deployments, including Helm charts, secrets management, and ingress configuration.
Extend and maintain the core dubbing library, covering all stages from audio extraction to video stitching.
Integrate and optimize ML model serving for both remote and local inference scenarios.
Build and improve quality control orchestration with automated scoring and pronunciation verification.
Design efficient, async-first pipelines optimized for CPU-bound audio processing.
Maintain and evolve LLM integration layers for various Studio services.
We are looking for a seasoned Backend Engineer with 4-6 years of experience, specializing in building and operating scalable production services. Essential skills include strong proficiency in Python and hands-on experience with async web services like FastAPI. A deep understanding of asynchronous programming patterns, concurrent execution, and high-throughput workload design is crucial.
Familiarity with distributed task systems, such as Celery, message brokers, and designing fault-tolerant job orchestration, is highly valued. You should be comfortable working with PostgreSQL and an async ORM like SQLAlchemy, including query optimization and schema design.
Experience with audio/media processing tools like FFmpeg and libraries such as soundfile or librosa is beneficial. You should also have experience integrating ML models into production environments. The ability to build reusable libraries or SDKs with clean APIs and manage backward compatibility is important.
Proficiency in Docker, Kubernetes, Helm charts, and at least one major cloud platform (Azure/GCP/AWS) is required. A strong commitment to testing, including comprehensive test writing and CI/CD pipeline maintenance, is essential.
Sarvam AI
Technology