Multimedia GPU teamPune, Bangalore
Senior Deep Learning Research Engineer
Location
Pune, Bangalore
Compensation Tier
Standard
Category
Multimedia GPU team
Company Overview
Our client is a Research lead organisation focusing on cutting-edge software, AI, and hardware innovation.
Position Overview
The role focuses on researching and developing advanced Generative AI models for Audio and Speech, including Diffusion, Transformers and Autoregressive models, and training them on large-scale GPU clusters. It also involves building multimodal Audio-Speech-Video solutions, such as speech transformation, spatial audio and real-time audio/video enhancement. The engineer will optimize and deploy these models on GPUs/edge platforms, collaborate with research and product teams, and mentor senior engineers.
Responsibilities
- Lead Generative AI and Deep Learning research for audio, speech, image and video.
- Build and scale Diffusion, Transformer and Autoregressive models on large GPU clusters.
- Develop audio-speech-visual multimodal and real-time AI solutions.
- Productize and optimize models for GPU/Edge deployment.
- Collaborate with research, hardware and product teams.
- Mentor senior engineers and drive technical innovation.
Skills & Experience
- PhD in CS, AI, Applied Mathematics or related field.
- 10+ years in Deep Learning / AI research and development.
- Strong expertise in Generative AI: Diffusion, GANs, Transformers, VAEs, NeRF/3D Gaussian Splatting, GRUs.
- Expert Python + PyTorch skills.
- Strong knowledge of audio DSP, spectrograms, video processing and multimodal pipelines.
- Strong publications or proven commercial AI product impact.
Apply for this Role
You must be signed in to apply for this position.
Sign In to Apply