Google logo

Senior Software Engineer, Gemini Audio Inference - Mountain View

Google

Mountain View, NY
Full Time
Senior
166k-244k
9 days ago

Job Description

About the Role

Our team is responsible for building the Productionization and Serving infrastructure for Gemini Audio Inference at Google DeepMind. We help land Gemini Audio capabilities into numerous clients including Astra, GeminiApp, YouTube, Search, Meet, Cloud, Geo, Assistant etc. Based on research Gemini model flavors, our tasks involve model latency/throughput optimizations, serializations, orchestration, evaluation & finally landing in production, at Google scale. Our team focuses deeply on Inference efficiencies for Gemini and its related components. We actively develop new infrastructure to make Gemini more accessible to new streaming use cases. We have deep collaboration with both the research team and production platform team, exposed with SOTA research work and their inference optimizations. Our team owns the infrastructure to serve the Audio tokenization & Audio generation around the Gemini models. Join us if you are interested in having a direct impact on making Google's products better for our users in over 100 languages!

Key Responsibilities

  • Work with research teams to build the serving solution of Gemini models for different clients.
  • Propose and test the best serving configurations based on client needs.
  • Build new infrastructure to serve Gemini in different ways, such as streaming.
  • Design and build audio-specific logic in the orchestration framework.
  • Ensure the Gemini model quality in the production environment.
  • Build and improve infrastructure to support streaming Gemini prompts at large scale.
  • Bring Gemini Audio related capabilities to clients, e.g., Long Context capabilities to Meet and YouTube; Audio Out capabilities to Maps.
  • Embrace a new orchestration infrastructure around Gemini model servings.

Requirements

  • BS, MS or PhD degree in computer science, mathematics, applied stats, machine learning or similar.
  • Experience working on software engineering projects from proof-of-concept through to implementation.
  • Required programming languages: C++; Python a plus.
  • Experience in performance engineering; GPU/TPU experience a plus.
  • Great communication and interpersonal skills.
  • Knowledge of machine learning and inference.
  • Experience productionizing state-of-the-art large language and multimodal models is a plus.

Nice to Have

  • Experience in performance engineering; GPU/TPU a plus.
  • Experience in productionizing state-of-the-art large language and multimodal models a plus.

Qualifications

  • BS, MS or PhD degree in relevant fields.

Benefits & Perks

  • US base salary range between $166,000 - $244,000 plus bonus, equity, and benefits.

Working at Google

At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunities regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition. We prioritize safety and ethics in our work and support accommodations for individuals with disabilities or additional needs.

Apply Now

Job Details

Posted AtJul 17, 2025
Salary166k-244k
Job TypeFull Time
ExperienceSenior

Job Skills

AI Insights

Key skills identified from this job posting

Sign upto access all insights for this job

About Google

Website

google.com

Location

Mountain View, NY

Industry

Web Search Portals and All Other Information Services

Get job alerts

Set up personalized alerts for your job search and get tailored job digests for close matches