Senior Software Engineer, Gemini Audio Inference - Mountain View at Google

Job Description

About the Role

Our team is responsible for building the Productionization and Serving infrastructure for Gemini Audio Inference at Google DeepMind. We help land Gemini Audio capabilities into numerous clients including Astra, GeminiApp, YouTube, Search, Meet, Cloud, Geo, Assistant etc. Based on research Gemini model flavors, our tasks involve model latency/throughput optimizations, serializations, orchestration, evaluation & finally landing in production, at Google scale. Our team focuses deeply on Inference efficiencies for Gemini and its related components. We actively develop new infrastructure to make Gemini more accessible to new streaming use cases. We have deep collaboration with both the research team and production platform team, exposed with SOTA research work and their inference optimizations. Our team owns the infrastructure to serve the Audio tokenization & Audio generation around the Gemini models. Join us if you are interested in having a direct impact on making Google's products better for our users in over 100 languages!

Key Responsibilities

Work with research teams to build the serving solution of Gemini models for different clients.
Propose and test the best serving configurations based on client needs.
Build new infrastructure to serve Gemini in different ways, such as streaming.
Design and build audio-specific logic in the orchestration framework.
Ensure the Gemini model quality in the production environment.
Build and improve infrastructure to support streaming Gemini prompts at large scale.
Bring Gemini Audio related capabilities to clients, e.g., Long Context capabilities to Meet and YouTube; Audio Out capabilities to Maps.
Embrace a new orchestration infrastructure around Gemini model servings.

Requirements

BS, MS or PhD degree in computer science, mathematics, applied stats, machine learning or similar.
Experience working on software engineering projects from proof-of-concept through to implementation.
Required programming languages: C++; Python a plus.
Experience in performance engineering; GPU/TPU experience a plus.
Great communication and interpersonal skills.
Knowledge of machine learning and inference.
Experience productionizing state-of-the-art large language and multimodal models is a plus.

Nice to Have

Experience in performance engineering; GPU/TPU a plus.
Experience in productionizing state-of-the-art large language and multimodal models a plus.

Qualifications

BS, MS or PhD degree in relevant fields.

Benefits & Perks

US base salary range between $166,000 - $244,000 plus bonus, equity, and benefits.

Working at Google

At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunities regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition. We prioritize safety and ethics in our work and support accommodations for individuals with disabilities or additional needs.

Senior Software Engineer, Gemini Audio Inference - Mountain View

Job Description

About the Role

Key Responsibilities

Requirements

Nice to Have

Qualifications

Benefits & Perks

Working at Google

Job Details

Job Skills

About Google

Get job alerts

Similar Jobs

Senior/Lead Search Engineer

Senior/Lead Search Engineer

Senior/Lead Search Engineer

Senior/Lead Search Engineer

Distinguished Engineer (MTD)

Senior/Lead Search Engineer