Google
Posted 1w ago

Research Engineer, Gemini API and Beyond Serving Compute Efficiency, DeepMind

Google
Mountain View or New York City
$207k-$300k/yrOnsiteFull Time
Responsibilities
  • reporting compute
  • researching efficiency
  • presenting findings
Requirements
  • Bachelor's in CS/math/statistics or equivalent
  • Experience in Python or C++
  • Industry productization experience
  • Statistical methods
  • Product teamwork and communication skills
Technical tools mentioned
PythonC++

Job description

Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Mountain View, CA, USA; New York, NY, USA.

Minimum qualifications:

  • Bachelor's degree in Computer Science, Mathematics, Statistics, Machine Learning, Behavioral Science, or equivalent practical experience.
  • Experience programming in Python or C++.
  • Experience working in industry, taking projects from proof-of-concept through to implementation.
  • Experience applying experimental ideas to applied problems using statistical methods.
  • Experience working on a product team.

Preferred qualifications:

  • Master's degree or PhD/DPhil in a related technical or scientific field.
  • Experience applying and productionizing state-of-the-art large visual, language, and multimodal research, coupled with a passion for improving developer experience and generative AI.
  • Proficiency or experience with front-end design skill.
  • Understanding of compute allocation and serving architecture.
  • Excellent data visualization and communication skills, with a track record of cross-functional collaboration and working effectively on a product team.

About the job

As a part of DeepMind's high-visibility AI Studio and Gemini API team, you will shape the future of AI development. You will do research and analysis that directly informs strategy for a product that empowers developers worldwide. You will collaborate closely with product managers, engineers, and researchers to design, develop, and deploy scalable, high-performance AI solutions. You will grow in a fast-paced, innovative environment where your technical expertise and passion for AI will directly impact Google's AI strategy.

You can see some of our work in aistudio.google.com.

In this role, you will enable us to understand usage of our systems in scalable ways, informing decisions around compute allocation and production strategy. Your goal is to identify and understand efficiency opportunities and gaps.Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer various learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits

Learn more about benefits at Google.

Responsibilities

  • Lead the reporting of compute costs and utilization, while implementing critical business, product, and engineering metrics for platforms like AI Studio, Gemini API, and Beyond.
  • Research and develop quantitative approaches to improve fleet-wide compute efficiency, and connect these usage patterns to product and business strategy.
  • Contribute to innovative reporting and analytical infrastructure, including the use of agentic workflows and advanced dashboarding techniques.
  • Present and communicate complex research findings in a compelling manner to a varied set of stakeholders.
  • Collaborate across DeepMind (GDM) organizations to support high-quality, cross-product journeys for developers.
Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.

About Google

Provides online search, advertising, cloud computing, and consumer electronics.

Year founded
1998
Employees
190000
Organization type
Public
Headquarters
US

Similar jobs

Research Engineer roles near Mountain View, California
21h
Save
Mark Applied
Hide
Research Engineer, Preference Data
San Francisco, California, United States
$250k-$400k/yr OnsiteFull Time
Vizcom
Vizcom: AI-powered tools for industrial designers to visualize concepts instantly.
Experience building training-data or large-scale data pipelines; experimental mindset. Model training, labeling, human feedback, evaluation operations, and privacy or contractual data constraints are preferred.
LeetCode
1d
Save
Mark Applied
Hide
Research Engineer – Benchmarking
San Francisco or New York City or London
$130k-$500k/yr OnsiteFull Time
Mercor
Mercor: Connecting expert human intelligence with frontier AI model development.
Applied AI research, model evaluation or benchmarking, strong coding and ML experience, data structures and algorithms, backend systems, APIs, SQL or NoSQL, cloud platforms, and model behavior analysis.
SQL, NoSQL, NeurIPS, ICML, ACL
3d
Save
Mark Applied
Hide
Research Engineer, Lab Automation
Menlo Park or San Francisco
$200k-$250k/yr OnsiteFull Time
Periodic Labs
Periodic Labs: Builds autonomous laboratories for AI-driven scientific discovery.
PhD or equivalent research experience in materials science, chemistry, chemical engineering, or related field; materials lab hardware expertise; Python proficiency; and ability to translate scientific workflows into automation requirements.
Python, Electronic Lab Notebooks, LIMS
4d
Save
Mark Applied
Hide
Research Engineer - New Grad (2027)
Sunnyvale or Washington, D.C. or San Diego or Fort Walton Beach or Ann Arbor or London or Stuttgart or Munich or Stockholm or Bangalore or Seoul or Tokyo
$140k-$200k/yr OnsiteFull Time
Applied Intuition
Applied Intuition: Developing software and simulation infrastructure for autonomous vehicles.
Recent MSc or PhD graduate in machine learning, computer vision, autonomy, robotics, or related field; experience with Python, PyTorch, computer vision, robotics, and distributed model training.
Python, PyTorch
4d
Save
Mark Applied
Hide
Research Engineer, LangSmith Engine
New York City or San Francisco
OnsiteFull Time
LangChain
LangChain: Tools for building and deploying production-ready AI agents.
4+ YOERequires 4+ years in ML/AI research, a relevant master's or Ph.D., LLM and AI agent experience, benchmark and experiment design, and strong software engineering skills.
LangSmith, LangChain, LangGraph, Deep Agents, LLMs, AI agents, GPU infrastructure, SFT, RLHF, RLAIF
4d
Save
Mark Applied
Hide
Research Engineer, Synthetic Data
San Francisco or Singapore
$150k-$250k/yr OnsiteFull Time
Clera
Clera: AI talent agent matching professionals with high-growth startup roles
2+ YOERequires 2–4 years in software, ML engineering, or AI research; Python, Linux, Docker, synthetic data pipelines, evaluation frameworks, structured datasets, and independent project ownership.
Python, Linux, Docker
5d
Save
Mark Applied
Hide
Lead Research Engineer, Search & Retrieval
New York City or Frisco or Toronto or Ann Arbor or Eagan or San Francisco or Los Angeles or Irvine or McLean or Washington
$137k-$255k/yr HybridFull Time
Thomson Reuters
Thomson ReutersNASDAQ: TRI: Provides professional software, data, and news services globally.
7+ YOEBachelor's or master's in computer science, engineering, or related field; 7+ years building production software; search and retrieval expertise; Python, AWS, OpenSearch or Vespa, distributed systems, and technical leadership.
OpenSearch, Vespa, Elasticsearch, Solr, Lucene, Python, AWS, Kafka, RAG, A/B tests
5d
Save
Mark Applied
Hide
New College Grad - AI Innovation Research Engineer
San Jose, California, United States
$113k-$242k/yr OnsiteFull Time
Micron Technology
Micron TechnologyNASDAQ: MU: Designs and manufactures semiconductor memory and data storage solutions.
Recent Master's or PhD in a technical field; AI research or project experience; understanding of generative AI, LLMs, agents, machine learning, or analytics; Python programming; strong analytical and communication skills.
Generative AI, Large Language Models (LLMs), Python, Microsoft Copilot, Azure AI, OpenAI, Anthropic, Google AI