logo

AI Job Summary

You'll work with ML infrastructure teams to optimize performance across frameworks like TensorFlow, JAX, and PyTorch, analyzing bottlenecks in GPU and TPU systems using tools like Xprof and Grafana. This role requires eight years of software engineering experience, with demonstrated people management and expertise building infrastructure for ML systems. You'll partner across teams on performance optimization, contribute to open-source frameworks like vLLM and SGLang, and debug tools in VSCode and Cursor environments.

Written from this posting by Neural Jobs AI. The full description is below.

AI Resume Tailoring Sign in to use this AI Cover Letter Sign in to use this

Job Description

  • Learn and build an intuitive understanding of existing data collection, analysis, and visualization workflows with deep introspection across Frameworks, Accelerated Linear Algebra (XLA) and runtime stack.
  • Support new and exciting ML paradigms (such as horizontal scaling for upcoming TPU chips) by making contributions across the end to end stack and analysis tools. Partner with ML Stack leads to understand model optimization use cases and bring debugging to feel native in 3P environments (VSCode, Cursor, Grafana, etc).
  • Work with OSS ML inference frameworks such as vLLM, SGLang to provide insights in Xprof about performance improvement opportunities.
  • Partner with other teams that own various parts of the ML stack to understand performance optimization use cases. Work with OSS ML inference frameworks such as TorchTPU, vLLM, SGLang to provide insights into performance bottlenecks.

Minimum qualifications:

  • Bachelor's degree or equivalent practical experience.
  • 8 years of experience with software engineering, machine learning infrastructure, computer architecture, distributed computing, people management, debugging tools, communication.
  • Experience with people management, and building infrastructure to improve performance of Machine Learning (ML) systems and applications.

Preferred qualifications:

  • Experience with Performance Optimization, Graphics Processing Unit (GPU) Programming, High Performance Computing, Large Language Model, Open Source Contributor.
  • Experience with the Machine Learning infra and frameworks and hands-on experience with GPU or Tensor Processing Unit (TPU) performance analysis.
  • Experience in building agentic workflows for performance debugging and optimization. Ability to generate ideas and resolve ambiguity.
  • Hands-on experience with Machine Learning (ML) frameworks such as TensorFlow, JAX, and PyTorch, Keras. Experience with ML Inference frameworks such as vLLM, SG Lang, Pathways and experience in open-source software development, including experience in releasing and supporting open-source projects.
  • Bachelor's degree or equivalent practical experience.
  • 8 years of experience with software engineering, machine learning infrastructure, computer architecture, distributed computing, people management, debugging tools, communication.
  • Experience with people management, and building infrastructure to improve performance of Machine Learning (ML) systems and applications.
Do you match this job?

Here is what this employer asked for. Sign in and we will fill in your half.

  • Role Software Engineering Manager, AI/ML Infrastructure and Performance Engineering
  • Experience 5-7 years
  • Education Bachelor Degree, or equivalent experience
  • Work type On-site
  • Location India
Check my match (free)
Google
AI Research Lab · 500+ Members · Mountain View, CA, United States

Google builds internet, software, cloud, and AI products used by consumers, developers, and organizations. Its portfolio includes Search, YouTube, Android, Chrome, Maps, Gmail, Workspace, Google Cloud, advertising platforms, devices, and Gemini AI products. The company develops large-scale computing infrastructure and research that power information retrieval, communication, productivity, media, navigation, and machine learning. Google is the largest operating business within Alphabet and earns a substantial share of its revenue from digital advertising.

All jobs at Google

37 more Machine Learning Engineer roles in Bengaluru

Job Overview
Eligibility
India Right to work in India required.
Workplace
On-site
Job Posted:
3 days ago
Job Type
Full Time
Education
Bachelor Degree, or equivalent experience
Experience
5-7 years

Share This Job: