Senior Software Engineer, AI Infrastructure
New
We sent a six-digit code to . Enter it below — or use the link in the same email.
Or click the link in the same email — either works.
Enter the email address on your account and we'll send you a link to set a new password.
Remembered it? Sign In
Microsoft
NewLead technical program management for AI infrastructure, owning the platform roadmap that scales generative AI workloads from research to production. You'll translate business goals into technical requirements, manage cross-functional teams across engineering, research, and product, and drive end-to-end product development cycles. Requires a bachelor's degree plus six years in engineering, technical program management, or product development, with three years managing cross-functional projects. Proficiency with distributed systems, cloud infrastructure, and ideally Python or ML platforms strongly preferred. Base pay ranges from $119,800–$234,700 annually, or $160,200–$261,000 in San Francisco Bay Area and New York City.
Written from this posting by Neural Jobs AI. The full description is below.
We are part of the Cloud & AI division at Microsoft. Our mission is to build and operate the trusted cloud and AI platform that enables customers and Microsoft to innovate and scale AI securely, reliably, and efficiently.
We are looking for a Principal TPM with technical depth, customer empathy, and an abundance of energy. The right candidate is highly effective, taking the initiative and thriving in building products that differentiate in the most competitive segment of the industry.
This role is in the AI Infrastructure team which powers the Microsoft AI Foundry, CoPilot and other AI services. We manage the GPU fleet to run Microsoft’s AI services on a planet scale. We are in the eye of the storm to accelerate the transition and scaling of Generative AI models to work with the latest AI aware application. Our team focuses on fleet efficiency, reliability and the agility to bring the latest AI innovations from research into production.
You will deeply understand customer problems, the rapidly evolving AI landscape and use a data driven approach combined with creativity, curiosity and AI expertise to lead the team and drive platform features through design, development, market fit and the relentless pursuit of scale adoption.
We bring together all the ingredients for a successful platform across Engineering, Research and Go To Market functions partnering as peers working towards a common goal. To thrive in this role, you will create clarity in ambiguity, lead and energize cross-functional teams, and deliver powerful features with measurable outcomes. Grow your career by joining one of the fastest growing products in Microsoft.
Microsoft’s mission is to empower every person and every organization on the planet to achieve more. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.
In alignment with our Microsoft values, we are committed to cultivating an inclusive work environment for all employees to positively impact our culture every day.
Join the AI Infrastructure team at Microsoft, where we accelerate the transition and scaling of Generative AI models. The Platform Infrastructure team builds and maintains the platform that powers the most demanding AI workloads at Microsoft. We focus on scalability for distributed computing systems, stability of the multi-node GPU infrastructure and efficiency of the fleet to run inference and training workloads.
Deliver a world-class AI Infrastructure stack that will host AI Foundry Services.
Accountable for the platform roadmap for platform to scale AI workloads from research to production
Own a product area and be responsible for understanding developer needs and behaviors, defining product requirements, managing end-to-end product development, launches and iterations.
Find a path to get things done despite roadblocks to get your work into the hands of customers quickly and iteratively.
Enjoy working in a fast-paced, customer-first, product development cycle.
Translate business goals into strategy, user experience and technical requirements in close collaboration with UX, Data Science, Engineering, AI research, and Product Marketing teams.
Define goals and performance indicators, set up and oversee experiments, measure success with data and research.
Collaborate effectively and communicate clearly with cross-functional teams, including product managers, designers, and other engineers, to build exceptional consumer-grade applications.
Embody our culture and values
Required Qualifications:
Other Requirements:
Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include but are not limited to the following specialized security screenings:
Microsoft Cloud Background Check: This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter.
Preferred Qualifications:
8+ years of experience designing and shipping complex products for developers, ML professionals, or similar
3+ years of experience with distributed platform ecosystem (multi-tenant / real-time processing, batch computing etc.)
Proven experience in technical program management, preferably in AI infrastructure or cloud services.
Strong understanding of GPUs, VM, OS and cloud infrastructure fundamentals.
Excellent communication and stakeholder management skills.
Ability to navigate ambiguity and drive clarity across complex, cross-functional initiatives.
Experience creating solutions using Azure, AWS, or Google Cloud
Experience writing Python, particularly for machine learning
Experience with ML platforms
Effective at driving complex multi-stakeholder processes and cross-team programs
Ability to build effective relationships, influence and collaborate at all organizational levels
Technical Program Management IC5 - The typical base pay range for this role across the U.S. is USD $142,800 - $274,800 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $188,000 - $304,200 per year.
Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay
Technical Program Management IC4 - The typical base pay range for this role across the U.S. is USD $119,800 - $234,700 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $160,200 - $261,000 per year.
Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay
This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.
Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.
Here is what this employer asked for. Sign in and we will fill in your half.
Microsoft builds Windows, Azure, Office and the Copilot family of AI assistants, and operates one of the largest AI training and inference fleets in the world. Microsoft Research and the AI platform teams work across foundation models, systems for large-scale training, and applied ML in every product line.
Founded in 1975 and headquartered in Redmond, Washington, the company is also OpenAI's principal compute partner and ships AI tooling for developers through GitHub, VS Code and Azure AI.
Already have an account? Sign in
Continue without an account and apply on the Microsoft website
Search by role, company, or anything a posting mentions.