logo

AI Job Summary

Lead technical strategy for optimizing frontier-scale AI models on Microsoft AI accelerators and cloud infrastructure. Partner with hardware teams on co-design decisions spanning compilers, kernels, frameworks, and distributed systems. Requires 6+ years of software engineering experience in C, C++, Python, or similar languages, plus hardware-software co-design and performance optimization background. Preferred: deep experience with CUDA, Triton, PyTorch, and large-scale AI workload optimization. Hybrid role, three days weekly in office. Base pay $142,800–$274,800 annually ($188,000–$304,200 in Bay Area or NYC).

Written from this posting by Neural Jobs AI. The full description is below.

AI Resume Tailoring Sign in to use this AI Cover Letter Sign in to use this

Job Description

Overview

Do you want to help shape the future of AI infrastructure and influence the hardware-software co-design decisions powering Microsoft's next generation of AI platforms?
Join the Systems Planning and Architecture (SPARC) team within Azure Hardware Systems and Infrastructure (AHSI), where we are building infrastructure for some of the world's most demanding AI workloads. Our team drives model enablement, performance optimization, and hardware-software co-design across Microsoft's custom AI accelerators and cloud-scale AI systems.
As a Principal Software Engineer, you will provide technical leadership for enabling and optimizing frontier-scale AI models on Microsoft AI accelerators. You will work across hardware architecture, compilers, kernels, frameworks, runtimes, and model teams to define performance strategies, influence future platform investments, and improve training and inference efficiency at scale. This opportunity places you at the intersection of AI systems, distributed computing, and accelerator architecture, with the ability to influence technical direction across multiple engineering organizations.
This role is flexible and offers a hybrid work model with three days per week in the office.
Microsoft's mission is to empower every person and every organization on the planet to achieve more. As employees, we come together with a growth mindset, innovate to empower others, and collaborate to achieve our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.


Responsibilities

  • Lead technical strategy for enabling and optimizing frontier-scale AI models across Microsoft AI accelerators and cloud-scale AI infrastructure.
  • Architect and optimize AI systems across hardware, compilers, kernels, frameworks, runtimes, and distributed infrastructure to improve training and inference performance.
  • Partner with hardware architecture teams on hardware-software co-design, influencing accelerator features, memory systems, interconnects, execution models, and future silicon roadmaps.
  • Drive performance optimization across kernels, communication, memory movement, quantization, attention, mixture-of-experts (MoE), and other critical AI workloads.
  • Architect distributed training and inference solutions spanning large accelerator clusters, including parallelism, communication, memory management, and scaling strategies.
  • Drive model enablement and performance improvements across AI frameworks and inference technologies such as PyTorch, Triton, vLLM, SGLang, and related ecosystems.
  • Lead architecture reviews and complex cross-organization technical initiatives, build alignment across engineering teams, mentor engineers, and provide technical recommendations that inform Microsoft's long-term AI infrastructure strategy.


Qualifications

Required/minimum qualifications

  • Bachelor's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python OR equivalent experience.
  • Experience partnering with hardware architecture teams on hardware-software co-design, accelerator enablement, or performance optimization.

Other Requirements

  • Ability to meet Microsoft, customer, and/or government security screening requirements, including Microsoft Cloud Background Check requirements.

Preferred Qualifications

  • 7+ years of experience developing and optimizing high-performance AI systems, kernels, or accelerator software using CUDA, ROCm, Triton, or similar programming models.
  • Deep experience optimizing large-scale AI workloads, including attention, mixture-of-experts (MoE), quantization, FP8, KV-cache management, memory efficiency, or related techniques.
  • Experience enabling and optimizing large language, reasoning, multimodal, or other foundation models on AI accelerators.
  • Experience designing distributed training or inference systems using techniques such as tensor, pipeline, expert, or sequence parallelism.
  • Deep knowledge of AI frameworks such as PyTorch and experience optimizing production-scale training or inference workloads.
  • Demonstrated experience leading complex technical initiatives across multiple engineering organizations and influencing technical strategy beyond immediate team boundaries.
  • Publications, patents, open-source contributions, or other recognized contributions in AI systems, distributed computing, machine learning infrastructure, or hardware acceleration.


Software Engineering IC5 - The typical base pay range for this role across the U.S. is USD $142,800 - $274,800 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $188,000 - $304,200 per year.

Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay


This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.




Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.

Do you match this job?

Here is what this employer asked for. Sign in and we will fill in your half.

  • Role Principal Software Engineer
  • Experience 8-9 years
  • Education Bachelor Degree, or equivalent experience
  • Salary 143K - 275K Yearly
  • Work type Hybrid
  • Location Canada
Check my match (free)
Microsoft
Hardware & Semiconductors · 500+ Members · Redmond, WA, United States

Microsoft builds Windows, Azure, Office and the Copilot family of AI assistants, and operates one of the largest AI training and inference fleets in the world. Microsoft Research and the AI platform teams work across foundation models, systems for large-scale training, and applied ML in every product line.

Founded in 1975 and headquartered in Redmond, Washington, the company is also OpenAI's principal compute partner and ships AI tooling for developers through GitHub, VS Code and Azure AI.

32 more Software Engineer roles in Vancouver

Job Overview
Salary
143K - 275K Yearly
Eligibility
Canada Right to work in Canada required.
Workplace
Hybrid
Job Posted:
1 week ago
Job Type
Full Time
Education
Bachelor Degree, or equivalent experience
Experience
8-9 years

Share This Job: