Role at a glance
- Job function
-
AI & Data Machine Learning Engineering
- Salary
- $220.8K – $331.2K/yr
- Location
- Multiple Locations, United States United States
- Work arrangement
- Remote
- Employment
- Full-time
- Experience
- Bachelor’s Degree in Computer Science or a related technical field and 8+ years of technical engineering experience coding in languages...
- Education
- Bachelor’s Degree in Computer Science or a related technical field and 8+ years of technical engineering experience coding in languages...
Spotted an issue?
We’ll check it against the original posting.
Role Summary
The Principal Software Engineer – AI Frameworks provides technical leadership for AI frameworks and the serving stack for frontier models, including OpenAI GPT models. The role sets technical direction and leads cross-team work to optimize model performance, reliability, and efficiency across GPU and Microsoft silicon platforms.
What You'll Do
- Own the technical vision, architecture, roadmap, and multi-release execution strategy for AI framework, performance, benchmarking, or...
- Align teams and partners on architecture, priorities, decisions, execution plans, and measurable outcomes.
- Lead cross-stack investigations and resolve technical trade-offs across models, frameworks, compilers, runtimes, systems, services, and...
- Provide hands-on leadership through prototypes, implementation, design and code reviews, debugging, and operational readiness.
- Develop measurement, automation, observability, and engineering mechanisms into scalable platform capabilities.
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
Required: 8+ years of technical engineering experience coding in languages such as C++ or Python, or equivalent experience. Preferred qualifications include technical leadership across complex initiatives, high-performance or distributed systems, software and computer architecture, accelerator-aware optimization, and LLM inference or AI accelerator kernel optimization.
Required
- Bachelor’s Degree in Computer Science or a related technical field, or equivalent experience.
- 8+ years of technical engineering experience coding in languages such as C++ or Python, or equivalent experience.
Preferred
- Master's Degree in Computer Science or related technical field AND 12+ years technical engineering experience with coding in languages...
- Demonstrated experience serving as a technical lead for complex cross-team initiatives, including setting strategy, making architectural...
- Robust/extensive experience designing and shipping complex, high-performance or distributed software systems.
- Deep foundation in software architecture, computer architecture, and accelerator-aware optimization.
- Advanced experience with LLM inference frameworks or AI accelerator kernel optimizations.
- Proven track record of creating reusable platforms, influencing senior technical and business stakeholders, developing other technical...
Original job description
Content provided by the employer
Original job description
Content provided by the employer
The Microsoft AI Frameworks team develops the software and performance systems that enable state-of-the-art AI models to run reliably and efficiently at cloud scale. We own the serving stack for OpenAI models and build optimizations across model architectures, frameworks, compilers, runtimes, libraries, observability, benchmarking, and hardware platforms—including NVIDIA and AMD GPUs and Microsoft silicon. Our engineers partner with model developers, researchers, hardware teams, and production services to accelerate model onboarding, improve performance and reliability, reduce deployment time and hardware footprint, and turn performance insights into durable platform capabilities.
This is a hands-on technical leadership role for an engineer who can establish direction and lead teams through ambiguous, end-to-end systems challenges. Successful candidates combine strong software engineering fundamentals with deep curiosity about AI workloads, use disciplined measurement to guide decisions, and create alignment across organizational boundaries to deliver measurable production impact.
As a Principal Software Engineer – AI Frameworks, you will serve as a technical lead for high-impact initiatives spanning AI frameworks and the serving stack for frontier models, including OpenAI GPT models. You will establish technical direction, make consequential architectural decisions, and lead engineers across teams through the design and delivery of optimizations for NVIDIA and AMD GPUs and Microsoft silicon. Your work will improve model velocity, platform efficiency, and production reliability. You will also identify systemic opportunities, align stakeholders around durable architectures and success measures, and remain hands-on in critical areas of implementation.
At Microsoft, our mission to empower every person and every organization on the planet to achieve more guides how we partner with customers to deliver trusted, impactful solutions. With a growth mindset culture, we innovate responsibly and measure success by shared progress people, teams, and customers. Join us to do meaningful work that changes the world and helps shape what’s next for everyone.
Responsibilities
- Own the technical vision, architecture, roadmap, and multi-release execution strategy for critical AI framework, performance, benchmarking, or developer-productivity capabilities.
- Lead technical alignment across teams by shaping architecture and priorities, resolving competing requirements, and holding researchers, product groups, infrastructure owners, and hardware partners accountable to clear decisions, execution plans, and measurable outcomes.
- Build technical leadership capacity across the organization by mentoring engineers, developing emerging technical leads, establishing high standards for quality and maintainability, and creating an inclusive engineering culture in which teams can execute effectively at scale.
- Steer ambiguous, cross-stack investigations and investments spanning models, frameworks, compilers, runtimes, systems, services, and silicon; set technical priorities, resolve trade-offs, and drive teams to decisions.
- Provide hands-on technical leadership by setting the engineering approach, delegating and unblocking critical work, and directly contributing through prototypes, critical-path implementation, design and code reviews, complex debugging, and operational readiness.
- Spearhead common measurement, automation, observability, and engineering mechanisms that turn one-off analyses into scalable platform capabilities.
- Drive measurable improvements in model onboarding velocity, runtime performance, reliability, hardware utilization, and Azure capacity efficiency.
- Embody Microsoft’s culture and values.
Qualifications
Qualifications
Required Qualifications:
- Bachelor’s Degree in Computer Science or a related technical field and 8+ years of technical engineering experience coding in languages such as C++ or Python,
- OR equivalent experience.
- OR equivalent experience.
Other Requirements:
Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include but are not limited to the following specialized security screenings: Microsoft Cloud Background Check: This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter.
Preferred Qualifications:
- Master's Degree in Computer Science or related technical field AND 12+ years technical engineering experience with coding in languages including, but not limited to, C++ or Python
- OR Bachelor's Degree in Computer Science or related technical field AND 15+ years technical engineering experience with coding in languages including, but not limited to, C++ or Python
- OR equivalent experience.
- Demonstrated experience serving as a technical lead for complex cross-team initiatives, including setting strategy, making architectural decisions, aligning stakeholders, and driving execution from concept through production.
- Robust/extensive experience designing and shipping complex, high-performance or distributed software systems
- Deep foundation in software architecture, computer architecture, and accelerator-aware optimization
- Advanced experience with LLM inference frameworks or AI accelerator kernel optimizations
- Proven track record of creating reusable platforms, influencing senior technical and business stakeholders, developing other technical leaders, and scaling engineering practices across teams.
#AIInfra
Software Engineering IC6 - The typical base pay range for this role across the U.S. is USD $165,600 - $296,400 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $220,800 - $331,200 per year.
Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay
This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.
Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.
About the company
Microsoft
Large Enterprise
Microsoft is a global technology leader that empowers individuals and organizations to achieve more through innovative software, services, and devices. Founded in 1975, the company is best known for its flagship products like the Windows operating system and Microsoft Office suite. In addition to personal computing, Microsoft is a leader in cloud computing with its Azure platform, providing a range of solutions for businesses to enhance productivity and efficiency. With a strong commitment to sustainability and accessibility, Microsoft continues to drive technological advancements that shape the future of work and learning.