Microsoft Logo

Microsoft

Principal Software Engineer

Posted 3 Days Ago
Be an Early Applicant
In-Office
Vancouver, BC, CAN
Senior level
In-Office
Vancouver, BC, CAN
Senior level
Lead technical strategy and architecture for enabling and optimizing frontier-scale AI models on Microsoft accelerators. Design performance improvements across hardware, compilers, kernels, frameworks, runtimes, and distributed infrastructure. Drive hardware-software co-design, distributed training and inference, workload optimization, and model enablement using frameworks such as PyTorch, Triton, vLLM, and SGLang. Lead cross-organization initiatives, influence platform roadmaps, conduct architecture reviews, mentor engineers, and shape AI infrastructure strategy.
The summary above was generated by AI
Overview
Do you want to help shape the future of AI infrastructure and influence the hardware-software co-design decisions powering Microsoft's next generation of AI platforms?
Join the Systems Planning and Architecture (SPARC) team within Azure Hardware Systems and Infrastructure (AHSI), where we are building infrastructure for some of the world's most demanding AI workloads. Our team drives model enablement, performance optimization, and hardware-software co-design across Microsoft's custom AI accelerators and cloud-scale AI systems.
As a Principal Software Engineer, you will provide technical leadership for enabling and optimizing frontier-scale AI models on Microsoft AI accelerators. You will work across hardware architecture, compilers, kernels, frameworks, runtimes, and model teams to define performance strategies, influence future platform investments, and improve training and inference efficiency at scale. This opportunity places you at the intersection of AI systems, distributed computing, and accelerator architecture, with the ability to influence technical direction across multiple engineering organizations.
This role is flexible and offers a hybrid work model with three days per week in the office.
Microsoft's mission is to empower every person and every organization on the planet to achieve more. As employees, we come together with a growth mindset, innovate to empower others, and collaborate to achieve our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.

Responsibilities
  • Lead technical strategy for enabling and optimizing frontier-scale AI models across Microsoft AI accelerators and cloud-scale AI infrastructure.
  • Architect and optimize AI systems across hardware, compilers, kernels, frameworks, runtimes, and distributed infrastructure to improve training and inference performance.
  • Partner with hardware architecture teams on hardware-software co-design, influencing accelerator features, memory systems, interconnects, execution models, and future silicon roadmaps.
  • Drive performance optimization across kernels, communication, memory movement, quantization, attention, mixture-of-experts (MoE), and other critical AI workloads.
  • Architect distributed training and inference solutions spanning large accelerator clusters, including parallelism, communication, memory management, and scaling strategies.
  • Drive model enablement and performance improvements across AI frameworks and inference technologies such as PyTorch, Triton, vLLM, SGLang, and related ecosystems.
  • Lead architecture reviews and complex cross-organization technical initiatives, build alignment across engineering teams, mentor engineers, and provide technical recommendations that inform Microsoft's long-term AI infrastructure strategy.

Qualifications
Required/minimum qualifications
  • Bachelor's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python OR equivalent experience.
  • Experience partnering with hardware architecture teams on hardware-software co-design, accelerator enablement, or performance optimization.

Other Requirements
  • Ability to meet Microsoft, customer, and/or government security screening requirements, including Microsoft Cloud Background Check requirements.
Preferred Qualifications
  • 7+ years of experience developing and optimizing high-performance AI systems, kernels, or accelerator software using CUDA, ROCm, Triton, or similar programming models.
  • Deep experience optimizing large-scale AI workloads, including attention, mixture-of-experts (MoE), quantization, FP8, KV-cache management, memory efficiency, or related techniques.
  • Experience enabling and optimizing large language, reasoning, multimodal, or other foundation models on AI accelerators.
  • Experience designing distributed training or inference systems using techniques such as tensor, pipeline, expert, or sequence parallelism.
  • Deep knowledge of AI frameworks such as PyTorch and experience optimizing production-scale training or inference workloads.
  • Demonstrated experience leading complex technical initiatives across multiple engineering organizations and influencing technical strategy beyond immediate team boundaries.
  • Publications, patents, open-source contributions, or other recognized contributions in AI systems, distributed computing, machine learning infrastructure, or hardware acceleration.

Software Engineering IC5 - The typical base pay range for this role across the U.S. is USD $142,800 - $274,800 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $188,000 - $304,200 per year.

Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay


This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.



Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.

Microsoft Vancouver, British Columbia, CAN Office

155 Water St, Vancouver, BC, Canada, V6B 1A7

Similar Jobs

4 Days Ago
In-Office
Vancouver, BC, CAN
Expert/Leader
Expert/Leader
Cloud • Fintech • HR Tech
Principal Software Engineer responsible for evolving Evisort’s AI-driven contract management platform. The role involves collaborating across engineering, data science, product, and design; shaping technical strategy and architecture; building scalable SaaS solutions; owning systems from development through production; mentoring teammates; and applying distributed systems, cloud, DevOps, containerization, testing, and observability practices.
Top Skills: Asp.NetAWSAws DynamodbCi/CdDevOpsDjangoDockerElasticsearchFastapiFlaskJavaKubernetesMs Sql ServerNestjsPostgresRuby On RailsSpring BootTypescript
One Month Ago
Remote or Hybrid
Burnaby, BC, CAN
Expert/Leader
Expert/Leader
Gaming • Information Technology • Mobile • Software • Esports
Lead architecture and implementation of cinematic and narrative systems in Unreal Engine 4/5, including Sequencer, camera, animation, dialog, and editor tools. Own runtime systems, authoring workflows, performance optimization, and integration with Animation, Audio, and Design. Mentor engineers, define technical standards, and research new engine features to improve narrative production and creator workflows.
Top Skills: Animation BlueprintsC++Control RigDialog SystemsLip-SyncLocalization WorkflowsMetahumanMoviesceneNarrative Scripting SystemsPerformance CaptureSkeletal MeshUnreal Engine 4Unreal Engine 5Unreal Sequencer
One Month Ago
Hybrid
Vancouver, BC, CAN
Expert/Leader
Expert/Leader
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Lead architecture and platform strategy for Security Solutions, defining reusable patterns and prototypes. Partner with product, engineering, data science, and security to translate needs into scalable, secure architectures. Build proofs-of-concept, use AI-assisted coding, evaluate trade-offs, and drive cross-functional alignment to production readiness. Focus areas include agentic trust data, AI agents, cloud-native platforms, and security integrations.
Top Skills: Agentic CommerceAi AgentsAi-Assisted CodingAPIsCloud-Native PlatformsDistributed SystemsFraud DetectionGenerative AiIdentity ManagementPaymentsPrototypingSecurity

What you need to know about the Vancouver Tech Scene

Raincouver, Vancity, The Big Smoke — Vancouver is known by many names, and in recent years, it has gained a reputation as a growing hub for both tech and sustainability. Renowned for its natural beauty, the city has become a magnet for professionals eager to create environmental solutions, and with an emphasis on clean technology, renewable energy and environmental innovation, it's attracted companies across various industries, all working toward a shared goal: advancing clean technology.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account