We use cookies. Find out more about it here. By continuing to browse this site you are agreeing to our use of cookies.
#alert
Back to search results
New

Senior Researcher - GPU Performance

Microsoft
United States, Washington, Redmond
Oct 26, 2025
OverviewGenerative AI is transforming how people create, collaborate, and communicate - redefining productivity across Microsoft 365 and our customers globally. At Microsoft, we run the biggest platform for collaboration and productivity in the world with hundreds of millions of consumer/enterprise users. Tackling AI efficiency challenges is crucial for delivering these experiences at scale.Within our Microsoft wide Systems Innovation initiative, we are working to advance efficiency across AI systems, where we look at novel designs and optimizations across AI stacks: models, AI frameworks, cloud infrastructure, and hardware. We are an Applied Research team driving mid- and long-term product innovations. We closely collaborate with multiple research teams and product groups across the globe who bring a multitude of technical knowledge in cloud systems, machine learning and software engineering. We communicate our research both internally and externally through academic publications, open-source releases, blog posts, patents, and industry conferences. Further, we also collaborate with academic and industry partners to advance the state of the art and target material product impact that will affect 100s of millions of customers.We are looking for a Senior Researcher - GPU Performance - Hardware/Software Codesign researcher to explore hardware/kernel-level optimizations to deliver significant efficiency gains for Large Language Models and Generative AI experiences.The qualified candidate will have a solid background in GPU architecture, accelerator design, machine learning, or systems research and the ambition to apply them to large scale production systems. This role combines deep technical knowledge in GPU architecture with practical implementation skills to create efficient, scalable computational kernels. Further, the qualified candidate is expected to demonstrate a history of solving hard technical problems and is motivated to tackle the hardest problems in building a full end-to-end AI stack. An entrepreneurial approach and ability to take initiative and move fast are essential. Have a look at this link for reading: Efficient AI - Microsoft Research Microsoft's mission is to empower every person and every organization on the planet to achieve more. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.
ResponsibilitiesDesign, implement, and optimize GPU kernels for complex computational workloads such as AI inferencing.Research and develop novel optimization techniques for generation of GPU kernels.Profile and analyze kernel performance using advanced diagnostic tools.Generate automated solutions for kernel optimization and tuning.Collaborate with other researchers to improve model performance.Document optimization strategies and maintain performance benchmarks.Contribute to the development of internal GPU computing frameworks.
Applied = 0

(web-675dddd98f-24cnf)