MNC InsiderMNC Insider
Microsoft logo

Principal Software Engineer(M365 Storage Fabric team)

Microsoft

Principal Software Engineer(M365 Storage Fabric team)

full-timePosted: Aug 24, 2026Updated: Aug 27, 2026Beijing, Beijing, CN

Job Description

OverviewMicrosoft 365 (M365) is at the center of Microsoft's cloud-first, AI-powered productivity strategy, bringing together services such as Teams, Exchange, SharePoint, Microsoft Search, Copilot, and Office to empower organizations around the world. The M365 Storage Fabric team builds foundational platform capabilities that enable Microsoft 365 services to operate reliably, efficiently, and at hyperscale. We develop systems that optimize resource utilization, manage rapidly changing workload demands, improve service resilience, and support the next generation of AI-powered experiences. Our mission is to deliver dependable, scalable, and efficient infrastructure that helps Microsoft 365 customers achieve more while maximizing operational excellence across the platform. As AI workloads continue to accelerate across Microsoft 365, we are driving the evolution of intelligent platform services that can automatically adapt to dynamic traffic patterns, optimize infrastructure efficiency, and protect customer experiences at massive scale. Microsoft’s mission is to empower every person and every organization on the planet to achieve more. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.ResponsibilitiesThe team is looking for an experienced engineer to:Lead the design and development of large-scale distributed systems that manage resource allocation, workload execution, and service protection.Define and drive technical strategy for platform capabilities that improve reliability, efficiency, scalability, and operational excellence.Build intelligent, signal-driven automation using telemetry, health indicators, and real-time platform insights.Develop solutions that balance customer experience, infrastructure utilization, operational cost, and service performance.Drive innovations that proactively identify, mitigate, and prevent service disruptions.Partner with engineering teams across Microsoft 365, Copilot, Azure, and infrastructure organizations to deliver end-to-end platform solutions.Influence architecture, design, and engineering best practices across multiple teams.Mentor engineers and contribute to a culture of technical excellence and continuous improvement.Help shape how future AI-powered services are scaled, managed, protected, and optimized across Microsoft 365.QualificationsRequired Qualifications:Bachelor's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or PythonOR equivalent experience.Preferred Qualifications:Master's Degree in Computer Science or related technical field AND 8+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or PythonOR Bachelor's Degree in Computer Science or related technical field AND 12+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or PythonOR equivalent experience.Deep understanding of distributed systems, concurrency, reliability engineering, scalability, and platform architecture.Experience developing and operating highly available backend services at scale.Demonstrated ability to lead technically complex initiatives across multiple teams and organizations.Solid problem-solving skills involving system performance, resiliency, resource management, and operational excellence.Demonstrated ability to effectively leverage AI-assisted engineering tools and autonomous coding agents to improve software development productivity, quality, and operational effectiveness.Ability to critically evaluate, validate, and refine AI-generated code, designs, tests, diagnostics, and recommendations while maintaining full engineering ownership and accountability.Experience building infrastructure platforms, resource management systems, scheduling systems, load balancing systems, storage platforms, or large-scale backend services.Solid understanding of workload management, traffic engineering, fault tolerance, admission control, and capacity planning.Experience using telemetry, monitoring, and service health signals to drive automated operational decisions.Experience supporting services with rapidly changing demand patterns and large-scale customer workloads.Hands-on experience with cloud-native architectures and distributed platforms such as Azure or similar cloud environments.Experience with:Microservices and service-oriented architecturesEvent-driven systemsLarge-scale telemetry systemsContainerized and cloud-native environmentsExperience building or supporting AI-powered services and high-throughput systems.Familiarity with AI workload characteristics, including bursty traffic, latency sensitivity, resource contention, and dynamic scaling requirements.Experience integrating AI-assisted engineering workflows into software development, testing, debugging, code review, and operational processes.Proven ability to drive technical alignment and influence engineering decisions across organizational boundaries.Solid communication skills with the ability to explain complex technical concepts to diverse audiences. This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.

Locations

  • Beijing, Beijing, CN
  • Suzhou, Jiangsu, CN
  • Shanghai, Shanghai, CN

Skills Required

  • coding in languages includingintermediate
  • infrastructure platformsintermediate
  • telemetryintermediate
  • cloud-native architecturesintermediate
  • or supporting AI-powered servicesintermediate
  • AI workload characteristicsintermediate

Required Qualifications

  • Bachelor's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or PythonOR equivalent experience. (experience, 6 years)

Preferred Qualifications

  • Master's Degree in Computer Science or related technical field AND 8+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or PythonOR Bachelor's Degree in Computer Science or related technical field AND 12+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or PythonOR equivalent experience. (experience, 8 years)
  • Deep understanding of distributed systems, concurrency, reliability engineering, scalability, and platform architecture. (experience)
  • Experience developing and operating highly available backend services at scale. (experience)
  • Demonstrated ability to lead technically complex initiatives across multiple teams and organizations. (experience)
  • Solid problem-solving skills involving system performance, resiliency, resource management, and operational excellence. (experience)
  • Demonstrated ability to effectively leverage AI-assisted engineering tools and autonomous coding agents to improve software development productivity, quality, and operational effectiveness. (experience)
  • Ability to critically evaluate, validate, and refine AI-generated code, designs, tests, diagnostics, and recommendations while maintaining full engineering ownership and accountability. (experience)
  • Experience building infrastructure platforms, resource management systems, scheduling systems, load balancing systems, storage platforms, or large-scale backend services. (experience)
  • Solid understanding of workload management, traffic engineering, fault tolerance, admission control, and capacity planning. (experience)
  • Experience using telemetry, monitoring, and service health signals to drive automated operational decisions. (experience)
  • Experience supporting services with rapidly changing demand patterns and large-scale customer workloads. (experience)
  • Hands-on experience with cloud-native architectures and distributed platforms such as Azure or similar cloud environments. (experience)
  • Experience with: (experience)
  • Microservices and service-oriented architectures (experience)
  • Event-driven systems (experience)
  • Large-scale telemetry systems (experience)
  • Containerized and cloud-native environments (experience)
  • Experience building or supporting AI-powered services and high-throughput systems. (experience)
  • Familiarity with AI workload characteristics, including bursty traffic, latency sensitivity, resource contention, and dynamic scaling requirements. (experience)
  • Experience integrating AI-assisted engineering workflows into software development, testing, debugging, code review, and operational processes. (experience)
  • Proven ability to drive technical alignment and influence engineering decisions across organizational boundaries. (experience)
  • Solid communication skills with the ability to explain complex technical concepts to diverse audiences. (experience)
  • This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled. (experience)

Responsibilities

  • The team is looking for an experienced engineer to:
  • Lead the design and development of large-scale distributed systems that manage resource allocation, workload execution, and service protection.
  • Define and drive technical strategy for platform capabilities that improve reliability, efficiency, scalability, and operational excellence.
  • Build intelligent, signal-driven automation using telemetry, health indicators, and real-time platform insights.
  • Develop solutions that balance customer experience, infrastructure utilization, operational cost, and service performance.
  • Drive innovations that proactively identify, mitigate, and prevent service disruptions.
  • Partner with engineering teams across Microsoft 365, Copilot, Azure, and infrastructure organizations to deliver end-to-end platform solutions.
  • Influence architecture, design, and engineering best practices across multiple teams.
  • Mentor engineers and contribute to a culture of technical excellence and continuous improvement.
  • Help shape how future AI-powered services are scaled, managed, protected, and optimized across Microsoft 365.
  • Lead the design and development of large-scale distributed systems that manage resource allocation, workload execution, and service protection.
  • Define and drive technical strategy for platform capabilities that improve reliability, efficiency, scalability, and operational excellence.
  • Build intelligent, signal-driven automation using telemetry, health indicators, and real-time platform insights.
  • Develop solutions that balance customer experience, infrastructure utilization, operational cost, and service performance.
  • Drive innovations that proactively identify, mitigate, and prevent service disruptions.
  • Partner with engineering teams across Microsoft 365, Copilot, Azure, and infrastructure organizations to deliver end-to-end platform solutions.
  • Influence architecture, design, and engineering best practices across multiple teams.
  • Mentor engineers and contribute to a culture of technical excellence and continuous improvement.
  • Help shape how future AI-powered services are scaled, managed, protected, and optimized across Microsoft 365.

Benefits

  • general: Flexibility: Balance what matters—your work, your life, and your team—through trust, autonomy, and shared accountability
  • general: Growth: Stretch your skills, expand your impact, and grow with support that meets you where you are
  • general: Wellbeing: Support for your body, mind, and financial future—so you can stay energized and do your best work
  • general: Community PCS: Find your people, build your network, and feel supported every step of the way

Travel Requirements

Less than 25%

Target Your Resume for "Principal Software Engineer(M365 Storage Fabric team)" , Microsoft

Get personalized recommendations to optimize your resume specifically for Principal Software Engineer(M365 Storage Fabric team). Takes only 15 seconds!

AI-powered keyword optimization
Skills matching & gap analysis
Experience alignment suggestions

Check Your ATS Score for "Principal Software Engineer(M365 Storage Fabric team)" , Microsoft

Find out how well your resume matches this job's requirements. Get comprehensive analysis including ATS compatibility, keyword matching, skill gaps, and personalized recommendations.

ATS compatibility check
Keyword optimization analysis
Skill matching & gap identification
Format & readability score

Tags & Categories

Software EngineeringSoftware EngineeringSoftware EngineeringSoftware Engineering

Answer 10 quick questions to check your fit for Principal Software Engineer(M365 Storage Fabric team) @ Microsoft.

Quiz Challenge
10 Questions
~2 Minutes
Instant Score

Related Books and Jobs

No related jobs found at the moment.