Data Engineer
Software Engineering, Data Science
Redmond, WA, USA
USD 102,100-202,200 / year
Commercial Engineering & AI (CEAI) partners closely with stakeholders to accelerate the transformation of Microsoft’s commercial business into a frontier organization. We bring together AI‑native engineering, modern platforms, and deep commercial insight to reimagine how work gets done - at scale and with impact. Our mission is to unlock new ways of operating through intelligent systems while creating the conditions for our teams to do the most meaningful work of their careers. If you’re excited to build, experiment, and shape the future of commercial execution with AI at the core, we invite you to apply.
Within this context, the Scale and Partner AI Experiences (SPAI) organization is at the forefront of transforming Microsoft’s sales ecosystem through AI-first innovation. Our mission is to build intelligent, scalable platforms that power partner and customer success, accelerating Microsoft’s commercial growth through modern, AI-driven experiences.
As a Software Engineer, you will join our AMP Platform Engineering team. In this role, you will help operate and evolve the AMP platform, ensuring reliability, security, compliance, cost efficiency, and operational excellence across AMP-Processing and AMP-Analytics environments.
You will work closely with platform engineers, solution teams, security organizations, compliance stakeholders, and Microsoft product groups to maintain a highly available and secure platform while enabling engineering teams to innovate and deliver business value.
Responsibilities
Platform Operations & Reliability
- Own day-to-day operational health of AMP-Processing and AMP-Analytics environments.
- Monitor platform availability, capacity utilization, performance, and service health.
- Drive incident management, troubleshooting, root cause analysis, and postmortem reviews.
- Partner with engineering teams to ensure production workloads meet platform standards and operational requirements.
- Participate in live-site operations and support critical production incidents.
Security, Compliance & SFI
- Drive Secure Future Initiative (SFI) and Security 360 (S360) remediation activities across the platform.
- Partner with security, privacy, compliance, and engineering teams to maintain platform compliance.
- Monitor security recommendations and vulnerability remediation efforts.
- Support SDL, TrIP, OSOT, privacy, and governance requirements for hosted workloads.
- Ensure platform services comply with Microsoft engineering and security standards.
Resource Provisioning & Platform Enablement
- Provision and maintain Azure subscriptions, resource groups, Fabric capacities, workspaces, identities, and platform resources.
- Support onboarding of new projects and engineering teams onto AMP.
- Configure access, entitlements, service tree components, deployment pipelines, and operational controls.
- Maintain platform architecture, service inventory, and ownership records.
Cost Optimization & Capacity Management
- Monitor platform spending, capacity utilization, and operational efficiency.
- Identify and implement cost-reduction opportunities without compromising reliability or performance.
- Build reporting and visibility into platform costs and utilization trends.
- Partner with solution owners to optimize resource consumption and forecast capacity requirements.
- Support budget planning and chargeback/showback processes.
Product Group Escalation & Vendor Coordination
- Serve as the primary platform engineering contact for complex platform issues requiring Product Group engagement.
- Coordinate escalations with Microsoft Fabric, Azure, Power BI, Security, and other engineering teams.
- Track issue resolution through ICMs, support engagements, and engineering workstreams.
- Drive timely communication and resolution of service-impacting issues.
Engineering Excellence
- Improve platform automation, observability, monitoring, and operational tooling.
- Develop scripts, dashboards, and operational processes that reduce manual effort and improve reliability.
- Contribute to architecture reviews, platform standards, and operational best practices.
- Help define and evolve the AMP Platform operating model.
Qualifications
Required/minimum qualifications
Master's Degree in Computer Science, Math, Software Engineering, Computer Engineering, or related field AND 1+ year(s) experience in business analytics, data science, software development, data modeling, or data engineering OR Bachelor's Degree in Computer Science, Math, Software Engineering, Computer Engineering, or related field AND 2+ years experience in business analytics, data science, software development, data modeling, or data engineering OR equivalent experience.
Other Requirements
Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include but are not limited to the following specialized security screenings: Microsoft Cloud Background Check. This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter.
Preferred Qualifications
- Experience supporting production cloud platforms running business-critical workloads.
- Experience with Microsoft Azure services, including Subscriptions, Resource Groups, Managed Identities, RBAC, Networking fundamentals, Monitoring and diagnostics
- Experience managing operational incidents, outages, escalations, and live-site support activities.
- Experience implementing or supporting security and compliance initiatives, including vulnerability remediation and security controls.
- Experience with infrastructure governance, environment provisioning, and access management.
- Communication skills and ability to collaborate across technical and business organizations.
- Experience with Microsoft Fabric, Power BI, Lakehouse architectures, or enterprise analytics platforms.
- Experience supporting Data Engineering, Analytics, AI, or Business Intelligence environments.
- Experience with CI/CD pipelines, GitHub, Azure DevOps, and automation frameworks.
- Experience with cost optimization and capacity planning for Azure or Fabric services.
- Knowledge of observability, monitoring, telemetry, and platform reliability practices.
#CEAIjobs
Data Engineering IC3 - The typical base pay range for this role across the U.S. is USD $102,100 - $202,200 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $133,800 - $219,200 per year.
Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay
This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.
Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.