Senior Software Engineer - Kubernetes & IAC

IndiaPosted Aug 7, 2026

We are looking for a Senior Software Engineer to join our team! Power Platform brings together multiple products designed to empower customers in their digital transformation journey. In this role, you'll help build scalable, reliable, secure, and compliant AI infrastructure that powers every product across Power Platform. You'll work alongside a passionate team of engineers who thrive on solving complex challenges at scale while delivering exceptional quality. If you're excited about driving innovation and shaping the future of AI-powered solutions, we'd love to have you on board. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond. Design, build, and operate cloud-scale, multi-tenant infrastructure platforms with a strong focus on reliability, security, scalability, and operational excellence. Lead the design and implementation of Kubernetes-based infrastructure, including cluster architecture, networking, storage, and workload isolation strategies. Build and evolve Infrastructure as Code (IaC) solutions (e.g., ARM, Bicep, Terraform, Helm) to enable repeatable, auditable, and automated infrastructure provisioning and lifecycle management. Apply systems thinking to design for failure domains, including regional isolation, availability zone strategies, dependency management, and blast-radius reduction. Drive resiliency and reliability improvements through proactive design reviews, fault modeling, chaos testing, and post-incident learning. Build tooling and automation to detect, diagnose, and self-heal infrastructure and platform issues, enabling customers and support teams to self-resolve problems. Identify recurring operational issues and escalation patterns, and drive engineering solutions such as self-healing mechanisms, automation, guardrails, and platform abstractions. Partner closely with product, SRE, and Azure platform teams to define regional deployment strategies, capacity planning, and safe rollout patterns. Contribute to product and platform improvements by filing impactful bugs, proposing design changes, and shipping fixes to production to prevent customer impact. Communicate complex technical issues and recommendations clearly and concisely, influencing cross-team decisions and driving measurable business outcomes. Bachelor's Degree in Computer Science or related technical field AND 8+ years of technical engineering experience Hands-on experience with Kubernetes or container orchestration platforms in production environments. Experience with Infrastructure as Code (IaC) tools such as ARM, Bicep, Terraform, or similar. Strong understanding of distributed systems concepts, including failure domains, consistency, availability, and fault tolerance. Experience designing for high availability, regional resiliency, and disaster recovery in cloud environments. Familiarity with cloud networking, storage, and security fundamentals, including identity, access control, and isolation boundaries. Experience operating large-scale services with an emphasis on reliability engineering, incident response, and postmortem-driven improvements. Ability to reason about blast radius, dependency management, and safe rollout strategies in complex systems.

Want jobs like this matched to you?

SimpleCareer scores fresh postings against your résumé so you only see the matches that matter.

Get started free