Staff Engineer (Cloud Platforms)
<p><strong>About the Role: </strong></p><p><br></p><p>We are looking for a <strong>Staff Engineer</strong> to join our <strong>Cloud Platform</strong> team and take ownership of the architecture and evolution of our multi-tenant SaaS platform. You will work at the intersection of backend engineering and cloud infrastructure, designing and building the systems that power our control plane and data plane at scale.</p><p><br></p><p>This is a high-impact, high-ownership role. You will be a technical anchor for the team: driving architecture decisions, mentoring engineers, and partnering with product and infrastructure leaders to shape the future of our cloud platform.</p><p><br></p><p><strong>What you'll do:</strong></p><p><br></p><p><strong>Architecture &amp; Design</strong></p><p><br></p><p>- Own end-to-end architecture for the cloud platform, spanning control plane and data plane</p><p>- Design multi-tenant systems with strong isolation, security, and resource governance</p><p>- Define platform abstractions that work across hybrid environments: AWS, GCP, Azure, and on-premises</p><p>- Drive architectural reviews, RFCs, and ensure decisions are well-reasoned, documented, and scalable</p><p><br></p><p><strong> Control Plane</strong></p><p><br></p><p>- Architect and build systems for tenant provisioning, lifecycle management, and configuration</p><p>- Design cluster orchestration and management systems that operate reliably at scale</p><p>- Build APIs and automation that enable self-service for operators and tenants</p><p>- Ensure control plane is highly available, auditable, and observable</p><p><br></p><p><strong> Data Plane</strong></p><p><br></p><p>- Design data path components for high-throughput, low-latency workloads</p><p>- Build and enforce isolation boundaries between tenants at the data layer</p><p>- Optimise for performance, reliability, and cost efficiency at scale</p><p><br></p><p><strong> Engineering Excellence</strong></p><p><br></p><p>- Set technical direction and coding standards for the platform team</p><p>- Identify and address systemic risks — reliability, scalability, security, and operability</p><p>- Partner with SRE and DevOps on observability, incident response, and capacity planning</p><p><br></p><p><strong> Mentorship &amp; Leadership</strong></p><p><br></p><p>- Mentor and grow senior and mid-level engineers</p><p>- Contribute to hiring — define bar, conduct interviews, help build the India platform team</p><p>- Collaborate closely with cross geo-based engineering and product teams</p><p><br></p><p><strong>What you must bring:</strong></p><p><br></p><p><strong>SaaS Platform &amp; Multi-Tenancy</strong></p><p><br></p><p>- Hands-on experience building or operating multi-tenant SaaS platforms at scale silo/pool/bridge models, noisy neighbour mitigation, tenant resource quotas, and lifecycle automation (onboarding, provisioning, off-boarding)</p><p>- Knowledge of data isolation strategies schema-per-tenant, database-per-tenant, row-level security, and per-tenant encryption at rest</p><p><br></p><p><strong>Control Plane &amp; Data Plane</strong></p><p><br></p><p>- Proven experience with control plane / data plane separation and cluster management — Kubernetes operators, CRDs, admission web-hooks, RBAC, and namespace isolation</p><p>- Understanding of configuration management at scale — GitOps workflows, feature flags, and dynamic config propagation across distributed environments</p><p><br></p><p><strong>Cloud &amp; Hybrid Infrastructure</strong></p><p><br></p><p>- Deep expertise in at least two of AWS, GCP, or Azure VPC, IAM, managed Kubernetes (EKS/GKE/AKS), IaC (Terraform/Pulumi), and cost optimisation across hybrid environments</p><p>- Familiarity with service mesh technologies — Istio, Linkerd, or Envoy — for traffic management, mTLS, and microservices observability</p><p><br></p><p><strong>Distributed Systems &amp; Security</strong></p><p><br></p><p>- Strong grasp of distributed systems fundamentals HA patterns, fault tolerance, observability (OpenTelemetry, Prometheus, Grafana), and resilience testing</p><p>- Understanding of zero-trust security, secrets management (Vault, AWS Secrets Manager), and compliance frameworks (SOC 2, ISO 27001, GDPR) at the infrastructure level</p><p><br></p><p><strong>AI-Augmented Engineering</strong></p><p><br></p><p>- Actively uses AI coding assistants — Claude, Cursor, Copilot — for infrastructure tasks, runbook generation, incident analysis, and ADR drafting</p><p>- Able to craft effective prompts for complex engineering problems and critically evaluate AI output; evaluates and drives AI tool adoption across the platform team</p><p><br></p><p><strong> Good to Haves </strong></p><p><br></p><p>- Exposure to FinOps practices: cost attribution per tenant, showback/chargeback models, and cloud cost anomaly detection</p><p>- Familiarity with eBPF or low-level networking for custom observability or performance optimisation</p><p>- Contributions to open-source infrastructure or platform projects</p><p><br></p><p><strong>What success looks like: </strong></p><p><br></p><p><strong>In 3 months: </strong></p><p><br></p><p>- Deep understanding of the current platform architecture, gaps, and roadmap</p><p>- Shipped at least one meaningful improvement to control or data plane</p><p>- Established strong working relationships across engineering and product</p><p><br></p><p><strong>In 6 months: </strong></p><p><br></p><p>- Own and drive architecture for one major platform initiative end-to-end</p><p>- Identified and mitigated at least one systemic risk</p><p>- Actively contributing to India team hiring and growth</p><p><br></p><p><strong>In 12 months: </strong></p><p><br></p><p>- Recognised as the technical authority for the cloud platform in India</p><p>- Multi-tenant and hybrid cloud architecture significantly evolved and documented</p><p>- A set of senior engineers measurably grown in scope and impact under your mentorship</p>