Software Engineer II and/or Senior Software Engineer
Design, implement, test, deploy, and operate highly available distributed services and automation used to configure and migrate large-scale telemetry workloads that will be leveraged across the fleet. Build and enhance APIs, tools, and subsystems for telemetry collection, routing, storage, and efficient data access. Integrate advanced capabilities (e.g., machine learning-based anomaly detection and data validation) to enhance platform intelligence and insights. Implement robust monitoring, alerting, and diagnostics and ensure production services run reliably, including participation in on-call rotations and incident response. Collaborate with partner teams to deliver end-to-end observability solutions and contribute to design reviews and best practices that uphold high engineering standards. Help evolve how the team builds with AI, championing agentic development practices and sharing repeatable patterns that improve engineering velocity and quality, as well as training and evolving AI to assist with non-coding scenarios such as site reliability engineering, support and administration. Bachelor's Degree in Computer Science or related technical field AND 2+ years technical engineering experience with coding in languages including, but not limited to, C#, Java, .NET, or Python OR equivalent experience. These requirements include, but are not limited to, the following specialized security screenings: Master's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C#, Java, .NET, or Python OR Bachelor's Degree in Computer Science or related technical field AND 8+ years technical engineering experience with coding in languages including, but not limited to, C#, Java, .NET, or Python OR equivalent experience. Experience with cloud-native infrastructure, distributed systems, or large-scale production services. Experience improving reliability, security, observability, networking, or operational excellence for live services. Experience partnering across teams to deliver infrastructure capabilities with clear ownership and supportability.