Senior Manager, Platform Operations
Raito
United States · Remote$168k–$210kPosted Jul 24, 2026
Skip to content Back to CareersRemote, East Coast USA Senior Manager, Platform Operations Beware of recruitment fraud: Collibra will never ask candidates for
payment or personal information through text messages or social media
(e.g., WhatsApp). The Collibra Talent Acquisition team does not request or
require details like bank account numbers, tax forms or credit card
information during the recruitment process. All official communication
comes from our company domain (@collibra.com), and job postings can be
verified on collibra.com/careers. Joining Collibra's Platform Operations team
You'll manage the team that keeps the Collibra Platform running and current for every enterprise customer, around the clock, spanning incident response, change management, release execution, and the health of the fleet that powers it.
You'll report to the Senior Director of Reliability Engineering & Operations and oversee the team directly, partnering closely with two technical leads who anchor deep expertise across the team's two domains.
This is a highly visible, business-critical function. When this team is at its best, customers don't notice it. Their environments are healthy, current, and secure. When it isn't, the whole company feels it. We're looking for someone who treats that responsibility as a craft, not just a job.
Our customers are our true north. Every alert answered, every release shipped, and every patch applied is in direct service of the enterprise customers running on this platform.
Senior Managers of Platform Operations at Collibra are responsible for
Own 24x7 incident response, SaaS customer operations, and change management in order to deliver consistent platform reliability for every enterprise customer
Direct deployment and fleet operations across thousands of virtual machines underpinning the Collibra Platform, including weekly release delivery, in order to keep customer environments healthy, current, and secure
Build succession and development plans across the team's two technical domains in order to create space for the team's strengths to surface and grow
Advance AI-powered automation across incident remediation, release delivery, and internal tooling in order to reduce manual toil and continuously modernize how the team operates
Partner across engineering, product, security, finance, and support organizations in order to align priorities, resolve cross-team dependencies, and maintain compliance in a regulated environment
You have
7+ years of experience in engineering, with at least 3+ years in a leadership or management role, overseeing incident response, release, or deployment operations for customer-facing SaaS production environments
Experience managing 24x7 production operations for customer-facing systems, including on-call rotation and escalation models
Experience operating within a FedRAMP or comparable regulated compliance environment, including continuous monitoring and audit cycles
Experience with cloud infrastructure at enterprise scale, including AWS, AWS GovCloud and GCP. Azure experience is a plus.
Experience managing distributed infrastructure fleets, virtual machines, containers, or equivalent, supporting production SaaS environments
Demonstrated proficiency in leveraging AI tools (e.g., Claude, Gemini, ChatGPT, Copilot) to solve real-world business challenges, drive measurable outcomes, or streamline workflows
A bachelor's degree or equivalent related working experience is required
Because this role supports the US government, it is required that this candidate be a US citizen who resides on US soil
You are able to
Balance hands-on technical depth with people leadership, staying credible in an incident or a release window while building growth paths for the team around you
Operate with urgency under production pressure while preserving rigor and rollback discipline
Think strategically about how the team and its practices should evolve, not only how to run today's process
Communicate...