Senior Staff DevOps Engineer - Data Discovery & AI Governance
Job description
Senior Staff DevOps Engineer - Data Discovery & AI Governance at OneTrust.
About the role
You will architect and own the infrastructure strategy for OneTrust's Detect & Discover and AI Governance platforms, shaping how Kubernetes, automation, and security controls power our responsible data and AI mission. You will lead the design of production-grade clusters across multi-cloud and on-premises environments while mentoring engineers and influencing technical direction through code reviews and design discussions. You will partner closely with product security and engineering squads to ensure platform reliability, scalability, and compliance with enterprise standards. You will drive container-hardening initiatives to eliminate vulnerabilities and strengthen the security posture of our distributed scanning and classification platform. You will define operational runbooks and observability practices that keep our services available and performant for thousands of enterprise customers. You will act as a hands-on technical leader who balances innovation with operational rigor to accelerate responsible AI adoption. You will collaborate with customer-facing teams to translate platform requirements into robust infrastructure solutions that support global deployments.
Key facts
What you'll do
Define and evolve the Kubernetes-based infrastructure strategy for on-premises and cloud environments, ensuring alignment with platform reliability and security goals.
Automate the deployment and scaling of worker nodes for Azure Marketplace, AWS Marketplace, and GCP using infrastructure-as-code tools to reduce manual intervention and accelerate provisioning.
Lead container-hardening initiatives by driving image scanning, policy enforcement, and vulnerability remediation to systematically reduce CVEs across the platform.
Design and maintain CI/CD pipelines and GitOps workflows that enable safe, auditable, and repeatable deployments for distributed scanning and classification services.
Build and operate observability, monitoring, and alerting solutions to provide end-to-end visibility into platform health, performance, and security incidents.
Partner with engineering squads to standardize platform components, improve developer experience, and remove operational bottlenecks for data discovery and AI governance products.
Mentor engineers by conducting code reviews, sharing best practices, and guiding the adoption of infrastructure patterns that support scalability and resilience.
Collaborate with product security and compliance teams to ensure infrastructure controls meet regulatory expectations and enterprise customer requirements.
Evaluate emerging technologies and run proof-of-concepts to assess their fit for secure, scalable, and efficient platform operations in multi-cloud and on-premises contexts.
Act as a technical escalation owner for critical platform incidents, driving root cause analysis and implementing preventative measures.
Requirements
Demonstrate extensive experience designing, operating, and securing Kubernetes clusters in production across multiple cloud providers and on-premises environments.
Show a proven track record with infrastructure-as-code tools such as Terraform, Helm, BICEP, CloudFormation, or equivalent technologies for automated provisioning and lifecycle management.
Bring strong expertise in container platforms, including container runtime hardening, image scanning, and policy enforcement using tools like OPA, Gatekeeper, or similar frameworks.
Have hands-on experience with CI/CD and GitOps toolchains, including Jenkins, GitHub Actions, ArgoCD, or similar systems that support secure deployment workflows.
Exhibit deep knowledge of networking, identity and access management, and security controls within cloud and on-premises infrastructures.
Display strong scripting and programming skills in at least one modern language such as Python, Go, or Bash to automate complex operational tasks.
Provide evidence of experience collaborating with cross-functional teams, including product, security, and customer success, in a regulated enterprise environment.
Possess excellent written and verbal communication skills to articulate technical trade-offs and lead design discussions with stakeholders at all levels.
Nice to have
Experience with marketplace deployments and supporting customer self-service provisioning flows in cloud marketplaces.
Background in regulated industries where data governance, auditability, and compliance are critical.
Practical notes
This role is based in Atlanta, Georgia.
Travel may be required as part of the position.
Visa sponsorship may be available for qualified candidates.