**LiveRamp is the data collaboration platform of choice for the world's most innovative companies. A groundbreaking leader in consumer privacy, data ethics, and foundational identity, LiveRamp is setting the new standard for building a connected customer view with unmatched clarity and context while protecting precious brand and consumer trust. LiveRamp offers complete flexibility to collaborate wherever data lives to support the widest range of data collaboration use cases-within organizations, between brands, and across its premier global network of top-quality partners.**
**Hundreds of global innovators, from iconic consumer brands and tech giants to banks, retailers, and healthcare leaders turn to LiveRamp to build enduring brand and business value by deepening customer engagement and loyalty, activating new partnerships, and maximizing the value of their first-party data while staying on the forefront of rapidly evolving compliance and privacy requirements.**
As a Senior DevOps Engineer, you will join a platform-focused team that builds and operates the infrastructure foundations for LiveRamp's global data collaboration platform. The team owns shared compute, networking, storage, and Kubernetes-based orchestration layers, along with the CI/CD and observability platforms used by product engineering teams. Together, you will solve complex problems in scalability, multi-cloud and hybrid environments, cost optimization, and reliability for large-scale data and streaming workloads.
Your team will partner with service owners across the company to provide paved roads for building, deploying, and running services-abstracting away infrastructure complexity while preserving flexibility where it matters. You will collaborate closely with Security, SRE, and Architecture groups to define standards, evolve our platform roadmap, and execute cross-cutting initiatives such as major migrations, resilience programs, and compliance-driven changes.
**You will:**
+ Lead the design, implementation, and operation of highly available, scalable, and secure infrastructure that powers LiveRamp's core platform and data collaboration products.
+ Enable AI Powered development for Engineering to accelerate engineering velocity by integrating AI assistants, Agentic frameworks, and Large Language Models (LLMs) into the Software Development Life Cycle (SDLC).
+ Own critical production services end to end, including capacity planning, performance tuning, observability, reliability, and incident response, with clear SLOs and error budgets.
+ Architect and evolve CI/CD pipelines, deployment strategies, and release automation to enable fast, safe, and repeatable delivery across dozens of microservices.
+ Drive our infrastructure-as-code strategy (e.g., Terraform, Helm, Kubernetes manifests) to ensure all environments are reproducible, auditable, and easy to change.
+ Partner closely with product engineering teams as a technical leader and advisor, helping them design resilient services, optimize resource usage, and adopt DevOps best practices.
+ Lead and participate in on-call rotations, improve incident management processes, and implement preventive measures that reduce MTTR and the frequency of high-severity incidents.
+ Champion security-by-design in our stack, collaborating with security and compliance teams to harden systems, manage secrets, and maintain audit-ready infrastructure.
+ Evaluate, introduce, and standardize tooling across logging, metrics, tracing, configuration management, and workflow orchestration to improve developer and operator productivity.
+ Mentor and grow other engineers through design reviews, pairing, technical talks, and thoughtful feedback, raising the bar for DevOps and reliability practices across the org.
+ Define and track success metrics for reliability, performance, cost efficiency, and operational excellence, and use data to prioritize and communicate tradeoffs and investments.
**About you:**
+ 5+ years of experience in DevOps, Site Reliability Engineering, Infrastructure Engineering, or a closely related role, with significant time spent operating production systems at scale.
+ Hands-on experience with at least one major cloud provider (e.g., AWS, GCP, or Azure) and with building secure, reliable, and cost-conscious cloud-native architectures.
+ Strong expertise with containers and orchestration (e.g., Docker, Kubernetes), including designing multi-tenant clusters, managing deployments, and troubleshooting complex runtime issues.
+ Proven track record implementing and maintaining CI/CD pipelines and release automation (e.g., using GitHub Actions, Jenkins, GitLab CI, or similar systems).
+ Proficiency with infrastructure-as-code tools (such as Terraform, CloudFormation, Helm, or similar) and configuration management (e.g., Ansible, Chef, or Puppet).
+ Experience building and operating robust observability stacks (metrics, logs, traces) and using them to diagnose and remediate prod