**LiveRamp is the data collaboration platform of choice for the world's most innovative companies. A groundbreaking leader in consumer privacy, data ethics, and foundational identity, LiveRamp is setting the new standard for building a connected customer view with unmatched clarity and context while protecting precious brand and consumer trust. LiveRamp offers complete flexibility to collaborate wherever data lives to support the widest range of data collaboration use cases-within organizations, between brands, and across its premier global network of top-quality partners.**
**Hundreds of global innovators, from iconic consumer brands and tech giants to banks, retailers, and healthcare leaders turn to LiveRamp to build enduring brand and business value by deepening customer engagement and loyalty, activating new partnerships, and maximizing the value of their first-party data while staying on the forefront of rapidly evolving compliance and privacy requirements.**
You will:
+ Design, build, deploy and optimize infrastructure of big data tools like streaming platform clusters(ex: kafka, Redpanda) and orchestration tools like Airflow, temporal on kubernetes and cloud.
+ Implement and harden operational tooling for observability (metrics, logs, traces), incident detection, and automated remediation to reduce MTTR and improve system resilience.
+ Maintain and enhance CI/CD tooling and terraform scripts to improve the efficiency of deployments and infra automation.
+ Automate infrastructure provisioning using Infrastructure as Code (IaC) to enable repeatable, self-service environment creation for engineering teams.
+ Lead and collaborate with technical, application, and security stakeholders to deliver reliable, secure Big Data infrastructure leveraging tools and platforms.
+ Own and participate in on-call responsibilities for Big Data infrastructure, triaging and resolving incidents, responding to tickets, and ensuring systems consistently meet defined SLAs for availability, performance, and data quality.
+ Mentor engineers on DevOps best practices, helping teams adopt platform capabilities, automation patterns, and operational standards
+ Define and implement monitoring, alerting, and runbooks to provide end‑to‑end observability and drive continuous improvement in reliability and operational excellence using tools like Grafana.
+ Onboard, train, and set up infra for tools by working with vendor teams and external partners so they can effectively support, operate, and extend the solutions owned by the Big Data Infrastructure (BDI) team.
+ Drive the technical roadmap for emerging data and infrastructure technologies by evaluating options, building proofs of concept (POCs), and authoring solution selection and design documents.
Your team will:
+ Build and operate the shared infrastructure, tooling, and platforms that power LiveRamp's core data and identity services, with a focus on reliability, security, and scale
+ Provide paved roads for service development and deployment, so product teams can onboard quickly to standardized CI/CD, runtime, and observability stacks.
+ Partner with application, data, and security teams to design infrastructure architectures that support high-throughput data processing and low-latency services.
+ Define and evolve best practices for operational readiness, production rollouts, configuration management, and runtime governance across the engineering organization.
+ Maintain and continuously improve shared environments (staging, performance, and production) to ensure predictable releases and consistent behavior across tiers.
+ Act as the escalation point for complex infrastructure and deployment issues, collaborating across teams to debug, root cause, and remediate systemic problems.
+ Drive cross-functional initiatives to modernize our platform (for example, container orchestration, secrets management, and cost optimization for cloud resources).
+ Build infra for large‑scale data processing tools, analytics, and AI/ML workloads across LiveRamp using tools like Dataproc, Redpanda, Airflow and Temporal.
+ Provide secure, self-service, and cost‑efficient data environments that enable product and data teams to experiment, ship, and scale data collaboration applications quickly.
+ Partner closely with application, security, SRE, and platform teams to ensure data systems meet LiveRamp's standards for privacy, compliance, reliability, and performance.
About you:
+ 7+ years of experience in DevOps, Site Reliability Engineering (SRE), or infrastructure-focused software engineering roles supporting production systems.
+ Bachelor's degree in Computer Science, Engineering, Mathematics, or a related technical field, or equivalent practical experience.
+ Significant hands-on experience with at least one major cloud provider (such as AWS, GCP, or Azure) and the associated networking, security, and managed services.
+ Strong proficiency with Infrastructure as Code toolin