PazaJobs

Site Reliability Engineer II, tvScientific

Pinterest · San Francisco, CA, US; Remote, US
Remote FULL TIME MID tvScientific United States 114.297–235.319 USD

Depaza has read the full posting at the employer and structured it for you.

Responsibilities

  • Sikre pålidelighed, tilgængelighed og ydeevne af produktionsinfrastruktur
  • Drive og skalere Kubernetes-platforme
  • Administrere GitOps-baserede deployment workflows med ArgoCD og Helm
  • Supportere infrastructure provisioning via Terraform/Terragrunt
  • Bygge og vedligeholde CI/CD automation med GitHub Actions
  • Deltage i incident response og root cause analysis
  • Reducere operational toil gennem scripting og automation
  • Avancere observability practices (logs, metrics, traces, dashboards, alerting)
  • Supportere secure secrets integration og IAM-aware operations
  • Samarbejde med application, security og platform teams

Requirements

  • 4+ års erfaring inden for SRE, DevOps, Platform Engineering eller Cloud Infrastructure
  • Produktionserfaring med AWS
  • Ekspertise i Kubernetes (cluster operations, troubleshooting, workload reliability)
  • Erfaring med Kubernetes multi-tenancy (namespaces, RBAC, quotas, policies)
  • Erfaring med ArgoCD og GitOps
  • Stærk erfaring med Helm
  • Erfaring med Terraform/Terragrunt
  • Scripting i Bash og/eller Python
  • CI/CD pipelines, gerne GitHub Actions
  • Troubleshooting på tværs af Linux, containers, IAM, networking og distributed systems
  • Monitoring, alerting og observability i produktion
  • Evne til at bruge AI til at forbedre hastighed og kvalitet i dagligt arbejde

Skills

  • AWS
  • Kubernetes
  • EKS
  • ArgoCD
  • GitOps
  • Helm
  • Terraform
  • Terragrunt
  • GitHub Actions
  • CI/CD
  • Python
  • Bash
  • Linux
  • IAM
  • Observability
  • AI collaboration

Benefits

  • Equity compensation
  • PinFlex working model
  • Remote-eligible position
  • Inclusive workplace culture
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here.About tvScientific tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. We leverage massive data and cutting-edge science to automate and optimize TV advertising to drive business outcomes. Our solution combines media buying, optimization, measurement, and attribution in one, efficient platform. Our platform is built by industry leaders with a long history in programmatic advertising, digital media, and ad verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business. We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows. This role will contribute to improving the reliability, scalability, automation, observability, and operational maturity of our infrastructure and delivery ecosystem. The ideal candidate is a hands-on engineer with solid production experience and a strong foundation in building and supporting resilient platforms using infrastructure as code, automation, and modern Kubernetes operational practices. What you’ll do: Ensuring the reliability, availability, and performance of production infrastructure and platform services Operating and scaling Kubernetes platforms, including governance and support for multi-tenant workloads Managing GitOps-based deployment workflows using ArgoCD and Helm Supporting infrastructure provisioning and change management through Terraform/Terragrunt Building and supporting CI/CD automation and deployment workflows using GitHub Actions Participating in incident response, root cause analysis, and post-incident improvement initiatives Reducing operational toil through scripting, tooling, and process automation Advancing observability practices across logs, metrics, traces, dashboards, and alerting Supporting secure secrets integration, IAM-aware operations, and platform guardrails Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes   What we're looking for: 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure Strong hands-on experience operating AWS in production environments Good expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration Experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns Experience implementing and operating ArgoCD within a GitOps delivery model Strong hands-on experience with Helm Experience with Terraform/Terragrunt for infrastructure provisioning and environment management Solid scripting and automation skills using Bash and/or Python Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems Experience with monitoring, alerting, and observability in production environments Demonstrated ownership mindset with experience handling incidents and resolving production issues Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams Bachelor’s degree in computer science, engineering, a related field or equivalent experience Demonstrated ability to use AI to improve speed and quality in your day-to-day workflow for relevant outputs Strong track record of critical evaluation and verification of AI-assisted work (e.g., testing, source-checking, data validation, peer review) High integrity and ownership: you protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables   In-Office Requirement Statement: We recognize that the ideal environment for work is situational and may differ across departments. What this looks like day-to-day can vary based on the needs of each organization or role. Relocation Statement: This position is not eligible for relocation assistance. Visit our PinFlex page to learn more about our working model. #LI-SM4 #LI-REMOTEAt Pinterest we believe the workplace should be equitable, inclusive, and inspiring for every employee. In an effort to provide greater transparency, we are sharing the base salary range for this position. The position is also eligible for equity. Final salary is based on a number of factors including location, travel, relevant prior experience, or particular skills and expertise. Information regarding the culture at Pinterest and benefits available for this position can be found here.US based applicants only$114,297—$235,319 USDOur Commitment to Inclusion: Pinterest is an equal opportunity employer and makes employment decisions on the basis of merit. We want to have the best qualified people in every job. All qualified applicants will receive consideration for employment without regard to race, color, ancestry, national origin, religion or religious creed, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, age, marital status, status as a protected veteran, physical or mental disability, medical condition, genetic information or characteristics (or those of a family member) or any other consideration made unlawful by applicable federal, state or local laws. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. If you require a medical or religious accommodation during the job application process, please complete this form for support.   By submitting this application, I certify that all information submitted in my application and throughout the hiring process is true, accurate, and complete to the best of my knowledge. I understand that any false statement, omission, or misrepresentation may disqualify me from employment consideration or result in termination if discovered after hire.

Source: greenhouse. PazaJobs aggregates publicly available job postings and links to the original posting. Is this your posting and would you like it removed? Write to takedown@pazajobs.profectify.com.