Skip to main content
Posted August 21, 2026

Site Reliability Engineer II

Mastercard
O Fallon, Missouri 63368, United States Full-Time
125000.00 - 165000.00
Reference: 3157050267

Mastercard is seeking a Site Reliability Engineer II to enhance the reliability, scalability, and performance of critical IT systems. In this role, you will design and implement robust monitoring, automation, and incident response processes to support innovative, user focused software products. You will collaborate closely with development and infrastructure teams to build and maintain highly available, secure cloud environments. The culture emphasizes innovation, collaboration, and continuous growth, offering opportunities to work with cutting edge technologies and drive meaningful improvements in global IT services.

Responsibilities

  • Design, implement, and maintain highly available, scalable production systems.
  • Develop and optimize monitoring, alerting, and observability solutions.
  • Automate deployments, configuration, and operational tasks using Infrastructure as Code and scripting.
  • Collaborate with development and infrastructure teams to improve reliability and performance.
  • Participate in on-call rotations, incident response, and post-incident reviews.
  • Identify and remediate reliability, capacity, and performance bottlenecks.
  • Implement and refine SRE best practices, including SLIs, SLOs, and error budgets.
  • Contribute to security, compliance, and resilience improvements across systems.

Required Skills

  • Site Reliability Engineering (SRE) practices
  • Cloud platforms (AWS, GCP, or Azure)
  • Linux system administration
  • Kubernetes and container orchestration
  • Infrastructure as Code (Terraform, Cloud
  • Formation, etc.)
  • CI/CD pipelines and automation
  • Monitoring and observability (Prometheus, Grafana, etc.)
  • Scripting (Python, Bash, or similar)
  • Incident management and on-call operations
  • Networking and security fundamentals

Sign up for Job Alerts