Mastercard Logo

Mastercard

Lead Site Reliability Engineer

Posted 14 Days Ago
Be an Early Applicant
Remote or Hybrid
Hiring Remotely in Mexico City, Ciudad De México
Expert/Leader
Remote or Hybrid
Hiring Remotely in Mexico City, Ciudad De México
Expert/Leader
Leads site reliability and DevOps transformation for Mastercard’s enterprise AI platforms and global services. Responsibilities include lifecycle design, capacity planning, monitoring, incident response, postmortems, CI/CD automation, operational gating, resilience improvements, and production support. The role scales systems through automation, optimizes recovery times, collaborates globally across development and product teams, and mentors junior engineers.
The summary above was generated by AI
Our Purpose
Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.
Title and Summary
Lead Site Reliability Engineer
Title and Summary
Lead Site Reliability Engineer
Who is Mastercard?
Mastercard is a global technology company in the payments industry. Our mission is to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential.
Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. With connections across more than 210 countries and territories, we are building a sustainable world that unlocks priceless possibilities for all.
Overview
The Business Operations team is seeking a highly motivated and experienced Lead Site Reliability Engineer (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor.
Tech Skills:
Unix, Shell Scripting, SQL, Python, Apache Nifi, Splunk, Dynatrace, Jenkins, GIT, XLR, AI and Agentic workflows etc.
This role provides the opportunity to influence and operate the foundational AI platforms that enable innovation across Mastercard. Rather than focusing solely on individual AI models or applications, you will help build and scale the enterprise platforms, tooling, operational practices, and cloud infrastructure that support the next generation of AI capabilities across the organization.
Role:
Business Operations is leading the DevOps transformation at Mastercard through our tooling and by being an advocate for change & standards throughout the development, quality, release, and product organizations. We need team members with an appetite for change and pushing the boundaries of what can be done with automation. Experience in working across development, operations, and product teams to prioritize needs and to build relationships is a must.• Engage in and improve the whole lifecycle of services-from inception and design, through deployment, operation and refinement.• Analyse ITSM activities of the platform and provide feedback loop to development teams on operational gaps or resiliency concerns• Support services before they go live through activities such as system design consulting, capacity planning and launch reviews.• Maintain services once they are live by measuring and monitoring availability, latency and overall system health.• Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity.• Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead Mastercard in DevOps automation and best practices.• Practice sustainable incident response and blameless post-mortems.• Take a holistic approach to problem solving, by connecting the dots during a production event thru the various technology stack that makes up the platform, to optimize mean time to recover• Work with a global team spread across tech hubs in multiple geographies and time zones• Share knowledge and mentor junior resources
Qualifications• BS degree in Computer Science or related technical field involving coding (e.g., physics or mathematics), or equivalent practical experience.• Experience with algorithms, data structures, scripting, pipeline management, and software design.• Systematic problem-solving approach, coupled with strong communication skills and a sense of ownership and drive.• Ability to help debug and optimize code and automate routine tasks.• We support many different stakeholders. Experience in dealing with difficult situations and making decisions with a sense of urgency is needed.• Experience in one or more of the following is preferred: C, C++, Java, Python, Go, Perl or Ruby.• Experience supporting production AI, machine learning, data platforms, or large-scale distributed systems.• We need team members with an appetite for change and pushing the boundaries of what can be done with automation. Experience in working across development, operations, and product teams to prioritize needs and to build relationships is a must.• Experience in industry standard CI/CD tools like Git/BitBucket, Jenkins, Maven, Artifactory, and Chef. Experience designing and implementing an effective and efficient CI/CD flow that gets code from dev to prod with high quality and minimal manual effort is desired.
Corporate Security Responsibility
Every person working for, or on behalf of, Mastercard is responsible for information security. All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and therefore, it is expected that the successful candidate for this position must:
• Abide by Mastercard's security policies and practices;• Ensure the confidentiality and integrity of the information being accessed;• Report any suspected information security violation or breach, and • Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.
Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
  • Abide by Mastercard's security policies and practices;
  • Ensure the confidentiality and integrity of the information being accessed;
  • Report any suspected information security violation or breach, and
  • Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.

Similar Jobs at Mastercard

2 Hours Ago
Remote or Hybrid
Expert/Leader
Expert/Leader
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Leads site reliability initiatives across design, performance engineering, chaos testing, capacity planning, monitoring, automation, incident prevention, and service health. Maintains reliability and availability of distributed systems, develops dashboards and technology roadmaps, manages technical priorities, and mentors engineers. Partners across software engineering, operations, and external teams to improve platform resiliency, velocity, latency, and operational efficiency.
Top Skills: BambooBitbucketBlazemeterCi/CdCloud-Native ApplicationsConcourseDistributed SystemsDockerDockerDynatraceGatlingGrafanaHelmJavaJenkinsKubernetesMavenPrometheusPythonScalaSplunkStatic Analysis Tools
27 Days Ago
Remote or Hybrid
Senior level
Senior level
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Lead ownership of service lifecycle including architecture, deployment, operations, and optimization. Define production readiness, monitoring, CI/CD governance, automation-first practices, incident response, capacity planning, and cross-team alignment. Mentor engineers, drive resiliency and reliability improvements, and influence roadmaps to reduce toil and improve system performance.
Top Skills: AWSCC++Ci/CdCloud-NativeDbaDevOpsDynatraceGoJavaLinuxMonitoringObservabilityOraclePerlPythonRubySplunkSQLUnix
2 Hours Ago
Remote or Hybrid
Senior level
Senior level
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Own production readiness and reliability for platform services: design for scalability, automate deployments and recovery, monitor availability and performance, run incident response and postmortems, consult on capacity and launch reviews, and mentor junior engineers while collaborating across global development and product teams.
Top Skills: AutomationCC++Ci/CdDevOpsDynatraceGoItsmJavaOraclePerlPythonRubyScriptingSplunkSQLUnix/Linux

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account