Role at a glance
- Job function
-
Software Engineering & IT DevOps & Site Reliability Engineering
- Salary
- Not Disclosed
- Location
- Mexico City, Mexico
- Employment
- Full-time
- Experience
- IT experience including demonstrating thought-leadership and relationship building across large-scale organizations.
- Education
- Bachelor’s Degree in Computer Science, Computer Systems, Information Technology or related. Equivalent experience is acceptable.
Spotted an issue?
We’ll check it against the original posting.
Role Summary
The Lead Site Reliability Engineer supports the MDES platform by connecting software engineering, operations, and internal and external partners. The role focuses on preparing services for launch and improving the reliability, availability, performance, and operational efficiency of live systems.
What You'll Do
- Lead site reliability initiatives before launch through system design consulting, performance engineering, chaos testing, capacity...
- Measure, monitor, and communicate service availability, latency, performance, and overall system health.
- Scale systems through automation and recommend changes that improve reliability and velocity or tune performance.
- Review production incidents and drive solutions to prevent recurrence.
- Manage project priorities, deadlines, and deliverables, and create and maintain technology roadmaps.
- Build and maintain system-health dashboards, and coach and mentor technical talent.
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
Bachelor’s degree in Computer Science, Computer Systems, Information Technology, or a related field, or equivalent experience. Required qualifications include advanced knowledge of web applications, distributed systems infrastructure and architecture, and software engineering concepts and methodologies; experience with Docker and Kubernetes, monitoring or observability tools such as Dynatrace, Splunk, Grafana, Prometheus, or similar, static analysis tools, and IT experience demonstrating thought-leadership and relationship building across large-scale organizations. Business-level English proficiency and excellent written and verbal communication are required. The posting also calls for initiative and ownership, attention to detail, customer focus, ability to multitask and prioritize, learn new technologies quickly, work with geographically distributed and cross-functional teams, and provide positive service to internal and external partners.
Required
- Experience with Java, Python, Scala or other Object-oriented programming languages.
- Experience with BitBucket, Maven, Jenkins and/or Gatling
- Experience with Blazemeter
- Experience with performance tuning of cloud-native applications
- Experience with root cause analysis and troubleshooting, working across teams to troubleshoot complex issues and providing guidance.
- In-depth knowledge of Kubernetes and Helm Chart deployments.
- Experience building and supporting on-soil solutions.
Original job description
Content provided by the employer
Original job description
Content provided by the employer
Our Purpose
Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.
Title and Summary
Lead Site Reliability EngineerMastercard is a global technology company in the payments industry. Our mission is to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential.Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. With connections across more than 210 countries and territories, we are building a sustainable world that unlocks priceless possibilities for all.
Overview:
The Mastercard Digital Enablement Services (MDES) team is looking for a strong, innovative Technical Lead to help lead Mastercard's next generation of Digital Payment products that change how our customers choose to pay. We are looking for a Technical Lead who can bring unique perspectives and innovative ideas to all areas of development and are interested in continuing to improve our platform through the ever-changing technology landscape.
In this role, our mission is to bridge gaps between Software Engineering, Operations, Internal and External Partners by strengthening relations, safely progressing engineering feature delivery and commitments. Through this role we shorten feedback loops, add collaborative focus on lowering operational overhead, provide exemplary capacity management, increasing availability & resiliency, and reducing latency to the MDES platform.
An ideal candidate is someone who works well with and can lead a cohesive team-oriented group of engineers, has a passion for developing, improving, and implementing great software, enjoys solving problems in a challenging environment, and has the desire to take their career to the next level. If any of these opportunities excite you, we would love to talk.
Role:
• Lead a technical initiatives of site reliability support services before they go live through activities such as system design consulting, performance engineering, chaos testing, capacity planning and launch reviews.
• Increase, maintain and communicate service metrics once live by measuring and monitoring availability, latency, performance and overall system health.
• Scale systems sustainably through mechanisms, like automation, and evolve systems by pushing for changes that improve reliability, velocity and recommend performance tuning approaches.
• Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
• Review production incidents to identifying and driving solutions to prevent recurrence.
• Manage individual project priorities, deadlines, and deliverables.
• Create and maintain technology roadmaps.
• Look at all tasks with an eye for automation; then work to automate them.
• Build, manage and maintain robust dashboards reflecting system health.
• Applies expert technical capabilities across discipline(s) to coach and mentor technical talent.
All About You
- Bachelor’s Degree in Computer Science, Computer Systems, Information Technology or related. Equivalent experience is acceptable.
- Advanced knowledge of web applications and distributed systems infrastructure and architecture.
- Excellent verbal and written communication to a variety of audience with various levels of technical acumen. Business-level English proficiency is a must.
- Experience working with Docker and Kubernetes.
- Experience with Dynatrace, Splunk, Grafana, Prometheus, or similar monitoring/observability tools.
- Interest and ability to learn new coding languages, frameworks, and paradigms as needed.
- IT experience including demonstrating thought-leadership and relationship building across large-scale organizations.
- Advanced knowledge and understanding of Software Engineering Concepts and Methodologies.
- Experience with static analysis tools to improve software quality.
- Knowledge of CI/CD platforms (Jenkins, Bamboo, Concourse, etc.)
- Displays initiative and ownership.
- Detailed oriented and customer obsessed.
- Excellent verbal and written communication skills
- Ability to multi-task and prioritize efforts.
- Ability to pick up new technologies at a quick pace and learn on-the-go.
- Ability to work within a geographically distributed team.
- Displays excellent collaborative skills with cross-functional teams.
- Ability to provide positive customer service to external and internal business partners.
Preferred:
- Experience with Java, Python, Scala or other Object-oriented programming languages.
- Experience with BitBucket, Maven, Jenkins and/or Gatling
- Experience with Blazemeter
- Experience with performance tuning of cloud-native applications
- Experience with root cause analysis and troubleshooting, working across teams to troubleshoot complex issues and providing guidance.
- In-depth knowledge of Kubernetes and Helm Chart deployments.
- Experience building and supporting on-soil solutions.
Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
Abide by Mastercard’s security policies and practices;
Ensure the confidentiality and integrity of the information being accessed;
Report any suspected information security violation or breach, and
Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.
About the company
Mastercard
Large Enterprise
Mastercard is a global technology company in the payments industry, committed to empowering individuals and businesses through secure and efficient payment solutions. With a presence in over 210 countries, Mastercard connects consumers, financial institutions, merchants, and governments, enabling seamless transactions across various platforms. The company is at the forefront of innovation, focusing on enhancing financial inclusion and expanding access to digital payment technologies. Through its advanced network and partnerships, Mastercard continues to revolutionize the way people engage in commerce worldwide.