Senior Site Reliability Engineer at Mastercard
- Company: Mastercard
- Location: Multiple Locations
- Job type: full time
- Workplace: onsite
- Posted: 2026-05-22
Job description
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Senior Site Reliability Engineer - SRE - Java/Spring Boot, Microservices Company: Mastercard Category: Software Engineering Experience: 8+ years Location: Pune Overview Mastercard is a technology company in global payments. We’re seeking a Senior Software Engineer for SRE Role with strong Java Microservices development experience to improve the reliability, performance, and operability of mission critical services. You will serve as a subject matter expert across development, testing and Support for production/non production environments, driving reliability, monitoring & observability, performance engineering, root cause analysis, and automation to increase production stability. Acts as a recognized technical authority with sound commercial awareness. What You’ll Do • Reliability & Operations • Improve service reliability through better architecture, automation, capacity modeling/planning, and sub linear operational scaling. • Lead incident handling, mitigation, RCA, and postmortems; clearly explain performance bottlenecks to stakeholders. • Build best in class monitoring/observability using Splunk and Dynatrace (or equivalent APM); create advanced dashboards, queries, and runbooks. • Maintain and improve program quality matrix like availability, MTTM, latency, capacity, scalability, reliability for the project • Engineering & Performance • Contribute to and review Java/Spring Boot microservices and RESTful applications for Reliability and Performance. • Drive performance testing/tuning/analysis, using applications like JMeter, Blaze meter (or similar), through thread/heap dump analysis, Java, SQL, network & application configuration tuning/fixing. • Partner as a co owner of resilient production services and CI/CD pipelines with software engineering teams. • Architecture & Tooling • Work across Java app servers, web servers, Docker, Kubernetes, Kafka, Redis, and cloud platforms (AWS/Azure/GCP); Cloud Foundry is a plus. • Lead POCs to incubate new features/capabilities; recommend product customization for system integration. • Collaboration & Governance • Collaborate with SRE, operations, production support, performance testing, architects, and application owners. • Identify patterns and implement reusable solutions that reduce complexity and operational risk. • Uphold compliance with applicable laws, policies, and risk standards; demonstrate ethical judgment and transparency. • Mentor junior engineers and influence engineering decisions through counsel and expertise. • working in large-scale agile software development environments utilizing the SAFe framework (scrum methodology) Required Qualifications • Education: BE/BTech in Computer Science or equivalent experience. • Experience: 8+ years in Engineering/IT; 3+ years in enterprise Java development; 3+ years with Splunk/ELK (or similar log analytics); 3+ years with Dynatrace or other APM. • Development & Architecture • Java, Spring, Spring Batch, Spring Boot, Microservices, RESTful APIs, MVC architecture. • Databases: Oracle, SQL (queries & stored procedures). • Platform/infra: Java app servers, web servers, Linux/Unix, Docker, Kubernetes, Kafka, Redis; Cloud Foundry (plus). • Expertise in building Dashboards/Views in Splunk and Dynatrace is a Plus • SDLC & Tooling • CI/CD: Jenkins, Git, Maven; Agile/DevOps practices; experience in SAFe at scale. • Reliability & Performance • Production monitoring, observability, RCA, capacity
Senior Site Reliability Engineer on JobPost.