Role
This role involves working within a high-performing, globally distributed Agile team to build, operate, and evolve low latency, high throughput services that power real-time fraud decisioning on the Mastercard Fraud Platform. In this role, you will be responsible for:
- Develop and maintain CI/CD pipelines to enable reliable, repeatable deployments.
- Deploy, operate, and troubleshoot services across hybrid environments, including AWS (EKS/Kubernetes), on-premise cloud platforms, and traditional on-premise infrastructure, ensuring consistency, reliability, and scalability across all environments.
- Work closely with platform teams to improve environment stability, scalability, and developer experience.
- Contribute to improving reliability and observability, aligning with SRE best practices.
- Work as part of a co-located Agile Scrum team.
- Work closely with Product Owners, Business Analyst, Systems Analyst, Technical leads and other developers to define user stories.
- Develop high-quality, scalable and secure software solutions.
- Assist with operational issues by troubleshooting incidents.
- Research alternative technical solutions to meet changing business needs.
- Work with project team to meet due dates, while working through emerging issues and recommending solutions.
- Produce design documentation in accordance with Mastercard documentation standards.
All about you
- Experience of building CI/CD pipelines using tools such as Jenkins, Github Actions, Python and shell scripting
- Ability to design and maintain end-to-end deployment pipelines (build, test, deploy, rollback).
- Hands-on experience with monitoring, logging, and alerting tools (e.g., Splunk, Dynatrace, OpenTelemetry).
- Experience with build tools (Maven/Gradle) and dependency management.
- Exposure to configuration management and automation tools (Chef or equivalent).
- Experience with cloud platforms (AWS, Azure)
- Experience with containerization (Docker) and image lifecycle management.
- Experience with Infrastructure as Code (IaC) concepts and tools (e.g., Terraform, CloudFormation or similar).
- Understanding of platform engineering principles, including building reusable deployment templates and shared services.
- Experience supporting multi-environment setups (dev, test, staging, production) with proper isolation and governance.
- Strong troubleshooting skills in Linux-based production environments.
- Awareness of SRE principles and reliability practices.
- Experience developing and integrating AI-driven workflows and automation for use cases such as incident analysis, anomaly detection, and intelligent observability or CI/CD enhancements.
Passionate about Agile software development and working with SAFe or Scrum.