Company
MasterCard
Description
Our Purpose
- Operational Readiness Architect:
- Serve as the primary contact responsible for the overall application health, performance, and capacity
- Support services before they go live through activities such as system design consulting, capacity planning and launch reviews.
- Partner with the development and product team of a new application to establish the right monitoring and alerting strategy and create the framework to achieve zero downtime during deployment.
- Site Reliability Engineering:
- Serve as the primary contact responsible for ensuring application scalability, performance, and resilience.
- Practice sustainable incident response and blameless post-mortems while taking a holistic approach to problem solving and optimizing time to recover.
- Automate data-driven alerts to proactively escalate issues. Work with development teams to establish SLOs and improve reliability.
- DevOps/Automation:
- Tackle complex development, automation, and business process problems. Engage in and improve the whole lifecycle of services—from inception and design, through deployment, operation, and refinement.
- Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead Mastercard in DevOps automation and best practices.
- Increase automation and tooling to reduce toil and manual intervention
- ITSM Practices:
- Analyses ITSM activities of the platform and provide feedback loop to development teams on operational gaps or resiliency concerns
- BS degree in Computer Science or related technical field involving coding (e.g., physics or mathematics), or equivalent practical experience.
- Coding or scripting exposure.
- Appetite for change and pushing the boundaries of what can be done with automation. Be curious about new technology, infrastructure, and practices to scale our architecture and prepare for future growth.
- Experience with algorithms, data structures, scripting, pipeline management, and software design
- Systematic problem-solving approach, coupled with strong communication skills and a sense of ownership and drive.
- Interest in designing, analysing, and troubleshooting large-scale distributed systems.
- Willingness and ability to learn and take on challenging opportunities and to work as a member of matrix based diverse and geographically distributed project team.
- Ability to balance doing things right with fixing things quickly. Flexible and pragmatic, while working towards improving the long-term health of the system.
- Comfortable collaborating with cross-functional teams to ensure that expected system behaviour is understood and monitoring exists to detect anomalies.
- Coding experience in one or more of the following: C++, Java, Python, Go
- Experience with algorithms, data structures, scripting, pipeline management, and software design.
- Experience in working across development, operations, and product teams to prioritize needs and to build relationships is a must.
- Experience in a SRE role or related field.
- Background on cloud native tooling and orchestration technologies (Kubernetes preferred).
- Experience in Monitoring tools such as Splunk, Dynatrace.
- Experience with Java, J2EE, WebServices (SOAP/REST), Spring/Spring Boot is a plus.
- Experience in production support environments and ITIL processes.
- Experience with industry standard CI/CD tools like Git/BitBucket, Jenkins, Maven, Artifactory, Groovy and Chef. Experience designing and implementing an effective and efficient CI/CD flow that gets code from dev to prod with high quality and minimal manual effort is required.
- Developing and maintaining cloud solutions on Azure, GCP, or AWS in accordance with best practices.
- Understanding of:
- Client-server relationships
- Network concepts (Layer 1 to Layer 3)
- Stack trace analysis (TCP dumps, heap dumps, CPU/memory analysis, thread dumps).
- Load balancers and application firewalls.
- Operating System navigation.
- Logging and monitoring methods, standards, and tools.
- High availability and business continuity planning
- Caching concepts
- Configuration management
- Hands-on experience in Modernization through the adoption of Kubernetes and containerization technologies like Docker and Azure Container Registry.
- Ability to speak about Kubernetes from different perspectives: Comfortable handling Kubernetes discussions at the technical, business, or financial level.
- Strategizing, designing, and supporting highly efficient solutions on Public Cloud (Amazon Web Services, Azure or GCP) for security, resilience, performant, networking, availability, Blue-green deployments in context of business application.
- Azure DevOps (AZ - 400), Azure Cloud Developer (AZ-203) certificate is preferred.
- Hands-on expertise in diverse DevSecOps concepts/tools, especially on Azure DevOps, Pipelines, GitHub, GitHub actions.
- Knowledge of emerging technologies, various platforms, tools and products and their respective applications.
- Awareness of security implementations, certificate management lifecycle, mutual TLS, SSL handshake, SSH keys, symmetric and asy"
- Abide by Mastercard’s security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.
Identifier
0befd11b06d5fd470a7ba756b1b2dd96
Show More
Ready to join the team? We'd love to have you!