Job Overview
Compensation
Hourly
Range $56.50 - $75.00
Benefits
Medical insurance
Dental Insurance
Vision Insurance
401(k)
Flexible spending account
Health savings account
Life insurance
Paid Time Off
Wellness Program
Job Description
Visa is a global leader in payments technology, serving as a critical facilitator of transactions among consumers, merchants, financial institutions, and government entities worldwide. Operating in over 200 countries and territories, Visa is dedicated to transforming the way people pay and get paid, making financial transactions more secure, convenient, and accessible for everyone, everywhere. The company emphasizes innovation and the empowerment of individuals and businesses by providing cutting-edge payment solutions and infrastructure that support global commerce and economic growth. As part of its commitment to creating impact at scale, Visa cultivates a culture that encourages employees to tackle meaningful challenges, grow professionally, and contribute to societal progress through their work. Employees at Visa are offered opportunities to develop their skills in a dynamic, inclusive environment where their contributions can affect lives on a global scale. Progress at Visa is driven by its people, who are encouraged to bring their unique perspectives and talents to solve complex problems and deliver innovative payment experiences.
The Site Reliability Engineering (SRE) role at Visa is a pivotal position within the company's Cloud Platform strategy. This role focuses on ensuring that Visa's development platform and tooling effectively support engineering teams by enabling them to focus more on innovation and less on managing infrastructure. As part of this function, the SRE will lead the adoption of best practices in observability, automation, and AI-AIOps to enhance platform reliability, reduce manual workload, and resolve recurring operational challenges. The successful candidate will collaborate closely with software engineering teams to maintain the security, availability, reliability, and performance of the cloud platform. This includes working with peer engineering teams that support the platform as well as internal customers who consume it.
Key responsibilities of this hands-on engineering role include triaging complex incidents, partnering with infrastructure and operations teams, and advancing monitoring and alerting practices. The SRE engineer will also implement automation and AI-powered solutions to improve incident response times, service health, and overall platform efficiency. Additional duties involve defining and validating operational requirements during release and service transition reviews, supporting platform maintenance, and addressing various technical challenges faced by internal stakeholders. Working within the Visa Cloud SRE team entails supporting a 24/7/365 operational model, including shift-based and on-call coverage as necessary. This role requires a minimum of three days in-office presence, with specific expectations managed by the hiring manager.
Visa offers a competitive salary range for this position from $88,000 to $136,900 annually, which may include incentive payments. The company provides a comprehensive benefits package, including medical, dental, vision coverage, 401(k), FSA/HSA accounts, life insurance, paid time off, and wellness programs. Employees enjoy working in a supportive, diverse, and equitable workplace that adheres to all applicable non-discrimination policies and employment laws. Travel requirements for this role are limited, generally between 5 to 10 percent, and the work environment is primarily office-based with typical physical and mental demands of a professional setting.
The Site Reliability Engineering (SRE) role at Visa is a pivotal position within the company's Cloud Platform strategy. This role focuses on ensuring that Visa's development platform and tooling effectively support engineering teams by enabling them to focus more on innovation and less on managing infrastructure. As part of this function, the SRE will lead the adoption of best practices in observability, automation, and AI-AIOps to enhance platform reliability, reduce manual workload, and resolve recurring operational challenges. The successful candidate will collaborate closely with software engineering teams to maintain the security, availability, reliability, and performance of the cloud platform. This includes working with peer engineering teams that support the platform as well as internal customers who consume it.
Key responsibilities of this hands-on engineering role include triaging complex incidents, partnering with infrastructure and operations teams, and advancing monitoring and alerting practices. The SRE engineer will also implement automation and AI-powered solutions to improve incident response times, service health, and overall platform efficiency. Additional duties involve defining and validating operational requirements during release and service transition reviews, supporting platform maintenance, and addressing various technical challenges faced by internal stakeholders. Working within the Visa Cloud SRE team entails supporting a 24/7/365 operational model, including shift-based and on-call coverage as necessary. This role requires a minimum of three days in-office presence, with specific expectations managed by the hiring manager.
Visa offers a competitive salary range for this position from $88,000 to $136,900 annually, which may include incentive payments. The company provides a comprehensive benefits package, including medical, dental, vision coverage, 401(k), FSA/HSA accounts, life insurance, paid time off, and wellness programs. Employees enjoy working in a supportive, diverse, and equitable workplace that adheres to all applicable non-discrimination policies and employment laws. Travel requirements for this role are limited, generally between 5 to 10 percent, and the work environment is primarily office-based with typical physical and mental demands of a professional setting.
Job Requirements
- Bachelor's degree or 3+ years of relevant work experience
- 2 or more years working in a platform, SRE or production engineering group for high availability-critical platforms-applications
- 2 years experience with CI-CD tooling in a large-scale environment
- 2 years experience with observability tooling in a large-scale environment
- 2 years experience supporting relational and non-relational databases
- Basic understanding of YAML, JSON, HTML, XML
- Hands on experience in Linux and/or Windows systems
- Beginner level programming and/or scripting in multiple languages including Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, Cloud Formation
- Experience with AI-enabled tools for automation and observability
- Experience managing container infrastructure and distributed container platforms
- Exposure to virtualization technologies
- Ability to support a 24/7/365 operational model including on-call duties
- Willingness to work in-office at least three days per week as required by the hiring manager
Job Qualifications
- Bachelor's degree or 3+ years of relevant work experience
- 2 or more years working in a platform, SRE or production engineering group for high availability-critical platforms-applications
- 2 years experience with CI-CD tooling such as Jenkins, Github, Bitbucket, ArgoCD, Artifactory, Azure DevOps in a large-scale environment
- 2 years experience with observability tooling such as Grafana, Prometheus, Splunk, Datadog, New Relic, DynaTrace, Sentry in a large-scale environment
- 2 years experience supporting relational and non-relational databases including creating and running queries, managing performance and scaling
- Basic understanding of YAML, JSON, HTML, XML
- Hands on experience in Linux and/or Windows systems and good understanding of distributed computing environments
- Beginner level programming and/or scripting in 3 or more of the following: Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, Cloud Formation
- Experience with AI-enabled solutions and tools such as Claude, ChatGPT, GitHub Copilot with practical application in automation, observability, incident response, or operational workflows
- Experience managing container infrastructure and supporting development transformation to a container first model
- Exposure to Virtualization (Hyper-V, VMware, Openshift Virtualization etc.)
- Experience managing a distributed container platform including deployment-release management, provisioning, capacity management, workload management
Job Duties
- Help maintain the platform's defined SLAs and SLOs by driving operational excellence, delivering value-added process and procedure improvements, and partnering with engineering and operations teams to eliminate manual touchpoints through automation and standardization
- Own and operationalize end-to-end observability, alerting, and monitoring for the Visa Cloud Platform across IaaS, PaaS, and Container-as-a-Service environments, ensuring telemetry, dashboards, alerts, SLIs, and operational workflows are meaningful, actionable, and effective in supporting production reliability
- Own and deliver automation and AI-AIOps initiatives that reduce toil, improve reliability, accelerate incident response, and enhance operational intelligence across the SRE organization
- Partner with development teams during release and service transition reviews to define and validate operational requirements, including SLIs-SLOs, monitoring, alerting, dashboards, runbooks, incident response procedures, capacity expectations, and production readiness criteria
- Partner with Operations & Infrastructure peers to support ongoing platform maintenance, enhancement, and reliability
- Support multiple internal stakeholders across a variety of technical challenges by analyzing recurring issues, identifying patterns, and proposing effective solutions
- Support the Visa Cloud SRE team’s 24-7-365 operating model, including shift-based and on-call coverage, with weekend support as required
Job Criteria
Experience
Mid Level (3-7 years)
Job Location
Your Profile Is Visible To Hiring Managers Across OysterLink.
We'll match you with best jobs
Get job offers faster


Search For More Opportunities:
How Candidates Get Hired Faster
Apply to 2–3 similar roles
Complete profile & get best matches
Check new opportunities daily

