| Location | Austin, TX |
| Industry | Business Services - Other |
| Year Founded | 1986 |
| Website | http://www.visa.com.hk/index.shtml |
About Us
Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.
At Visa, you'll have the opportunity to create impact at scale — tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.
Join Visa and do work that matters – to you, to your community, and to the world. Progress starts with you.
Job Description
Every time someone taps, swipes, or clicks to pay- Visa infrastructure makes it happen in milliseconds, across 200+ countries. As a Software Development Engineer on the Product Reliability Engineering (PRE) team, you won’t just watch those systems run- you’ll be one of the engineers building, automating, and evolving them.
PRE is not a traditional ops team. We are a software engineering organization that treats infrastructure as code, reliability as a product, and automation as a strategic advantage. You’ll write Python, build agentic AI tools, manage data platforms, and contribute to the distributed systems that process billions of real-time transactions. From day one, you are an engineer- and from day one, your work matters.
If you are endlessly curious about how large-scale systems stay resilient, obsess over elegant automation, and want to launch your career at the intersection of AI, infrastructure, and global financial technology — this role was built for you.
Build Automation That Scales
▪ Design, develop, and deploy end-to-end automation for deployment pipelines, infrastructure provisioning, platform operations, and release orchestration across complex production environments.
▪ Write clean, production-grade Python (and Go or Bash where it counts) to eliminate toil, reduce operational risk, and improve the reliability and scalability of critical engineering workflows.
▪ Design and implement reusable frameworks for release scheduling, validation, rollback, reporting, and configuration management that support the software delivery lifecycle.
▪ Drive automation initiatives that improve engineering efficiency, standardization, and operational excellence across teams.
Manage & Evolve Data Platforms
▪ Design, build, operate, and continuously improve relational database platforms supporting critical payment systems and high-volume transaction processing.
▪ Contribute to architecture decisions, platform enhancements, and engineering solutions that improve scalability, resiliency, and performance.
▪ Lead database health and lifecycle operations including upgrades, patching, backup and recovery strategies, and platform modernization efforts.
▪ Analyze and optimize database performance through index tuning, execution plan analysis, replication monitoring, and capacity management.
▪ Develop automation for database operations, configuration management, and schema deployments using tools such as Ansible, Liquibase, and CI/CD pipelines.
▪ Build proactive monitoring, observability, and reporting solutions that identify reliability risks before they impact production services.
Ship Agentic AI & ML-Powered Tools
▪ Design and build GenAI-powered engineering solutions that automate deployment orchestration, operational workflows, release governance, and platform management.
▪ Integrate LLM-driven capabilities into observability, incident response, troubleshooting, and developer productivity workflows to improve operational effectiveness.
▪ Evaluate and implement emerging AI, automation, and machine learning technologies that improve reliability, efficiency, and engineering velocity.
▪ Contribute to agentic automation strategies that help evolve PRE into an increasingly intelligent and autonomous engineering organization.
Own Observability & Platform Health
▪ Design and build dashboards, alerts, telemetry pipelines, and health indicators using tools such as Prometheus, Grafana, Splunk, or ELK to provide visibility across globally distributed systems.
▪ Analyze platform performance, reliability, utilization, and availability data to identify trends and implement long-term improvements.
▪ Lead troubleshooting efforts across infrastructure, applications, databases, and platform services, performing root cause analysis and driving durable corrective actions.
▪ Design and implement self-healing, automated remediation, and auto-scaling capabilities that improve system resilience and reduce operational overhead.
Engineer for Reliability & Security
▪ Design and implement highly available, scalable infrastructure solutions that support business-critical payment systems operating at global scale.
▪ Ensure platforms and services meet security, compliance, governance, and resiliency requirements across cloud-native and hybrid environments.
▪ Drive vulnerability remediation, configuration hardening, patch management, and security automation efforts to improve platform security posture.
▪ Partner with engineering teams to build reliability and security practices directly into the software development lifecycle.
Collaborate, Learn & Grow Fast
▪ Partner with software engineers, product managers, platform teams, and global PRE peers to design, deliver, and operate reliable engineering solutions.
▪ Participate in architecture reviews, design discussions, code reviews, and technical planning activities, contributing engineering expertise and best practices.
▪ Create and maintain technical documentation, runbooks, operational procedures, and engineering standards that improve team effectiveness and knowledge sharing.
▪ Participate in on-call rotations and incident response activities, driving operational improvements and helping teams learn from production events.
▪ Take ownership of assigned initiatives from design through implementation, deployment, and operational support while continuously seeking opportunities to improve systems and processes.
Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager.Qualifications
Basic Qualifications
Preferred Qualifications
Information for US Applicants
Work Hours
Varies upon the needs of the department.
Travel Requirements
This position requires travel 5-10% of the time.
Mental/Physical Requirements
This position will be performed in an office setting. The position will require the incumbent to sit and stand at a desk, communicate in person and by telephone, frequently operate standard office equipment, such as telephones and computers.
Visa is an EEO Employer
Qualified applicants will receive consideration for employment without regard to race, color religion, sex, national origin, sexual orientation, gender identity, disability or protect veteran status. Visa will also consider for employment qualified applicants with criminal histories in a manner consistent with the EEOC guidelines and applicable local law.Visa has been a proud sponsor of the Olympic Games since 1986. Sport unites people, communities and nations. It enriches people’s lives and creates economic development opportunities. In today’s world, where brand and trust mean so much, the Olympic Games reflect those equities found at Visa — worldwide acceptance, reliability, versatility and leadership. Sponsoring the Olympic Games makes good business sense for Visa and our clients.