Senior Site Reliability Engineer
Software Engineering · Full-time
Bengaluru, Karnataka, India
Position Summary...
Role Summary:As a Senior Site Reliability Engineer, you will apply advanced expertise to design, develop, and enhance scalable, resilient systems that meet business and customer needs. This role involves creating modular, extensible solutions, ensuring disaster recovery readiness, and optimizing performance across distributed environments. You will collaborate with stakeholders to implement robust infrastructure, automate processes, and maintain high-quality code while adhering to security and compliance standards. Your contributions will support continuous improvement, reliability monitoring, and incident management to uphold service excellence and operational stability within Walmart’s technology ecosystem.
What you'll do...
Job Description
About the team:
The Cloud Powered Checkout team develops and supports a scalable omni-channel checkout solution handling millions of daily transactions worldwide. The team delivers reliable, high-performance checkout experiences across in-store channels by building modern software deployed in cloud and edge environments. Collaborating closely with product management, architecture, and engineering teams, they drive product strategy and roadmap execution through data-driven decisions. Team members apply advanced technologies and engineering best practices to enhance capabilities, improve operational excellence, and deliver innovative solutions that meet evolving customer and business needs globally.
What you'll do:
- Develop and refine scalable, modular, and extensible system designs aligned with business requirements and disaster recovery standards.
- Convert high-level designs into detailed functional modules, incorporating telemetry and security policies.
- Build and maintain robust, high-quality code following coding standards and best practices.
- Monitor system performance, troubleshoot issues, and optimize reliability using established metrics and tools.
- Collaborate with stakeholders to identify business needs and implement effective infrastructure automation and compliance solutions.
- Support disaster recovery planning and execution to ensure operational continuity.
- Foster continuous improvement by researching emerging technologies and applying innovative solutions.
What you'll bring:
- Proven expertise in software architecture, distributed systems, and scalability design patterns.
- Strong knowledge of disaster recovery planning and implementation within complex environments.
- Experience in coding with languages such as JavaScript, Python, or C, adhering to coding standards and best practices.
- Ability to design modular, extensible, and functional solutions aligned with business requirements.
- Familiarity with monitoring, alerting tools, and performance optimization techniques for Unix/Linux and Java-based systems.
- Skilled in root cause analysis, troubleshooting, and continuous integration/continuous delivery automation.
- Commitment to maintaining security standards and compliance throughout development and deployment processes.
- Master's degree in site reliability engineering, site and system administration, infrastructure management, or related area and 1 year’ experience in experience in site reliability engineering, site and system administration, infrastructure management, or related area.; SRE certification (for example, IBM Cloud Site Reliability Engineer)
- Bachelor's degree in computer science, computer engineering, computer information systems, software engineering, or related area and 3 years’ experience in site reliability engineering, site and system administration, infrastructure management, or related area.
- 5 years’ experience in site reliability engineering, site and system administration, infrastructure management, or related area.
About Walmart Global Tech
Imagine working in an environment where one line of code can make life easier for hundreds of millions of people. That’s what we do at Walmart Global Tech. We’re a team of software engineers, data scientists, cybersecurity expert's and service professionals within the world’s leading retailer who make an epic impact and are at the forefront of the next retail disruption. People are why we innovate, and people power our innovations. We are people-led and tech-empowered.
We train our team in the skillsets of the future and bring in experts like you to help us grow. We have roles for those chasing their first opportunity as well as those looking for the opportunity that will define their career. Here, you can kickstart a great career in tech, gain new skills and experience for virtually every industry, or leverage your expertise to innovate at scale, impact millions and reimagine the future of retail.
Walmart’s culture sets us apart, and we know being together helps us innovate, learn and grow great careers. This role is based in our [Bangalore/Chennai] office for daily work, with the flexibility for associates to manage their personal lives.
Benefits
Beyond our great compensation package, you can receive incentive awards for your performance. Other great perks include a host of best-in-class benefits maternity and parental leave, pto, health benefits, and much more.
Belonging
We aim to create a culture where every associate feels valued for who they are, rooted in respect for the individual. Our goal is to foster a sense of belonging, to create opportunities for all our associates, customers and suppliers, and to be a Walmart for everyone.
At Walmart, our vision is "everyone included." by fostering a workplace culture where everyone is—and feels—included, everyone wins. Our associates and customers reflect the makeup of all 19 countries where we operate. By making Walmart a welcoming place where all people feel like they belong, we’re able to engage associates, strengthen our business, improve our ability to serve customers, and support the communities where we operate.
Equal opportunity employer
Walmart, inc., is an equal opportunities employer – by choice. We believe we are best equipped to help our associates, customers and the communities we serve live better when we really know them. That means understanding, respecting and valuing unique styles, experiences, identities, ideas and opinions – while being inclusive of all people.
Minimum Qualifications...
Outlined below are the required minimum qualifications for this position. If none are listed, there are no minimum qualifications.
Option 1: Bachelor's degree in computer science, computer engineering, computer information systems, software engineering, or related area and 3 years’ experience in site reliability engineering, site and system administration, infrastructure management, or related area.Option 2: 5 years’ experience in site reliability engineering, site and system administration, infrastructure management, or related area.
Preferred Qualifications...
Outlined below are the optional preferred qualifications for this position. If none are listed, there are no preferred qualifications.
5 years’ experience in experience in site reliability engineering, site and system administration, infrastructure management, or related area., Master's degree in site reliability engineering, site and system administration, infrastructure management, or related area and 1 year’ experience in experience in site reliability engineering, site and system administration, infrastructure management, or related area., SRE certification (for example, IBM Cloud Site Reliability Engineer).