Hero Image

AnitaB.org Talent Network

Connecting women in tech with the best professional opportunities!

Senior Site Reliability Developer

Oracle

Oracle

Software Engineering
Mexico
Posted on Feb 19, 2026

Key Responsibilities:

  • Deploy, monitor, maintain, and optimize production server hardware, software, and databases within a cloud environment.
  • Deliver escalated technical support for complex and critical issues, engaging in problem management and communicating status updates to stakeholders.
  • Lead resolution of escalated support cases, collaborating with internal technical resources and third-party vendors when necessary.
  • Manage storage infrastructure and coordinate upgrades, bug fixes, and patching for Oracle systems and database appliances.
  • Support standardization and automation initiatives across hardware and software within the Oracle technology stack to enhance reliability and supportability.
  • Maintain exceptional system uptime and ensure systems meet the rigorous performance and availability expectations of cloud-native environments.
  • Participate in an on-call rotation to provide after-hours support for production systems and respond to urgent incidents.

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.


Description: As a Site Reliability Developer, you will be responsible for ensuring the reliability, performance, and availability of mission-critical production environments within Oracle’s Cloud Infrastructure. You will manage and optimize systems and databases that underpin our core business operations, leveraging your deep technical expertise to identify innovative solutions that drive continuous improvement. This role is an exciting opportunity to tackle complex challenges involving high growth, extreme scalability, and stringent high-availability requirements. You will serve as an escalation point for advanced production support issues, applying your knowledge to deliver effective resolutions and impactful recommendations. This position is a hybrid role which does require employees to be in office 3 day a week at our Guadalajara office.

Career Level - IC3


Preferred Qualifications:

  • Strong experience in administration and analysis of cloud-based production environments, preferably within Oracle Cloud Infrastructure.
  • Advanced experience with Linux systems administration.
  • Demonstrated expertise in troubleshooting complex technical problems related to scalability and high availability.
  • Proficiency in managing server operating systems, storage environments, and database appliances.
  • Effective communicator with strong problem-solving skills and the ability to lead cross-functional resolution efforts.
  • Familiarity with automation tools, standardization projects, and best practices for maintaining production environments in the cloud.
  • Willingness to participate in an on-call rotation as required.
  • This position is a hybrid role which does require employees to be in office 3 day a week at our Guadalajara office.

Responsibilities:

Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible for the design and delivery of the mission critical stack, with focus on security, resiliency, scale, and performance. Authority for end-to-end performance and operability. Partner with development teams in defining and implementing improvements in service architecture. Articulate technical characteristics of services and technology areas and guide Development Teams to engineer and add premier capabilities to the Oracle Cloud service portfolio. Understand and communicate the scale, capacity, security, performance attributes, and requirements of the service and technology stack. Demonstrate clear understanding of automation and orchestration principles. Act as ultimate escalation point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs). Utilize a deep understanding of service topology and their dependencies required to troubleshoot issues and define mitigations. Understand and explain the affect of product architecture decisions on distributed systems. Professional curiosity and a desire to a develop deep understanding of services and technologies.