Site Reliability Engineering

hace 4 semanas


Zapopan, México Oracle A tiempo completo

**Job Description**:Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of Database cloud services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible for the design and delivery of the mission critical stack, with focus on security, resiliency, scale, and performance. Authority for end-to-end performance and operability. Partner with development teams in defining and implementing improvements in service architecture. Articulate technical characteristics of services and technology areas and guide Development Teams to engineer and add premier capabilities to the Oracle Cloud service portfolio. Understand and communicate the scale, capacity, security, performance attributes, and requirements of the service and technology stack. Demonstrate clear understanding of automation and orchestration principles. Act as ultimate escalation point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs). Utilize a deep understanding of service topology and their dependencies required to troubleshoot issues and define mitigations. Understand and explain the effect of product architecture decisions on distributed systems. Professional curiosity and a desire to a develop deep understanding of services and technologies.(This is the default posting for this role in recruitment portal)**Skills**:**Mandatory**1) Experience on SRE role2) Terraform3) GIT (Version control)4) JIRA5) Experience in operation supportNice to have:1) Dockers2) Kubernetes3) Cloud services (OCI/AWS or anything equivalent)Career Level - IC3Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible for the design and delivery of the mission critical stack, with focus on security, resiliency, scale, and performance. Authority for end-to-end performance and operability. Partner with development teams in defining and implementing improvements in service architecture. Articulate technical characteristics of services and technology areas and guide Development Teams to engineer and add premier capabilities to the Oracle Cloud service portfolio. Understand and communicate the scale, capacity, security, performance attributes, and requirements of the service and technology stack. Demonstrate clear understanding of automation and orchestration principles. Act as ultimate escalation point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs). Utilize a deep understanding of service topology and their dependencies required to troubleshoot issues and define mitigations. Understand and explain the affect of product architecture decisions on distributed systems. Professional curiosity and a desire to a develop deep understanding of services and technologies.



  • Zapopan, Jalisco, México Oracle A tiempo completo

    We are looking for a skilled and motivated Cloud Region Build Site Reliability Engineer (SRE) to join our Oracle Cloud Infrastructure Region Build team. In this role, you will be responsible for building, deploying, and maintaining compute cloud infrastructure services across multiple regions to ensure high availability, scalability, and performance. You...


  • Zapopan, México Oracle A tiempo completo

    As part of the Site Reliability Engineering (SRE) team, you’ll contribute to designing, automating, and evolving mission-critical systems. You'll combine deep systems expertise with modern software engineering practices to reduce operational toil and build resilient, self-healing services.This is a high-impact role where your work directly affects the...


  • Zapopan, México Oracle A tiempo completo

    **Job Description**: Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of Database cloud services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible for the design and delivery of the mission...


  • Zapopan, Jalisco, México Oracle A tiempo completo

    DescriptionWe are looking for a skilled and motivated Cloud Region Build Site Reliability Engineer (SRE) to join our Oracle Cloud Infrastructure Region Build team. In this role, you will be responsible for building, deploying, and maintaining compute cloud infrastructure services across multiple regions to ensure high availability, scalability, and...


  • Zapopan, México Oracle A tiempo completo

    **Responsibilities**- Solve complex problems related to Linux infrastructure and Oracle Cloud Infrastructure- Act as escalation point for critical issues that may not have a documented procedure and provide Root Cause Analysis (RCA)- Understand the end-to-end configuration, technical dependencies, characteristics of production infrastructure and services-...


  • Zapopan, Jalisco, México Oracle A tiempo completo

    Job DescriptionWe are looking for a skilled and motivated Cloud Region Build Site Reliability Engineer (SRE) to join our Oracle Cloud Infrastructure Region Build team. In this role, you will be responsible for building, deploying, and maintaining compute cloud infrastructure services across multiple regions to ensure high availability, scalability, and...

  • Site Reliability Engineer

    hace 3 semanas


    Zapopan, México Oracle A tiempo completo

    Solve complex problems related to infrastructure cloud services and build automation to prevent problem recurrence. Design, write, and deploy software to improve the availability, scalability, and efficiency of Oracle products and services. Design and develop designs, architectures, standards, and methods for large-scale distributed systems. Facilitate...


  • Zapopan, México Oracle A tiempo completo

    We are hiring for OCI Corporate Network and Security operation. You will be a member of the team responsible for network and security Incident/change/capacity management for supporting the Oracle Corporate Network. Resolve the complex problems related to network infrastructure services and build automation to prevent problem recurrence. Design, write, and...


  • Zapopan, México Oracle A tiempo completo

    We are hiring for OCI Corporate Network and Security operation. You will be a member of the team responsible for network and security Incident/change/capacity management for supporting the Oracle Corporate Network. Resolve the complex problems related to network infrastructure services and build automation to prevent problem recurrence. Design, write, and...


  • Zapopan, Jalisco, México Oracle A tiempo completo

    DescriptionAs a senior member of the Site Reliability Engineering (SRE) team, you'll take ownership of highly available systems, influence service design, and work across teams to drive resiliency, automation, and operational excellence. This is a hands-on engineering role where deep infrastructure knowledge meets software engineering expertise, ideal for...