Oracle India Private Limited

Manager, Site Reliability Engineering

Oracle India Private Limited
Bengaluru/Bangalore
Not disclosed
Work from OfficeWork from Office
Full TimeFull Time
Min. 8 yearsMin. 8 years

Job Description

Manager, Site Reliability Engineering

The Autonomous Recovery Service (RCV) team delivers highly available, secure, and resilient cloud services that protect Oracle Cloud Infrastructure customer data. Our mission is to provide reliable recovery capabilities that customers can depend on during planned and unplanned events.

As a hands-on technical manager, you will lead a team of engineers while also managing the reliability, operation, and continuous improvement of the RCV service. This is not a role focused solely on people management. You will remain actively engaged in production operations, technical escalations, incident response, service architecture, capacity planning, automation, and operational readiness.

You will apply your technical expertise in Oracle Database, Recovery Manager (RMAN), Zero Data Loss Recovery Appliance (ZDLRA), Exadata, and cloud infrastructure to guide technical decisions and resolve complex service issues. You will work closely with software development, database, infrastructure, and cross-functional teams to improve service reliability, scalability, security, and customer experience.

You will coach and develop engineers, establish execution priorities, review technical work, and help team members build strong operational and site reliability practices. At the same time, you will serve as a trusted escalation point for complex incidents and participate directly in troubleshooting, root cause analysis, post-incident reviews, and permanent corrective actions.

The successful candidate combines strong people leadership with deep technical judgment, operational ownership, and a willingness to stay hands-on. You should be comfortable moving between team development, customer-impacting escalations, service-level decisions, and hands-on engineering work. You will also encourage responsible use of automation and AI-assisted tools to reduce operational toil, improve productivity, and continuously evolve the RCV service.

Technical Leadership & Service Ownership

  • Lead and develop a team responsible for the reliability, operation, and continuous improvement of the Autonomous Recovery Service (RCV).
  • Remain hands-on in service operations, technical escalations, architecture reviews, incident response, and production troubleshooting.
  • Provide day-to-day technical direction, establish priorities, delegate work, and ensure team commitments are delivered.
  • Manage the end-to-end service lifecycle, including provisioning, deployments, patching, upgrades, security updates, backup and recovery, maintenance, and decommissioning.
  • Apply deep technical expertise in Oracle Database, RMAN, ZDLRA, Exadata, OCI, and related infrastructure to guide service decisions and resolve complex issues.

Capacity, Architecture & Reliability

  • Guide the team in designing and architecting reliable, secure, scalable, and highly available infrastructure and services.
  • Forecast demand, evaluate capacity requirements, and ensure RCV has sufficient resources to support current and future workloads.
  • Review service architecture, dependencies, performance characteristics, scalability, and operational readiness.
  • Partner with software development teams to deliver reliable features, infrastructure, and deployment solutions.
  • Use operational data, health reports, and performance trends to recommend improvements to service reliability, efficiency, and customer experience.

Incident Management & Escalation

  • Serve as a technical escalation point for complex incidents and issues affecting Oracle services and RCV customers.
  • Participate directly in incident response, troubleshooting, mitigation, service restoration, and cross-functional coordination.
  • Guide engineers in data collection, triage, technical analysis, debugging, and resolution of issues spanning multiple services.
  • Lead or review root cause analyses, postmortems, and corrective actions to prevent incident recurrence.
  • Ensure incidents, operational conditions, known issues, and resolutions are accurately documented.
  • Coach team members to independently perform post-incident reviews and implement permanent improvements.

Automation & Operational Excellence

  • Identify opportunities to reduce operational toil through automation, orchestration, self-service workflows, and intelligent tooling.
  • Coach team members in evaluating automation opportunities, estimating benefits, and selecting practical solutions.
  • Review automation tools, scripts, and software developed by team members for quality, safety, maintainability, and operational value.
  • Ensure automation is tested thoroughly and produces reliable, repeatable results.
  • Encourage responsible use of AI-assisted engineering and operational tools to accelerate analysis, troubleshooting, documentation, and development while maintaining sound engineering judgment.
  • Drive improvements to monitoring, alerting, observability, deployment processes, and operational readiness.

Technical Communication & Guidance

  • Enable team members to communicate the scale, capacity, security, performance, dependencies, and requirements of RCV services.
  • Review proposed infrastructure, feature, and tooling changes and help the team understand their operational and customer impact.
  • Communicate technical risks, service health, capacity needs, and improvement plans to engineering leadership and cross-functional stakeholders.
  • Build strong partnerships with software development, database, infrastructure, security, product, and business teams.
  • Promote clear documentation, effective runbooks, standard operating procedures, and knowledge sharing.

Innovation & Continuous Improvement

  • Enable engineers to experiment with new technologies and evaluate their potential impact on performance, reliability, security, and operational efficiency.
  • Supervise the execution of improvements addressing performance bottlenecks, deployment risks, resource usage, and scalability.
  • Encourage the team to challenge existing processes and recommend more effective approaches.
  • Share emerging Site Reliability Engineering, database recovery, cloud operations, automation, and AI practices with the team.
  • Use production analysis and clear data to support service design changes, investment decisions, and broader business development decisions.

Planning, Execution & Resource Management

  • Create and own execution plans for team deliverables, projects, and operational initiatives.
  • Monitor timelines, priorities, dependencies, budgets, and resource needs to ensure work is completed effectively.
  • Adjust plans as business priorities, customer needs, or operational risks change.
  • Work with leadership to identify staffing, financial, and technical resource requirements.
  • Manage operating budgets and project financials where applicable.

People Leadership & Development

  • Coach team members through technical challenges, production incidents, architectural decisions, and career development opportunities.
  • Set clear goals and expectations aligned with RCV and broader organizational objectives.
  • Identify skill gaps and provide training, mentoring, documentation, and hands-on learning opportunities.
  • Foster a culture of ownership, accountability, inclusion, continuous learning, and knowledge sharing.
  • Provide performance guidance and feedback in accordance with management processes.
  • Lead candidate interviews, contribute to talent acquisition, assess promotion readiness, and support talent planning.
  • Develop engineers who can build, test, deploy, operate, and continuously improve mission-critical cloud services.
Minimum Job Qualifications
Education and/or Experience:
8 years of experience in software engineering, infrastructure management, or related field

OR

Bachelor's Degree in Computer Science, Engineering, or related field AND 4 years of experience in software engineering, infrastructure management, or related field

OR

Master's Degree in Computer Science, Engineering, or related field AND 2 year of experience in software engineering, infrastructure management, or related field.

Job Skills:
Data Analysis Demonstrated ability to analyze and interpret data to produce actionable business insights.

Automation Experience:
3 years of experience in automation.

Programming Experience:
3 years of experience in programming and/or scripting.

Preferred Job Qualifications
Education and/or Experience:
9 years of experience in software engineering, infrastructure management, or related field

OR

Bachelor's Degree in Computer Science, Engineering, or related field AND 5 years of experience in software engineering, infrastructure management, or related field

OR

Master's Degree in Computer Science, Engineering, or related field AND 3 years of experience in software engineering, infrastructure management, or related field

OR

Doctorate in Computer Science, Engineering, or related field AND 1 year of experience in software engineering, infrastructure management, or related field.

People Leadership / Management Experience:
1 year of experience in a leadership role with or without direct reports.

Budget Experience:
1 year of experience working with operating budgets and/or project financials.

Automation Experience:
5 years of experience in automation.

Programming Experience:
5 years of experience in programming and/or scripting.

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Experience Level

Senior Level

Job role

Work location
Work locationBENGALURU, KARNATAKA, India
Department
DepartmentSoftware Engineering
Role / Category
Role / CategorySoftware Development
Employment type
Employment typeFull Time
Shift
ShiftDay Shift

Job requirements

Experience
ExperienceMin. 8 years

About company

Name
NameOracle India Private Limited
Job posted by Oracle India Private Limited

Similar jobs you can apply for

Accounts / Finance
Ciao Green Private Limited

MEP Design Engineer

Ciao Green Private Limited
HSR Layout, Bengaluru/Bangalore
₹60,000 - ₹83,000
Work from Office
Full Time
Min. 3 years
Good (Intermediate / Advanced) English

Telecaller

Cnj financial solutions
Rajaji Nagar, Bengaluru/Bangalore
₹12,000 - ₹18,000
Work from Office
Full Time
Any experience
Basic English
Classic Infracon Bangalore Private Limited

Accountant

Classic Infracon Bangalore Private Limited
Rajaji Nagar, Bengaluru/Bangalore
₹20,000 - ₹30,000
Work from Office
Full Time
Min. 5 years
Basic English
Pvr Cinemas

Service Associate

Pvr Cinemas
White Field, Bengaluru/Bangalore
₹18,000 - ₹25,000*
Work from Office
Full Time
Night Shift
Freshers only
Basic English
Skyline Electronetworks

Electrical Site Engineer

Skyline Electronetworks
Bengaluru/Bangalore
₹22,000 - ₹30,000
Field Job
Full Time
Min. 2 years
Good (Intermediate / Advanced) English
Thyrocare

Team Leader - Operations

Thyrocare
Sahakara Nagar, Bengaluru/Bangalore
₹20,000 - ₹25,000
Work from Office
Full Time
Min. 1 year
Basic English

You can expect a minimum salary of 0 INR. The salary offered will depend on your skills, experience and performance in the interview.

The candidate should have completed the required education and people who have 8 to 31 years are eligible to apply for this job. You can apply for more jobs in Bengaluru/Bangalore to get hired quickly.

The candidate should have sound communication skills and sound communication skills for this job.

Both Male and Female candidates can apply for this job.

No, it's not a work from home job and can't be done online. You can explore and apply for other work from home jobs in Bengaluru/Bangalore at apna.

No work-related deposit needs to be made during your employment with the company.

Go to the apna app and apply for this job. Click on the apply button and call HR directly to schedule your interview.

The last date to apply for this job is . For more details, download apna app and find Full Time jobs in Bengaluru/Bangalore . Through apna, you can find jobs in 64 cities across India. Join NOW!