Manager, Site Reliability Engineering

🏢 Oracle
📍 Reston, United StatesOn-site
📅 Posted: 1mo ago🔄 Updated: 1mo ago
CV%
✨ AI Summary
The Manager, Site Reliability Engineering will lead a team in designing and architecting reliable and scalable infrastructure and services. Responsibilities include forecasting demands, ensuring adequate system resources, collaborating with software development teams, advising on data collection and optimization, assisting in incident response, reviewing performance reports, and training team members on automation and communication of changes. The role serves as an escalation point for incidents and promotes experimentation with new technologies to drive improvements and site reliability knowledge.
Required Skills
Soft Skills & Professional Competencies
Data Analysis
Nice to have:
Productivity & Workplace Tools
RPA
Information Technology
Concurrent ProgrammingScripting
Soft Skills & Professional Competencies
LeadershipCollaborationProblem SolvingContinuous Learning
Business, Sales & Management
Process ImprovementPerformance ManagementTalent Acquisition
🎁 Benefits & Perks
flexible medical, life insurance, and retirement options, volunteer programs
Requirements
Requires 8 years of experience in software engineering, infrastructure management, or related field, or a Bachelor's degree with 4 years of experience, or a Master's degree with 2 years of experience. Must have 3 years of experience in automation, programming, and scripting, along with demonstrated data analysis skills. Preferred qualifications include additional years of experience, leadership/management experience, budget experience, and more extensive automation/programming/scripting experience.
Description
Supports team members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help forecast demands and ensure systems have adequate resources. Assists with collaboration between team members and the software development team to create reliable, scalable infrastructures. Advises on data collection, optimizing operations and infrastructure reliability. Aids in incident response activities to ensure service reliability. Reviews health and performance reports. Trains team members to identify automation. Trains team members to communicate and understand the impact of changes. Serves as an escalation point for incidents and reviews documentation. Enables team members to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data.
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00