Senior Manager, Core Infrastructure Engineering

🏢 Oracle
📍 Nashville, United StatesFull-timeOn-site
📅 Posted: 3w ago🔄 Updated: 3w ago
CV%
✨ AI Summary
The Senior Manager, Core Infrastructure Engineering role is responsible for managing a team that delivers scalable distributed systems and components. Key responsibilities include standardizing engineering practices, optimizing for high-throughput and hyper-scale workloads, and ensuring effective use of distributed state tools and data plane platforms. The role also involves guiding teams to design fault-tolerant, in-service-upgradable systems, setting SLO-aligned targets, and implementing resiliency mechanisms like load-shedding and throttling. The manager will oversee KPIs, telemetry, and dashboards, direct the design of functional and correctness requirements, and ensure proactive incident management, operational readiness, and on-call coverage. Security, compliance, and automation through Infrastructure as Code (IaC) are also critical aspects, along with managing change-management plans for safe patching and rollbacks. The position requires strong project management, cross-functional collaboration, problem-solving, and continuous improvement skills. The role is based in Nashville, TN.
Required Skills
Information Technology
Distributed SystemsScalabilityHyper-VData PlatformsKPI ReportingOpenTelemetryDashboardsData ReplicationIncident ManagementEncryptionInfrastructure as Code
Engineering, Construction & Trades
Site Management
Other
fault tolerancein-service updatesSLOsload sheddingthrottlingrate limitingcorrectness requirementsfault injectionsynchronizationaccess control
Soft Skills & Professional Competencies
ResilienceCollaborationTeam LeadershipProblem SolvingContinuous Learning
Business, Sales & Management
Change ManagementProject ManagementProcess ImprovementPerformance Management
Finance, Legal & Governance
MediationHR Compliance
Productivity & Workplace Tools
RPA
Education & Training
Training Content Development
🎁 Benefits & Perks
Medical, dental, and vision insurance, including expert medical opinion, Short term disability and long term disability, Life insurance and AD&D, Supplemental life insurance (Employee/Spouse/Child), Health care and dependent care Flexible Spending Accounts, Pre-tax commuter and parking benefits, 401(k) Savings and Investment Plan with company match, Paid time off (13 days annually for the first three years of employment and 18 days annually for subsequent years of employment, prorated for part-time), 11 paid holidays, 72 hours of paid sick leave upon date of hire (refreshes each calendar year, up to a maximum cap of 112 hours), Paid parental leave, Adoption assistance, Employee Stock Purchase Plan, Financial planning and group legal, Voluntary benefits including auto, homeowner and pet insurance.
Requirements
Manages team delivering scalable distributed systems and components on a 2–4 quarter horizon. Standardizes engineering practices and scalability requirements across teams; oversees optimization for high‑throughput, hyper‑scale workloads; and ensures effective use of distributed state tools and data plane platforms. Guides teams to design fault‑tolerant, in‑service‑upgradable systems, set SLO‑aligned durability/availability targets, and implement resiliency mechanisms.
Description

This position is based at our Nashville, TN location.


Manages team delivering scalable distributed systems and components on a 2–4 quarter horizon. Standardizes engineering practices and scalability requirements across teams; oversees optimization for high‑throughput, hyper‑scale workloads; and ensures effective use of distributed state tools and data plane platforms. Guides teams to design fault‑tolerant, in‑service‑upgradable systems, set SLO‑aligned durability/availability targets, and implement resiliency mechanisms (load‑shedding, throttling, rate‑limiting). Provides oversight for KPIs, telemetry, and moderately complex dashboards; directs design of functional/correctness requirements, fault‑injection tests, and replication/synchronization strategies. Ensures proactive incident management, operational readiness, and on‑call coverage; drives encryption/access control practices, remediation plans, and compliance documentation. Oversees development and maintenance of automation/IaC and partners with teams on change‑management plans enabling safe patching, updates, and rollbacks.

✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00