Lead Principal Software Engineer, Core Infrastructure

🏢 Oracle
📍 Nashville, United StatesFull-timeOn-site
📅 Posted: Today🔄 Updated: Today
CV%
✨ AI Summary
The Lead Principal Software Engineer, Core Infrastructure will mentor teams and lead the architecture of highly scalable, interdependent distributed systems. This role involves identifying and resolving performance and scalability bottlenecks for hyper-scale workloads, defining scalability requirements, and designing elastic, high-impact systems. The engineer will also oversee fault-tolerant designs, optimize resilience mechanisms, and set SLO-aligned durability and availability standards. Responsibilities include establishing KPIs, applying formal verification, developing replication strategies, advising on production issue resolution, setting operational readiness standards, directing incident response, architecting security controls, and delivering enterprise-level automation (IaC) and change strategies.
Required Skills
Information Technology
Distributed SystemsScalabilityPerformance OptimizationSystem DesignOpenTelemetryData ReplicationIncident ResponseSecurity ArchitectureInfrastructure as Code
Other
fault toleranceformal verificationsynchronization
Soft Skills & Professional Competencies
ResilienceRoot Cause AnalysisLeadership
Finance, Legal & Governance
Tax Compliance
Productivity & Workplace Tools
RPA
Business, Sales & Management
Mentoring
🎁 Benefits & Perks
Medical, dental, and vision insurance, including expert medical opinion; Short term disability and long term disability; Life insurance and AD&D; Supplemental life insurance (Employee/Spouse/Child); Health care and dependent care Flexible Spending Accounts; Pre-tax commuter and parking benefits; 401(k) Savings and Investment Plan with company match; Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.; 11 paid holidays; Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.; Paid parental leave; Adoption assistance; Employee Stock Purchase Plan; Financial planning and group legal; Voluntary benefits including auto, homeowner and pet insurance.
Requirements
Mentors teams and leads the architecture of highly scalable, interdependent distributed systems. Identifies and removes performance/scalability bottlenecks for hyper-scale workloads; defines scalability requirements with stakeholders; and designs elastic, high-impact systems while advancing innovation in data plane platforms. Engineers and oversees fault-tolerant, in-service-upgradable designs; optimizes resilience mechanisms (load-shedding, throttling, rate-limiting); and sets SLO-aligned durability and availability standards across dependent services. Establishes KPIs and advanced telemetry; applies formal verification for complex features; and develops robust replication/synchronization strategies. Advises and leads resolution of complex production issues, sets operational readiness and SOP standards, and directs incident response and RCAs. Architects advanced security controls, drives remediation and compliance, and delivers enterprise-level automation (IaC) and change strategies enabling safe, automated patching, updates, and rollbacks.
Description

Mentors teams and leads the architecture of highly scalable, interdependent distributed systems. Identifies and removes performance/scalability bottlenecks for hyper‑scale workloads; defines scalability requirements with stakeholders; and designs elastic, high‑impact systems while advancing innovation in data plane platforms. Engineers and oversees fault‑tolerant, in‑service‑upgradable designs; optimizes resilience mechanisms (load‑shedding, throttling, rate‑limiting); and sets SLO‑aligned durability and availability standards across dependent services. 

Establishes KPIs and advanced telemetry; applies formal verification for complex features; and develops robust replication/synchronization strategies. Advises and leads resolution of complex production issues, sets operational readiness and SOP standards, and directs incident response and RCAs. Architects advanced security controls, drives remediation and compliance, and delivers enterprise‑level automation (IaC) and change strategies enabling safe, automated patching, updates, and rollbacks.

✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00