Principal Software Engineer, Core Infrastructure

🏢 Oracle
📍 Nashville, United StatesFull-timeOn-site
📅 Posted: 3w ago🔄 Updated: 3w ago
CV%
✨ AI Summary
This Principal Software Engineer role focuses on leading the development and architecture of scalable, elastic distributed systems. Key responsibilities include optimizing code for high-throughput workloads, designing fault-tolerant and in-service-upgradable systems, and implementing robust security controls and automation. The role also involves proactive production issue diagnosis and resolution, mentoring peers, and ensuring operational readiness and compliance. Qualifications for this position include experience in designing and implementing scalable and reliable distributed systems, handling network unreliability, establishing KPIs and telemetry, and developing Infrastructure as Code. The candidate will also be responsible for compliance, security, and automation tasks, as well as contributing to team development and process improvements.
Required Skills
Information Technology
Distributed SystemsScalabilityHyper-VData PlatformsData ReplicationKPI ReportingOpenTelemetryDashboardsAlertingInfrastructure as CodeCybersecuritySystem DesignSolution ArchitecturePerformance TestingIncident ResponseEncryptionCloud Architecture
Other
elasticityfault tolerancein-service upgradesredundancyfailovernetwork partitionsload sheddingthrottlingrate limitingservice level objectivesfault injectionsynchronizationdurabilityproduction issue resolutionsecurity controlsupdatesrollbacksload testingaccess controlstechnical oversightcandidate interviewing
Soft Skills & Professional Competencies
Time ManagementIntegrityRoot Cause AnalysisTeam LeadershipDelegationPrioritizationCollaborationStakeholder ManagementProblem SolvingContinuous Learning
Engineering, Construction & Trades
ATS Systems
Business, Sales & Management
MentoringChange ManagementProcess Improvement
Finance, Legal & Governance
MediationHR Compliance
Productivity & Workplace Tools
RPA
Education & Training
Training Content Development
🎁 Benefits & Perks
Medical, dental, and vision insurance, including expert medical opinion; Short term disability and long term disability; Life insurance and AD&D; Supplemental life insurance (Employee/Spouse/Child); Health care and dependent care Flexible Spending Accounts; Pre-tax commuter and parking benefits; 401(k) Savings and Investment Plan with company match; Flexible Vacation (13-18 days annually); 11 paid holidays; 72 hours of paid sick leave upon date of hire, carrying over up to 112 hours; Paid parental leave; Adoption assistance; Employee Stock Purchase Plan; Financial planning and group legal; Voluntary benefits including auto, homeowner and pet insurance.
Requirements
Leads development and begins architecting components of scalable, elastic distributed systems. Defines and enforces scalability requirements for owned components; optimizes code and data paths for high-throughput, hyper-scale workloads; and leverages data plane platforms for large-scale retrieval, storage, and processing. Designs fault-tolerant, in-service-upgradable systems using redundancy, replication, failover, and policies for partitions, applying load-shedding, throttling, and rate-limiting to handle network unreliability while meeting SLOs. Establishes KPIs and telemetry; builds proactive dashboards and alerts; and designs complex validation (fault injection, brownouts), replication, and synchronization for correctness and durability. Proactively diagnoses and resolves production issues, mentors peers, and ensures operational readiness. Implements robust security controls, executes remediation, maintains compliance documentation, and develops IaC and automation that enable safe patching, updates, and rollbacks within change-management plans.
Description

Leads development and begins architecting components of scalable, elastic distributed systems. Defines and enforces scalability requirements for owned components; optimizes code and data paths for high‑throughput, hyper‑scale workloads; and leverages data plane platforms for large‑scale retrieval, storage, and processing. Designs fault‑tolerant, in‑service‑upgradable systems using redundancy, replication, failover, and policies for partitions, applying load‑shedding, throttling, and rate‑limiting to handle network unreliability while meeting SLOs. Establishes KPIs and telemetry; builds proactive dashboards and alerts; and designs complex validation (fault injection, brownouts), replication, and synchronization for correctness and durability. Proactively diagnoses and resolves production issues, mentors peers, and ensures operational readiness. Implements robust security controls, executes remediation, maintains compliance documentation, and develops IaC and automation that enable safe patching, updates, and rollbacks within change‑management plans.

✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00