Principal Core Infrastructure Engineer

🏢 Oracle
📍 Nashville, United StatesFull-timeOn-site
📅 Posted: 4w ago🔄 Updated: 4w ago
CV%
✨ AI Summary
The Principal Core Infrastructure Engineer will lead the development and architecture of scalable, elastic distributed systems. This role involves defining and enforcing scalability requirements, optimizing code and data paths for high-throughput workloads, and leveraging data plane platforms for large-scale data retrieval, storage, and processing. The engineer will design fault-tolerant, in-service-upgradable systems, implement strategies to handle network unreliability while meeting SLOs, and establish KPIs and telemetry. Responsibilities include proactive diagnosis and resolution of production issues, mentoring peers, ensuring operational readiness, implementing robust security controls, and developing Infrastructure as Code (IaC) and automation for safe patching, updates, and rollbacks within change-management plans.
Required Skills
Information Technology
System DesignCloud ArchitectureScalabilityDistributed SystemsData ReplicationObjective-COpenTelemetryDashboardsAlertingDebuggingIncident ManagementInfrastructure as CodeCybersecurity
Other
elasticityfault toleranceredundancyfailoverload sheddingthrottlingrate limitingkey performance indicators (KPIs)fault injectionbrownoutsdata synchronizationsecurity controlsupdatesrollbacks
Finance, Legal & Governance
MediationHR Compliance
Productivity & Workplace Tools
RPA
Business, Sales & Management
Change ManagementMentoringProcess Improvement
Soft Skills & Professional Competencies
CollaborationProblem SolvingContinuous Learning
🎁 Benefits & Perks
Medical, dental, and vision insurance, including expert medical opinion; Short term disability and long term disability; Life insurance and AD&D; Supplemental life insurance (Employee/Spouse/Child); Health care and dependent care Flexible Spending Accounts; Pre-tax commuter and parking benefits; 401(k) Savings and Investment Plan with company match; Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.; 11 paid holidays; Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.; Paid parental leave; Adoption assistance; Employee Stock Purchase Plan; Financial planning and group legal; Voluntary benefits including auto, homeowner and pet insurance.
Requirements
Leads development and begins architecting components of scalable, elastic distributed systems. Defines and enforces scalability requirements for owned components; optimizes code and data paths for high‑throughput, hyper‑scale workloads; and leverages data plane platforms for large‑scale retrieval, storage, and processing. Designs fault‑tolerant, in‑service‑upgradable systems using redundancy, replication, failover, and policies for partitions, applying load‑shedding, throttling, and rate‑limiting to handle network unreliability while meeting SLOs. Establishes KPIs and telemetry; builds proactive dashboards and alerts; and designs complex validation (fault injection, brownouts), replication, and synchronization for correctness and durability. Proactively diagnoses and resolves production issues, mentors peers, and ensures operational readiness. Implements robust security controls, executes remediation, maintains compliance documentation, and develops IaC and automation that enable safe patching, updates, and rollbacks within change‑management plans.
Description

Leads development and begins architecting components of scalable, elastic distributed systems. Defines and enforces scalability requirements for owned components; optimizes code and data paths for high‑throughput, hyper‑scale workloads; and leverages data plane platforms for large‑scale retrieval, storage, and processing. Designs fault‑tolerant, in‑service‑upgradable systems using redundancy, replication, failover, and policies for partitions, applying load‑shedding, throttling, and rate‑limiting to handle network unreliability while meeting SLOs. Establishes KPIs and telemetry; builds proactive dashboards and alerts; and designs complex validation (fault injection, brownouts), replication, and synchronization for correctness and durability. Proactively diagnoses and resolves production issues, mentors peers, and ensures operational readiness. Implements robust security controls, executes remediation, maintains compliance documentation, and develops IaC and automation that enable safe patching, updates, and rollbacks within change‑management plans.

✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00