Senior Core Infrastructure Engineer

🏢 Oracle
📍 Austin, United StatesFull-timeOn-site
📅 Posted: 2mo ago🔄 Updated: 2mo ago
CV%
✨ AI Summary
Oracle is seeking a Senior Core Infrastructure Engineer to design, implement, and optimize components in distributed systems, focusing on scalability, resiliency, and operability. The role involves developing features for a critical Tier 0 service within Oracle Cloud Infrastructure (OCI), ensuring its resilience and efficiency as OCI grows. Responsibilities include building fault-tolerant systems, implementing recovery-oriented principles, and proactively detecting and mitigating issues. The engineer will also be responsible for operational excellence, including automation, security controls, incident response, and compliance.
Required Skills
Information Technology
ScalabilityPerformance TestingData PlatformsStorage SolutionsData ReplicationDashboardsOpenTelemetryIncident ResponseInfrastructure as CodeDebuggingEncryptionTechnical DocumentationSystem DesignSolution ArchitectureModel OptimizationMonitoringHigh AvailabilityCloud ArchitectureCybersecurity
Soft Skills & Professional Competencies
ResilienceTime ManagementIntegrityProblem SolvingRelationship BuildingPlanningExecutionPrioritizationCollaborationStakeholder ManagementActive ListeningData AnalysisContinuous Learning
Other
operabilityload testinghigh-volume retrievalfault-tolerant pathsredundancyautomatic failoverrecovery-oriented principlesretriescircuit breakerstimeoutstestsrunbooksRCAssynchronizationsecurity controlsaccess controlshorizontal scalingvertical scalingreliability designservice disruptionsnetwork unreliabilityalarm configurationoperational procedurescomponent healthfunctional requirementsfeature developmentsystem correctnesspreventing interruptionssecurity measuresupdating applicationsrolling back applicationstimelinespartnership buildingknowledge sharingbest practicesskill developmentindustry trendsfeedback seekingprocess efficiencyworkflow effectiveness
Hospitality, Retail & Customer Service
Payment Processing
Engineering, Construction & Trades
Fire AlarmEquipment MaintenanceGas ProcessingAutomation
Productivity & Workplace Tools
RPA
Finance, Legal & Governance
MediationTax ComplianceRegulatory Compliance
Business, Sales & Management
Change ManagementProcess Improvement
Science & Research
Optimization
Education & Training
EAL Support
🎁 Benefits & Perks
Medical, dental, and vision insurance, including expert medical opinion; Short term disability and long term disability; Life insurance and AD&D; Supplemental life insurance (Employee/Spouse/Child); Health care and dependent care Flexible Spending Accounts; Pre-tax commuter and parking benefits; 401(k) Savings and Investment Plan with company match; Paid time off: Flexible Vacation (13 days annually for the first three years, 18 days annually thereafter for full-time employees); 11 paid holidays; Paid sick leave (72 hours upon hire, carrying over up to 112 hours); Paid parental leave; Adoption assistance; Employee Stock Purchase Plan; Financial planning and group legal; Voluntary benefits including auto, homeowner and pet insurance.
Requirements
The ideal candidate will have experience designing, implementing, and optimizing components in distributed systems with an emphasis on scalability, resiliency, and operability. This includes experience with data plane platforms, distributed state tools, building fault-tolerant paths, and implementing recovery-oriented principles. Proficiency in detecting and mitigating issues via tests, alarms, dashboards, and telemetry, as well as authoring runbooks and participating in incident response is crucial. Experience with automation/IaC for troubleshooting and maintenance, and applying advanced security controls is also required.
Description

Designs, implements, and optimizes components in distributed systems with an emphasis on scalability, resiliency, and operability. Delivers features and load/performance tests; leverages data plane platforms and distributed state tools for high-volume retrieval, storage, and processing; and reviews peers’ implementations for scalability compliance. Builds fault-tolerant paths (redundancy, replication, automatic failover), applies recovery‑oriented principles, and implements retries, circuit breakers, and timeouts. Proactively detects and mitigates issues via tests, alarms, dashboards, and telemetry; authors runbooks and participates in incident response and RCAs. Implements standard replication and synchronization, develops automation/IaC for troubleshooting and maintenance, and applies advanced security controls (encryption, access, remediation) while ensuring change, compliance, and documentation standards are met.

Oracle Cloud Infrastructure Workflow is a Tier 0 service that is critical to the smooth functioning of ALL basic OCI services by enabling their execution of distributed, multi-step work in a fault tolerant manner. An engineer on this team is responsible for the development of features making the platform more resilient and efficient including the launch of a brand new V2 version, as well as the operational excellence of the service, as we scale to meet OCI's exponential growth needs.

✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00