This Director, Core Infrastructure Engineering role at Oracle focuses on leading multiple teams to implement strategies for the architecture and delivery of interdependent, scalable distributed systems. The position involves orchestrating cross-group optimization for high-throughput data processing, aligning stakeholders on scalability requirements, and overseeing elastic designs and effective use of data plane platforms. The role also provides strategic oversight for fault-tolerant, in-service-upgradable architectures, sets direction for partition-aware design choices, and leads initiatives to harden networks via load-shedding, throttling, and rate-limiting. Furthermore, it establishes expectations for formal verification and peer reviews, sets SLO-aligned durability and availability standards, drives KPI and telemetry strategies, and ensures functional/correctness validation, data replication, and synchronization meet organizational needs. The Director will guide organization-wide incident management and operational readiness, eliminating customer maintenance windows and ensuring consistent SOPs, while also providing strategic security guidance and sponsoring automation (IaC) and change-management alignment.
Required Skills
Information Technology
○Distributed Systems○Scalability○KPI Reporting○OpenTelemetry○Dashboards○Alerting○Data Replication○Incident Management○Encryption○Infrastructure as Code
Medical, dental, and vision insurance, expert medical opinion, short term and long term disability, life insurance and AD&D, supplemental life insurance, health care and dependent care Flexible Spending Accounts, pre-tax commuter and parking benefits, 401(k) Savings and Investment Plan with company match, Flexible Vacation, 11 paid holidays, 72 hours of paid sick leave, paid parental leave, adoption assistance, Employee Stock Purchase Plan, financial planning and group legal, voluntary benefits including auto, homeowner and pet insurance.
Requirements
The role requires leadership of multiple teams implementing strategies for scalable distributed systems, focusing on architecture, delivery, data processing, fault-tolerant designs, network hardening, and security. Key responsibilities include driving KPI and telemetry strategies, overseeing incident management, and sponsoring automation and change management for safe system updates.
Description
Leads a multple teams to implement strategies for the architecture and delivery of interdependent, scalable distributed systems that meet organizational and customer demands. Orchestrates cross-group optimization for high‑throughput, large‑scale data processing; aligns stakeholders on scalability requirements; and oversees elastic designs and effective use of data plane platforms. Provides strategic oversight for fault‑tolerant, in‑service‑upgradable architectures, sets direction for partition‑aware design choices, and leads initiatives to harden networks via load‑shedding, throttling, and rate‑limiting. Establishes expectations for formal verification and peer reviews, and sets SLO‑aligned durability and availability standards across the department. Drives KPI and telemetry strategies; directs creation of complex dashboards and alerting for proactive health assurance; and ensures functional/correctness validation, data replication, and synchronization meet organizational needs. Guides organization‑wide incident management and operational readiness, eliminating customer maintenance windows and ensuring consistent SOPs. Provides strategic security guidance (encryption, access controls), oversees remediation and compliance documentation, and sponsors automation (IaC) and change‑management alignment so systems can be safely patched, updated, and rolled back at scale.