Principal Core Infrastructure Engineer

🏢 Global Corporation
📍 Dubai, United Arab EmiratesFull-timeHybrid
📅 Posted: 6d ago🔄 Updated: 6d ago
CV%
✨ AI Summary
As a Principal Core Infrastructure Engineer, you will be responsible for designing, implementing, and maintaining the infrastructure that supports customer AI and machine learning initiatives. This role involves working closely with customer technical teams and internal cloud services to ensure efficient, secure, and scalable AI/ML solutions. You will troubleshoot proof of concept and production deployments, optimize infrastructure for performance, reliability, and cost-effectiveness, and provide technical guidance to junior engineers. Key qualifications include experience with scripting and automation tools (Ansible, Terraform, Python, Kubernetes), containerization (Docker, Kubernetes), orchestration tools (Slurm, PBS), networking, security, and strong Linux system administration skills. Preferred qualifications include proficiency in programming languages like Python, Rust, Go, Java, or Scala, and experience with AI/ML or HPC workloads, DevOps practices, and high-performance computing/GPU systems.
Required Skills
Other
AnsibleSlurmPBSRHELCentOSUbuntuDebian
Information Technology
TerraformPythonKubernetesDockerCore NetworkingCloud SecurityTechnical DocumentationLinuxSystems Administration
Soft Skills & Professional Competencies
Problem SolvingCommunicationCollaborationStrategic ThinkingPlanningExecutionDelegationPrioritization
Business, Sales & Management
Mentoring
Nice to have:
Information Technology
RustGoJavaScalabilityTensorFlowPyTorchScikit-learnJenkinsCI/CDPrometheus
🎁 Benefits & Perks
flexible medical, life insurance, and retirement options, volunteer programs
Requirements
The ideal candidate will have experience in scripting and automation using tools like Ansible, Terraform, Python, and/or Kubernetes. They should also have experience with containerization technologies (Docker, Kubernetes) and orchestration tools (Slurm, PBS), solid understanding of networking and security principles, and strong Linux skills. Preferred qualifications include proficiency in programming languages like Python, Rust, Go, Java, or Scala, and experience designing/managing infrastructure for AI/ML or HPC workloads.
Description
Job description / Role Job Type Full Time Job Location Dubai, UAE Nationality Any Nationality Salary Not Specified Gender Not Specified Arabic Fluency Not Specified Job Function IT - Software & Web Development Company Industry Software & Internet Services Job description As a principal core infrastructure engineer (AI/ML forward deployed infrastructure engineer), you will play a critical role in designing, implementing, and maintaining the infrastructure that supports our customers' AI and machine learning initiatives. You will work closely with customer technical teams (data scientists, software engineers, and Infra/IT professionals) and internal cloud services team to ensure their AI/ML solutions are deployed efficiently, securely, and at scale. Your expertise will be crucial in optimizing our infrastructure for performance, reliability, and cost-effectiveness. In this hands-on technical role you'll troubleshoot issues for proof of concept (POC) and production deployments. Responsibilities Lead the design and architecture of AI/ML infrastructure, considering performance, scalability, and security. Implement and maintain AI/ML infrastructure, ensuring efficient and reliable operations. Collaborate with customer technical teams to understand their AI/ML requirements and provide tailored solutions. Work closely with our cloud services team to integrate AI/ML solutions into our cloud platform. Ensure the secure deployment of AI/ML models and data, adhering to industry best practices. Monitor and optimize AI/ML infrastructure performance, identifying and resolving bottlenecks. Provide technical guidance and mentorship to junior engineers, fostering a culture of knowledge sharing. Stay updated with the latest AI/ML technologies and trends, driving innovation within the team. Document and communicate infrastructure designs, ensuring clear and concise documentation. Engage with customers and stakeholders to gather feedback and ensure their satisfaction with our AI/ML offerings. Qualifications Experience in scripting and automation using tools like Ansible, Terraform, Python and/or Kubernetes. Experience with containerization technologies (e.g., Docker, Kubernetes) and orchestration tools (like Slurm, PBS, etc.) for managing distributed systems. Solid understanding of networking concepts, security principles, and best practices. Excellent problem-solving skills, with the ability to troubleshoot complex issues and drive resolution in a fast-paced environment. Strong communication and collaboration skills, with the ability to work effectively in cross-functional teams and convey technical concepts to non-technical stakeholders. Strong documentation skills with experience documenting infrastructure designs, configurations, procedures, and troubleshooting steps to facilitate knowledge sharing, ensure maintainability, and enhance team collaboration. Strong Linux skills with hands-on experience in Oracle Linux/RHEL/CentOS, Ubuntu, and Debian distributions, including system administration, package management, shell scripting, and performance optimization. Preferred qualifications Strong proficiency in at least one of the programming languages such as Python, Rust, Go, Java, or Scala. Proven experience designing, implementing, and managing infrastructure for AI/ML or HPC workloads. Understanding machine learning frameworks and libraries such as TensorFlow, PyTorch, or sci-kit-learn and their deployment in production environments is a plus. Familiarity with DevOps practices and tools for continuous integration, deployment, and monitoring (e.g., Jenkins, GitLab CI/CD, Prometheus). Strong experience with high-performance computing/GPU systems. Additional information Core responsibilities Strategic thought leadership: You should also have a demonstrated ability to think strategically about business, products, and technical challenges. Business experience: Understand the challenges in working with large customers. Planning & execution: Manages and coordinates moderately complex tasks, monitoring timelines and deliverables to ensure timely completion and adherence to requirements for a moderately sized project or initiative. Efficiently delegates, monitors, and prioritizes work across multiple projects, providing technical oversight and adjusting plans to address shifts in resources or timelines. Collaboration & partnership: Collaborates across the organization to align on expectations and achieve shared objectives. Leverages understanding of business leaders, stakeholders, and/or customers to ensure proposed solutions meet their needs. Supports inclusivity by actively seeking and listening to diverse perspectives, ensuring others feel heard and respected. Problem solving: Identifies and addresses moderately complex issues by analyzing a wide range of data and/or information to identify solutions in accordance with standard practices. Proactively escalates unresolved or critical issues with a thorough assessment and suggests potential solutions. Reviews, contributes to, and documents problem solving strategies. Continuous learning: Pursues learning opportunities to expand knowledge and skills and/or tools in new areas and stays abreast of the latest industry trends and best practices. Proactively seeks and leverages ongoing feedback and training to improve skills. Coaches and mentors junior team members, fostering continuous learning and knowledge sharing within and across teams. Continuous improvement: Develops ideas, recommends updates, and/or collaborates on the implementation of process improvements to increase the efficiency and effectiveness of processes, protocols, and workflows across teams, and evaluates the impact on key stakeholders. Solicits feedback from others on ideas for alternative approaches and methods for continued improvement. Performance and development: Contributes to the talent development pipeline by participating in candidate interviews, assessing candidates, and providing hiring recommendations. Qualifications Career level - IC4 About us Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We're committed to including people with disabilities at all stages of the employment process. Oracle is an equal employment opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law. Apply Now
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00