Machine Learning Engineer AI Architecture Research

🏢 Jobgether
📍 Saudi ArabiaFull-timeHybrid
📅 Posted: 6d ago🔄 Updated: 6d ago
CV%
✨ AI Summary
This role is for a Machine Learning Engineer specializing in AI Architecture Research in Saudi Arabia. The position focuses on researching and developing next-generation AI model architectures, moving from experimental concepts to scalable production systems. You will work at the intersection of machine learning research, model engineering, and real-world deployment, challenging established architectural assumptions and exploring alternatives to conventional Transformer-based designs. Responsibilities include designing experiments, prototyping new neural networks, evaluating trade-offs, and collaborating with inference and systems engineers for efficient deployment. The role also involves contributing to research reproduction, benchmarking, and potentially open-source work, offering a chance to significantly influence AI architecture in a fast-moving, research-oriented environment.
Required Skills
Other
Neural Network ArchitecturesHybrid ArchitecturesRNNsAttention MechanismsState-Space Models
🎁 Benefits & Perks
Competitive compensation and meaningful equity. Opportunity to work directly on core AI model architecture rather than focusing primarily on fine-tuning. Significant influence over technical and research direction within a rapidly growing organization. Small, high-caliber team with fast feedback loops and a strong research-oriented environment. Opportunity to take research concepts from experimentation through to production deployment.
Requirements
Requires a strong foundation in machine learning and deep learning with practical experience in model development. Must have hands-on experience implementing neural network or model architectures from scratch and a strong understanding of attention mechanisms, RNNs, state-space models, hybrid architectures, or related approaches. Solid knowledge of training dynamics, optimization, scaling behavior, and architecture-level performance considerations is essential, along with proficiency in PyTorch or JAX. Ability to evaluate architectural ideas through theoretical reasoning and empirical experimentation, coupled with strong communication skills, is required.
Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Machine Learning Engineer — AI Architecture Research based in Saudi Arabia.This role focuses on researching and building next-generation AI model architectures that can move from experimental concepts to scalable production systems.You will work at the intersection of machine learning research, model engineering, and real-world deployment.The position offers the opportunity to challenge established architectural assumptions and explore alternatives to conventional Transformer-based designs.You will design experiments, prototype new neural networks, and evaluate trade-offs across compute, memory, latency, and model performance.The role involves close collaboration with inference and systems engineers to make research ideas efficient and deployable.You will also contribute to research reproduction, benchmarking, technical exploration, and potentially open-source work.This is an opportunity to have meaningful influence on AI architecture while working in a fast-moving, research-oriented environment.AccountabilitiesResearch and develop novel neural network architectures, including alternatives or extensions to Transformers, recurrent and hybrid models, and long-context systems.Design and execute architecture-level experiments focused on scaling laws, memory mechanisms, training behavior, and compute-performance trade-offs.Prototype models end-to-end, translating research concepts into robust, training-ready implementations.Analyze model behavior, failure modes, inductive biases, and architectural strengths and limitations.Collaborate with inference and systems engineering teams to ensure new architectures are efficient, scalable, and suitable for deployment.Read, reproduce, evaluate, and extend cutting-edge machine learning research papers.Contribute to internal research notes, benchmarks, experiments, and open-source initiatives where applicable.Move fluidly between theoretical investigation, rapid experimentation, and production-oriented engineering.RequirementsStrong foundation in machine learning and deep learning fundamentals, with practical experience applying them to model development.Hands-on experience implementing neural network or model architectures from scratch.Strong understanding of attention mechanisms, RNNs, state-space models, hybrid architectures, or related approaches.Solid knowledge of training dynamics, optimization, scaling behavior, and architecture-level performance considerations.Understanding of model-level memory, latency, compute, and efficiency constraints.Proficiency with PyTorch or JAX and the ability to develop and experiment with research-oriented ML code.Ability to evaluate architectural ideas through both theoretical reasoning and empirical experimentation.Strong communication skills, with the ability to clearly explain technical concepts and architectural trade-offs.Preferred experience with non-Transformer architectures such as RNN variants, state-space models, or long-context systems.Preferred background in research-driven startups, open-source machine learning projects, large-scale training, or custom training loops.Publications, preprints, notable research contributions, or experience with inference optimization and deployment constraints are advantageous.BenefitsCompetitive compensation and meaningful equity.Opportunity to work directly on core AI model architecture rather than focusing primarily on fine-tuning.Significant influence over technical and research direction within a rapidly growing organization.Small, high-caliber team with fast feedback loops and a strong research-oriented environment.Opportunity to take research concepts from experimentation through to production deployment.Full-time position with a globally distributed work environment.How Jobgether WorksWe use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.We appreciate your interest and wish you the best! Why Apply Through JobgetherData Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00