Sr. Engineer, ML Platform
4 days ago
Since launching in Kuwait in 2004, talabat, the leading on-demand food and Q-commerce app for everyday deliveries, has been offering convenience and reliability to its customers. talabat's local roots run deep, offering a real understanding of the needs of the communities we serve in eight countries across the region.
We harness innovative technology and knowledge to simplify everyday life for our customers, optimize operations for our restaurants and local shops, and provide our riders with reliable earning opportunities daily.
Here at talabat, we are building a high performance culture through engaged workforce and growing talent density. We're all about keeping it real and making a difference. Our 6,000+ strong talabaty are on an awesome mission to spread positive vibes. We are proud to be a multi great place to work award winner.
Job DescriptionSummary
As the leading delivery platform in the region, we have a unique responsibility and opportunity to positively impact millions of customers, restaurant partners, and riders. To achieve our mission, we must scale and continuously evolve our machine learning capabilities, including cutting-edge Generative AI (genAI) initiatives. This demands robust, efficient, and scalable ML platforms that empower our teams to rapidly develop, deploy, and operate intelligent systems.
As an ML Platform Engineer, your mission is to design, build, and enhance the infrastructure and tooling that accelerates the development, deployment, and monitoring of traditional ML and genAI models at scale. You'll collaborate closely with data scientists, ML engineers, genAI specialists, and product teams to deliver seamless ML workflows—from experimentation to production serving—ensuring operational excellence across our ML and genAI systems.
QualificationsResponsibilities
Design, build, and maintain scalable, reusable, and reliable ML platforms and tooling that support the entire ML lifecycle, including data ingestion, model training, evaluation, deployment, and monitoring for both traditional and generative AI models.
Develop standardized ML workflows and templates using MLflow and other platforms, enabling rapid experimentation and deployment cycles.
Implement robust CI/CD pipelines, Docker containerization, model registries, and experiment tracking to support reproducibility, scalability, and governance in ML and genAI.
Collaborate closely with genAI experts to integrate and optimize genAI technologies, including transformers, embeddings, vector databases (e.g., Pinecone, Redis, Weaviate), and real-time retrieval-augmented generation (RAG) systems.
Automate and streamline ML and genAI model training, inference, deployment, and versioning workflows, ensuring consistency, reliability, and adherence to industry best practices.
Ensure reliability, observability, and scalability of production ML and genAI workloads by implementing comprehensive monitoring, alerting, and continuous performance evaluation.
Integrate infrastructure components such as real-time model serving frameworks (e.g., TensorFlow Serving, NVIDIA Triton, Seldon), Kubernetes orchestration, and cloud solutions (AWS/GCP) for robust production environments.
Drive infrastructure optimization for generative AI use-cases, including efficient inference techniques (batching, caching, quantization), fine-tuning, prompt management, and model updates at scale.
Partner with data engineering, product, infrastructure, and genAI teams to align ML platform initiatives with broader company goals, infrastructure strategy, and innovation roadmap.
Contribute actively to internal documentation, onboarding, and training programs, promoting platform adoption and continuous improvement.
Requirements
Technical Experience
Strong software engineering background with experience in building distributed systems or platforms designed for machine learning and AI workloads.
Expert-level proficiency in Python and familiarity with ML frameworks (TensorFlow, PyTorch), infrastructure tooling (MLflow, Kubeflow, Ray), and popular APIs (Hugging Face, OpenAI, LangChain).
Experience implementing modern MLOps practices, including model lifecycle management, CI/CD, Docker, Kubernetes, model registries, and infrastructure-as-code tools (Terraform, Helm).
Demonstrated experience working with cloud infrastructure, ideally AWS or GCP, including Kubernetes clusters (GKE/EKS), serverless architectures, and managed ML services (e.g., Vertex AI, SageMaker).
Proven experience with generative AI technologies: transformers, embeddings, prompt engineering strategies, fine-tuning vs. prompt-tuning, vector databases, and retrieval-augmented generation (RAG) systems.
Experience designing and maintaining real-time inference pipelines, including integrations with feature stores, streaming data platforms (Kafka, Kinesis), and observability platforms.
Familiarity with SQL and data warehouse modeling; capable of managing complex data queries, joins, aggregations, and transformations.
Solid understanding of ML monitoring, including identifying model drift, decay, latency optimization, cost management, and scaling API-based genAI applications efficiently.
Qualifications
Bachelor's degree in Computer Science, Engineering, or a related field; advanced degree is a plus.
3+ years of experience in ML platform engineering, ML infrastructure, generative AI, or closely related roles.
Proven track record of successfully building and operating ML infrastructure at scale, ideally supporting generative AI use-cases and complex inference scenarios.
Strategic mindset with strong problem-solving skills and effective technical decision-making abilities.
Excellent communication and collaboration skills, comfortable working cross-functionally across diverse teams and stakeholders.
- Strong sense of ownership, accountability, pragmatism, and proactive bias for action.
-
ML Platform Engineer
6 days ago
London, Greater London, United Kingdom Isomorphic Labs Full time £80,000 - £120,000 per yearIsomorphic Labs is applying frontier AI to help unlock deeper scientific insights, faster breakthroughs, and life-changing medicines with an ambition to solve all disease.The future is coming. A future enabled and enriched by the incredible power of machine learning. A future in which diseases are curtailed or cured starting with better and faster drug...
-
ML Platform
1 week ago
London, Greater London, United Kingdom Isomorphic Labs Full time £120,000 - £180,000 per yearIsomorphic Labs is applying frontier AI to help unlock deeper scientific insights, faster breakthroughs, and life-changing medicines with an ambition to solve all disease.The future is coming. A future enabled and enriched by the incredible power of machine learning. A future in which diseases are curtailed or cured starting with better and faster drug...
-
Platform Engineer
3 days ago
London, Greater London, United Kingdom Carbon3 - Building the UK's AI Solution Platform Full time £40,000 - £80,000 per yearWe are building the UK's next generation AI platform, powered by renewable energy, rooted in sovereign capability, and designed to give enterprises and innovators the compute they need.AI Platform OperationsSupport Engineer / Cluster Administrator to provide Level 1 and Level 2 support for AI platform. This role will be customer facing, involve technical...
-
ML Platform Manager, London
1 week ago
London, Greater London, United Kingdom Isomorphic Labs Full time £120,000 - £180,000 per yearIsomorphic Labs is applying frontier AI to help unlock deeper scientific insights, faster breakthroughs, and life-changing medicines with an ambition to solve all disease.The future is coming. A future enabled and enriched by the incredible power of machine learning. A future in which diseases are curtailed or cured starting with better and faster drug...
-
ML Ops Engineer
1 week ago
London, Greater London, United Kingdom Vertexsearch Full time £100,000 - £120,000 per yearJob DescriptionWe're looking for an experienced ML Ops engineer to join a newly formed team in a leading quant firm, responsible for ML Operations across a next-generation research platform.This is a high-impact, greenfield role where you'll help design and build the future of ML infrastructure—from how data is shared, to how models are trained, deployed,...
-
Sr. Data Engineer – Industry 4.0
3 days ago
London, Greater London, United Kingdom Cognizant Full time £80,000 - £150,000 per yearJD: Sr. Data Engineer – Industry 4.0We are hiring a senior Data Engineer to lead the development of intelligent, scalable data platforms for Industry 4.0 initiatives. This role will drive integration across OT/IT systems, enable real-time analytics, and ensure robust data governance and quality frameworks. The engineer will collaborate with...
-
Sr. Data Engineer – Industry 4.0
2 days ago
London, Greater London, United Kingdom Cognizant Technology Solutions Full time £80,000 - £120,000 per yearJD: Sr. Data Engineer – Industry 4.0We are hiring a senior Data Engineer to lead the development of intelligent, scalable data platforms for Industry 4.0 initiatives. This role will drive integration across OT/IT systems, enable real-time analytics, and ensure robust data governance and quality frameworks. The engineer will collaborate with...
-
Senior ML Engineer
2 days ago
London, Greater London, United Kingdom Grid Dynamics Full time £80,000 - £120,000 per yearJoin us to contribute to a cutting-edge platform that transforms corporate learning through advanced ML translation and search capabilitiesCertainly Here's a professional job ad including the duties for the Senior ML Engineer role based on the provided information:We are seeking an experiencedSenior Machine Learning Engineerto join our client's innovative...
-
Junior ML Engineer
2 days ago
London, Greater London, United Kingdom Advai Full time £45,000 - £60,000 per yearJob Title:Junior ML Engineer (AI Safety and Security Research)Salary:£45,000 per annumLocation:London, UKType:Full-Time, PermanentWork Model:Hybrid (Minimum 2 days a week in the office)About Us:Advai are at the leading edge of AI innovation, focusing on AI testing and evaluation to help our customers deploy safe and secure AI with confidence. Our mission is...
-
ML Engineer
5 days ago
London, Greater London, United Kingdom Vortexa Full time £90,000 - £120,000 per yearProcessing thousands of energy data points per second from diverse operational sources, handling massive volumes of energy data while running sophisticated classification and anomaly detection models in real-time, maintaining comprehensive data lineage, and delivering insights through high-performance platforms used by energy operators globally requires...