This is a hybrid role with a mixed Anchoring of Data Engineering and Machine Learning leadership. You will be the technical authority responsible for scaling ingestion pipelines across diverse formats (PDFs, images, Exl, Doc, Video and audio, raw text, logs) and integrating state-of-the-art ML/NLP/Vision extraction models into robust, enterprise-grade production pipelines.
As a Technical Anchor, you will balance hands-on architecture and coding with technical mentorship, driving engineering standards, MLOps best practices, Platform reliability, and cross-team alignment.
Key Responsibilities
1. Technical Anchor & Architecture Leadership
Direct the end-to-end technical strategy, design, and architecture for the unstructured data ingestion and ML extraction platform.
Serve as the primary technical contact and subject matter expert across engineering teams, product managers, and enterprise stakeholders.
Lead architectural reviews, define design patterns, establish coding standards, and enforce data security, privacy, and governance protocols.
Mentor and coach senior and mid-level software engineers, data engineers, and ML engineers to foster technical excellence.
2. Unstructured Data Engineering & Pipeline Infrastructure
Architect scalable, fault-tolerant batch and real-time streaming ingestion pipelines for processing complex unstructured data sources at scale. (PDFs, PPT, Excel, Word, Video and Audio Formats)
Maintain robust data orchestration, transformation, parsing, chunking, deduplication, and metadata enrichment strategies.
Build automated data quality gates, validation frameworks, and observability pipelines to ensure high data integrity.
3. ML Extraction & AI Integration
Integrate advanced Machine Learning, Computer Vision, OCR, NLP, and LLM-based models into production extraction pipelines (e.g., document intelligence, entity extraction, semantic parsing).
Establish scalable model serving and inference patterns (low-latency streaming inference and high-throughput batch extraction).
Partner with Data Science and ML research teams to optimize model deployment, quantization, prompt engineering, fine-tuning workflows, and output validation.
4. Platform Reliability
Drive platform scalability, performance tuning, infrastructure cost optimization, and high availability across cloud and hybrid environments.
Required Qualifications & Technical Experience
Experience: 8+ years of progressive experience in Software Engineering, Data Engineering, or ML Engineering, with at least 2+ years as a Technical Anchor, Lead Engineer, or Principal Architect.
Preferred Qualifications
Copyright © 2026 Grabjobs Pte.Ltd. All Rights Reserved.