Location
Helsinki, Finland
I'm an
ML Engineer II at Smartly in Helsinki, building AI-powered products. Previously led GenAI platform development at Nokia — LLMs, RAG, agentic workflows, and Kubernetes-native microservices with observability-first design.
ML Engineer II at Smartly in Helsinki, with 6+ years of experience in machine learning and software engineering. Previously an AI Tech Lead at Nokia, where I led a team building a production GenAI platform — RAG pipelines, agentic workflows, and distributed LLM training on private Kubernetes clusters, with observability-first design using MLFlow, OpenTelemetry, and Grafana.
I hold a fully funded Erasmus Mundus dual Master’s in Engineering of Data-Intensive Intelligent Software Systems (EDISS) — one of 23 selected worldwide — from Åbo Akademi University (Finland) and UIB (Spain), specialising in computer vision. I enjoy mentoring engineers, driving technical strategy, and shipping systems built to last.
I love meeting new people — feel free to say hello!
Companies Worked At
ML/GenAI Projects
Years Experience
Programming Competitions Won
Technical skills developed across production AI systems, cloud-native infrastructure, and NLP research.
A selection of work across GenAI platforms, LLM research, and NLP engineering.
Led end-to-end development of a production GenAI platform with RAG pipelines, agentic workflows, and distributed LLM training deployed on private on-prem Kubernetes clusters.
Systematic evaluation of Llama 3, Mistral, and Gemma across summarization, code generation, and conversational tasks. Benchmarked using ROUGE, BLEU, BERTScore, and LLM-as-judge — surfacing key accuracy and consistency trade-offs to guide model selection.
Built NLP tagging pipelines to classify and filter user reviews on travel platforms including KAYAK, enabling hotel-specific tag extraction (cleanliness, service) for smarter personalised search and improved SEO discoverability.
Built a governance and evidence layer for multi-agent AI systems. Handles output conflicts between agents using arbitration strategies (weighted voting, LLM-as-judge, human deferral), policy enforcement, explainability, and audit logging.
Contributed the LEPOR machine translation evaluation metric to NLTK — one of the most widely used NLP libraries in Python. Implemented sentence_lepor and corpus_lepor with customisable tokenisation options.
Areas where I can help you build, evaluate, and scale AI systems from prototype to production.
Designing and deploying Retrieval-Augmented Generation pipelines and conversational AI systems for enterprise search and knowledge management.
Building multi-tool and multi-agent orchestration systems for complex, autonomous enterprise tasks.
Full experiment tracking with MLFlow, distributed tracing with OpenTelemetry, and dashboards via Prometheus and Grafana.
Cloud-native microservice architecture, containerisation with Docker, and on-prem Kubernetes cluster management.
Rigorous model assessment across summarization, code generation, and conversational tasks — guiding model selection for production use.
Leading engineering teams, running agile sprints, mentoring junior engineers, and aligning technical roadmaps with business strategy.
Interested in collaborating on AI products, GenAI systems, or research? Let’s talk.
Based in Helsinki. Open to consulting, collaboration, and interesting AI challenges.
Helsinki, Finland
ulhaqi12@gmail.com