AI / Computer Vision Engineer

Hoang Tan DuyDaniel — 7+ years, Computer Vision · Generative AI · LLM Agents

  • Good command of English
  • Real-world AI product building experience, not just POCs or side projects
  • In-depth understanding of production: evaluation, monitoring, scaling, deployment...
  • Capability to drive the solution rather than just implementing tasks
HANOI, VIETNAM  ·  AI TEAM LEADER  ·  OPEN TO NEW ROLES
Multimodal AI agent — data, training, deployment, LLMOps, application
01 · About

Senior AI/ML engineer, generalist by necessity

I design and ship end-to-end AI systems across computer vision, generative AI, and LLM-based agents — not research demos, production systems with real users and real acceptance criteria. That means owning the full path: translating a client's business requirement into technical architecture, preparing datasets and fine-tuning models (LoRA and beyond), and getting the result running reliably on AWS, VastAI, or RunPod.

Most of the interesting problems I've worked on sit at the boundary between "the model can do this in principle" and "the client needs this to work every time" — which usually means multi-agent orchestration with LangChain/LangGraph on the generative side, and confronting domain gap and strict false-negative bars on the detection side.

I also lead: mentoring engineers and interns, running technical and model reviews, and translating between business stakeholders and the engineering team.

02 · Experience

Where I've worked

APR 2024 —
PRESENT
AI Team Leader / AI Engineer · SotaTek

Leading AI sub-teams across generative and computer-vision client projects — from technical architecture and model strategy through to production deployment. Includes Flickrz, DrawMind, TryNectar, and Bloom.

FEB 2019 —
MAR 2024
AI Engineer · BHSoft

Five years building and deploying early AI/ML systems — the foundation the later computer-vision and generative-AI specialization was built on.

03 · Selected Work

Featured projects

Four production systems, reverse-chronological by how central they are to my current focus: agent orchestration and computer vision under real constraints. Flickrz, DrawMind, and TryNectar open into a full case study.

AI LEADER

Webtoon AI image-generation platform for a Korean distributor with 60M+ users. Designed a multi-agent pipeline — a Script Writer Agent, a Script Reviewer Agent, dedicated agents for character and per-scene image prompts, and an Image Quality Supervisor Agent — with two human-in-the-loop hard gates. LoRA fine-tuning keeps each character's identity consistent across hundreds of panels.

Multi-Agent OrchestrationLangGraphLoRAComfyUIFastAPI
AI ENGINEER

An AI agent system that reads and reasons deeply over complex technical engineering drawings — orchestrating LLM/VLM capabilities and purpose-built tools to extract information and answer detailed questions. A production RT-DETR detection and segmentation pipeline serves as one of those tools, locating View, Note, and Table regions with mAP above 0.95 despite limited data, compute constraints, and heavy visual overlap between classes.

Object DetectionRT-DETRDomain AdaptationCVAT
Bloom
AI ENGINEER

Production multi-agent LLM system for biological agriculture, giving farmers data-driven treatment recommendations. LangGraph state machines route crop-lifecycle and environmental data through a RAG pipeline built and validated with agronomists.

LangGraphRAGMulti-AgentFastAPI
AI ENGINEER

Live, profitable multimodal AI companion product — text, image, and video generation orchestrated through ComfyUI and LangGraph. Solved latency and character-consistency problems at scale via dynamic persona injection and context-window management.

Multimodal AIComfyUILangGraphRunPod
2D Drawing Generation
CAD-intelligence system that learns 3D geometric representations and auto-generates structured 2D engineering drawings — GNN-based feature recognition over CAD topology graphs, plus a Point Cloud Transformer for shape embeddings. ~90%+ accuracy on major feature classes internally.
RAG Legal Lookup
Legal research assistant for traffic-safety law — full RAG pipeline with chunking strategy, embeddings, metadata filtering, and LLM re-ranking for citation-grounded answers. Adopted into daily legal-advisor workflows.
Marketing Image Gen
Internal on-premise Stable Diffusion tool that cut campaign visual turnaround from days to minutes, integrated directly into the company CMS via FastAPI.
04 · Skills

Where the years went

Years of hands-on experience per area, out of 7 — the length of my AI/ML career so far.

Core AI / ML
Computer Vision
7y
Evaluation
5y
AI Image Generation
4y
LLM (Multimodal & Multilingual)
4y
Prompt Engineering
4y
AI Agent Orchestration
3y
Tools & Infra
ComfyUI
3y
FastAPI
3y
RunPod / VastAI / AWS
3y
LangChain / LangGraph
2y

Education

Hanoi University of Science and Technology — Information Technology (Global ICT)

Languages

Vietnamese — Native
English — Fluent (C1)
05 · Contact

Let's talk.

Open to new roles in computer vision and applied AI. The fastest way to reach me is LinkedIn.

Location
Hanoi, Vietnam