I’m an undergraduate researcher working on multimodal AI and small Vision-Language Models (VLMs), with a focus on improving reasoning faithfulness, grounding, and reliability. My current research explores reinforcement learning and preference-based post-training, including GRPO and DPO, with an emphasis on data-, label-, and memory-efficient methods. I’m also interested in extending this work toward 3D vision, computational imaging, and scientific AI.

Alongside my research, I’m completing a B.E. in Electrical Engineering at Chungnam National University and work as an AI/ML/DL Engineer Intern at AIdanBio, where I build LLM/RAG systems, AI backends, and biomedical AI applications.

Research Interests

Multimodal AI, Vision-Language Models, Efficient Post-Training, Grounded Reasoning, Computational Imaging, and 3D Vision

Thura Win Kyaw

Thura Win Kyaw

Undergraduate Researcher in Multimodal AI

Dec 2024 – Present

AI/ML/DL Engineer Intern

AIdanBio · Daejeon, KR

  • Developed AI backend systems, APIs, model-serving infrastructure, and LLM-powered applications.
  • Built RAG pipelines using vector databases, semantic search, prompt engineering, and domain-specific knowledge bases.
  • Fine-tuned Vision-Language Models (VLMs) for medical AI using multimodal SFT, grounding, GRPO, and evaluation pipelines.
  • Worked on knowledge distillation and teacher–student training to transfer reasoning from large multimodal models to smaller models.
Oct 2025 – Nov 2025

AI Software Engineer Intern

GRINDA AI · Daejeon, KR

Built an AI-powered Slack bot that automatically converts issue reports into structured GitHub issues using Claude AI and FastAPI, with auto-labeling, translation, and monitoring.

B.E. in Electrical Engineering

Chungnam National University · Daejeon, KR

Sept 2022 – Present

Relevant Coursework: Sensors and Measurement, Electromagnetics, Electronic Circuits, Linear Algebra, Computer Programming, Artificial Intelligence and Data Utilization

Deep Learning / ML
PyTorch, Transformers, PEFT/LoRA, NumPy, Scikit-learn
VLM / Multimodal AI
Vision-Language Models, Visual Grounding, Hallucination Mitigation, Multimodal Reasoning
Post-Training
SFT, GRPO, DPO, LoRA, Preference Data Generation, Knowledge Distillation
Compute / Infrastructure
Multi-GPU Training, Linux, Git, GitHub
Applied AI Systems
RAG, Vector Databases, FastAPI, Model Serving
Programming Languages
Python, C, MATLAB
Engineering Simulation
Ansys
Languages
Burmese (Native), English (Fluent), Korean (TOPIK 5)