Skip to content
View eastha10's full-sized avatar

Block or report eastha10

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
eastha10/README.md

김동하 | Dongha Kim

AI Engineer & Researcher

Building efficient learning systems — from language models to low-latency robot control.

Email LinkedIn GitHub


About Me

  • 🎓 Konkuk University — Computer Science & Engineering
  • 🔬 Interested in Reinforcement Learning, NLP, Model Compression, and Physical AI
  • 🧩 I enjoy turning research ideas into reproducible experiments and working systems
  • 🌱 Currently studying policy distillation and compression for real-time robot control

Current Research

A research pipeline for distilling a high-performance PPO teacher into a compact student policy and measuring real CPU latency gains from quantization and structured pruning.

Stage Status
Minimal PPO baseline ✅ Completed — Reacher-v5, 2.5M environment steps
Reproducible evaluation ✅ Completed — 30 seeded episodes
Improved PPO teacher 🚧 In progress
Teacher–student distillation 🧱 Pipeline scaffolded
Quantization & structured pruning 🧱 Evaluation interfaces scaffolded

Research focus: PPO · Knowledge Distillation · PTQ/QAT · Structured Pruning · CPU Latency Benchmarking

Featured Work

Implemented an encoder–decoder Transformer directly in PyTorch and compared full fine-tuning with LoRA ranks 4, 8, and 16 on approximately 1.6M Korean–English sentence pairs.

  • Built the Transformer architecture and training pipeline from scratch
  • Evaluated with BLEU, chrF, BERTScore, perplexity, throughput, and GPU memory
  • LoRA rank 8 trained only 0.4842% of total parameters
  • Reduced training time by about 9.14% and peak GPU memory by about 20% compared with full fine-tuning

PyTorch Transformer LoRA SentencePiece NLP Parameter-Efficient Fine-Tuning

✈️ TripTailor

AI-powered personalized travel curation platform developed as a five-person team project.

  • Recommends destinations using user preferences, tags, and vector search
  • Supports itinerary creation, sharing, reviews, and recommendation feedback loops
  • Contributed to AI recommendation features and frontend integration
  • Built and deployed as a Django service with Docker, Gunicorn, Nginx, and GitHub Actions

Django Python JavaScript Vector Search AI Recommendation Docker

Tech Stack

AI / Research

PyTorch Transformers NumPy Pandas MuJoCo

Development

Python C++ Java Django Docker Git

Algorithm

Solved.ac profile

GitHub Activity

Dongha's GitHub stats Most used languages

Open to conversations about AI research, efficient models, and collaborative projects.

Pinned Loading

  1. TripTailor TripTailor Public

    Forked from pirogramming/TripTailor

    TripTailor - 특색 있는 여행지 AI 큐레이션 플랫폼

    Python 1

  2. enko-transformer-lora enko-transformer-lora Public

    Python 1