Retrieval-Augmented Dense Video Captioning and QA System
End-to-end pipeline for automated highlight extraction, dense captioning, and natural-language QA over long-form (1hr+) videos. Funded by Piaspace (2025).
Industry-funded research and applied projects.
End-to-end pipeline for automated highlight extraction, dense captioning, and natural-language QA over long-form (1hr+) videos. Funded by Piaspace (2025).
VLM-based risk-state prediction system fine-tuned to detect unsafe worker behavior and equipment hazards. Funded by Doosan Enerbility (2025).
Zero-shot image captioning for in-vehicle driver monitoring, generating natural-language descriptions of driver states. Funded by Hyundai NGV (2024).