liu PRO
che111
AI & ML interests
None yet
Organizations
VideoForMed
-
Distilling Vision-Language Models on Millions of Videos
Paper • 2401.06129 • Published • 19 -
Koala: Key frame-conditioned long video-LLM
Paper • 2404.04346 • Published • 6 -
MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Paper • 2404.05726 • Published • 22 -
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
Paper • 2406.07471 • Published • 2
AlphaMed
VideoForMed
-
Distilling Vision-Language Models on Millions of Videos
Paper • 2401.06129 • Published • 19 -
Koala: Key frame-conditioned long video-LLM
Paper • 2404.04346 • Published • 6 -
MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Paper • 2404.05726 • Published • 22 -
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
Paper • 2406.07471 • Published • 2