[ICCV 2023] Accurate and Fast Compressed Video Captioning
-
Updated
Jul 28, 2025 - Python
[ICCV 2023] Accurate and Fast Compressed Video Captioning
[CVPR 2023] Efficient Semantic Segmentation by Altering Resolutions for Compressed Videos
VOCA: Visual Odometry with Codec Awareness (ECCV 2026 - Spotlight)
Hav-Cocap: Hybrid Audio-Visual Compressed Video Captioning framework. Extends CoCap with an Audio Encoder and evaluated on the AVCaps dataset.
An extension of CoCap for fast and accurate compressed video captioning. FocalCap introduces Distilled Motion MAE pretraining and an AGDTR module to selectively enrich visual patches from H.264 encoded videos, operating entirely without an audio encoder.
To associate your repository with the compressed-video topic, visit your repo's landing page and select "manage topics."