[CVPR 2023] Efficient Semantic Segmentation by Altering Resolutions for Compressed Videos
-
Updated
Feb 23, 2024 - Python
[CVPR 2023] Efficient Semantic Segmentation by Altering Resolutions for Compressed Videos
[ICCV 2023] Accurate and Fast Compressed Video Captioning
An extension of CoCap for fast and accurate compressed video captioning. FocalCap introduces Distilled Motion MAE pretraining and an AGDTR module to selectively enrich visual patches from H.264 encoded videos, operating entirely without an audio encoder.
Hav-Cocap: Hybrid Audio-Visual Compressed Video Captioning framework. Extends CoCap with an Audio Encoder and evaluated on the AVCaps dataset.
Add a description, image, and links to the compressed-video topic page so that developers can more easily learn about it.
To associate your repository with the compressed-video topic, visit your repo's landing page and select "manage topics."