[CVPR20] Video Object Grounding using Semantic Roles in Language Description (https://arxiv.org/abs/2003.10606)
-
Updated
Jun 10, 2020 - Python
[CVPR20] Video Object Grounding using Semantic Roles in Language Description (https://arxiv.org/abs/2003.10606)
[NeurIPS 2025 Spotlight] Talk2Event: Grounded Understanding of Dynamic Scenes from Event Cameras
Zero-shot language-guided driving perception with Grounding DINO, CLIP, and Qwen2.5-VL.
Add a description, image, and links to the object-grounding topic page so that developers can more easily learn about it.
To associate your repository with the object-grounding topic, visit your repo's landing page and select "manage topics."