Highlight offline video understanding
Browse files
README.md
CHANGED
|
@@ -5,7 +5,7 @@ pipeline_tag: video-text-to-text
|
|
| 5 |
|
| 6 |
# JoyAI-VL-Interaction
|
| 7 |
|
| 8 |
-
**The first open, vision-driven real-time interaction model — it watches a live video stream and decides on its own when to speak, stay silent, or delegate.**
|
| 9 |
|
| 10 |
[📄 Paper](https://arxiv.org/abs/2606.14777) · [🌐 Project Page & Demos](https://joyai-vl-video-future-academy-jd.github.io/JoyAI-VL-Interaction/) · [💻 GitHub](https://github.com/jd-opensource/JoyAI-VL-Interaction) · [🤗 Paper Page](https://huggingface.co/papers/2606.14777)
|
| 11 |
|
|
|
|
| 5 |
|
| 6 |
# JoyAI-VL-Interaction
|
| 7 |
|
| 8 |
+
**The first open, vision-driven real-time interaction model — it watches a live video stream and decides on its own when to speak, stay silent, or delegate. While enabling online, real-time interaction, this release also delivers powerful offline video understanding, making it the most comprehensive open-source model for video-related capabilities in the 8B parameter class.**
|
| 9 |
|
| 10 |
[📄 Paper](https://arxiv.org/abs/2606.14777) · [🌐 Project Page & Demos](https://joyai-vl-video-future-academy-jd.github.io/JoyAI-VL-Interaction/) · [💻 GitHub](https://github.com/jd-opensource/JoyAI-VL-Interaction) · [🤗 Paper Page](https://huggingface.co/papers/2606.14777)
|
| 11 |
|