Video-Text-to-Text
Safetensors
qwen3_vl
QingyiSi commited on
Commit
86124c6
·
verified ·
1 Parent(s): 3eba857

Highlight offline video understanding

Browse files
Files changed (1) hide show
  1. README.md +1 -1
README.md CHANGED
@@ -5,7 +5,7 @@ pipeline_tag: video-text-to-text
5
 
6
  # JoyAI-VL-Interaction
7
 
8
- **The first open, vision-driven real-time interaction model — it watches a live video stream and decides on its own when to speak, stay silent, or delegate.**
9
 
10
  [📄 Paper](https://arxiv.org/abs/2606.14777) · [🌐 Project Page & Demos](https://joyai-vl-video-future-academy-jd.github.io/JoyAI-VL-Interaction/) · [💻 GitHub](https://github.com/jd-opensource/JoyAI-VL-Interaction) · [🤗 Paper Page](https://huggingface.co/papers/2606.14777)
11
 
 
5
 
6
  # JoyAI-VL-Interaction
7
 
8
+ **The first open, vision-driven real-time interaction model — it watches a live video stream and decides on its own when to speak, stay silent, or delegate. While enabling online, real-time interaction, this release also delivers powerful offline video understanding, making it the most comprehensive open-source model for video-related capabilities in the 8B parameter class.**
9
 
10
  [📄 Paper](https://arxiv.org/abs/2606.14777) · [🌐 Project Page & Demos](https://joyai-vl-video-future-academy-jd.github.io/JoyAI-VL-Interaction/) · [💻 GitHub](https://github.com/jd-opensource/JoyAI-VL-Interaction) · [🤗 Paper Page](https://huggingface.co/papers/2606.14777)
11