We introduce Vidi, a family of Large Multimodal Models (LMMs) for a wide range of video understanding and editing (VUE) scenarios. The first release focuses on temporal retrieval (TR), i.e., identifying the time ranges in input videos corresponding to a given text query.
This model is the Vidi1.5 model version for temporal retrieval.
Vidi1.5-9B huggingface.co is an AI model on huggingface.co that provides Vidi1.5-9B's model effect (), which can be used instantly with this bytedance-research Vidi1.5-9B model. huggingface.co supports a free trial of the Vidi1.5-9B model, and also provides paid use of the Vidi1.5-9B. Support call Vidi1.5-9B model through api, including Node.js, Python, http.
Vidi1.5-9B huggingface.co is an online trial and call api platform, which integrates Vidi1.5-9B's modeling effects, including api services, and provides a free online trial of Vidi1.5-9B, you can try Vidi1.5-9B online for free by clicking the link below.
bytedance-research Vidi1.5-9B online free url in huggingface.co:
Vidi1.5-9B is an open source model from GitHub that offers a free installation service, and any user can find Vidi1.5-9B on GitHub to install. At the same time, huggingface.co provides the effect of Vidi1.5-9B install, users can directly use Vidi1.5-9B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.