You can use the raw model for video feature extraction.
How to use
Here is how to use this model to extract a video feature:
from transformers import VideoMAEImageProcessor, AutoModel, AutoConfig
import numpy as np
import torch
config = AutoConfig.from_pretrained("OpenGVLab/VideoMAEv2-Base", trust_remote_code=True)
processor = VideoMAEImageProcessor.from_pretrained("OpenGVLab/VideoMAEv2-Base")
model = AutoModel.from_pretrained('OpenGVLab/VideoMAEv2-Base', config=config, trust_remote_code=True)
video = list(np.random.rand(16, 3, 224, 224))
# B, T, C, H, W -> B, C, T, H, W
inputs = processor(video, return_tensors="pt")
inputs['pixel_values'] = inputs['pixel_values'].permute(0, 2, 1, 3, 4)
with torch.no_grad():
outputs = model(**inputs)
BibTeX entry and citation info
@InProceedings{wang2023videomaev2,
author = {Wang, Limin and Huang, Bingkun and Zhao, Zhiyu and Tong, Zhan and He, Yinan and Wang, Yi and Wang, Yali and Qiao, Yu},
title = {VideoMAE V2: Scaling Video Masked Autoencoders With Dual Masking},
booktitle = {Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)},
month = {June},
year = {2023},
pages = {14549-14560}
}
@misc{videomaev2,
title={VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking},
author={Limin Wang and Bingkun Huang and Zhiyu Zhao and Zhan Tong and Yinan He and Yi Wang and Yali Wang and Yu Qiao},
year={2023},
eprint={2303.16727},
archivePrefix={arXiv},
primaryClass={cs.CV}
}
Runs of OpenGVLab VideoMAEv2-Base on huggingface.co
26.4K
Total runs
0
24-hour runs
1.4K
3-day runs
9.7K
7-day runs
9.7K
30-day runs
More Information About VideoMAEv2-Base huggingface.co Model
VideoMAEv2-Base huggingface.co is an AI model on huggingface.co that provides VideoMAEv2-Base's model effect (), which can be used instantly with this OpenGVLab VideoMAEv2-Base model. huggingface.co supports a free trial of the VideoMAEv2-Base model, and also provides paid use of the VideoMAEv2-Base. Support call VideoMAEv2-Base model through api, including Node.js, Python, http.
VideoMAEv2-Base huggingface.co is an online trial and call api platform, which integrates VideoMAEv2-Base's modeling effects, including api services, and provides a free online trial of VideoMAEv2-Base, you can try VideoMAEv2-Base online for free by clicking the link below.
OpenGVLab VideoMAEv2-Base online free url in huggingface.co:
VideoMAEv2-Base is an open source model from GitHub that offers a free installation service, and any user can find VideoMAEv2-Base on GitHub to install. At the same time, huggingface.co provides the effect of VideoMAEv2-Base install, users can directly use VideoMAEv2-Base installed effect in huggingface.co for debugging and trial. It also supports api for free installation.