Categories

AI Models Marketplace

Explore the latest and most popular AI models — LLM, vision, audio and more — to pick and integrate the right one.

ms eff gcvit deepfake b0 kodf

mit
The ms-eff-gcvit-deepfake-b0 is a specialized video classification model designed to detect synthetic media and deepfake...
KoreaPeter video-classification
17.6K 0

ms eff gcvit deepfake b5 kodf

mit
The ms-eff-gcvit-deepfake-b5-kodf is a specialized video classification model designed for high-accuracy deepfake detect...
KoreaPeter video-classification
16.3K 0

ms eff gcvit deepfake b0 ff plus plus

mit
The ms-eff-gcvit-deepfake-b0-ff++ is a specialized video classification model designed for high-precision deepfake detec...
KoreaPeter video-classification
13.7K 0

ms eff gcvit deepfake b0 celeb df v2

mit
The ms_eff_gcvit_deepfake_b0_celeb_df_v2 is a specialized video classification model designed to detect synthetic facial...
KoreaPeter video-classification
13.4K 0

ms eff gcvit deepfake b5 celeb df v2

mit
The ms-eff-gcvit-deepfake-b5 model is a specialized video classification tool designed to detect synthetic facial manipu...
KoreaPeter video-classification
13.4K 0

ms eff gcvit deepfake b5 ff plus plus

mit
The MS-EFF-GCVit-Deepfake-B5-FF++ is a specialized video classification model designed for high-precision deepfake detec...
KoreaPeter video-classification
10.5K 0

vjepa2 vitl fpc64 256

mit
V-JEPA (ViT-L/fpc64/256) is a self-supervised video representation model designed for efficient spatial-temporal feature...
facebook video-classification
792 1

vivit b 16x2 kinetics400

mit
ViViT-B/16x2 is a Video Vision Transformer designed specifically for high-accuracy video classification. Unlike traditio...
google video-classification
448 0

vivit b 16x2

mit
ViViT-B/16x2 is a Video Vision Transformer designed for efficient video classification by treating video clips as sequen...
google video-classification
423 0

vjepa2 vitg fpc64 256

apache-2.0
V-JEPA (ViT-G FPC64 256) is a sophisticated video representation model designed for self-supervised learning. Unlike tra...
facebook video-classification
398 0

xclip base patch32

mit
xclip base patch32 is a specialized vision-language model optimized for video classification tasks. Unlike standard imag...
microsoft video-classification
363 0

vjepa2 vitl fpc16 256 ssv2

mit
V-JEPA (Video Joint-Embedding Predictive Architecture) represents a shift from generative to predictive pre-training for...
facebook video-classification
325 0

xclip base patch32 16 frames

mit
The xclip-base-patch32 model is a specialized video classification tool designed to extend CLIP's visual-language capabi...
microsoft video-classification
172 0
Join our Telegram