Valkyrie by AzumoSchedule a call
Vision trends · updated Oct 2, 2026

Which open-weight vision models get downloaded most

A daily read of Hugging Face's top 5,000 models, focused on vision models such as image classification, object detection and segmentation: the most downloaded, new uploads in the top 5,000, what they are used for, and who makes them.

Most downloaded

The most-downloaded open-weight vision models

Ranked by downloads over the last 30 days. Fine-tuned versions count; most format conversions and re-uploads are left out. Counts include automated fetches, which lift models that tools and test suites pull by default. How models are placed is explained under About this data.

#ModelMakerTaskLicenceLast 30 daysDownloadsValkyrie
1clip-vit-base-patch32openaiZero-shot image classification-21.4M
2mobilenetv3_small_100.lamb_in1ktimmImage classificationApache-2.021.3MFine-tune one like it →Fine-tune one →
3vit-base-patch16-224googleImage classificationApache-2.010.0MFine-tune one like it →Fine-tune one →
4clip-vit-large-patch14openaiZero-shot image classification-8.9M
5siglip2-base-patch16-256googleZero-shot image classificationApache-2.03.9MFine-tune one like it →Fine-tune one →
6CLIP-ViT-B-32-laion2B-s34B-b79KlaionZero-shot image classificationMIT3.8MFine-tune one like it →Fine-tune one →
7CLIP-ViT-L-14-laion2B-s32B-b82KlaionZero-shot image classificationMIT3.1MFine-tune one like it →Fine-tune one →
8dinov2-smallfacebookImage featuresApache-2.02.8MFine-tune one like it →Fine-tune one →
9dinov2-basefacebookImage featuresApache-2.02.8MFine-tune one like it →Fine-tune one →
10Depth-Anything-V2-Small-hfdepth-anythingDepth estimationApache-2.02.7MFine-tune one like it →Fine-tune one →
11fashion-clippatrickjohncyhZero-shot image classificationMIT2.2MFine-tune one like it →Fine-tune one →
12clip-vit-large-patch14-336openaiZero-shot image classification-2.2M
13nsfw_image_detectionFalconsaiImage classificationApache-2.02.2M
14sam3facebookSegmentationCustom licence2.2M
15siglip2-giant-opt-patch16-384googleZero-shot image classificationApache-2.02.1MFine-tune one like it →Fine-tune one →
16resnet18.a1_in1ktimmImage classificationApache-2.02.0MFine-tune one like it →Fine-tune one →
17owlv2-base-patch16-ensemblegoogleZero-shot object detectionApache-2.02.0MFine-tune one like it →Fine-tune one →
18siglip-base-patch16-224googleZero-shot image classificationApache-2.02.0MFine-tune one like it →Fine-tune one →
19GLM-OCRzai-orgImage to textMIT1.8MFine-tune one like it →Fine-tune one →
20tf_efficientnetv2_s.in21k_ft_in1ktimmImage classificationApache-2.01.8MFine-tune one like it →Fine-tune one →
21vit-base-patch16-224-in21kgoogleImage featuresApache-2.01.8MFine-tune one like it →Fine-tune one →
22blip-image-captioning-baseSalesforceImage to textBSD-3-Clause1.6MFine-tune one like it →Fine-tune one →
23siglip2-base-patch16-224googleZero-shot image classificationApache-2.01.6MFine-tune one like it →Fine-tune one →
24resnet50.a1_in1ktimmImage classificationApache-2.01.6MFine-tune one like it →Fine-tune one →
25grounding-dino-baseIDEA-ResearchZero-shot object detectionApache-2.01.5MFine-tune one like it →Fine-tune one →
New releases

New vision models in the top 5,000

Uploaded to Hugging Face in the last six months with at least 25 likes or a trending score of 10, ranked by downloads over the last 30 days. Most format conversions and re-uploads are left out.

tipsv2-b14googleZero-shot image classification · uploaded Apr 9, 2026172K downloads in 30 daysFine-tune one like it →Fine-tune one →
PP-OCRv6_medium_detPaddlePaddleImage to text · uploaded Jun 10, 202678K downloads in 30 daysFine-tune one like it →Fine-tune one →
PP-OCRv6_medium_recPaddlePaddleImage to text · uploaded Jun 10, 202666K downloads in 30 daysFine-tune one like it →Fine-tune one →
NuExtract3numindImage to text · uploaded Apr 29, 202634K downloads in 30 daysFine-tune one like it →Fine-tune one →
LeVJEPA-VideoMix-Largegalilai-groupImage features · uploaded Aug 27, 202629K downloads in 30 days
What they're used for

Downloads by task

Share of downloads of vision models in Hugging Face's top 5,000 on Oct 2, 2026, leaving out most format conversions and re-uploads.

Zero-shot image classification36.1%
Image classification32.2%
Image features10.0%
Image to text5.4%
Segmentation5.3%
Object detection3.8%
Depth estimation3.0%
Zero-shot object detection2.8%
Video classification0.7%
Keypoint detection0.6%
Who's winning

Share of downloads by maker

Each repo counts under the account that uploaded it, leaving out most format conversions and re-uploads. Change is in percentage points since Jul 4, 2026.

timm25.1% (−0.4 pts)
openai18.7% (−1.3 pts)
google16.5% (+7.3 pts)
facebook8.7% (+0.2 pts)
laion5.5% (+2.3 pts)
microsoft2.6% (−0.8 pts)
depth-anything2.4% (+0.6 pts)
PaddlePaddle1.9% (+0.3 pts)
Other18.6% (−8.2 pts)

Licences of the most-downloaded vision models

Apache-2.0 57.9%MIT 12.5%Other or none 29.6%

As declared on each repo, leaving out most format conversions and re-uploads.