huggingface/pytorch-image-models

The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT,...

View on GitHub ↗Jump to charts ↓Open shareable report →

Data as of . Signed-in members get hourly updates — create a free account.

Summary Information

Updated 2 hours ago
Added to GitGenius on December 4th, 2024
Created on February 2nd, 2019
Open Issues & Pull Requests: 36 (+0)
GitHub issues: Enabled
Number of forks: 5,214
Total Stargazers: 37,204 (+0)
Total Subscribers: 321 (+0)

Repository Insights (GitGenius)

Median issue/PR response: 3.9 hours
Mean response time: 175.8 days
90th percentile: 717.3 days
Tracked items: 255

Maintainer activity

2 people did triage or write work on this repository in the last 12 months.

Counts unlabeled, assigned, unassigned, milestoned, demilestoned, locked, unlocked over the last 12 months. These are issue and pull request events that require triage or write permission. Commits and code review are not counted. labeled and renamed are excluded because GitHub issue forms record the issue author as the actor. Figures from October 7, 2026. This count is not comparable across projects: each project's automation decides which of these events a person emits.

How this project is maintained

Practically every issue opened in the past year has drawn a reply. 93% of issues opened in the past year have since been closed. Three people close 91% of everything that gets resolved.

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Open issues: 19
New in 7 days: 0
Closed in 7 days: 1
Avg open age: 910 days
Stale 30+ days: 18
Stale 90+ days: 17

Recent activity

Opened in 7 days: 0
Closed in 7 days: 1
Comments in 7 days: 0
Events in 7 days: 1

Top labels

  • bug (405)
  • enhancement (349)
  • help wanted (22)
  • good first issue (1)

Most active issues this week

Sign in to see which issues are moving.
Sign in

Detailed Description

The repository serves as a comprehensive model zoo encompassing a wide range of architectures including ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer variants, MobileNetV4, MobileNet-V3 and V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, and ConvNeXt. Beyond model definitions, the repository provides complete training, evaluation, inference, and export scripts alongside pretrained weights, making it a practical resource for practitioners implementing computer vision tasks.

The codebase is classified across multiple computer vision domains including semantic segmentation, object detection, image classification, and feature extraction, reflecting its broad applicability to diverse vision tasks.

Recent development activity shows intensive focus on Vision Transformer variants and emerging architectures. As of May 2026, the repository added EUPE ViT models with DINOv3-style training and ConvNeXt variants, along with TIPSv2 model definitions for DINOv2-style Vision Transformers. Earlier updates in 2025 introduced DINOv3 support for both ConvNeXt and ViT models, MobileCLIP-2 vision encoders, MetaCLIP-2 Worldwide ViT weights, and SigLIP-2 NaFlex ViT encoders. The repository also integrated support for Naver ROPE-ViT models and added MobileNetV5 backbone variants designed for Google Gemma 3n image encoding.

The codebase maintains active optimization efforts across multiple fronts. Recent releases introduced the Muon optimizer with customizations for convolutional weights and fallback mechanisms, alongside improvements to AdaMuon and NAdaMuon variants. Security enhancements include improved pickle checkpoint handling with weights_only=True as default and safe_global support for argument parsing. The repository added device and dtype factory keyword argument support across all models and modules, enabling flexible initialization strategies including meta-device model creation.

Benchmark coverage has expanded significantly, with new inference timing results added for RTX Pro 6000, 5090, and 4090 graphics cards using PyTorch 2.9.1. The repository maintains compatibility across PyTorch versions from 1.13 through 2.9.1 and Python versions from 3.10 through 3.13. Recent architectural additions include differential attention mechanisms, pooling modules like LsePlus and SimPool, and various normalization variants including Fp32 LayerNorm and RMSNorm options.