facebookresearch/sapiens

High-resolution models for human tasks.

View on GitHub ↗Jump to charts ↓Open shareable report

Summary Information

Updated 2 minutes ago
Added to GitGenius on September 12th, 2026
Created on August 13th, 2024
Open Issues & Pull Requests: 19 (+0)
GitHub issues: Enabled
Number of forks: 321
Total Stargazers: 5,423 (+0)
Total Subscribers: 44 (+0)

Repository Insights (GitGenius)

Median issue/PR response: 5.2 hours
Mean response time: 3.7 days
90th percentile: 3.5 days
Tracked items: 160

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Open issues: 16
New in 7 days: 0
Closed in 7 days: 0
Avg open age: 401 days
Stale 30+ days: 16
Stale 90+ days: 16

Recent activity

Opened in 7 days: 0
Closed in 7 days: 0
Comments in 7 days: 0
Events in 7 days: 0

Top labels

No label distribution available yet.

Most active issues this week

No issue events were indexed in the last 7 days.

Detailed Description

Sapiens is a foundation model family for human-centric computer vision tasks.

The project addresses the need for high-quality, generalizable models across multiple human vision problems including 2D pose estimation, part segmentation, depth prediction, and surface normal estimation. Rather than building separate specialized models, Sapiens provides a unified pretrained foundation trained on a large corpus of in-the-wild human images that generalizes well to unconstrained conditions. The models are designed specifically for high-resolution feature extraction, trained natively at 1024 by 1024 image resolution with a 16-pixel patch size to preserve fine-grained spatial information.

Teams working on human-centric vision applications should consider Sapiens when they need to handle multiple related tasks or require strong performance on diverse, unconstrained imagery. The high-resolution training makes it particularly suitable for applications demanding detailed spatial understanding of human bodies and their properties. The foundation model approach means developers can adapt these pretrained representations to their specific tasks rather than training from scratch.

The project maintains active development with a newer version available, indicating ongoing refinement of the model family and approach.