github Actively maintained

lucidrains/vit-pytorch

Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch

1 awesome list

Quick read

Stars
25,457
Forks
3,494
Open issues
141
Commits
417

Activity and growth

Latest capture 2026-08-03 03:05

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+252
Stored snapshots
6

Classification

Metadata

Language
Python
License
MIT
Default branch
main
Created
2020-10-03
First commit
2020-10-03
Last pushed
2026-08-02
GitHub updated
2026-08-03
Last synced
2026-08-03 03:05
Stack scanned
2026-08-03 03:05
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

6 observed captures since 2026-05-25. Observed captures are shown by default.

Stars from first capture +252

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

huggingface/transformers.js

State-of-the-art Machine Learning for the web. Run 🤗 Transformers directly in your browser, with no need for a server!

16,224 stars
JavaScript 1 awesome list

NVlabs/VILA

VILA is a family of state-of-the-art vision language models (VLMs) for diverse multimodal AI tasks across the edge, data center, and cloud.

3,846 stars
Python 1 awesome list

FoundationVision/VAR

[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!

8,726 stars
Jupyter Notebook 1 awesome list

huggingface/transformers

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

163,267 stars
Python 6 awesome lists