transformers
https://github.com/huggingface/transformers
Python
🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
Python not yet supported55 Subscribers
View all SubscribersAdd a CodeTriage badge to transformers
Help out
- Issues
- QA: mtp models are gone from hub
- fix(zamba): add use_associative_scan config flag to avoid torch.compile slowdown
- Use the shared attention interface in encoder-decoder attention blocks
- Native TP does not shard qwen3_vl_moe experts (weights fully replicated per rank)
- [Umbrella] Continuous batching fixes for training-while-generating on one weight copy
- Continuous batching CUDA-graph capture crashes under distributed runs (allocator internal assert)
- Fix Moshi output fields
- [`Kernels`] Add swiglu and geglu MLP across models
- CLI: transformers env fails with ModuleNotFoundError: No module named 'requests' when run in clean environment
- Parakeet TDT generate shares request state across concurrent calls
- Docs
- Python not yet supported