transformers
https://github.com/huggingface/transformers
Python
π€ Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
Python not yet supported55 Subscribers
View all SubscribersAdd a CodeTriage badge to transformers
Help out
- Issues
- Add MOSS-Transcribe-Diarize model
- qwen4_exp: fp8-quantized n-gram (PLE) embedding rows are gathered without dequantization
- Shape cross-attention keys and values by the states they come from
- Modernize FSMT: shared attention interface and remove legacy fairseq code
- Standardise head_dim onto the config and drop the guards that stood in for it
- Add --experts-implementation flag to transformers serve
- Unpin the bert copy markers, and delete the ones CI never enforced
- ιι ιΈΏθ
- NVFP4 MoE experts are dequantized without weight_global_scale, silently
- Carry weight_global_scale when decompressing NVFP4 MoE experts
- Docs
- Python not yet supported