Skip to content

Pull requests: bitsandbytes-foundation/bitsandbytes

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

[ROCm] Restore Wave64 warp size for all gfx9 targets
#2059 opened Aug 25, 2026 by 0xDELUXA Loading…
Add Apple MPS NF4 benchmark
#2057 opened Aug 25, 2026 by hamedrabah Loading…
Tune SM103 fused 4-bit GEMM SIMT dispatch
#2054 opened Aug 22, 2026 by heiheiha798 Loading…
Fuse nested 4-bit scale reconstruction on SM103
#2051 opened Aug 22, 2026 by heiheiha798 Loading…
Fix CPU dequantize_4bit output shape for 1-D inputs
#2048 opened Aug 20, 2026 by 2sumtech Loading…
ci: expand ROCm architecture coverage ROCm
#2046 opened Aug 20, 2026 by sstamenk Contributor Loading… v0.50.2
Fix mixed INT8 FakeTensor output metadata
#2044 opened Aug 19, 2026 by tandede Loading…
Fix decoupled weight decay ordering in the CUDA Adam/AdEMAMix kernels CUDA Issues and PRs related to the CUDA backend, excluding installation/support help. Optimizers Issues or feature requests relating to optimizers
#2040 opened Aug 14, 2026 by yentur Loading… v0.51.0
Fix int8 GEMM by using dedicated BLAS Lt handle on ROCm/CUDA ROCm
#2018 opened Jul 23, 2026 by zjin-lcf Contributor Loading…
2 of 3 tasks
Add Experts4bit for 4-bit quantization of fused MoE experts
#1965 opened Jun 5, 2026 by pjordanandrsn Contributor Loading…
refactor: Rewrite assert statements as exceptions
#1931 opened Apr 23, 2026 by rapsealk Contributor Loading…
5 of 7 tasks
docs: add an energy-efficiency FAQ entry for quantization Documentation Improvements or additions to documentation
#1882 opened Feb 24, 2026 by hongping-zh Loading…
ProTip! Type g p on any issue or pull request to go back to the pull request listing page.