transformers/__init__.py at d3d835d4fc145e5062d2153ac23ccd4b3e2c2cbd

mirror of https://github.com/huggingface/transformers.git synced 2025-07-03 21:00:08 +06:00

Mohamed Mekkouri efe72fe21f

Adding FP8 Quantization to transformers (#36026 )

* first commit

* adding kernels

* fix create_quantized_param

* fix quantization logic

* end2end

* fix style

* fix imports

* fix consistency

* update

* fix style

* update

* udpate after review

* make style

* update

* update

* fix

* update

* fix docstring

* update

* update after review

* update

* fix scheme

* update

* update

* fix

* update

* fix docstring

* add source

* fix test

---------

Co-authored-by: Marc Sun <57196510+SunMarc@users.noreply.github.com>

2025-02-13 13:01:19 +01:00

0 lines Python Raw Blame History

0 lines

Python

Raw Blame History