auto-round
fe5e5976
- Enhance model caching and warnings for NVFP4 E5M3 scheme in model_free.py and model.py; update documentation for llm_compressor support
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
23 days ago
Enhance model caching and warnings for NVFP4 E5M3 scheme in model_free.py and model.py; update documentation for llm_compressor support Signed-off-by: Xin He <xin3.he@intel.com>
References
#2123 - [Experimental] Add NVFP4 E5M3 support and related components for model-free quantization
Author
xin3he
Parents
dfe6c605
Loading