refine AutoScheme readme/code #958

wenhuach21 · 2025-10-29T06:01:12Z

No description provided.

for more information, see https://pre-commit.ci

…o refine_1129

Copilot

Pull Request Overview

This pull request refactors the AutoScheme functionality by moving it to a dedicated module structure and updating documentation. The changes improve code organization by separating concerns and enhance the README with better examples and parameter documentation.

Moved AutoScheme class from schemes.py to a dedicated auto_scheme module with proper registration system
Updated import statements across multiple files to reflect the new module structure
Enhanced README documentation with AutoScheme usage examples and reorganized hyperparameter descriptions

Reviewed Changes

Copilot reviewed 9 out of 10 changed files in this pull request and generated 2 comments.

Show a summary per file

File	Description
docs/step_by_step.md	Added GGUF quantization scheme documentation and removed redundant torch compile recommendation
auto_round/schemes.py	Removed AutoScheme class and related imports, updated all exports
auto_round/compressors/base.py	Updated imports and moved parameters to function signature
auto_round/autoround.py	Updated AutoScheme import to use new module location
auto_round/auto_scheme/register.py	New registration system for auto scheme methods
auto_round/auto_scheme/gen_auto_scheme.py	New home for AutoScheme class definition
auto_round/auto_scheme/init.py	Module initialization with simplified imports
auto_round/init.py	Updated AutoScheme import path
README.md	Enhanced documentation with AutoScheme examples and reorganized parameter descriptions

💡 Add Copilot custom instructions for smarter, more guided reviews. Learn how to get started.

auto_round/auto_scheme/gen_auto_scheme.py

auto_round/compressors/base.py

for more information, see https://pre-commit.ci

Signed-off-by: n1ck-guo <[email protected]>

…o refine_1129

for more information, see https://pre-commit.ci

…o refine_1129

for more information, see https://pre-commit.ci

* Fix rtn tuning_device issue (#893) Signed-off-by: Kaihui-intel <[email protected]> * fix vlm gguf ut (#895) Signed-off-by: n1ck-guo <[email protected]> * update alg_ext.abi3.so with python compatible version (#894) * move ste from quant to round for nvfp4 (#889) Signed-off-by: He, Xin3 <[email protected]> * Add GPT-OSS quant support (#887) * better help printing information (#883) * better help printing information Signed-off-by: n1ck-guo <[email protected]> * speedup quant and evaluation, fix recompile issue (#897) * rewrite the implementation for ease-of-maintain Signed-off-by: He, Xin3 <[email protected]> * fix bug Signed-off-by: He, Xin3 <[email protected]> * fix quant performance Signed-off-by: He, Xin3 <[email protected]> * Update auto_round/compressors/base.py --------- Signed-off-by: He, Xin3 <[email protected]> * fix nvfp act quantization bug (#891) * fix nvfp act quantization bug Signed-off-by: Zhang, Weiwei1 <[email protected]> * add cuda ut for moe nvfp quantize Signed-off-by: Zhang, Weiwei1 <[email protected]> * add cpu UT, refine cuda UT Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * fix ut typo Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * fix cpu ut Signed-off-by: Zhang, Weiwei1 <[email protected]> * enhance experts amax match, refine UT Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci --------- Signed-off-by: Zhang, Weiwei1 <[email protected]> Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com> * support automatic mixed bits assignment (#851) * try to fix gguf issue (#886) * remove numba from requirments (#905) Signed-off-by: yiliu30 <[email protected]> * Extend mxfp loading dtypes (#907) * block dataset logger info (#908) Signed-off-by: n1ck-guo <[email protected]> * fix torch compile issue in AutoScheme (#909) * Revert "Extend mxfp loading dtypes (#907)" (#915) This reverts commit 0c2619c. * support disable_opt_rtn in auto-scheme (#913) * fix llama 4 ut (#896) * fix ut of llama 4 Signed-off-by: n1ck-guo <[email protected]> * add numba for cpu lib (#919) Signed-off-by: yiliu30 <[email protected]> * Loosen the packing restrictions for mxfp&nvfp (#911) * Loosen the packing restrictions for mxfp&nvfp, enable Qwen1.5-MoE-A2.7B quantize Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * fix UT Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * refine mxfp&nvfp layer checker Signed-off-by: Zhang, Weiwei1 <[email protected]> * fix pylint Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci --------- Signed-off-by: Zhang, Weiwei1 <[email protected]> Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com> * Extend mxfp loading dtypes (#916) Signed-off-by: root <[email protected]> Signed-off-by: yiliu30 <[email protected]> Co-authored-by: root <[email protected]> Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com> * Fix act config exporting for mixed schemes (#903) * fp8 exporting bugfix Signed-off-by: Zhang, Weiwei1 <[email protected]> * fix act related config saving Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * add ut for act_config check Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * refine extra_config saving, add UTs Signed-off-by: Zhang, Weiwei1 <[email protected]> * fix ut typo Signed-off-by: Zhang, Weiwei1 <[email protected]> * fix ut typo Signed-off-by: Zhang, Weiwei1 <[email protected]> * fixtypo Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * fix CI Signed-off-by: Zhang, Weiwei1 <[email protected]> * fix scan issue Signed-off-by: Zhang, Weiwei1 <[email protected]> * fix scan issue Signed-off-by: Zhang, Weiwei1 <[email protected]> * rm global variable Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * rerun ut Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * refine ut Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci --------- Signed-off-by: Zhang, Weiwei1 <[email protected]> Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com> * optimize rtn for int woq (#924) * fix bug of gguf and support for LiquidAI/LFM2-1.2B (#927) Signed-off-by: n1ck-guo <[email protected]> * remove numpy<2.0 limitation (#921) * enable regex quantization config saving for mixed bits (#825) * enable dynamic quantization config saving Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * fixtypo Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * rebase code, refine config saving Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * refine ut Signed-off-by: Zhang, Weiwei1 <[email protected]> * fix UT Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * enable hf loading for regex, add UTs Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * refine export, enhance gptq UT Signed-off-by: Zhang, Weiwei1 <[email protected]> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci --------- Signed-off-by: Zhang, Weiwei1 <[email protected]> Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com> * Fix Flux tuning issue (#936) Signed-off-by: Mengni Wang <[email protected]> * gguf support for inclusionAI/Ling-flash-2.0 (#940) * remove low_cpu_mem (#934) * Add compatibility test (#918) * Add commit hash to version (#941) Signed-off-by: Sun, Xuehao <[email protected]> * gguf weight type align with original, output.weight, token_embed (#900) * support attention mask in user's dataset (#930) * Add diffusion README (#923) * update readme (#949) * refactor utils file (#943) * refact utils Signed-off-by: n1ck-guo <[email protected]> * update readme for sglang support (#953) * update readme for sglang support Signed-off-by: Zhang, Weiwei1 <[email protected]> * refine doc Signed-off-by: Zhang, Weiwei1 <[email protected]> * Update README.md --------- Signed-off-by: Zhang, Weiwei1 <[email protected]> Co-authored-by: Wenhua Cheng <[email protected]> * update gguf and support for CompressedLinear (#950) * Reduce AutoSchem VRAM usage by up to 10X (#944) * add self attribution and fix avg_bits error (#956) * add self attribution and fix avg_bits error --------- Signed-off-by: He, Xin3 <[email protected]> Co-authored-by: Wenhua Cheng <[email protected]> * add logo (#960) * refine AutoScheme readme/code (#958) * update readme (#962) * fix critic disable_opt_rtn regression (#963) * [1/N] Initial vllm-ext evaluation support (MXFP4 MOE) (#935) Signed-off-by: yiliu30 <[email protected]> * fix bug of imatrix contains 0 (#955) * fix rtn bug (#966) * enhance flux doc (#967) * clean code (#968) * support for model scope (#957) * support for model scope Signed-off-by: n1ck-guo <[email protected]> * merge main branch to alg_ext (#970) * fix cuda CI backend issue, fixtypo (#974) * disable compile packing by default (#975) Signed-off-by: yiliu30 <[email protected]> * enhance auto device map and support XPU (#961) * enhance auto device map and support XPU --------- Signed-off-by: He, Xin3 <[email protected]> * refine readme (#978) * cli support for positional arguments model (#979) Signed-off-by: n1ck-guo <[email protected]> * update bits (#986) Signed-off-by: He, Xin3 <[email protected]> * fix guff scheme and device_map bug (#969) * add support for Magistral-Small (#980) * support model_dtype and fix bug of scheme contains quotes, mllm eval (#985) --------- Signed-off-by: Kaihui-intel <[email protected]> Signed-off-by: n1ck-guo <[email protected]> Signed-off-by: He, Xin3 <[email protected]> Signed-off-by: Zhang, Weiwei1 <[email protected]> Signed-off-by: yiliu30 <[email protected]> Signed-off-by: root <[email protected]> Signed-off-by: Mengni Wang <[email protected]> Signed-off-by: Sun, Xuehao <[email protected]> Co-authored-by: Tang Kaihui <[email protected]> Co-authored-by: Heng Guo <[email protected]> Co-authored-by: Xin He <[email protected]> Co-authored-by: Yi Liu <[email protected]> Co-authored-by: Weiwei <[email protected]> Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com> Co-authored-by: Wenhua Cheng <[email protected]> Co-authored-by: root <[email protected]> Co-authored-by: Wang, Mengni <[email protected]> Co-authored-by: Sun, Xuehao <[email protected]>

wenhuach21 and others added 7 commits October 29, 2025 13:56

refine

84b0545

Merge branch 'main' into refine_1129

9c660c8

mv AutoScheme class

6460958

[pre-commit.ci] auto fixes from pre-commit.com hooks

302bb18

for more information, see https://pre-commit.ci

Add autoscheme usage in homepage

e0b7301

update

6ddec8a

Merge branch 'refine_1129' of https://github.com/intel/auto-round int…

92d4350

…o refine_1129

wenhuach21 changed the title ~~refine code~~ refine AutoScheme readme/code Oct 29, 2025

fix preci

ab882cf

wenhuach21 requested review from WeiweiZhang1, Copilot, n1ck-guo and yiliu30 October 29, 2025 06:51

Copilot AI reviewed Oct 29, 2025

View reviewed changes

auto_round/auto_scheme/gen_auto_scheme.py Show resolved Hide resolved

auto_round/compressors/base.py Show resolved Hide resolved

wenhuach21 and others added 15 commits October 29, 2025 14:56

update

3206b2d

update

1c31fcc

fix post_init

3e43b48

[pre-commit.ci] auto fixes from pre-commit.com hooks

4c96eac

for more information, see https://pre-commit.ci

fix import error

a96737b

Signed-off-by: n1ck-guo <[email protected]>

fix

3fe7e65

Merge branch 'refine_1129' of https://github.com/intel/auto-round int…

e9b893b

…o refine_1129

Merge branch 'main' into refine_1129

b687410

[pre-commit.ci] auto fixes from pre-commit.com hooks

5e24487

for more information, see https://pre-commit.ci

try to fix import issue in another way

b545e18

Merge branch 'refine_1129' of https://github.com/intel/auto-round int…

d4bd21e

…o refine_1129

try to fix preci issue

4487fc2

[pre-commit.ci] auto fixes from pre-commit.com hooks

7158730

for more information, see https://pre-commit.ci

try to fix preci issue

6b6246f

[pre-commit.ci] auto fixes from pre-commit.com hooks

9a75aae

for more information, see https://pre-commit.ci

yiliu30 approved these changes Oct 29, 2025

View reviewed changes

update readme

b9d5824

wenhuach21 merged commit 8ac82a4 into main Oct 29, 2025
16 of 22 checks passed

wenhuach21 deleted the refine_1129 branch October 29, 2025 14:00

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

refine AutoScheme readme/code #958

refine AutoScheme readme/code #958

Uh oh!

wenhuach21 commented Oct 29, 2025

Uh oh!

Copilot AI left a comment

Uh oh!

Uh oh!

Uh oh!

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

4 participants

refine AutoScheme readme/code #958

refine AutoScheme readme/code #958

Uh oh!

Conversation

wenhuach21 commented Oct 29, 2025

Uh oh!

Copilot AI left a comment

Choose a reason for hiding this comment

Pull Request Overview

Reviewed Changes

Uh oh!

Uh oh!

Uh oh!

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

4 participants