Go
Home
Pricing
FAQ
Install
Home
Pricing
FAQ
Install
Login
via GitHub
intel/auto-round
Pull Requests
Commits
suyue/ai2civ3
AutoAdamRound_bugfix
Chinesization
ZaneMark-patch-1
ZaneMark-patch-3
acp
actvation_quant
add_lag_solver
add_task_args_for_lmeval
agent/fp8-per-head-kv-attn-merge
ar_agent
ark_v0.13.4
ark_zp
autoround_support_qbits_backend
bf16_scale
chore/claude-init
copilot/adapt-to-v5-chat-template
copilot/add-bandwidth-metrics
copilot/convert-script-improvements
copilot/copilotoptimize-int4-moe-performance
copilot/fix-2206-onednn-stream-compatibility
copilot/fix-corner-case-in-auto-round
copilot/fix-deprecated-fp-layers-handling
copilot/fix-issue-with-auto-rounding
copilot/fix-llm-type-70b-bits-setting
copilot/fix-occasional-test-failure
copilot/fix-typeerror-wrapped-fn
copilot/fix-vllm-model-inference-issue
copilot/improve-pr-template-type-of-change
copilot/investigate-quantization-group-and-ffn
copilot/optimize-xpu-moe-bf16-fp16
copilot/replace-getset-module-torch-api
copilot/sageattention
copilot/speedup-fp8-linear-convert
copilot/speedup-fp8-linear-convert-again
copilot/speedup-fp8-linear-convert-another-one
copilot/sub-pr-1237-again
copilot/sub-pr-1237
copilot/sub-pr-1324
copilot/sub-pr-1522-again
copilot/sub-pr-1532
copilot/update-phase-auto-dispatch-logic
copilot/update-user-settings-page
copilot/vscode-mo3shmf8-8qa6
cosmos
ddp
debug_time_cost
debug/usable_rotation
debug-hang
debug-nvfp4
deepseekv3
ds-qwen
ds-v5
ds-v32
dsv4
enable_glm4_moe_lite_quantization
enable_llama4_int8_baseline
enable_llama4_quant
enable_mxfp_exporting
feat/activation-checkpointing
feat/ark-xpu-int3-woq-gemm
feat/ark-xpu-int3-woq-gemv
feat/auto-code-calibration-dataset-v2
feat/autoround-quarot
feat/cosmos3-nano-quant
feat/fp8-per-head-kv-attn
feat/rrq-phase1
feat/sage-int4
feat/sparse-bf16-prefill-v2-bk-818
feat/sparse-bf16-prefill-v2
feature/overlap_for_nblocks
fix_bug0627
fix_bug_0722
fix_bug_1105
fix_compile
fix_disable_act_dynamic_usage_in_mxfp.py
fix_dq
fix/fp-layers-deprecation-mapping
fix_gemma3_issue
fix_gguf_fp8
fix_gptqmodel
fix_int4_acc
fix/llm-compressor-mixed-targets
fix/llm-compressor-targets-main
fix/llm-compressor-targets-only
fix/llmc-ut-skips
fix_low_cpu
fix_rotation
fix_save_quantized_func_nvfp_checker
fix_0107
fix_0109
fix_0113
fix-attn-mask-b60
fix-ds
fix-flashinfer
fix-gpt-oss
fix-hpu
fix-to-meta-assertion-error-1499
fixbug_0717
fp4_v2
fp4_v3
fp8-cache
fp8-cache-based-export
fp8-static-quant-patch
fp8_export_backup_stable
fp8_export_for_test
good-flux
hengguo/fix_cuda_ut
hengguo/fix_gguf_ds
hengguo/fix_offload
hengguo/quantizers
hengguo/reduce_cuda_ut
hengguo/refactor_algs
hengguo/refactor_init
hengguo/refactor_quant_step1
hengguo/smoothquant
hengguo/w4afp8_sim
henguo/update_so
hpu_only_kg
hpu_only_pkg
hpu/only/v1
hpu-limit-tran
improve_doc
kaihui/torch_dtype
lazy-model-replace
leq_opub
lib/pre-4.4.0
llama/new/9-610
llama/new/9
llm-main
llmc
llmc-backup
llmc-test
lm-head-quant
load-kv
load-w8a8-replace-mod
load-w8a8
lvl/autoscheme_ram_opt
lvl/cpu_ram_optimization
lvl/fix_gguf_convert_issue
lvl/fix_no_init_weights
lvl/fix_perf_moe
lvl/general_moe_replacement
lvl/offload_cleanup
lvl/sa_fp8
lvl/support_diffusiongemma
lvl/support_fp8_with_ark
lvl/support_hunyuan_image
lvl/support_turbo_quant
lyt/numpy_fix
lyt/omni
main
marlin_modify
mengni/bug_fix
mengni/expert
mengni/mengni/block_wise
mengni/vlm
mengniwang95-patch-1
minimax-h3-w4a16-clean
more-ar-ext
mxfp8
opt_moe_kernel
origin/block_wise
patch/for/ao/581/stable
patch-for-ao-2
pr1775-followup
pre-release/internal-inc/w4a8
quant-attn-hpu
quant-attn-hpu-o-scale
quant-attn-hpu-pr
quant-llama
quarot-llama
qwen3-vl
qwen3_vl_moe
qwen-split
qwen-v5
refine_calib
refine_device_tmp
refine_device
refine_quantizer2
refine-doc-table
replace-lm-head
revert_order
revert-318-fix/hpu/check
revert-1231-set_disable_opt_rtn_default_2_none
revert-1562-suyue/ut
revert-2113-suyue/xpu
sage_int4
save_memory
scheme_awq_bk
set_disable_opt_rtn_default_2_none
sparse-attn
sparse-attn-clean
sparse-attn-prefill-clean
sparse-attn-v0
static_quant
suport_fuse_moe
support_gemma4
suyue/ai2civ3
suyue/ai4ci
suyue/ci
suyue/ci-bak
suyue/hpu
suyue/xpu
teq_algo
test-git
try_new_optimizer
try_to_fix_hadamard_regression
update_fp_compile
update_0522
update_0819
upstream-ao
use-ep
ut-time
v0.7.0rc
v0.7.1rc
v0.8.0rc
v0.8.0rc2
v0.9.1rc
v0.9.2-release
v0.9.2rc
v0.9.3rc
v0.9.4rc
v0.9.5rc
v0.9.6rc
v0.9.7rc
v0.10.0rc
v0.10.1rc
v0.10.2rc
v0.10.3rc
v0.12.0rc
v0.12.1rc
v0.12.2rc
v0.12.3rc
v0.13.0rc
v0.13.1rc
v0.14.0rc
v0.14.1rc
v0.14.2rc
v0.15.0rc
v0.15.1rc
vllm-sharing-deck-2026
w4a4_int_quaro
w4int8dynamic
wangchang/agent
wangchang/docs
wangchang/fix_oom
wangchang/vllm_loading
wangchang/vllm
wangchang/vllmmixin
wfp8-afp8-bk
xinhe/8-19
xinhe/8-26a
xinhe/8-26b
xinhe/8-27-mxfp
xinhe/8-31-mxfp
xinhe/9-2-mx_fp_even
xinhe/9-8
yi/random-init-diffusion-handoff
zhenzhong/toolkit_release
zhenzhong/w4a8_debug
zhenzhong/woqgemm_s8_update
fix report format
chensuyue
committed
7 minutes ago
5bf64d2b
[pre-commit.ci] auto fixes from pre-commit.com hooks
pre-commit-ci[bot]
committed
3 hours ago
de19189d
optimize report format & update AI model
chensuyue
committed
3 hours ago
eeaa78a8
optimize format
chensuyue
committed
7 hours ago
e9d9679e
for test only, large mount test failed
chensuyue
committed
8 hours ago
c5578220
for test only
chensuyue
committed
10 hours ago
9c92090c
revert test code
chensuyue
committed
10 hours ago
63378e4d
minor update
chensuyue
committed
11 hours ago
0c20d226
optimize
chensuyue
committed
11 hours ago
2c24d933
bug fix
chensuyue
committed
22 hours ago
466a5848
feat: implement reusable AI failure analysis stage and enhance comment handling
chensuyue
committed
1 day ago
730d938c
refactor: update AI analysis scripts to use PR merge commit SHA instead of diff file
chensuyue
committed
1 day ago
8c602fc9
fix report format
chensuyue
committed
1 day ago
6d0c8df9
[pre-commit.ci] auto fixes from pre-commit.com hooks
pre-commit-ci[bot]
committed
1 day ago
0549c459
optimize known issue match
chensuyue
committed
1 day ago
d8020d04
[pre-commit.ci] auto fixes from pre-commit.com hooks
pre-commit-ci[bot]
committed
1 day ago
0807312c
feat: enhance AI analysis with per-workflow comment markers and update pipeline template
chensuyue
committed
1 day ago
ad6b876f
report format update
chensuyue
committed
1 day ago
dcd1604b
feat: enhance AI analysis scripts with improved logging and report formatting
chensuyue
committed
1 day ago
0a387ceb
for test only, need to revert before merge
sys-lpot-val
committed
1 day ago
f509dc80
Revert "For test only, need to revert before merge."
sys-lpot-val
committed
1 day ago
f2f0d581
feat: add JSON parsing for Copilot output and improve AI analysis script
chensuyue
committed
1 day ago
1399d07d
Merge branch 'main' into suyue/ai2civ3
chensuyue
committed
3 days ago
Verified
b7c1e6d8
bug fix
chensuyue
committed
3 days ago
69bd4685
[risky]reduce ram usage of loading moe without monkey patch and fix qwen3-flash-next group_size issue (#2291)
wenhuach21
committed
3 days ago
Verified
7d8905ef
fix: update regex for test header and improve AI analysis script
chensuyue
committed
3 days ago
f67752c1
install copilot
chensuyue
committed
4 days ago
064c8d73
Optimize INT4 MoE prefill/decode (sym) XPU kernels with dedicated w4a16 tiling (#2111)
Copilot
committed
4 days ago
Verified
21bab556
For test only, need to revert before merge
chensuyue
committed
4 days ago
0be44231
[pre-commit.ci] auto fixes from pre-commit.com hooks
pre-commit-ci[bot]
committed
4 days ago
d941e032
Older