llama.cpp
quantize: cap working memory size to avoid loading big tensors onto RAM
#27795
Merged
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Overview
Commits
1
Changes
View On
GitHub
quantize: cap working memory size to avoid loading big tensors onto RAM
#27795
ngxson
merged 1 commit into
master
from
xsn/quant_cap
quantize: cap working memory size to avoid loading big tensors onto RAM
c0c7fa93
ngxson
requested a review
from
ggerganov
6 days ago
github-actions
added
examples
ngxson
requested a review
from
ServeurpersoCom
6 days ago
ServeurpersoCom
approved these changes on 2026-08-27
ngxson
merged
732707df
into master
6 days ago
Login to write a write a comment.
Login via GitHub
Reviewers
ServeurpersoCom
ggerganov
Assignees
No one assigned
Labels
examples
Milestone
No milestone
Login to write a write a comment.
Login via GitHub