FazBrowse GitHub Viewer
|
Trending
|
URL:
|
Home
Tools:
[Download Repo ZIP]
[Original HTTPS Page]
History for tools/quantize/quantize.cpp - allozaur/llama.cpp · GitHub
allozaur
/
llama.cpp
Public
forked from
ggml-org/llama.cpp
Notifications
You must be signed in to change notification settings
Fork
0
Star
2
Code
Pull requests
0
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Pull requests
Actions
Projects
Security and quality
Insights
Commits
Breadcrumbs
History for
llama.cpp
tools
quantize
quantize.cpp
on
master
User selector
Datepicker
Commit history
Commits on Jul 7, 2026
Add Q2_0 quantization: type definition and CPU backend (#24448)
khosravipasha
authored
bec4772
View commit details
Copy full SHA for bec4772
View code at this point
Browse repository at this point
Commits on Jun 4, 2026
Move duplicated imatrix code into single common imatrix-loader.cpp (#22445)
Show description for e7bcf1c
bartowski1182
authored
e7bcf1c
View commit details
Copy full SHA for e7bcf1c
View code at this point
Browse repository at this point
Commits on May 21, 2026
app : add batched-bench, fit-params, quantize & perplexity (#23459)
Show description for 1d7ab2b
angt
authored
1d7ab2b
View commit details
Copy full SHA for 1d7ab2b
View code at this point
Browse repository at this point
Commits on Apr 17, 2026
libs : rename libcommon -> libllama-common (#21936)
Show description for 6990e2f
ggerganov
authored
6990e2f
View commit details
Copy full SHA for 6990e2f
View code at this point
Browse repository at this point
Commits on Apr 6, 2026
ggml: add Q1_0 1-bit quantization support (CPU) (#21273)
Show description for 2e1f0a8
khosravipasha
and
CISC
authored
2e1f0a8
View commit details
Copy full SHA for 2e1f0a8
View code at this point
Browse repository at this point
Commits on Apr 1, 2026
llama : refactor llama_model_quantize_params to expose a pure C interface (#20346)
Show description for 4951250
EAddario
and
ggerganov
authored
4951250
View commit details
Copy full SHA for 4951250
View code at this point
Browse repository at this point
Commits on Mar 10, 2026
llama-quant : fail early on missing imatrix, refactor type selection, code cleanup (#19770)
Show description for 1dab5f5
ddh0
authored
1dab5f5
View commit details
Copy full SHA for 1dab5f5
View code at this point
Browse repository at this point
Commits on Mar 4, 2026
Fix locale-dependent float printing in GGUF metadata (#17331)
Show description for cb8f4fa
ssam18
and
JohannesGaessler
authored
cb8f4fa
View commit details
Copy full SHA for cb8f4fa
View code at this point
Browse repository at this point
Commits on Feb 20, 2026
quantize : add --dry-run option (#19526)
Show description for 492bc31
ddh0
and
CISC
authored
492bc31
View commit details
Copy full SHA for 492bc31
View code at this point
Browse repository at this point
Commits on Feb 8, 2026
llama-quantize : cleanup `--help` output (#19317)
Show description for 5999b50
ddh0
authored
5999b50
View commit details
Copy full SHA for 5999b50
View code at this point
Browse repository at this point
Commits on Jan 31, 2026
quantize: add option --tensor-type-file to llama-quantize (#18572)
Show description for 3dd9591
authored
3dd9591
View commit details
Copy full SHA for 3dd9591
View code at this point
Browse repository at this point
Commits on Dec 31, 2025
quantize: prevent input/output file collision (#18451)
Show description for 33ded98
Anri-Lombard
authored
33ded98
View commit details
Copy full SHA for 33ded98
View code at this point
Browse repository at this point
Commits on Aug 5, 2025
llama : add gpt-oss (#15091)
Show description for fd1234c
authored
fd1234c
View commit details
Copy full SHA for fd1234c
View code at this point
Browse repository at this point
Commits on Aug 4, 2025
quantize : fix confusing error message if ftype is invalid (#15071)
CISC
authored
2721257
View commit details
Copy full SHA for 2721257
View code at this point
Browse repository at this point
Commits on Jul 30, 2025
quantize : fix using combined imatrix GGUFs (multiple datasets) (#14973)
EAddario
authored
e9192be
View commit details
Copy full SHA for e9192be
View code at this point
Browse repository at this point
Commits on Jul 19, 2025
imatrix : use GGUF to store importance matrices (#9400)
Show description for 9008328
authored
9008328
View commit details
Copy full SHA for 9008328
View code at this point
Browse repository at this point
Commits on Jun 22, 2025
quantize : handle user-defined pruning of whole layers (blocks) (#13037)
EAddario
authored
fa4a9f2
View commit details
Copy full SHA for fa4a9f2
View code at this point
Browse repository at this point
Commits on May 13, 2025
quantize : improve tensor-type pattern matching (#13033)
EAddario
authored
e5c834f
View commit details
Copy full SHA for e5c834f
View code at this point
Browse repository at this point
Commits on May 2, 2025
llama : move end-user examples to tools directory (#13249)
Show description for 1d36b36
slaren
and
ngxson
authored
1d36b36
View commit details
Copy full SHA for 1d36b36
View code at this point
Browse repository at this point
Loading
Back
|
FazBrowse Home
|
New Git URL