lostmsu
lostmsu
AI & ML interests
None yet
Organizations
None yet
mxfp4 QAT variant
#12 opened 2 months ago
by
lostmsu
Can thinking be toggled per response in a multi-turn conversation?
5
#2 opened 2 months ago
by
lostmsu
Was Q8_0 quantized from original model or is it a copy of -FP8 variant?
1
#3 opened 6 months ago
by
lostmsu
Is Q8_0 requantized from original model or simply a copy of FP8?
#4 opened 6 months ago
by
lostmsu
Is the Q8_0 derived from the official FP8?
1
#13 opened 12 months ago
by
lostmsu
Any benchmarks on public datasets?
1
#1 opened about 1 year ago
by
lostmsu
Meta-Llama-3.1-8B-Instruct-bf16.gguf is still old version?
5
#4 opened about 2 years ago
by
Firepin