DavidAU
/

LemonadeRP-4.5.3-11B-GGUF-Plus

Text Generation

Inference Endpoints

Model card Files Files and versions Community

Edit model card

GGUFs PLUS:

Q8 and Q6 GGUFs with critical parts of the model in F16 / Full precision.

File sizes will be slightly larger than standard, but should yeild higher quality results under all tasks and conditions.

Downloads last month: 39

GGUF

Model size

10.7B params

Architecture

llama

6-bit

8-bit

Inference Examples

Text Generation

Unable to determine this model's library. Check the docs .

Collection including DavidAU/LemonadeRP-4.5.3-11B-GGUF-Plus

D_AU - Higher Precision GGUFs / Imatrix Plus

Models compressed in higher precision with parts of the model compression remaining in F16/full precision. Increases overall quality in all tasks. • 35 items • Updated 7 days ago • 6