Joseph717171
/

DeepHermes-3-Llama-3.2-3B-Preview-OQ8_0-F32.EQ8_0-F32.IQ4_K-Q8_0-GGUF

GGUF

imatrix

conversational

Model card Files Files and versions

xet

Community

DeepHermes-3-Llama-3.2-3B-Preview-OQ8_0-F32.EQ8_0-F32.IQ4_K-Q8_0-GGUF / README.md

Joseph717171

Create README.md

b54b5e6 verified about 1 month ago

preview code

raw

history blame

281 Bytes

Custom GGUF quants of NousResearch/DeepHermes-3-Llama-3-3B-Preview, where the Output Tensors are kept at F32 or quantized to Q8_0, and the Embeddings are kept at F32 or Quantized to Q8_0. Enjoy! 🧠🔥🚀