A quantized version of Granite Guardian 3.1 2B model from IBM.

Quantization is done by llama.cpp.

Model Summary (from original repository)

Granite Guardian 3.1 2B is a fine-tuned Granite 3.1 2B Instruct model designed to detect risks in prompts and responses. It can help with risk detection along many key dimensions catalogued in the IBM AI Risk Atlas. It is trained on unique data comprising human annotations and synthetic data informed by internal red-teaming. It outperforms other open-source models in the same space on standard benchmarks.

Downloads last month
44
GGUF
Model size
2.53B params
Architecture
granite
Hardware compatibility
Log In to view the estimation

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for ktoprakucar/granite-guardian-3.1-2b-Q8-GGUF

Quantized
(4)
this model