QuantFactory/glm-4-9b-chat-abliterated-GGUF
This is quantized version of byroneverson/glm-4-9b-chat-abliterated created using llama.cpp
Original Model Card
GLM 4 9B Chat - Abliterated
Check out the jupyter notebook for details of how this model was abliterated from glm-4-9b-chat.
The python package "tiktoken" is required to quantize the model into gguf format. So I had to create a fork of GGUF My Repo (+tiktoken).
- Downloads last month
- 253
Inference Providers
NEW
This model is not currently available via any of the supported third-party Inference Providers, and
the model is not deployed on the HF Inference API.
Model tree for QuantFactory/glm-4-9b-chat-abliterated-GGUF
Base model
THUDM/glm-4-9b-chat