|
--- |
|
tags: |
|
- uqff |
|
- mistral.rs |
|
base_model: google/gemma-3-270m-it |
|
base_model_relation: quantized |
|
--- |
|
|
|
<!-- Autogenerated from user input. --> |
|
|
|
# `google/gemma-3-270m-it`, UQFF quantization |
|
|
|
|
|
Run with [mistral.rs](https://github.com/EricLBuehler/mistral.rs). Documentation: [UQFF docs](https://github.com/EricLBuehler/mistral.rs/blob/master/docs/UQFF.md). |
|
|
|
1) **Flexible** π: Multiple quantization formats in *one* file format with *one* framework to run them all. |
|
2) **Reliable** π: Compatibility ensured with *embedded* and *checked* semantic versioning information from day 1. |
|
3) **Easy** π€: Download UQFF models *easily* and *quickly* from Hugging Face, or use a local file. |
|
3) **Customizable** π οΈ: Make and publish your own UQFF files in minutes. |
|
|
|
## Examples |
|
|Quantization type(s)|Example| |
|
|--|--| |
|
|AFQ4|`./mistralrs-server -i plain -m EricB/gemma-3-270m-it-UQFF -f gemma3-270m-it-afq4-0.uqff`| |
|
|AFQ6|`./mistralrs-server -i plain -m EricB/gemma-3-270m-it-UQFF -f gemma3-270m-it-afq6-0.uqff`| |
|
|AFQ8|`./mistralrs-server -i plain -m EricB/gemma-3-270m-it-UQFF -f gemma3-270m-it-afq8-0.uqff`| |
|
|Q4K|`./mistralrs-server -i plain -m EricB/gemma-3-270m-it-UQFF -f gemma3-270m-it-q4k-0.uqff`| |
|
|Q5K|`./mistralrs-server -i plain -m EricB/gemma-3-270m-it-UQFF -f gemma3-270m-it-q5k-0.uqff`| |
|
|Q8_0|`./mistralrs-server -i plain -m EricB/gemma-3-270m-it-UQFF -f gemma3-270m-it-q8_0-0.uqff`| |
|
|