huangang commited on
Commit
96e35a0
·
verified ·
1 Parent(s): 78ce920

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +38 -0
README.md ADDED
@@ -0,0 +1,38 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: other
3
+ license_name: katanemo-research
4
+ license_link: https://huggingface.co/katanemo/Arch-Agent-32B/blob/main/LICENSE
5
+ base_model: katanemo/Arch-Agent-32B
6
+ language:
7
+ - en
8
+ pipeline_tag: text-generation
9
+ library_name: transformers
10
+ tags:
11
+ - mlx
12
+ ---
13
+
14
+ # huangang/Arch-Agent-32B-mlx-4Bit
15
+
16
+ The Model [huangang/Arch-Agent-32B-mlx-4Bit](https://huggingface.co/huangang/Arch-Agent-32B-mlx-4Bit) was converted to MLX format from [katanemo/Arch-Agent-32B](https://huggingface.co/katanemo/Arch-Agent-32B) using mlx-lm version **0.22.3**.
17
+
18
+ ## Use with mlx
19
+
20
+ ```bash
21
+ pip install mlx-lm
22
+ ```
23
+
24
+ ```python
25
+ from mlx_lm import load, generate
26
+
27
+ model, tokenizer = load("huangang/Arch-Agent-32B-mlx-4Bit")
28
+
29
+ prompt="hello"
30
+
31
+ if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
32
+ messages = [{"role": "user", "content": prompt}]
33
+ prompt = tokenizer.apply_chat_template(
34
+ messages, tokenize=False, add_generation_prompt=True
35
+ )
36
+
37
+ response = generate(model, tokenizer, prompt=prompt, verbose=True)
38
+ ```