mark-arts commited on
Commit
ce807ed
·
verified ·
1 Parent(s): abb4ce3

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +38 -0
README.md ADDED
@@ -0,0 +1,38 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ license_link: https://huggingface.co/Qwen/QWQ-32B/blob/main/LICENSE
4
+ language:
5
+ - en
6
+ pipeline_tag: text-generation
7
+ base_model: Qwen/QwQ-32B
8
+ tags:
9
+ - chat
10
+ - mlx
11
+ - mlx-my-repo
12
+ ---
13
+
14
+ # mark-arts/QwQ-32B-Q4-mlx
15
+
16
+ The Model [mark-arts/QwQ-32B-Q4-mlx](https://huggingface.co/mark-arts/QwQ-32B-Q4-mlx) was converted to MLX format from [Qwen/QwQ-32B](https://huggingface.co/Qwen/QwQ-32B) using mlx-lm version **0.21.5**.
17
+
18
+ ## Use with mlx
19
+
20
+ ```bash
21
+ pip install mlx-lm
22
+ ```
23
+
24
+ ```python
25
+ from mlx_lm import load, generate
26
+
27
+ model, tokenizer = load("mark-arts/QwQ-32B-Q4-mlx")
28
+
29
+ prompt="hello"
30
+
31
+ if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
32
+ messages = [{"role": "user", "content": prompt}]
33
+ prompt = tokenizer.apply_chat_template(
34
+ messages, tokenize=False, add_generation_prompt=True
35
+ )
36
+
37
+ response = generate(model, tokenizer, prompt=prompt, verbose=True)
38
+ ```