vinhnx90 commited on
Commit
df6a1e6
·
verified ·
1 Parent(s): 27ff128

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +43 -0
README.md ADDED
@@ -0,0 +1,43 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ base_model: vinhnx90/gemma-3-1b-thinking-v2
3
+ tags:
4
+ - text-generation-inference
5
+ - transformers
6
+ - unsloth
7
+ - gemma3_text
8
+ - trl
9
+ - grpo
10
+ - mlx
11
+ - mlx-my-repo
12
+ license: apache-2.0
13
+ language:
14
+ - en
15
+ datasets:
16
+ - openai/gsm8k
17
+ ---
18
+
19
+ # vinhnx90/gemma-3-1b-thinking-v2-mlx-4Bit
20
+
21
+ The Model [vinhnx90/gemma-3-1b-thinking-v2-mlx-4Bit](https://huggingface.co/vinhnx90/gemma-3-1b-thinking-v2-mlx-4Bit) was converted to MLX format from [vinhnx90/gemma-3-1b-thinking-v2](https://huggingface.co/vinhnx90/gemma-3-1b-thinking-v2) using mlx-lm version **0.22.1**.
22
+
23
+ ## Use with mlx
24
+
25
+ ```bash
26
+ pip install mlx-lm
27
+ ```
28
+
29
+ ```python
30
+ from mlx_lm import load, generate
31
+
32
+ model, tokenizer = load("vinhnx90/gemma-3-1b-thinking-v2-mlx-4Bit")
33
+
34
+ prompt="hello"
35
+
36
+ if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
37
+ messages = [{"role": "user", "content": prompt}]
38
+ prompt = tokenizer.apply_chat_template(
39
+ messages, tokenize=False, add_generation_prompt=True
40
+ )
41
+
42
+ response = generate(model, tokenizer, prompt=prompt, verbose=True)
43
+ ```