AkshayCoder48 shuhulx commited on
Commit
ce2fb03
0 Parent(s):

Duplicate from shuhulx/Qwopus3.5-4B-Coder-Fable5-v1-GGUF

Browse files

Co-authored-by: Shuhul Razdan <shuhulx@users.noreply.huggingface.co>

.gitattributes ADDED
@@ -0,0 +1,38 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ *.7z filter=lfs diff=lfs merge=lfs -text
2
+ *.arrow filter=lfs diff=lfs merge=lfs -text
3
+ *.bin filter=lfs diff=lfs merge=lfs -text
4
+ *.bz2 filter=lfs diff=lfs merge=lfs -text
5
+ *.ckpt filter=lfs diff=lfs merge=lfs -text
6
+ *.ftz filter=lfs diff=lfs merge=lfs -text
7
+ *.gz filter=lfs diff=lfs merge=lfs -text
8
+ *.h5 filter=lfs diff=lfs merge=lfs -text
9
+ *.joblib filter=lfs diff=lfs merge=lfs -text
10
+ *.lfs.* filter=lfs diff=lfs merge=lfs -text
11
+ *.mlmodel filter=lfs diff=lfs merge=lfs -text
12
+ *.model filter=lfs diff=lfs merge=lfs -text
13
+ *.msgpack filter=lfs diff=lfs merge=lfs -text
14
+ *.npy filter=lfs diff=lfs merge=lfs -text
15
+ *.npz filter=lfs diff=lfs merge=lfs -text
16
+ *.onnx filter=lfs diff=lfs merge=lfs -text
17
+ *.ot filter=lfs diff=lfs merge=lfs -text
18
+ *.parquet filter=lfs diff=lfs merge=lfs -text
19
+ *.pb filter=lfs diff=lfs merge=lfs -text
20
+ *.pickle filter=lfs diff=lfs merge=lfs -text
21
+ *.pkl filter=lfs diff=lfs merge=lfs -text
22
+ *.pt filter=lfs diff=lfs merge=lfs -text
23
+ *.pth filter=lfs diff=lfs merge=lfs -text
24
+ *.rar filter=lfs diff=lfs merge=lfs -text
25
+ *.safetensors filter=lfs diff=lfs merge=lfs -text
26
+ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
27
+ *.tar.* filter=lfs diff=lfs merge=lfs -text
28
+ *.tar filter=lfs diff=lfs merge=lfs -text
29
+ *.tflite filter=lfs diff=lfs merge=lfs -text
30
+ *.tgz filter=lfs diff=lfs merge=lfs -text
31
+ *.wasm filter=lfs diff=lfs merge=lfs -text
32
+ *.xz filter=lfs diff=lfs merge=lfs -text
33
+ *.zip filter=lfs diff=lfs merge=lfs -text
34
+ *.zst filter=lfs diff=lfs merge=lfs -text
35
+ *tfevents* filter=lfs diff=lfs merge=lfs -text
36
+ Qwopus3.5-4B-Coder-Fable5-v1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
37
+ Qwopus3.5-4B-Coder-Fable5-v1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
38
+ Qwopus3.5-4B-Coder-Fable5-v1-mmproj-BF16.gguf filter=lfs diff=lfs merge=lfs -text
Qwopus3.5-4B-Coder-Fable5-v1-Q4_K_M.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:3d3621cf4d7745b196d7bd6795eb452d777e6ebb910d049ab151339f32d29a00
3
+ size 2783446432
Qwopus3.5-4B-Coder-Fable5-v1-Q5_K_M.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:8f48d43eca68c80ec2455bd482ca4c0efab411402edc4dcee48383585b8fedc8
3
+ size 3161425312
Qwopus3.5-4B-Coder-Fable5-v1-mmproj-BF16.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:263043e99496ce01a6a050defcd3ded5a6d64e4638e5291a1b48926587c1fe50
3
+ size 675568672
README.md ADDED
@@ -0,0 +1,201 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ base_model: shuhulx/Qwopus3.5-4B-Coder-Fable5-v1
4
+ datasets:
5
+ - Glint-Research/Fable-5-traces
6
+ language:
7
+ - en
8
+ pipeline_tag: text-generation
9
+ library_name: gguf
10
+ tags:
11
+ - gguf
12
+ - llama-cpp
13
+ - lm-studio
14
+ - qwen3_5
15
+ - fable5
16
+ - reasoning
17
+ - agent
18
+ - tool-use
19
+ - function-calling
20
+ - coder
21
+ - coding
22
+ - debugging
23
+ - local-inference
24
+ - quantized
25
+ - conversational
26
+ ---
27
+
28
+ <div align="center">
29
+
30
+ # 馃捇 Qwopus3.5-4B-Coder-Fable5-v1 GGUF
31
+
32
+ ### GGUF builds for llama.cpp, LM Studio, and local inference
33
+
34
+ <p><b>Fable-5 traces</b> 路 <b>agentic coding</b> 路 <b>tool use</b> 路 <b>debugging</b></p>
35
+
36
+ </div>
37
+
38
+ ---
39
+
40
+ ## Overview
41
+
42
+ **Qwopus3.5-4B-Coder-Fable5-v1** is a Fable-5 trace continuation of [`Jackrong/Qwopus3.5-4B-Coder`](https://huggingface.co/Jackrong/Qwopus3.5-4B-Coder).
43
+
44
+ The base model, Qwopus3.5-4B-Coder, is a compact Qwen3.5-based coding model trained for reasoning, tool use, function calling, coding workflows, and agent-style behavior.
45
+
46
+ This release continues that model on [`Glint-Research/Fable-5-traces`](https://huggingface.co/datasets/Glint-Research/Fable-5-traces), a dataset of Claude Fable 5 local coding-agent traces. The dataset is heavily oriented around tool-use trajectories, repository work, local command context, code editing, debugging loops, and `<think>`-style reasoning completions.
47
+
48
+ The result is a small local coding-agent model intended for:
49
+
50
+ | Area | Description |
51
+ |---|---|
52
+ | Tool-use workflows | Bash, Read, Write, Edit, repo inspection, and action traces. |
53
+ | Debugging | Failing tests, stack traces, root-cause analysis, and patch planning. |
54
+ | Trace-style reasoning | Long-form planning and `<think>` style reasoning traces. |
55
+ | Local agents | Hermes-style, Claude-Code-style, OpenCode-style, and LM Studio workflows. |
56
+
57
+ ## Files
58
+
59
+ Typical GGUF files:
60
+
61
+ - `Qwopus3.5-4B-Coder-Fable5-v1-Q4_K_M.gguf`
62
+ - `Qwopus3.5-4B-Coder-Fable5-v1-Q5_K_M.gguf`
63
+ - `Qwopus3.5-4B-Coder-Fable5-v1-mmproj-BF16.gguf`
64
+
65
+ ## Which file should I use?
66
+
67
+ | File | Use case |
68
+ |---|---|
69
+ | `Q4_K_M` | Best default. Small, fast, good quality. |
70
+ | `Q5_K_M` | Better quality while still compact. |
71
+ | `Q8_0` | Higher quality, larger memory use, if included. |
72
+ | `mmproj-BF16` | Multimodal projector for compatible runtimes. |
73
+
74
+ ## llama.cpp
75
+
76
+ ```bash
77
+ llama-cli \
78
+ -m Qwopus3.5-4B-Coder-Fable5-v1-Q5_K_M.gguf \
79
+ -p "Write a Bash/Read/Edit style plan for debugging a failing Python repo." \
80
+ -n 768 \
81
+ --temp 0.7 \
82
+ --top-p 0.95
83
+ ```
84
+
85
+ ## llama.cpp Server
86
+
87
+ ```bash
88
+ llama-server \
89
+ -m Qwopus3.5-4B-Coder-Fable5-v1-Q5_K_M.gguf \
90
+ --host 0.0.0.0 \
91
+ --port 8080 \
92
+ --ctx-size 8192
93
+ ```
94
+
95
+ Then call it with an OpenAI-compatible client:
96
+
97
+ ```bash
98
+ curl -X POST "http://localhost:8080/v1/chat/completions" \
99
+ -H "Content-Type: application/json" \
100
+ --data '{
101
+ "model": "Qwopus3.5-4B-Coder-Fable5-v1-Q5_K_M.gguf",
102
+ "messages": [
103
+ {"role": "user", "content": "Write a tool-use plan for debugging a Python repo."}
104
+ ],
105
+ "temperature": 0.7,
106
+ "top_p": 0.95
107
+ }'
108
+ ```
109
+
110
+
111
+ ## About the Fable-5 Traces
112
+
113
+ [`Glint-Research/Fable-5-traces`](https://huggingface.co/datasets/Glint-Research/Fable-5-traces) contains Claude Fable 5 coding traces.
114
+
115
+ The dataset includes fields such as:
116
+
117
+ ```text
118
+ uid
119
+ source_file
120
+ session
121
+ model
122
+ context
123
+ cot
124
+ output_type
125
+ output
126
+ completion
127
+ origin
128
+ ```
129
+
130
+ The examples are not simple chat pairs. They are multi-step agent trajectories with local development context, reasoning traces, and tool-use outputs.
131
+
132
+ Common patterns in the dataset include:
133
+
134
+ - user coding requests
135
+ - local-command caveats
136
+ - repository inspection
137
+ - Bash command usage
138
+ - file reads
139
+ - file writes
140
+ - edits
141
+ - debugging passes
142
+ - playtesting / validation loops
143
+ - `<think>...</think>` reasoning traces
144
+ - tool-use completions
145
+
146
+ A large portion of the dataset is `tool_use` style data, which makes it especially relevant for local coding agents and developer automation.
147
+
148
+ ## Capabilities
149
+
150
+ ### Agentic coding
151
+
152
+ Designed for coding-agent loops where the model must inspect a repo, plan work, call tools, edit files, and validate changes.
153
+
154
+ ### Tool-use style outputs
155
+
156
+ Works well with prompts that expose structured tools such as:
157
+
158
+ ```text
159
+ Bash
160
+ Read
161
+ Write
162
+ Edit
163
+ Search
164
+ Grep
165
+ ```
166
+
167
+ ### Debugging and repair
168
+
169
+ Useful for:
170
+
171
+ - finding likely failing files
172
+ - explaining stack traces
173
+ - planning test commands
174
+ - proposing minimal patches
175
+ - iterating after errors
176
+
177
+ ### Local-first deployment
178
+
179
+ The release includes Transformers, GGUF, MLX, and MLX 4-bit formats so it can run in Python, llama.cpp, LM Studio, and Apple Silicon workflows.
180
+
181
+
182
+ ## Available Releases
183
+
184
+ | Release | Repo | Best for |
185
+ |---|---|---|
186
+ | Transformers / Safetensors | [`shuhulx/Qwopus3.5-4B-Coder-Fable5-v1`](https://huggingface.co/shuhulx/Qwopus3.5-4B-Coder-Fable5-v1) | Python, Transformers, custom inference. |
187
+ | GGUF | [`shuhulx/Qwopus3.5-4B-Coder-Fable5-v1-GGUF`](https://huggingface.co/shuhulx/Qwopus3.5-4B-Coder-Fable5-v1-GGUF) | llama.cpp, LM Studio, local CPU/GPU inference. |
188
+ | MLX | [`shuhulx/Qwopus3.5-4B-Coder-Fable5-v1-MLX`](https://huggingface.co/shuhulx/Qwopus3.5-4B-Coder-Fable5-v1-MLX) | Apple Silicon full MLX inference. |
189
+ | MLX 4-bit | [`shuhulx/Qwopus3.5-4B-Coder-Fable5-v1-MLX-4bit`](https://huggingface.co/shuhulx/Qwopus3.5-4B-Coder-Fable5-v1-MLX-4bit) | Apple Silicon low-memory inference. |
190
+
191
+ ## Credits
192
+
193
+ Built on:
194
+
195
+ - [`Jackrong/Qwopus3.5-4B-Coder`](https://huggingface.co/Jackrong/Qwopus3.5-4B-Coder) by Jackrong
196
+ - [`Glint-Research/Fable-5-traces`](https://huggingface.co/datasets/Glint-Research/Fable-5-traces) by Glint-Research
197
+ - Qwen / Qwen3.5 model family
198
+ - Unsloth
199
+ - Hugging Face
200
+ - llama.cpp
201
+ - mlx-lm