MLX
jinja
chat-template
qwen
qwen3.5
qwen3.6
qwen3.8
llama.cpp
lm-studio
vllm
tool-calling
thinking
token-efficient
Instructions to use peculiar-ragdoll/Qwen-Sharp-Chat-Templates with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use peculiar-ragdoll/Qwen-Sharp-Chat-Templates with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir Qwen-Sharp-Chat-Templates peculiar-ragdoll/Qwen-Sharp-Chat-Templates
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
compatible with Qwen3.8-Flash-Next too???
❤️ 1
1
#12 opened 5 days ago
by
vanSamstroem
Production adoption report (vLLM + MTP): terseness confirmed, quality up at n=4 — and the MTP think-leak did NOT reproduce
❤️ 1
1
#11 opened 10 days ago
by
joelafrite
Would this work on Kay coder dev?
1
#10 opened 12 days ago
by
datayoda
reasoning_effort info is outdated
#9 opened 13 days ago
by
brunocasado
Fix KV cache reuse when changing reasoning_effort
🔥 2
2
#8 opened 16 days ago
by
dormosh
Usage on Ornith-1.5-35B-A3B
1
#7 opened 17 days ago
by
nonitis
Add a `_default_reasoning_effort` knob and remove the dead `_initial_effort` line (discussion #3)
#6 opened 18 days ago
by
gdevenyi
Add a `terse` template kwarg to opt out of the terseness block
❤️ 1
2
#5 opened 18 days ago
by
gdevenyi
Make reasoning effort steering more intuitive in the Jinja template
3
#3 opened 20 days ago
by
extrabigmehdi
Froggeric just release the new Qwen-Fixed-Chat-Templates v22
❤️ 2
3
#2 opened 25 days ago
by
lawlietr
Possible enhancemet to avoid problems with tool calling
❤️ 1
1
#1 opened 25 days ago
by
NeoHuggingF