The short version
Community Jinja chat templates tuned for running Qwen models with vLLM.
From the HackerLinks archive
Community Jinja chat templates tuned for running Qwen models with vLLM.
The short version
Community Jinja chat templates tuned for running Qwen models with vLLM.
Why it caught our attention
A production user reported good agentic results and emphatically recommended the templates.
Where it surfaced on Hacker News
“Don't sleep on the froggeric templates.”
hadlock · recommendation · production use
Direct HN comment
Show HN: Running 104GB Qwen3.8-Flash-Next on 48GB Mac with at ~12 tok/s