Pi v0.79.9: chat-template thinking for DeepSeek/vLLM, GLM-5.2 routing fixes
OpenAI-compatible custom providers can now map Pi thinking levels into chat_template_kwargs, enabling vLLM/Hugging Face chat-template models such as DeepSeek to use provider-native thinking controls. GLM-5.2 gets corrected Fireworks OpenAI-compatible routing and OpenRouter xhigh thinking support. Multiple fixes include session extension reuse, deep branch quadratic-time fix, streaming code fence flicker, fuzzy edit file rewrite, and /model ranking improvements.
PUBLISHED2026-06-20
OBSERVED2026-08-21
AGE2mo
SOURCES1
- Chat-template thinking: OpenAI-compatible providers can map Pi thinking levels into
chat_template_kwargsfor vLLM/HF models like DeepSeek - GLM-5.2: corrected Fireworks routing to OpenAI-compatible Chat Completions endpoint with
reasoning_effort - GLM-5.2: OpenRouter
xhighreasoning effort exposed via nativexhighparameter - Fixed same-directory session switches reusing imported extension modules while preserving fresh instances
- Fixed deep session branches taking quadratic time to build context/branch paths
- Fixed Markdown streaming code fence rendering — partial closing fences no longer shrink/flicker code blocks
- Fixed fuzzy
editmatches preserving untouched line blocks instead of rewriting whole file - Fixed bash commands through legacy WSL
bash.exepassing scripts over stdin for variable expansion - Fixed
/modelhiding GitHub Copilot models unavailable to authenticated account - Fixed
/modelsearch ranking exact provider-prefixed matches before proxy-provider model ID matches - Fixed RPC unknown-command errors including request id to prevent client hangs
COMMUNITY
No curated reactions recorded for this event. Facts and takes are kept in separate layers — community context is added by hand, never blended into the record above.