Normalize GLM-style tool-call responses for OpenAI-compatible Python clients, including sync, async and streaming adapters.
-
Updated
Sep 22, 2026 - Python
Normalize GLM-style tool-call responses for OpenAI-compatible Python clients, including sync, async and streaming adapters.
v0.2.0: detected model id now sent in probe payloads; malformed endpoint responses classify cleanly (no raw tracebacks); clean exit-2 usage errors for bad --out; new no_tool probe for over-eager tool-call detection (suite cn-tc-v2).
Per-task reasoning-budget governor for Qwen3.8 coding agents: profile think traces, learn class budgets, and advise via an OpenAI-compatible streaming proxy.
国产大模型工具调用契约差分测试器
本地 CN 模型 tool-call JSON 运行期修复代理,让 Claude Code 接 Qwen3/DeepSeek 不再静默崩溃
Run-level exposure gauge for coding agents on local models — tracks wall-clock, state mutations, and token burn since last human review, forces a pause at thresholds.
为 coding agent 标出每周影响配置的国产模型 drop 变更
一条命令对比 DeepSeek/Qwen/Kimi/GLM 对你输入的分词数、上下文适配率与含缓存折扣的成本。
ctxfeed is a local MCP project-context backend that shards a whole repo into GLM-5.2's 1M-token window with cache-aware ingest ordering
Capability-gap diagnostician for locally-served CN open-weights models (DeepSeek/Qwen3/Kimi K3/GLM) — attributes your local model's felt 'dumbness' to a specific misconfiguration layer so you fix one layer, not the whole stack.
To associate your repository with the cn-model-devtools topic, visit your repo's landing page and select "manage topics."