Commit Graph
134 Commits
Author SHA1 Message Date
Can BölükandGitHub ed1259f855 Merge branch 'main' into fix/check-spoofed-versions-source 2026-06-17 11:28:43 +02:00
cagedbird043 aa586c4ec3 fix(catalog): route google-antigravity default baseUrl to primary daily endpoint 2026-06-17 17:15:47 +08:00
can1357 94a813788f chore: bump version to 16.0.4 2026-06-17 09:44:33 +02:00
can1357 fb3534740f feat(catalog): added GLM-5.2 reasoning support for ZAI and zhipu completions
- Added ZAI GLM-5.2 reasoning-effort mapping, translating minimal to none and xhigh to max.
- Enabled ZAI and zhipu GLM-5.2 completion requests to send reasoning_effort and tool_stream.
- Added provider token clamping so GLM-5.2 completion requests use capped max_tokens.
- Updated catalog policies to route GLM-5.2 max-token and reasoning support through ZAI/zhipu hosts.
- Removed synthetic HF model entries and aligned GLM-5.2 catalog specs with real providers.

Fixes #2833
2026-06-17 09:38:43 +02:00
lycaon 9a39b0093f fix(scripts): check current Gemini CLI version source 2026-06-17 00:16:36 -06:00
can1357 8eeb707387 chore: bump version to 16.0.3 2026-06-17 01:58:36 +02:00
can1357 dc14689fc1 chore: bump version to 16.0.2 2026-06-16 15:33:47 +02:00
oldschoola 251d7dbb86 fix umans gateway websearch handling 2026-06-16 01:14:43 -07:00
oldschoola 6382896115 fix umans max token cap 2026-06-15 23:38:48 -07:00
can1357 e24f789de4 chore: bump version to 16.0.1 2026-06-15 21:30:30 +02:00
can1357 b27737df9e chore: fix changelogs 2026-06-15 19:51:53 +02:00
can1357 fcde870bc0 Merge PR #2636: Add Umans AI Coding Plan provider 2026-06-15 19:46:11 +02:00
can1357 44f9867e0e chore: bump version to 16.0.0 2026-06-15 17:27:48 +02:00
oldschoola 1c4bda29ff Drop Xiaomi ASR models from catalog 2026-06-15 05:35:19 -07:00
oldschoola 7bbe2d89b1 Surface Umans discovery fetch errors 2026-06-15 05:05:36 -07:00
can1357 e1814d9a08 refactor: renamed grammar module to dialect with unified transcript rendering
- Renamed ToolCallSyntax type to Dialect and Grammar interface to DialectDefinition across all packages.
- Moved grammar directory to dialect and updated all import paths in agent, ai, catalog, and coding-agent packages.
- Added renderTranscript and renderThinking methods to DialectDefinition, enabling native dialect-aware conversation serialization.
- Consolidated rendering utilities into new dialect/rendering.ts with shared helpers for ChatML, legacy text, and dialect-specific formatting.
- Updated conversation serialization in agent and coding-agent to use dialect.renderTranscript() for native turn envelope rendering.
2026-06-15 13:58:12 +02:00
oldschoola a18242cb3d Simplify Umans discovery error handling 2026-06-15 04:43:25 -07:00
oldschoola d7df0b6a09 Address Umans provider review feedback 2026-06-15 04:33:05 -07:00
oldschoola 23e180d5cb Resolve changelog merge conflicts 2026-06-15 04:00:37 -07:00
oldschoola ac8a8fe057 Merge remote-tracking branch 'origin/main' into oldschoola/umans 2026-06-15 03:59:48 -07:00
oldschoola 7d721dc420 Add Umans AI Coding Plan provider 2026-06-15 03:58:20 -07:00
can1357 789d4f6948 chore: bump version to 15.13.3 2026-06-15 12:47:58 +02:00
can1357 2de2926320 feat: added Gemini and Gemma in-band tool syntax support in runtime
- Added Gemini and Gemma syntax routing by model family and owned syntax env values.
- Added Gemini and Gemma in-band parsers for tool_code and token-based tool_call streams.
- Added rendering support for Gemini fenced tool_code/tool_outputs and Gemma tool tokens.
- Fixed parsing edge cases for comments, string escapes, nested args, and truncated blocks.
2026-06-15 11:40:55 +02:00
can1357 4d95a0e1dc feat(catalog/provider-models): added OpenAI model provider descriptors
- Updated default model identifiers across many catalog providers to newer model versions.
- Renamed a couple OpenAI compatibility provider descriptors, including Together and Zhipu coding-plan identifiers.
- Added multiple new OpenAI-compatible specialized provider descriptors for additional model provider families.
2026-06-15 10:40:48 +02:00
can1357 e680bc0ca3 feat(catalog): added Azure OpenAI support to registry and catalog compatibility
- Added `azure` provider registration in the AI registry with API key env mapping.
- Added Azure provider descriptors with default model `gpt-4o` and catalog discovery metadata.
- Enabled Azure-specific OpenAI compatibility for developer roles and strict responses pairing.
- Added Azure models namespace using OpenAI-family IDs with `models.dev` filtering and responses transport.
2026-06-15 10:06:26 +02:00
can1357 b7b1b54f6a chore: bump version to 15.13.2 2026-06-15 08:49:41 +02:00
can1357 7687810c5d feat: added syntax-aware tool example rendering across catalog, AI, and agent modules
- Added model-to-syntax mapping in catalog with preferred tool-call syntax API.
- Added `ToolExample` typing and `ToolCallSyntax` exports across tool/grammar interfaces.
- Added syntax-aware tool example rendering through provider-specific grammar invocations.
- Added `exampleSyntax` context flow and example metadata so rendered prompts include examples.
2026-06-15 07:33:25 +02:00
can1357 fbba331f8a feat(cross-cutting): added multi-syntax in-band tool-call support for runtime tool conversion
- Added optional Agent and SDK tool-call syntax controls (`toolCallSyntax`, `PI_OWNED_TOOLS`) for owned calls.
- Added in-band grammar scanners and renderers for Anthropic, DeepSeek, GLM, Hermes, Kimi, PI, and Qwen3.
- Added supportsTools propagation and model schema updates to route unsupported models to fallback syntax.
- Replaced stream-markup parsing with syntax-specific in-band scanners and event conversion.
2026-06-15 07:33:24 +02:00
can1357 6c48b6224d chore: bump version to 15.13.1 2026-06-15 04:36:40 +02:00
can1357 ecd9f3220f feat(scripts): added changelog recovery mode for historical duplicate cleanup
- Added --recover CLI option to rebuild changelog fixes from tagged history.
- Pruned Unreleased bullets that match historical released items across all tags.
- Normalized output by sorting release sections by version and compacting bullet spacing.
2026-06-15 03:43:24 +02:00
roboomp d142652fbe style: bun run fix 2026-06-14 18:26:27 +00:00
roboomp 4efdefa398 Revert "style: bun run fix"
This reverts commit 60b93ea5c1.
2026-06-14 16:50:54 +00:00
roboomp 60b93ea5c1 style: bun run fix 2026-06-14 16:45:38 +00:00
roboomp caad59e526 fix(catalog): pinned minimax m3 context
Pinned MiniMax-M3 contextWindow to 1,000,000 for the minimax and minimax-cn bundled catalog entries during generation.

Added policy and bundled catalog regression coverage while leaving MiniMax coding-plan providers on upstream limits.

Fixes #2576
2026-06-14 16:45:23 +00:00
can1357 7e8122000a chore: bump version to 15.13.0 2026-06-14 18:21:59 +02:00
can1357 3709ae0ca9 Merge PR #2451: fix(coding-agent): preserve discovered context windows 2026-06-14 17:54:50 +02:00
can1357 e28269c235 docs(changelog): move vLLM discovery entries to [Unreleased]
Rebasing onto current main landed both the coding-agent and catalog
vLLM entries under the released [15.12.6] section. Per AGENTS.md, new
entries go under [Unreleased] and released sections are immutable, so
the release tooling would otherwise miss them.

Addresses review feedback on #2451.
2026-06-14 17:50:16 +02:00
Gerben Meijerandcan1357 4a6eec624b fix(coding-agent): preserve vLLM discovered context windows
Read vLLM max_model_len and OpenAI-compatible context_length metadata during model discovery, route providers.vllm.baseUrl into built-in discovery before cached models exist, and avoid sending local placeholder bearer tokens.

Scope the vLLM model cache to the discovery base URL so endpoint changes refetch immediately, and add focused regression coverage for configured and built-in vLLM discovery.
2026-06-14 17:50:16 +02:00
can1357 b0ab7ed28c Merge PR #2410: feat(extensions): expose model resolve() and family() via ctx.models 2026-06-14 17:44:28 +02:00
can1357 6d459c88a1 fix(catalog): classify GLM in modelFamilyToken family tokens
modelFamilyToken's fallback chain checked kimi/qwen/minimax/gpt-oss/
deepseek/mimo but omitted GLM, even though parseGlmModel is already
imported here and GLM is a first-class catalog family. GLM ids fell
through to "", so ctx.models.family() split same-lineage provider
mirrors (zai/glm-5.2 vs zhipu-coding-plan/glm-5.2) by provider instead
of folding them. Add the GLM check plus a cross-mirror regression test.

The mandated fmt pass over this file also wraps the ./classify import,
which the PR's added parseKnownModel pushed to 122 cols (biome lineWidth
is 120 -> would otherwise fail `biome check` in CI).

Addresses review feedback on #2410.
2026-06-14 17:31:59 +02:00
can1357 9a5e44df68 test(catalog): update zenmux default model expectation 2026-06-14 17:27:26 +02:00
Asaf Mahlevandcan1357 d12ed9135b fix(changelog): place ctx.models entry under [Unreleased] 2026-06-14 17:24:48 +02:00
Asaf Mahlevandcan1357 57ea171859 address #2406 review: full-priority role resolve, family JSDoc, catalog changelog, role + canonical tests 2026-06-14 17:24:48 +02:00
Asaf Mahlevandcan1357 3c53218e19 feat(extensions): add read-only ctx.models query facade
Expose `ctx.models` to extensions: list() / current() / resolve(spec) /
family(model). Lets extension tools select models the same way core does
(settings-backed aliases, match preferences, canonical-identity family
classification) without reaching into the mutable registry.

- types.ts: ExtensionModelQuery interface + `models` on ExtensionContext
- model-api.ts: createExtensionModelQuery facade
- runner.ts: thread optional Settings; build `models` in createContext()
- sdk.ts + agent-session.ts + extension-ui-controller.ts: pass settings / build models on the direct context literals
- catalog identity: modelFamilyToken() — coarse canonical-backed lineage token
- docs + changelog + tests

Implements #2406.
2026-06-14 17:24:48 +02:00
can1357 1b436b775b Merge PR #2543: fix(providers/google): avoid outputting system rules and preambles in Antigravity 2026-06-14 17:09:41 +02:00
can1357 baf753bab7 Merge remote-tracking branch 'origin/farm/bdf5fe70/strip-eager-input-streaming-copilot' 2026-06-14 16:10:52 +02:00
roboomp 69e9e66d03 test(catalog): sourced Model and ModelSpec types from pi-catalog
Followed AGENTS.md by importing the catalog model-spec types from @oh-my-pi/pi-catalog/types in the issue #2558 regression. The Context/Tool/TJsonSchema signatures used by streamAnthropic still come from pi-ai.
2026-06-14 11:14:18 +00:00
roboomp e9551a1312 test(catalog): decoupled copilot eager-stream regression from bundle
Replaced the issue #2558 regression's bundled github-copilot model lookup with a minimal ModelSpec resolved through buildModel, so the coverage exercises the Anthropic compat resolver without depending on generated models.json retaining a specific Copilot SKU.
2026-06-14 11:07:00 +00:00
roboomp d856821f51 style: bun run fix 2026-06-14 11:00:20 +00:00
roboomp 67a5cf454f fix(anthropic): drop eager_input_streaming on github-copilot transport
GitHub Copilot's Anthropic-compatible proxy (api.githubcopilot.com/v1/messages) validates tool definitions with additionalProperties: false and rejected every tool-bearing turn with `400 tools.0.custom.eager_input_streaming: Extra inputs are not permitted`. The catalog's Anthropic compat builder defaulted supportsEagerToolInputStreaming to true for all hosts, so convertTools added the flag to every tool sent to Copilot.

Gate the flag on host in buildAnthropicCompat (false for github-copilot, matching the existing supportsLongCacheRetention: official pattern), and stop pushing the legacy fine-grained-tool-streaming-2025-05-14 beta header on the Copilot transport — the proxy doesn't whitelist Anthropic beta features either, so the fallback path would 400 too.

Fixes #2558
2026-06-14 11:00:01 +00:00