714051d795
GitHub Copilot's /models response advertises supports.vision = true for Claude/GPT chat models on every host, but only the canonical personal endpoint (https://api.githubcopilot.com) actually accepts image inputs; the business (api.business.githubcopilot.com) and enterprise (copilot-api.{domain}) hosts respond '400 vision is not supported'. snapcompact then injected rasterized transcript frames after compaction and permanently broke every business-Copilot session. - Catalog discovery (githubCopilotModelManagerOptions.mapModel) now forces input=['text'] whenever the resolved baseUrl is not the canonical personal-Copilot host, so the upstream's vision flag is honoured only where it actually works. - mergeDynamicModel honours the dynamic input value (instead of OR-upgrading with the bundled reference) when the merged baseUrl differs from the bundled one, so a bundled spec pinned to the personal host can no longer taint a business-resolved merge. - snapcompact-inline's canSendImages helper short-circuits the rasterizer for any github-copilot model whose baseUrl is non-personal, catching stale cached specs that still advertise vision. - Helper isPersonalGitHubCopilotBaseUrl exported from pi-catalog/wire/github-copilot so catalog and coding-agent share one canonical check. Regression coverage in github-copilot-model-limits.test.ts (vision endpoint policy + full merge) and snapcompact-inline.test.ts (#3387 business/enterprise case). Fixes #3387
@oh-my-pi/pi-coding-agent
Core implementation package for the omp coding agent in the oh-my-pi monorepo.
For installation, setup, provider configuration, model roles, slash commands, and full CLI reference, see:
Package-specific references:
Memory backends
The agent supports three mutually-exclusive memory backends, selected via the memory.backend setting (Settings → Memory tab, or ~/.omp/config.yml):
off(default) — no memory subsystem runs.local— existing rollout-summarisation pipeline; writesmemory_summary.mdand consolidated artifacts under the agent dir.hindsight— talks to a Hindsight server (Cloud or self-hosted Docker), retains transcripts every Nth user turn, recalls memories on the first turn of a session, and exposesretain,recall, andreflect.
Hindsight quickstart
- Run a Hindsight server (Cloud or
docker run -p 8888:8888 ghcr.io/vectorize-io/hindsight:latest). - Set
memory.backend = "hindsight"andhindsight.apiUrl = "http://localhost:8888"(or your Cloud URL). - Optional environment overrides (env wins over settings):
HINDSIGHT_API_URL,HINDSIGHT_API_TOKEN— connectionHINDSIGHT_BANK_ID,HINDSIGHT_DYNAMIC_BANK_ID,HINDSIGHT_AGENT_NAME— bank addressingHINDSIGHT_AUTO_RECALL,HINDSIGHT_AUTO_RETAIN,HINDSIGHT_RETAIN_MODE— lifecycleHINDSIGHT_RECALL_BUDGET,HINDSIGHT_RECALL_MAX_TOKENS— recall sizingHINDSIGHT_BANK_MISSION,HINDSIGHT_DEBUG
Switching backends mid-session is honoured on the next system-prompt rebuild and the next /memory slash command. Existing users with memories.enabled = true|false are migrated to memory.backend = "local"|"off" exactly once on first launch.