docs(providers): documented novita support
This commit is contained in:
@@ -72,7 +72,7 @@ For each candidate issue, read the title, body, and **all comments** (comments o
|
||||
| `providers` | Provider-related behavior (generic provider scope) |
|
||||
|
||||
**Provider labels** (apply only when a specific provider is explicitly involved):
|
||||
`provider:anthropic`, `provider:bedrock`, `provider:brave`, `provider:cerebras`, `provider:cloudflare`, `provider:codex`, `provider:copilot`, `provider:cursor`, `provider:exa`, `provider:gemini`, `provider:gitlab`, `provider:groq`, `provider:huggingface`, `provider:jina`, `provider:kimi`, `provider:litellm`, `provider:minimax`, `provider:mistral`, `provider:moonshot`, `provider:nanogpt`, `provider:nvidia`, `provider:openai`, `provider:opencode`, `provider:openrouter`, `provider:perplexity`, `provider:qianfan`, `provider:qwen`, `provider:synthetic`, `provider:together`, `provider:venice`, `provider:vercel`, `provider:xai`, `provider:xiaomi`, `provider:zai`
|
||||
`provider:anthropic`, `provider:bedrock`, `provider:brave`, `provider:cerebras`, `provider:cloudflare`, `provider:codex`, `provider:copilot`, `provider:cursor`, `provider:exa`, `provider:gemini`, `provider:gitlab`, `provider:groq`, `provider:huggingface`, `provider:jina`, `provider:kimi`, `provider:litellm`, `provider:minimax`, `provider:mistral`, `provider:moonshot`, `provider:nanogpt`, `provider:novita`, `provider:nvidia`, `provider:openai`, `provider:opencode`, `provider:openrouter`, `provider:perplexity`, `provider:qianfan`, `provider:qwen`, `provider:synthetic`, `provider:together`, `provider:venice`, `provider:vercel`, `provider:xai`, `provider:xiaomi`, `provider:zai`
|
||||
|
||||
**Platform labels** (apply only when platform materially affects reproduction/root cause):
|
||||
| Label | Signals |
|
||||
|
||||
@@ -292,7 +292,7 @@ Anthropic `oauth` · OpenAI · OpenAI Codex `oauth` · Google Gemini · Google A
|
||||
|
||||
Subscription-routed. `/login` attaches the session.
|
||||
|
||||
Cursor `oauth` · GitHub Copilot `oauth` · GitLab Duo · Kimi Code `plan` · Moonshot · MiniMax Coding Plan `plan` · MiniMax Coding Plan CN `plan` · Alibaba Coding Plan `plan` · Qwen Portal · Z.AI / GLM Coding Plan `plan` · Xiaomi MiMo · Qianfan · NanoGPT · Venice · Kilo · ZenMux · OpenCode Go · OpenCode Zen
|
||||
Cursor `oauth` · GitHub Copilot `oauth` · GitLab Duo · Kimi Code `plan` · Moonshot · MiniMax Coding Plan `plan` · MiniMax Coding Plan CN `plan` · Alibaba Coding Plan `plan` · Qwen Portal · Z.AI / GLM Coding Plan `plan` · Xiaomi MiMo · Qianfan · NanoGPT · Novita · Venice · Kilo · ZenMux · OpenCode Go · OpenCode Zen
|
||||
|
||||
### Run it yourself
|
||||
|
||||
|
||||
@@ -49,6 +49,7 @@ These are consumed via `getEnvApiKey()` (`packages/ai/src/stream.ts`) unless not
|
||||
| `SYNTHETIC_API_KEY` | Synthetic auth | Using Synthetic models | |
|
||||
| `NVIDIA_API_KEY` | NVIDIA auth | Using `nvidia` provider | |
|
||||
| `NANO_GPT_API_KEY` | NanoGPT auth | Using `nanogpt` provider | |
|
||||
| `NOVITA_API_KEY` | Novita auth | Using `novita` provider | |
|
||||
| `VENICE_API_KEY` | Venice auth | Using `venice` provider | |
|
||||
| `LITELLM_API_KEY` | LiteLLM auth | Using `litellm` provider | OpenAI-compatible LiteLLM proxy key |
|
||||
| `LM_STUDIO_API_KEY` | LM Studio auth (optional) | Using `lm-studio` provider with authenticated hosts | Local LM Studio usually runs without auth; any non-empty token works when a key is required |
|
||||
|
||||
@@ -105,6 +105,7 @@ Each provider has one or more environment variables that supply a key when no st
|
||||
| `huggingface` | `HUGGINGFACE_HUB_TOKEN`, then `HF_TOKEN` |
|
||||
| `moonshot` | `MOONSHOT_API_KEY` |
|
||||
| `nanogpt` | `NANO_GPT_API_KEY` |
|
||||
| `novita` | `NOVITA_API_KEY` |
|
||||
| `venice` | `VENICE_API_KEY` |
|
||||
| `vercel-ai-gateway` | `AI_GATEWAY_API_KEY` (also `VERCEL_AI_GATEWAY_API_KEY` for catalog discovery) |
|
||||
| `cloudflare-ai-gateway` | `CLOUDFLARE_AI_GATEWAY_API_KEY` |
|
||||
|
||||
@@ -7,6 +7,7 @@
|
||||
- Added model-driven Codex Responses Lite: `responsesLite` now defaults to the catalog `useResponsesLite` flag (codex-rs `use_responses_lite`, set on the GPT-5.6 family), so lite requests are sent without per-call opt-in.
|
||||
- Added the full Responses Lite wire contract: lite requests move tools into a leading `{type: "additional_tools", role: "developer"}` input item and the base instructions into a developer message, omit top-level `instructions`/`tools`, and force `parallel_tool_calls: false`, mirroring codex-rs `build_responses_request`.
|
||||
- Added concurrent reasoning summaries on Codex Responses: requests with a reasoning summary send `stream_options: { reasoning_summary_delivery: "sequential_cutoff" }`, and the stream decoder consumes the matching atomic `response.reasoning_summary_text.done` events (resolved by `item_id`/`output_index`, stale dones dropped, incremental `.delta`/`.part.*` events ignored under the cutoff contract). The cutoff gate reads the post-`onPayload` wire body on both transports, and `response.reasoning_summary_text.done` now counts as websocket watchdog progress.
|
||||
- Added Novita API-key login with authenticated key validation and `NOVITA_API_KEY` discovery ([#4917](https://github.com/can1357/oh-my-pi/pull/4917) by [@jason-wu-ai](https://github.com/jason-wu-ai)).
|
||||
|
||||
### Changed
|
||||
|
||||
|
||||
@@ -59,6 +59,7 @@ Unified LLM API with automatic model discovery, provider configuration, token an
|
||||
- **Qianfan** (requires `QIANFAN_API_KEY`)
|
||||
- **NVIDIA** (requires `NVIDIA_API_KEY`)
|
||||
- **NanoGPT** (requires `NANO_GPT_API_KEY`)
|
||||
- **Novita** (requires `NOVITA_API_KEY`)
|
||||
- **Hugging Face Inference**
|
||||
- **xAI**
|
||||
- **Venice** (requires `VENICE_API_KEY`)
|
||||
@@ -943,6 +944,7 @@ In Node.js environments, you can set environment variables to avoid passing API
|
||||
| Synthetic | `SYNTHETIC_API_KEY` |
|
||||
| NVIDIA | `NVIDIA_API_KEY` |
|
||||
| NanoGPT | `NANO_GPT_API_KEY` |
|
||||
| Novita | `NOVITA_API_KEY` |
|
||||
| Venice | `VENICE_API_KEY` |
|
||||
| Moonshot | `MOONSHOT_API_KEY` |
|
||||
| xAI | `XAI_API_KEY` |
|
||||
@@ -981,6 +983,7 @@ Provider endpoint defaults for the current OpenAI-compatible integrations:
|
||||
- Qianfan: `https://qianfan.baidubce.com/v2`
|
||||
- NVIDIA: `https://integrate.api.nvidia.com/v1`
|
||||
- NanoGPT: `https://nano-gpt.com/api/v1`
|
||||
- Novita: `https://api.novita.ai/openai/v1`
|
||||
- Hugging Face Inference: `https://router.huggingface.co/v1`
|
||||
- Venice: `https://api.venice.ai/api/v1`
|
||||
- Xiaomi MiMo: `https://api.xiaomimimo.com/anthropic`
|
||||
@@ -1082,7 +1085,7 @@ Credentials are saved to `agent.db` in the agent directory. `/login qianfan` ope
|
||||
|
||||
`login` supports OAuth providers (Anthropic, OpenAI Codex, GitHub Copilot, Gemini CLI, Antigravity) and API-key onboarding flows.
|
||||
|
||||
For the current API-key onboarding flows, the library covers Together, Moonshot, Qianfan, NVIDIA, NanoGPT, Hugging Face, Venice, Xiaomi, vLLM, LiteLLM, Cloudflare AI Gateway, Qwen Portal, and Ollama Cloud. Ollama remains the local runtime integration; set `OLLAMA_API_KEY` only when your local or self-hosted deployment enforces bearer auth.
|
||||
For the current API-key onboarding flows, the library covers Together, Moonshot, Qianfan, NVIDIA, NanoGPT, Novita, Hugging Face, Venice, Xiaomi, vLLM, LiteLLM, Cloudflare AI Gateway, Qwen Portal, and Ollama Cloud. Ollama remains the local runtime integration; set `OLLAMA_API_KEY` only when your local or self-hosted deployment enforces bearer auth.
|
||||
|
||||
### Programmatic OAuth
|
||||
|
||||
@@ -1114,7 +1117,7 @@ import {
|
||||
getOAuthApiKey, // (provider, credentialsMap) => { newCredentials, apiKey } | null
|
||||
|
||||
// Types
|
||||
type OAuthProvider, // includes 'anthropic', 'openai-codex', 'github-copilot', 'google-gemini-cli', 'google-antigravity', 'together', 'moonshot', 'qianfan', 'nvidia', 'nanogpt', 'huggingface', 'venice', 'xiaomi', 'vllm', 'litellm', 'cloudflare-ai-gateway', 'qwen-portal', ...
|
||||
type OAuthProvider, // includes 'anthropic', 'openai-codex', 'github-copilot', 'google-gemini-cli', 'google-antigravity', 'together', 'moonshot', 'qianfan', 'nvidia', 'nanogpt', 'novita', 'huggingface', 'venice', 'xiaomi', 'vllm', 'litellm', 'cloudflare-ai-gateway', 'qwen-portal', ...
|
||||
type OAuthCredentials,
|
||||
} from "@oh-my-pi/pi-ai";
|
||||
```
|
||||
|
||||
@@ -8,6 +8,7 @@
|
||||
- Added support for Dolphin Mistral 24b Venice Edition
|
||||
- Added GLM5.2-Fast model
|
||||
- Added Zenmux variants for GPT-5.6 (Luna, Sol, and Terra)
|
||||
- Added Novita as a model provider with authoritative public catalog discovery and generated pricing, limits, modality, reasoning, and tool metadata ([#4917](https://github.com/can1357/oh-my-pi/pull/4917) by [@jason-wu-ai](https://github.com/jason-wu-ai)).
|
||||
|
||||
- Added `useResponsesLite` to `Model`/`ModelSpec` and Codex discovery parsing of the upstream `use_responses_lite` flag; regenerated `models.json` marks the GPT-5.6 family (`sol`/`terra`/`luna` and their pro aliases) for the Responses Lite transport. Added the `x-openai-internal-codex-responses-lite` marker to `OPENAI_HEADERS`.
|
||||
|
||||
|
||||
Reference in New Issue
Block a user