docs: document FutureInfra as an openai-compat provider

Add a FutureInfra example to the openai-compat section of docs/install.md, list it in docs/llm-providers.md, and add an [Unreleased] CHANGELOG entry. Docs only.

Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
This commit is contained in:
uxidev
2026-10-02 00:24:59 +09:00
co-authored by Claude Opus 5.5
parent 7580b74d0f
commit 2a9f12deae
3 changed files with 26 additions and 2 deletions
+4
View File
@@ -7,6 +7,10 @@ and this project adheres to [Semantic Versioning](https://semver.org/spec/v2.0.0
## [Unreleased]
### Changed
- Documented FutureInfra as an endpoint for the existing `openai-compat`
provider.
## [2.5.2] - 2026-10-01
### Added
+20
View File
@@ -2035,6 +2035,26 @@ hosted compatibility endpoints, no dedicated ai-memory provider is required.
Configure embeddings separately if your chosen API Route model does not
provide an OpenAI-compatible embeddings endpoint.
[FutureInfra](https://futureinfra.ai/ai/) is an OpenAI-compatible AI API
router and uses the same provider; no dedicated ai-memory provider is needed.
Its OpenAI-compatible base is `https://futureinfra.ai/v1/ai`. That path does
not end in a version segment, so pass the full Chat Completions URL (ai-memory
uses a base URL that already ends in `/chat/completions` as-is). Pass its API
key through the generic compatibility credential:
```bash
-e AI_MEMORY_LLM_PROVIDER=openai-compat
-e AI_MEMORY_LLM_BASE_URL=https://futureinfra.ai/v1/ai/chat/completions
-e AI_MEMORY_LLM_MODEL=openai/gpt-4o-mini
-e LLM_API_KEY="$FUTUREINFRA_API_KEY"
```
Model ids use the `provider/model` format, e.g. `anthropic/claude-sonnet-4` or
`deepseek/deepseek-chat`; `GET https://futureinfra.ai/v1/ai/models` lists the
current ids. Keys are created in the
[FutureInfra console](https://futureinfra.ai/console/?screen=ai-router). This
example configures the LLM only; configure embeddings separately.
OpenAI-compatible structured calls use the operation's JSON Schema by default:
```bash
+2 -2
View File
@@ -50,7 +50,7 @@ Recommended defaults:
| `copilot` | `gpt-5.5` | GitHub Copilot Chat backend via `ai-memory auth login copilot` or `COPILOT_GITHUB_TOKEN`; requires a Copilot subscription. |
| `gemini` | `gemini-3.5-flash` | Google-hosted option with a generous free tier. |
| `opencode` | `claude-sonnet-4-6` | OpenCode Go or Zen via `OPENCODE_API_KEY`. Go is the default endpoint; `AI_MEMORY_LLM_BASE_URL` selects Zen. Set `AI_MEMORY_LLM_MODEL` to an id the chosen endpoint serves. |
| `openai-compat` | no default | OpenRouter, Atlas Cloud, OrcaRouter, Cheaper Inference, API Route, Ollama, vLLM, LM Studio, and other compatible endpoints. |
| `openai-compat` | no default | OpenRouter, Atlas Cloud, OrcaRouter, Cheaper Inference, API Route, FutureInfra, Ollama, vLLM, LM Studio, and other compatible endpoints. |
| `openai-compat` + `AI_MEMORY_LLM_BASE_URL=https://openrouter.ai/api/v1` | no default (recommended: `anthropic/claude-haiku-4.5`) | Hosted access to a large model catalogue through one key. See [OpenRouter](#openrouter) below and the empirical comparison in [`llm-provider-comparison.md`](llm-provider-comparison.md). |
`openai-oauth` stores a refresh token in `<data_dir>/auth.json` and talks to
@@ -424,7 +424,7 @@ sentence embeddings run in-process (pure-Rust `all-MiniLM-L6-v2`,
pinned checksums — see [`docs/local-embeddings.md`](local-embeddings.md).
See [`docs/install.md#llm-provider-tiers`](install.md#llm-provider-tiers)
for env vars and Ollama/OpenRouter/Atlas Cloud/OrcaRouter/Cheaper Inference/API Route
for env vars and Ollama/OpenRouter/Atlas Cloud/OrcaRouter/Cheaper Inference/API Route/FutureInfra
examples, and
[`docs/llm-provider-comparison.md`](llm-provider-comparison.md)
for the empirical model comparison.