mirror of
https://github.com/akitaonrails/ai-memory.git
synced 2026-10-02 03:24:46 +08:00
docs: document FutureInfra as an openai-compat provider
Add a FutureInfra example to the openai-compat section of docs/install.md, list it in docs/llm-providers.md, and add an [Unreleased] CHANGELOG entry. Docs only. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
This commit is contained in:
co-authored by
Claude Opus 5.5
parent
7580b74d0f
commit
2a9f12deae
@@ -7,6 +7,10 @@ and this project adheres to [Semantic Versioning](https://semver.org/spec/v2.0.0
|
||||
|
||||
## [Unreleased]
|
||||
|
||||
### Changed
|
||||
- Documented FutureInfra as an endpoint for the existing `openai-compat`
|
||||
provider.
|
||||
|
||||
## [2.5.2] - 2026-10-01
|
||||
|
||||
### Added
|
||||
|
||||
@@ -2035,6 +2035,26 @@ hosted compatibility endpoints, no dedicated ai-memory provider is required.
|
||||
Configure embeddings separately if your chosen API Route model does not
|
||||
provide an OpenAI-compatible embeddings endpoint.
|
||||
|
||||
[FutureInfra](https://futureinfra.ai/ai/) is an OpenAI-compatible AI API
|
||||
router and uses the same provider; no dedicated ai-memory provider is needed.
|
||||
Its OpenAI-compatible base is `https://futureinfra.ai/v1/ai`. That path does
|
||||
not end in a version segment, so pass the full Chat Completions URL (ai-memory
|
||||
uses a base URL that already ends in `/chat/completions` as-is). Pass its API
|
||||
key through the generic compatibility credential:
|
||||
|
||||
```bash
|
||||
-e AI_MEMORY_LLM_PROVIDER=openai-compat
|
||||
-e AI_MEMORY_LLM_BASE_URL=https://futureinfra.ai/v1/ai/chat/completions
|
||||
-e AI_MEMORY_LLM_MODEL=openai/gpt-4o-mini
|
||||
-e LLM_API_KEY="$FUTUREINFRA_API_KEY"
|
||||
```
|
||||
|
||||
Model ids use the `provider/model` format, e.g. `anthropic/claude-sonnet-4` or
|
||||
`deepseek/deepseek-chat`; `GET https://futureinfra.ai/v1/ai/models` lists the
|
||||
current ids. Keys are created in the
|
||||
[FutureInfra console](https://futureinfra.ai/console/?screen=ai-router). This
|
||||
example configures the LLM only; configure embeddings separately.
|
||||
|
||||
OpenAI-compatible structured calls use the operation's JSON Schema by default:
|
||||
|
||||
```bash
|
||||
|
||||
@@ -50,7 +50,7 @@ Recommended defaults:
|
||||
| `copilot` | `gpt-5.5` | GitHub Copilot Chat backend via `ai-memory auth login copilot` or `COPILOT_GITHUB_TOKEN`; requires a Copilot subscription. |
|
||||
| `gemini` | `gemini-3.5-flash` | Google-hosted option with a generous free tier. |
|
||||
| `opencode` | `claude-sonnet-4-6` | OpenCode Go or Zen via `OPENCODE_API_KEY`. Go is the default endpoint; `AI_MEMORY_LLM_BASE_URL` selects Zen. Set `AI_MEMORY_LLM_MODEL` to an id the chosen endpoint serves. |
|
||||
| `openai-compat` | no default | OpenRouter, Atlas Cloud, OrcaRouter, Cheaper Inference, API Route, Ollama, vLLM, LM Studio, and other compatible endpoints. |
|
||||
| `openai-compat` | no default | OpenRouter, Atlas Cloud, OrcaRouter, Cheaper Inference, API Route, FutureInfra, Ollama, vLLM, LM Studio, and other compatible endpoints. |
|
||||
| `openai-compat` + `AI_MEMORY_LLM_BASE_URL=https://openrouter.ai/api/v1` | no default (recommended: `anthropic/claude-haiku-4.5`) | Hosted access to a large model catalogue through one key. See [OpenRouter](#openrouter) below and the empirical comparison in [`llm-provider-comparison.md`](llm-provider-comparison.md). |
|
||||
|
||||
`openai-oauth` stores a refresh token in `<data_dir>/auth.json` and talks to
|
||||
@@ -424,7 +424,7 @@ sentence embeddings run in-process (pure-Rust `all-MiniLM-L6-v2`,
|
||||
pinned checksums — see [`docs/local-embeddings.md`](local-embeddings.md).
|
||||
|
||||
See [`docs/install.md#llm-provider-tiers`](install.md#llm-provider-tiers)
|
||||
for env vars and Ollama/OpenRouter/Atlas Cloud/OrcaRouter/Cheaper Inference/API Route
|
||||
for env vars and Ollama/OpenRouter/Atlas Cloud/OrcaRouter/Cheaper Inference/API Route/FutureInfra
|
||||
examples, and
|
||||
[`docs/llm-provider-comparison.md`](llm-provider-comparison.md)
|
||||
for the empirical model comparison.
|
||||
|
||||
Reference in New Issue
Block a user