0d3cb498a3
CI / Shell Format Check (push) Has been cancelled
CI / Check Ruby (3.4) (push) Has been cancelled
CI / CI Config (push) Has been cancelled
CI / Test on Node ${{ matrix.node }} and ${{ matrix.os }}${{ matrix.shard && format(' (shard {0}/3)', matrix.shard) || '' }} (push) Has been cancelled
CI / Build on Node ${{ matrix.node }} (push) Has been cancelled
CI / Style Check (push) Has been cancelled
CI / Generate Assets (push) Has been cancelled
CI / Check Python (3.14) (push) Has been cancelled
CI / Check Python (3.9) (push) Has been cancelled
CI / Build Docs (push) Has been cancelled
CI / Code Scan Action (push) Has been cancelled
CI / Site tests (push) Has been cancelled
CI / webui tests (push) Has been cancelled
CI / Run Integration Tests (push) Has been cancelled
CI / Run Smoke Tests (push) Has been cancelled
CI / Go Tests (push) Has been cancelled
CI / Share Test (push) Has been cancelled
CI / Redteam (Production API) (push) Has been cancelled
CI / Redteam (Staging API) (push) Has been cancelled
CI / GitHub Actions Lint (push) Has been cancelled
CI / Check Ruby (3.0) (push) Has been cancelled
release-please / release-please (push) Has been cancelled
release-please / build (push) Has been cancelled
release-please / publish-npm (push) Has been cancelled
release-please / publish-npm-backfill (push) Has been cancelled
release-please / docker (push) Has been cancelled
release-please / publish-code-scan-action (push) Has been cancelled
release-please / attest-code-scan-action (push) Has been cancelled
Deploy local.promptfoo.app / Deploy to Cloudflare Pages (push) Has been cancelled
Test and Publish Multi-arch Docker Image / test (push) Has been cancelled
Test and Publish Multi-arch Docker Image / build-docker-and-push-digests (map[digest-suffix:linux-amd64 platform:linux/amd64 runner:ubuntu-latest]) (push) Has been cancelled
Test and Publish Multi-arch Docker Image / build-docker-and-push-digests (map[digest-suffix:linux-arm64 platform:linux/arm64 runner:ubuntu-24.04-arm]) (push) Has been cancelled
Test and Publish Multi-arch Docker Image / merge-docker-digests (push) Has been cancelled
Test and Publish Multi-arch Docker Image / Attest Multi-arch Image (push) Has been cancelled
Validate Renovate Config / Validate Renovate Configuration (push) Has been cancelled
39 lines
1.9 KiB
Markdown
39 lines
1.9 KiB
Markdown
---
|
|
sidebar_label: LocalAI
|
|
description: 'Run self-hosted OpenAI-compatible APIs locally with LocalAI for private, offline LLM deployment and testing environments'
|
|
---
|
|
|
|
# Local AI
|
|
|
|
LocalAI is an API wrapper for open-source LLMs that is compatible with OpenAI. You can run LocalAI for compatibility with Llama, Alpaca, Vicuna, GPT4All, RedPajama, and many other models compatible with the ggml format.
|
|
|
|
View all compatible models [here](https://github.com/go-skynet/LocalAI#model-compatibility-table).
|
|
|
|
Once you have LocalAI up and running, specify one of the following based on the model you have selected:
|
|
|
|
- `localai:chat:<model name>`, which invokes models using the
|
|
[LocalAI chat completion endpoint](https://localai.io/features/text-generation/#chat-completions)
|
|
- `localai:completion:<model name>`, which invokes models using the
|
|
[LocalAI completion endpoint](https://localai.io/features/text-generation/#completions)
|
|
- `localai:<model name>`, which defaults to chat-type model
|
|
- `localai:embeddings:<model name>`, which invokes models using the
|
|
[LocalAI embeddings endpoint](https://localai.io/features/embeddings/)
|
|
|
|
The model name is typically the filename of the `.bin` file that you downloaded to set up the model in LocalAI. For example, `ggml-vic13b-uncensored-q5_1.bin`. LocalAI also has a `/models` endpoint to list models, which can be queried with `curl http://localhost:8080/v1/models`.
|
|
|
|
## Configuring parameters
|
|
|
|
You can set parameters like `temperature` and `apiBaseUrl` ([full list here](https://github.com/promptfoo/promptfoo/blob/main/src/providers/localai.ts#L16)). For example, using [LocalAI's lunademo](https://localai.io/docs/getting-started/models/):
|
|
|
|
```yaml title="promptfooconfig.yaml"
|
|
providers:
|
|
- id: localai:lunademo
|
|
config:
|
|
temperature: 0.5
|
|
```
|
|
|
|
Supported environment variables:
|
|
|
|
- `LOCALAI_BASE_URL` - defaults to `http://localhost:8080/v1`
|
|
- `REQUEST_TIMEOUT_MS` - maximum request time, in milliseconds. Defaults to 60000.
|