Replace the direct Anthropic and OpenAI integrations with a single provider that talks to Switchboard, an OpenAI-compatible gateway that routes each request to the best available model. The app no longer pins a model id anywhere: it sends switchboard/auto and lets the gateway choose, then logs which model answered and what it cost. Routing levers are set per feature in src/lib/ai/routing.ts. Three of those choices came from measuring against the live gateway: - category and prefer_free are set explicitly on every request. An API key carries its own routing defaults, and anything left unset inherits them - drink prompts were being sent to a free coding model. - Token budgets are generous because the router may pick a reasoning model, and reasoning tokens come out of the same max_tokens budget as the answer. At 512 tokens a request returned null content; at 4096 the same request returned correct JSON. - No tier lever on text features. tier "cheap" pinned a slow reasoning model (42-180s, two timeouts and one truncated response in five trials) and tier "frontier" escalated as far as Opus at $0.02 a call, while unconstrained routing answered in about a second. Vision keeps "frontier", where the accuracy is worth a few tenths of a cent. Gateway failures are mapped to actionable messages rather than passed through: a 401 relayed as 401 would read as an expired session and bounce the user to login, and a 429 would collide with the app's own rate limiter. Also collapses the key lookup that was duplicated across ten call sites into getUserProvider(), which fixes a latent bug where a bare findFirst with no ordering let different features pick different providers. Existing claude/openai key rows are ignored at runtime and offered for removal in Settings, so no migration is needed before deploying. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
33 lines
1.0 KiB
Plaintext
33 lines
1.0 KiB
Plaintext
# Database
|
|
DATABASE_URL="postgresql://drinktracker:YOUR_PASSWORD@localhost:5432/drinktracker"
|
|
POSTGRES_USER="drinktracker"
|
|
POSTGRES_PASSWORD="YOUR_PASSWORD"
|
|
POSTGRES_DB="drinktracker"
|
|
|
|
# NextAuth
|
|
NEXTAUTH_URL="http://localhost:3000"
|
|
NEXTAUTH_SECRET="generate-with: openssl rand -base64 32"
|
|
AUTH_TRUST_HOST="true" # Set to true when behind a reverse proxy
|
|
|
|
# OAuth Providers
|
|
GOOGLE_CLIENT_ID=""
|
|
GOOGLE_CLIENT_SECRET=""
|
|
GITHUB_CLIENT_ID=""
|
|
GITHUB_CLIENT_SECRET=""
|
|
|
|
# MinIO / S3-compatible storage
|
|
MINIO_ENDPOINT="localhost"
|
|
MINIO_PORT="9000"
|
|
MINIO_ACCESS_KEY="generate-a-strong-access-key"
|
|
MINIO_SECRET_KEY="generate-a-strong-secret-key"
|
|
MINIO_BUCKET="drink-images"
|
|
MINIO_USE_SSL="false"
|
|
|
|
# Encryption (for API key storage)
|
|
ENCRYPTION_KEY="generate-with: openssl rand -hex 32"
|
|
|
|
# AI Gateway (Switchboard)
|
|
# OpenAI-compatible router that picks the best model per request. LAN-only, plain HTTP.
|
|
# Each user adds their own gateway API key in Settings; this is only the endpoint.
|
|
SWITCHBOARD_BASE_URL="http://192.168.2.11:8787/v1"
|