Changing ratelimit
This commit is contained in:
@@ -35,7 +35,7 @@ npm run keys:env
|
||||
|
||||
The global context limit defaults to 256k tokens and can be changed with `MAX_CONTEXT_TOKENS`.
|
||||
|
||||
Inference requests are limited per client IP to 2 requests per second, 100 requests per five hours, and $10 of reported upstream cost per five hours. Railway's `X-Real-IP` header is used to identify clients.
|
||||
Inference requests are limited per client IP to 2 requests per second, 100 requests per five hours, and $15 of reported upstream cost per five hours. Railway's `X-Real-IP` header is used to identify clients.
|
||||
|
||||
To exempt trusted clients from those proxy limits, set `UNLIMITED_API_KEYS` to a JSON array and have the client send a configured key using `Authorization: Bearer <key>` or `x-api-key`. These keys only bypass this proxy's rate and spend limits; they do not bypass upstream OpenCode Go limits.
|
||||
|
||||
|
||||
Reference in New Issue
Block a user