nai-degen
|
ff0d3dfdcd
|
prevents overwriting anthropic-version header if it's already provided
|
2024-09-19 00:55:17 -05:00 |
|
nai-degen
|
81a3ae1746
|
maybe fixes missing anthropic version header in some cases
|
2024-09-19 00:50:17 -05:00 |
|
nai-degen
|
4dfd57fcb4
|
updates render dockerfile to correctly copy patches dir into build context
|
2024-09-16 23:39:43 -05:00 |
|
khanon
|
d21e274358
|
Add configurable network interface or SOCKS/HTTP proxy for outgoing requests (khanon/oai-reverse-proxy!80)
|
2024-09-16 15:17:57 +00:00 |
|
nai-degen
|
7a4a16dd2f
|
fixes chatgpt-latest missing from models endpoint
|
2024-09-15 06:02:35 -05:00 |
|
nai-degen
|
f1cfa644c5
|
maybe fixes openai sk-svcacct keys
|
2024-09-13 00:55:29 -05:00 |
|
nai-degen
|
6a908b09cb
|
adds preliminary openai o1 support and some improvements to openai keychecker
|
2024-09-12 23:03:33 -05:00 |
|
honeytree
|
bd87ca60f7
|
Implement priority queue by tokens (khanon/oai-reverse-proxy!79)
|
2024-09-09 16:48:46 +00:00 |
|
nai-degen
|
ac1897fd17
|
returns more clear proxy_note hint on AWS 503 error
|
2024-09-09 09:56:18 -05:00 |
|
nai-degen
|
2a6f85e2e2
|
Revert "handles AWS HTTP 503 ServiceUnavailableException similarly to 429s"
This reverts commit ffcaa23511.
|
2024-09-09 09:43:59 -05:00 |
|
nai-degen
|
ffcaa23511
|
handles AWS HTTP 503 ServiceUnavailableException similarly to 429s
|
2024-09-09 08:07:53 -05:00 |
|
khanon
|
96fe974ad0
|
Use AWS Inference Profiles for higher rate limits (khanon/oai-reverse-proxy!78)
|
2024-09-01 22:55:07 +00:00 |
|
nai-degen
|
ee61f9be2b
|
removes unnecessary log from last commit
|
2024-08-27 23:58:32 -05:00 |
|
nai-degen
|
0c448cb59d
|
fixes azure dalle using wrong rate limit and out-of-spec Retry-After header
|
2024-08-27 23:53:28 -05:00 |
|
nai-degen
|
51a9ccceb2
|
supports alternate claude system prompt format
|
2024-08-27 23:27:20 -05:00 |
|
nai-degen
|
5000e59a61
|
fix for google makersuite prompt validation/transformation
|
2024-08-22 14:19:48 -05:00 |
|
nai-degen
|
d54acad6ad
|
adds support for sonnet 8192 output tokens on anthropic api
|
2024-08-15 11:55:13 -05:00 |
|
nai-degen
|
5e1fffe07d
|
adds chatgpt-4o-latest
|
2024-08-15 11:54:42 -05:00 |
|
nai-degen
|
6d323f6ea1
|
do not transform mistral chat prompts to text when using la plateforme
|
2024-08-14 12:26:27 -05:00 |
|
nai-degen
|
b58e7cb830
|
always applies Mistral prompt fixes on messages input
|
2024-08-14 10:48:55 -05:00 |
|
khanon
|
f531272b00
|
Refactor AWS service code and add AWS Mistral support (khanon/oai-reverse-proxy!75)
|
2024-08-14 04:40:41 +00:00 |
|
nai-degen
|
b7cd326d2a
|
handles 'invalid subscription' 403 errors from Mistral API
|
2024-08-07 14:14:53 -05:00 |
|
nai-degen
|
6c9f302fb9
|
minor gultra fix
|
2024-08-06 18:46:49 -05:00 |
|
nai-degen
|
9ab1e7d0ce
|
adds new gpt4o id
|
2024-08-06 13:08:25 -05:00 |
|
nai-degen
|
81f8dc2613
|
updates README.md
|
2024-08-05 11:33:16 -05:00 |
|
khanon
|
0c936e97fe
|
Merge GCP Vertex AI implementation from cg-dot/oai-reverse-proxy (khanon/oai-reverse-proxy!72)
|
2024-08-05 14:27:51 +00:00 |
|
nai-degen
|
29ed07492e
|
fixes info page display for gemini flash/ultra
|
2024-08-03 22:18:05 -05:00 |
|
nai-degen
|
2f7315379c
|
adds gemini/makersuite keychecker, native endpoint, and streaming fixes
|
2024-08-03 21:53:32 -05:00 |
|
nai-degen
|
e91532f4f7
|
handle dead makersuite keys triggering 400 error instead of 401/403
|
2024-08-03 19:09:50 -05:00 |
|
nai-degen
|
9a3cca6b80
|
adds new mistral models and updates older model lists/context limits
|
2024-07-28 13:15:03 -05:00 |
|
nai-degen
|
f242777596
|
fixes token index used as msg idx in anthropic chat-to-openai SSE transformer
|
2024-07-07 13:33:33 -05:00 |
|
nai-degen
|
edc0d094e2
|
tries to disable quarantined aws keys
|
2024-06-30 05:08:27 -05:00 |
|
nai-degen
|
994b30dcce
|
adjusts gemini pro model assignment
|
2024-06-26 13:37:23 -05:00 |
|
nai-degen
|
b4fb97ca5c
|
fixes model id typo
|
2024-06-20 10:42:48 -05:00 |
|
nai-degen
|
eb700d3da6
|
adds untested claude 3.5 model ids and model assignment
|
2024-06-20 10:34:48 -05:00 |
|
nai-degen
|
d706d4c59d
|
adds USER_CONCURRENCY_LIMIT environment variable
|
2024-06-14 22:52:16 -05:00 |
|
nai-degen
|
7660ed8b94
|
allows enabling vision prompts on a per-service basis
|
2024-06-07 12:09:43 -05:00 |
|
nai-degen
|
57fd17ede0
|
makes it easier for clients to detect proxy errors programatically
|
2024-05-27 15:30:28 -05:00 |
|
nai-degen
|
9d00b8a9de
|
adjusts max IP error message wording
|
2024-05-27 08:24:56 -05:00 |
|
scrappyanon
|
2d82e55d72
|
Sqlite backend with user event logging (khanon/oai-reverse-proxy!69)
|
2024-05-26 17:31:12 +00:00 |
|
nai-degen
|
68b48428de
|
adjusts gatekeeper module to send auth errors as fake chat completions
|
2024-05-21 12:44:43 -05:00 |
|
nai-degen
|
6dabc82bcf
|
adds preliminary gpt4o
|
2024-05-13 12:43:39 -05:00 |
|
nai-degen
|
d3e7ef3c14
|
prevents leaking headers to upstream API when serving via Tailscale
|
2024-05-01 11:26:15 -05:00 |
|
nai-degen
|
32b623d6bc
|
partial googleai fixes; adds jsonl file backend for promptlogger stolen from fiz
|
2024-04-23 03:43:38 -05:00 |
|
nai-degen
|
c15f07c0d8
|
adds OpenAI-to-AWS Claude3 compat endpoint
|
2024-04-17 21:23:30 -05:00 |
|
nai-degen
|
db28e90c51
|
adds proper Opus model check to aws claude keychecker
|
2024-04-17 21:09:00 -05:00 |
|
nai-degen
|
c0cd2c7549
|
adds aws opus maybe, idk cannot test
|
2024-04-16 11:33:44 -05:00 |
|
nai-degen
|
9445110727
|
adds gpt-4-turbo stable
|
2024-04-09 16:31:42 -05:00 |
|
nai-degen
|
34a673a80a
|
adds option to disable multimodal prompts
|
2024-03-23 14:30:14 -05:00 |
|
nai-degen
|
8cb960e174
|
fixes incorrect model assignment when requesting Haiku from AWS
|
2024-03-21 23:21:27 -05:00 |
|