Skip to main content

What Proxium does not do

This page lists what proxium.tech does not do. For each gap, it says what to do instead, if anything.

No audio translation​

Proxium transcribes audio with /v1/audio/transcriptions, in the language of the audio. It has no route that translates audio to English: /v1/audio/translations does not exist.

Instead: send translation calls to the vendor directly. Proxium does not record their cost. See Generate images, speech and video, and transcribe audio.

Data stays until you delete it​

Proxium keeps the attempt records, the usage, the stored prompts and answers, and the memories until a person deletes them. Two things expire on their own:

  • Memory conversations and closed memories, when the project sets Keep conversations for.
  • Cached answers, after a while, or the time the project sets in Settings › Cache.

Instead: in Settings, set Stored prompts and answers to the narrowest value you need. Keep conversations for is at the top of the Memory screen. To delete data, erase one end user or erase the project. Data handling explains both.

/v1/models lists the models of the project, not of the key​

/v1/models returns every provider/model id that the project can route to. A virtual key can have its own list of allowed models. /v1/models does not apply that list.

Instead: keep the list of allowed models of each key in your application. A call with a model that the key does not allow gets 400.

No response cache on most routes​

The response cache serves /v1/chat/completions and /v1/messages only. These routes send every request to a provider:

RouteResponse cache
/v1/responsesNo
/v1/embeddings, /v1/moderations, /v1/rerankNo
/v1/images/generations, /v1/audio/speechNo
/v1/audio/transcriptionsNo
/v1/videos, /video/generations, /video/operations, /video/downloadNo

Instead: if repeated requests cost you money, send them as chat or Messages calls. The response cache explains how it works.

No way to skip the cache for one call​

No request header makes Proxium skip the cache. A project cannot turn the cache off. The answer does not say that it came from the cache.

Instead: change the body. The cache key holds every field of the body except stream, stream_options and user. A request that differs in another field misses the cache. The Cache panel on Requests counts the hits.

No spend limit for a project, a tier or a month​

A limit belongs to a key or to an app, and it counts a rolling hour, minute or day. No limit refuses the calls of a whole project, of one tier or of one model. No limit counts a calendar month.

Instead: give each key and each app its own limits, and add spend levels to get an alert for the project. Set budgets and limits shows each case.

Some routes have no /v1 form​

These routes are at the origin, https://proxium.tech, not under https://proxium.tech/v1:

RouteUse
/video/generationsStart a video
/video/operationsPoll a video
/video/downloadDownload a finished video
/status/peekRead the budget state before a call
/usage/meRead the spend of the project this month

Instead: call them with the full origin URL, for example https://proxium.tech/usage/me.

No failover and no timeout header for images and speech​

/v1/images/generations and /v1/audio/speech call one provider. They do not fail over to another provider, and they do not read x-proxium-timeout-ms.

Instead: if an image or speech call fails, retry it in your application.

Messages requests lose some fields​

Proxium translates an Anthropic Messages request to the chat format, so it can route it to any provider. It translates these fields: system, messages, tools, tool_choice, temperature, top_p, top_k, stop_sequences, stream and metadata.user_id. It drops all other fields. It also drops thinking, redacted_thinking and document blocks.

Instead: if you need extended thinking or document blocks, send the call to the vendor directly.

Memory keeps conversations only​

Memory keeps only chat, Messages and Responses calls. On /v1/embeddings, /v1/moderations and /v1/rerank, a valid x-proxium-memory value has no effect. recall puts memories into chat and Messages calls only. On /v1/responses, recall acts as write.

Instead: use /v1/chat/completions or /v1/messages when a call needs the memories of the project. Use project memory explains the modes.