From the source
Cua launched the Cua VLM Router, a managed inference API providing unified access to multiple vision-language model providers through a single API key.
The initial release supports Anthropic's Claude Sonnet 4.5 and Haiku 4.5 models, with automatic routing to the best available provider among Anthropic, AWS Bedrock, or Microsoft Foundry.
Requests are billed in credits deducted from the Cua account balance, with each response including both the gateway cost and the upstream API cost.





