Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Microsoft Foundry blog post discusses four strategies to optimize agent costs: selecting the right model per request, caching, prompt and agent optimization, and observability. It describes capabilities like model router, deployment options, fine-tuning, and prompt optimizer to reduce cost per successful outcome.
From the source
Microsoft Foundry gives you four levers for making those tradeoffs deliberately, rather than accepting the ones your prototype happened to choose.
azure.microsoft.com