A catalog of AI models in Microsoft Foundry that you can discover, compare, and deploy using Azure’s built‑in tools for evaluation, fine‑tuning, and inference
Several Microsoft Foundry Issues - Cache, open-source, pricing
Hi,
I am writing after working for a long time now with many Azure Foundry models and have come to the conclusion they are not reliable and difficult to work with.
Cache issues and lack of transparency
The only model whose cache seems to be working fine are OpenAI models, and maybe Claude. For the rest, you cannot rely on a proper cache. While I was getting 90% cache hits on GPT I was getting ~50-60% on DeepSeek. This only happening last month as DeepSeek had no cache at all while the main provider had it implemented for a long time now.
Pricing not updated
For every price change the main providers do, we have to wait for a long time to see that change reflected on our bills. While GPT 5.6 Luna price reductions happened the 30th of July we are still being billed the old pricing. Some examples of people complaining.
- https://www.reddit.com/r/AZURE/comments/1vkrsj0/price_of_gpt56luna/
- https://learn.microsoft.com/en-us/answers/questions/5972276/pricing-gpt-5-6-luna-modle
The main Microsoft blog mentions the prices are updated https://azure.microsoft.com/en-us/blog/gpt-5-6-now-available-in-microsoft-foundry/ but they are neither reflected on the Azure Foundry pricing page https://azure.microsoft.com/en-us/pricing/details/azure-openai/ nor our bills.
Why is this taking so long ? All other AI providers updated this in no more than 1 day...
Open-source models we don't even get to try
Many people use Foundry models to keep their data somehow private within their own Microsoft Tenant. That means they use Foundry as their main source of AI in their organisations. However, Foundry seems to have delegated providing the latest open-source models purely to Fireworks. If I use Fireworks my data leaves Microsoft and other data policies apply. So in terms of privacy it makes no sense to use those models compared to using other providers such as Openrouter or Opencode.
- The latest DeepSeek V4 Flash update is still not present while, again, other providers are giving access at very little pricing, nearly instantly and with no cache issues.
- How much time will we have to wait to get accesss to the new DeepSeek V4 Pro if we do not even have the new Flash ?
- We have been waiting for Kimi K3 for a long time and is still not available, only through Fireworks...
- Today GLM 5.3 released an I am not expecting to even see that model in Foundry.
Amount of time to update things
I am very worried Azure Foundry falls apart due to how long everything takes to be part of the speed the market is in. Only models that are deployed quickly are OpenAI's and Anthropic's but in terms of new models. Not for pricing as we could see.
Question
So my questions here are:
- Am I right with my assumptions ?
- Why is all of this happening ?
- Is this a global feeling ?
- Will this continue to be like this for a long time ?