OpenRouter saved 22,000 users over $100,000 on GLM 5.2 inference by leveraging provider discounts during low demand periods. This resulted in a price drop of over 50% from typical weekday pricing for users.
Dynamic inference pricing is now a viable mechanism to reduce AI compute costs, shifting demand to off-peak hours and impacting cloud GPU utilization.