# SambaNova adds prompt caching for MiniMax M3 on SambaCloud

SambaNova has launched prompt caching for MiniMax M3 on SambaCloud, reducing input costs by 90% to $0.06 per million cached tokens and improving time to first token by 35% to 88% without code changes.

Language: en
Time zone: America/Los_Angeles

HTML: https://didcodexreset.com/news/cf918a32c1cf40839c80d8be.html

Content language: en
Localization state: sameLanguage

Source: SambaNova · Published 10/7/2026, 13:01:41

SambaNova has enabled prompt caching for MiniMax M3 on SambaCloud without requiring any code modifications.

According to SambaNova, reusing cached prefixes across long-context agent workflows delivers a 35% to 88% faster time to first token (TTFT), up to 4.7x overall speed improvements, and 90% lower input costs at $0.06 per million cached tokens, as outlined in its [announcement](https://bit.ly/4emDQ71).

Tags: SambaNova, MiniMax, Prompt Caching, SambaCloud

[View original post](https://x.com/SambaNovaAI/status/2107920016821702971)
