Mira Murati's Thinking Machines Lab released the open-weight AI model Inkling under an Apache 2.0 license. LMCache achieved up to a 10.7x speedup on LLM inference by reusing expensive parts of long prompts. Meta's Muse Spark 1.1 is pressuring competitors with cut-rate pricing in agentic coding.