Together AI's Thinking Machines Lab launched Inkling, a multimodal Mixture-of-Experts (MoE) model designed for token-efficient reasoning, supporting native text, image, and audio inputs. The model features controllable inference effort and is optimized with Together's FlashAttention-4-based kernel.