All tags
Model: "inkling-small"
not much happened today
gpt-5.6-luna gpt-5.6-terra gpt-5.6-sol arc-agi-3 inkling-small inkling gemini-robotics-2 openai thinking-machines lmsys modal unsloth artificial-analysis google price-optimization agent-systems memory-retention context-compaction multimodality mixture-of-experts model-compression benchmarking open-weights multimodal-models model-efficiency model-deployment embodied-ai robotics long-context sama fchollet kimmonismus gneubig scaling01 mervenoyann
OpenAI aggressively cut prices for GPT-5.6 Luna by 80% and Terra by 20%, introducing a faster Sol Fast tier with up to 2.5× lower latency at double the price, improving agent workflow costs by roughly 10×. The ARC-AGI-3 debate highlighted that the complete agent system, including memory retention and tool orchestration, is critical beyond just the base model. Thinking Machines released Inkling-Small, an open-weights, multimodal MoE model with 276B parameters (12B active), delivering performance comparable to the original Inkling at a quarter of the size, supporting audio, images, and Python-based image inspection. Benchmarks show Inkling-Small excels in coding and multimodality tasks, with 1M-context support and broad open inference stack adoption. The news also mentions Google's Gemini Robotics 2 advancing embodied AI from tabletop to full-body control.