All tags
Person: "jensenhuang"
not much happened today
qwen-3.8-max qwen-image-3.0-pro alpamayo-2-super shieldstral pokee-isaac-28b maple-preview deepseek-v4-flash alibaba nvidia mistral-ai pokee-ai deepgrove-ai nous-research clinepass vllm_project togethercompute cognition cursor_ai deepseek ollama epoch-ai-research multimodality vision long-context model-quantization model-efficiency inference routing model-serving moe training-systems open-source cost-reduction jensenhuang skalskip92 arena thsottiaux kimmonismus andrewcurran_ tomas_hk
Alibaba launched Qwen3.8-Max, enhancing multimodal capabilities and agent ecosystem integration. NVIDIA introduced Alpamayo 2 Super for autonomous vehicle reasoning, while Mistral AI released Shieldstral, a 3B parameter open-weights safety model for on-device moderation. Pokee AI unveiled Pokee-Isaac 28B with a 10M-token context and single-GPU deployability, and DeepGrove AI presented Maple-Preview, an open-source 20B ternary-weight reasoning model optimized for Mac Mini M4. Pricing shifts, notably with Luna and DeepSeek-V4-Flash, are influencing product design and serving economics. Routing innovations like Not Diamond Code and Devin Fusion are reducing costs significantly without quality loss. Infrastructure advances include Cursor AI's open-sourced MoK megakernel for MoE training.
not much happened today
kimi-k3 moonshot vllm baseten modal together-ai ollama dell nvidia mixture-of-experts model-scaling numerical-stability model-architecture open-models model-distribution model-licensing agentic-ai vision scaling-efficiency open-source-infrastructure commercial-restrictions ai-security kimi_moonshot jensenhuang natolambert petergostev artificialanlys
Moonshot released the Kimi K3 open-weights model, a 2.8T-parameter MoE with 104B active parameters, 896 experts, and 1M-token context featuring native visual understanding. The release includes open-source infrastructure like FlashKDA, MoonEP, and AgentENV, enabling large-scale agentic post-training and serving. The technical report highlights a ~2.5× scaling-efficiency improvement over K2 with innovations in numerical stability and MoE routing. Licensing is source-available with commercial-use restrictions, signaling a trend towards open-weight models with business carve-outs. Distribution was broad and immediate via platforms like vLLM, Baseten, Modal, Together, and Ollama Cloud. Separately, NVIDIA launched the Open Secure AI Alliance to build an ecosystem combining open and closed frontier models for AI security, emphasizing defense against attackers already equipped with strong AI.