All tags
Model: "qwen3.8-max"
not much happened today
grok-4.6 grok-4.7 qwen3.8-max deepseek-v4-pro mai-thinking-1 solar-pro-4 xai alibaba deepseek microsoft upstage agentic-ai intelligence-index model-training open-weights long-context reasoning pricing reinforcement-learning tool-use pawelhuryn kimmonismus mustafasuleyman elonmusk yuchenjin finbarrtimbers
xAI's Grok 4.6 advances frontier pricing and performance, scoring 61 on the Intelligence Index and showing strong agentic results, with Grok 4.7 already in training. Alibaba's Qwen3.8-Max open weights release features a 2.4T parameter model with 95B active MoE, notable for day-0 serving and long-context capabilities but initially text-only. DeepSeek V4 Pro GA offers significant cost advantages, priced at $0.435/M input tokens, with mixed capability reviews. Microsoft's MAI-Thinking-1 debuts as a practical reasoning model focused on tool use, available in Foundry. Upstage's Solar Pro 4 improved its Intelligence Index ranking from 14 to 42.
Qwen 3.8 Max
qwen3.8-max qwen3.8-27b kimi-k3 deepseek-v4-flash claude-opus-4.7 alibaba deepseek databricks multimodality model-quantization model-performance benchmarking reinforcement-learning model-deployment cost-efficiency inference-speed model-optimization agent-models alibaba_qwen zhihufrontier jaminball kimmonismus jonathanross321 _micah_h clementdelangue tonychenxyz yuchenj_uw casper_hansen_ htihle skalskip92
Alibaba launched Qwen3.8-Max, a 2.4T-parameter open-weight model emphasizing autonomous coding, long-horizon execution, and multimodal feedback, with aggressive pricing. Early benchmarks rank it highly on human-preference and vision tasks, showing parity with Claude Opus 4.7 and strong object-detection capabilities. However, operational demands remain high, especially for large MoE models like Qwen3.8-Max and Kimi K3, highlighting the strategic importance of smaller open models like the upcoming 27B variant. The open-weight frontier is increasingly led by Chinese labs including Kimi, DeepSeek, GLM, and MiniMax, narrowing the gap with US labs. DeepSeek V4 Flash is noted as a cost/performance disruptor in agent models. "Chinese labs are setting the pace in open models" and "inference provider materially changed leaderboard outcomes" are key insights from the community.