All tags
Topic: "api-pricing"
not much happened today
qwen3.8-27b deepseek-v4-pro gpt-5.6-luna openai nvidia stripe openrouter vercel cursor langchain vanta deepseek ai-infrastructure power-management model-routing api-pricing developer-platforms agentic-coding multi-agent-systems evaluation-tools harness-level-evaluation sandboxing permissioning model-compression local-models markchen90 kimmonismus hamelhusain tonbistudio teknium omarsar0 cline
OpenAI is advancing its power-and-compute infrastructure with a 4+ GW NVIDIA capacity commitment and an 8 GW Ohio campus buildout through 2032, emphasizing vertical integration across power, data centers, and chips. The model access and routing API layer is becoming a competitive pricing battlefield, highlighted by the Stripe–OpenRouter deal and recent price cuts by OpenRouter and Vercel. Cursor launched Origin, an AI-native IDE aiming for full control over coding workflows, signaling a shift toward agentic coding platforms. Multi-agent orchestration is evolving from demos to operational patterns with specialized, persistent-context agents, as seen in projects by Hermes Desktop, Bot Mode, and Codex orchestration. Evaluation tools like Hamel Husain’s eval-skills plugin and Agent Arena are advancing harness-level measurement with data from over 1.7M sessions. Enterprise agent tooling is improving with sandboxed, permissioned execution environments from Vanta and LangChain. Open models like Qwen3.8-27B are compressing the capability frontier, reaching performance comparable to DeepSeek V4-Pro and GPT-5.6 Luna on the Artificial Analysis Intelligence Index, marking a milestone for local models.
MiniMax M2 230BA10B — 8% of Claude Sonnet's price, ~2x faster, new SOTA open model
minimax-m2 hailuo-ai huggingface baseten vllm modelscope openrouter cline sparse-moe model-benchmarking model-architecture instruction-following tool-use api-pricing model-deployment performance-evaluation full-attention qk-norm gqa rope reach_vb artificialanlys akhaliq eliebakouch grad62304977 yifan_zhang_ zpysky1125
MiniMax M2, an open-weight sparse MoE model by Hailuo AI, launches with ≈200–230B parameters and 10B active parameters, offering strong performance near frontier closed models and ranking #5 overall on the Artificial Analysis Intelligence Index v3.0. It supports coding and agent tasks, is licensed under MIT, and is available via API at competitive pricing. The architecture uses full attention, QK-Norm, GQA, partial RoPE, and sigmoid routing, with day-0 support in vLLM and deployment on platforms like Hugging Face and Baseten. Despite verbosity and no tech report, it marks a significant win for open models.
not much happened today
veo-3 deepseek-r1t2 deepseek-tng-r1t2-chimera o3-deep-research o4-mini-deep-research deepswe-agent safe-superintelligence-inc perplexity-ai meta-ai-fair midjourney sakana-ai cohere google-deepmind deepseek openai together-ai video-generation assembly-of-experts model-licenses api-pricing research-roles product-expansion corporate-leadership model-release team-expansion ilya_sutskever daniel_levy daniel_gross aravsrinivas zeyuanallenzhu nat_friedman davidsholz fp_champagne demishassabis reach_vb
Ilya Sutskever confirmed his role as CEO of Safe Superintelligence Inc. (SSI) with Daniel Levy as President, dismissing acquisition rumors and emphasizing their strong team and compute resources. Perplexity AI expanded its data integrations by adding Morningstar's financial research and hinted at new product features for Pro users. Meta AI FAIR clarified its research structure, distinguishing its small lab from larger model training groups, and welcomed Nat Friedman to enhance AI product development. Midjourney and Sakana AI announced hiring for research and applied engineering roles. Cohere expanded its presence in Montréal, receiving praise from Canadian officials. On the model front, Google DeepMind's Gemini Pro released the Veo 3 video generation model globally. DeepSeek launched the faster DeepSeek R1T2 model using an Assembly of Experts approach, available under an MIT license. Kling AI showcased cinematic video generation capabilities. OpenAI introduced a high-cost Deep Research API with pricing up to $30 per call. Together AI announced the release of the DeepSWE agent.