All tags
Person: "stochasticchasm"
not much happened today
gemini-3.7-flash google-deepmind google deepseek arcee agentic-workflows coding knowledge-work benchmarking runtime-systems open-source long-running-processes asynchronous-computation software-architecture developer-tools price-performance _philschmid koraykv officiallogank tianyi eliebakouch bookwormengr 0xlogicrw teortaxestex latkins stochasticchasm code_star fujikanaeda
Google rapidly released Gemini 3.7 Flash just three weeks after 3.6 Flash, targeting coding, web development, knowledge work, and agentic workflows with a 50% introductory price cut and improved benchmark scores like DeepSWE 65.3% and Code Arena Elo 1588. The update quickly integrated across multiple platforms including Gemini API and Android Studio, with independent benchmarks confirming performance gains. Meanwhile, DeepSeek open-sourced DeepSeek Harness under MIT license as a developer preview, focusing on architecture innovations like KV-cache-aware append-only history semantics and treating the harness as an OS/runtime substrate for recursive improvement. Arcee also open-sourced NAC under Apache 2.0, designed for long-running asynchronous tasks and powering significant code pipelines, enabling orchestration from phones or delegation via Codex/Claude.
not much happened today
gemma-4 google huggingface intel ollama unsloth reasoning agentic-workflows multimodality on-device-ai local-inference model-benchmarking moe vision audio-processing memory-optimization open-source model-performance fchollet demishassabis clementdelangue quixiai googlegemma ggerganov osanseviero maartengr basecampbernie prince_canuma measure_plan kimmonismus anemll arena stochasticchasm reach_vb zeneca everlier erick_lindberg_ anomalistg
Gemma 4 was launched by Google under an Apache 2.0 license, marking a significant open-model release focused on reasoning, agentic workflows, multimodality, and on-device use. It outperforms models 10x larger and has immediate ecosystem support including vLLM, llama.cpp, Ollama, Intel hardware, Unsloth, and Hugging Face Inference Endpoints. Local inference benchmarks showed strong performance on consumer hardware, including RTX 4090 and Mac mini M4. Early benchmarking praised its efficiency and ranking improvements over previous versions. Meanwhile, Hermes Agent emerged as a popular open-source agent harness, noted for stability and capability on long tasks, with users switching from OpenClaw to Hermes.