All tags
Topic: "open-datasets"
not much happened today
flux-3 flux-mimic qwen-audio-3.0-tts hugging-face black-forest-labs mimicrobotics alibaba open-datasets code-datasets distillation multimodality robotics video-modeling audio-generation tts model-training model-architecture anton_lozhkov loubnabenallal1 lvwerra eliebakouch gergelyorosz schmidhuberai suhail garrytan bfl_ai hila_chefer robrombach mimicrobotics generalistai alibaba_qwen
The Stack v3 is released as the largest open code dataset with 114 TB raw data, 224M repositories, and 5T deduplicated tokens, significantly expanding data for open code models and cyber-defense. The debate on distillation continues as a key ideological fault line, with calls for stronger investment in open-weight domestic models. Black Forest Labs launched FLUX 3, a unified multimodal model covering image, video, audio, and action prediction, with robotics transfer demonstrated by FLUX-mimic for general-purpose dexterity on a single GPU. Alibaba introduced Qwen-Audio-3.0-TTS supporting 16 languages and advanced control features, claiming the top spot on the Artificial Analysis TTS leaderboard.
Not much happened today
claude-3 claude-3-opus claude-3-sonnet gpt-4 gemma-2b anthropic perplexity langchain llamaindex cohere accenture mistral-ai snowflake together-ai hugging-face european-space-agency google gpt4all multimodality instruction-following out-of-distribution-reasoning robustness enterprise-ai cloud-infrastructure open-datasets model-deployment model-discoverability generative-ai image-generation
Anthropic released Claude 3, replacing Claude 2.1 as the default on Perplexity AI, with Claude 3 Opus surpassing GPT-4 in capability. Debate continues on whether Claude 3's performance stems from emergent properties or pattern matching. LangChain and LlamaIndex added support for Claude 3 enabling multimodal and tool-augmented applications. Despite progress, current models still face challenges in out-of-distribution reasoning and robustness. Cohere partnered with Accenture for enterprise AI search, while Mistral AI and Snowflake collaborate to provide LLMs on Snowflake's platform. Together AI Research integrates Deepspeed innovations to accelerate generative AI infrastructure. Hugging Face and the European Space Agency released a large earth observation dataset, and Google open sourced Gemma 2B, optimized for smartphones via the MLC-LLM project. GPT4All improved model discoverability for open models. The AI community balances excitement over new models with concerns about limitations and robustness, alongside growing enterprise adoption and open-source contributions. Memes and humor continue to provide social commentary.