ThursdAI - Dec 4, 2025 - DeepSeek V3.2, Mistral 3 Apache 2.0, OpenAI Code Red & US-Trained MOEs!
Quick Overview
The Thursday AI news roundup highlighted major open-source releases including DeepSeek V3.2 Special rivaling frontier models, Mistral 3 Apache 2.0 models with mixed reception on reasoning capabilities, and RC AI's US-trained Trinity MoE models, alongside OpenAI's reported 'Code Red' in response to competitive pressures from Gemini and other major lab developments.
Key Points: DeepSeek V3.2 Special achieved a gold medal on the Olympiad with 685 billion parameters, ranking as the second most intelligent open-weights model, even surpassing Claude 4.5 on some stats, and costs only 28 cents per million tokens on OpenRouter. Mistral released Mistral 3, including Mistral Large (675B parameters, 256k context window, 41B active parameters) and smaller multimodal models (3B, 8B, 14B), all under the Apache 2.0 license. The Mistral Large model is noted as a non-reasoning instruction model, leading to lower scores on reasoning evaluations like the Artificial Analysis intelligence score compared to reasoning models. RC AI released Trinity, a family of fully US-trained Mixture of Experts (MoE) models under the Apache 2.0 license, including Trinity Mini (26B) and Trinity Nano (6B preview), with Trinity Large (420B parameters, 13 experts) targeting mid-January 2026 release. OpenAI declared a 'Code Red' internally following a reported 6% daily active user drop after Gemini 3's launch, pausing side projects to focus on speed and personalization. Whisper Thunder, revealed to be Runway's Gen 4.5, supposedly beats V3, Sora 2 Pro, and Cling on ELO scores with superior physics but lacks native audio. Weights & Biases launched a preview LLM evaluation service allowing users to evaluate any OpenAI-compatible API hosted model directly using standard evaluation sets.
Context: The hosts Alex Volov, Wolf, Niston, and Yam Pelleg discussed the top AI releases for the week of December 4th, 2025, covering significant advancements in both open-source and closed-source large language models, as well as breaking news concerning major lab strategies and new multimodal video generation capabilities.