Open Weights
Curated collection of thoughts and builds centered around Open Weights.
Daily Briefings 39
- Salesforce-NVIDIA reasoning model targets enterprise tasks; Apple ships on-device Siri with Gemini
- Iris-mini and Iris-pro open-weight search agents lead benchmarks; ElevenLabs Music v2.5 reaches production
- Google TimesFM-3 forecasting model; Anthropic adds plugin evaluation framework for Claude Code
- IFM releases K2 Horizon fleet of six open-weight models (0.9B–375B); Meta FAIR introduces research preference models to rank GPU experiments
- Nvidia acquires Hugging Face for $12.9B; GPT-6 Astra benchmarks diverge while Astra shifts to on-device compute routing
- Perplexity open-sources Lily inference engine; Qwen releases local search layer
- Google's Gemini Omni 1.1 Flash extends video generation; agent sandbox pricing comparison emerges
- OpenAI's Jalapeño inference chip outperforms Nvidia on throughput and efficiency; IBM releases Granite 4.2 open-weight models
- DeepSeek V4-Flash-Vision rivals Opus 4.8 on agent benchmarks; safety testing reveals benchmark gaming
- OpenAI patches Codex file-deletion bug; Anthropic demonstrates agent-driven protein design
- Alibaba Qwen 3.8 open-weights release; GLM-5.3 claims strongest coding model via post-training
- Google Gemini 3.7 Flash undercuts predecessor 50%; DeepSeek open-sources agent harness
- SpaceXAI's Grok 4.6 matches frontier performance at lower cost; Dyna-2 scales robot learning to 1M video hours
- NVIDIA Releases Nemotron 3.5 Lightning MoE and LTX-2.5 Open Video Model
- Meta releases Muse Glimmer 30B agentic model; webAI open-sources formal-logic models for local inference
- NVIDIA Releases NemotronLabs VoiceChat 11B Open Model; ByteDance Introduces SeedRealtime Multimodal LLM
- Liquid AI Releases On-Device Agentic Model; Microsoft Open-Sources Unit-Test Agent
- Meta launches Muse Code agent for large codebases; Mistral's 3B Shieldstral matches larger safety models
- NVIDIA Releases Alpamayo 2 Super Open Vision-Language-Action Model; CopilotKit Open-Sources Channels SDK for Agent Deployment
- Y Combinator open-sources QM multiplayer agent harness; MiniMax H3 becomes first open model to top video ranking
- Alibaba Qwen3.8-Max reaches general availability; Thinking Machines releases 12B active MoE model
- AMD Open-Sources 16B MoE Model; NVIDIA Releases Molt Agentic RL Framework
- DeepSeek V4 Flash Update Matches GPT-5.6 Luna at 60% Lower Cost; Thinking Machines Releases Smaller Inkling Model
- Liquid AI releases 8K-context encoders optimized for CPU inference; Fireworks launches routing layer for open-weight coding models
- Microsoft releases MAI-Cyber-1-Flash model; Moonshot opens Kimi K3 weights and AgentENV infrastructure
- Open Dreamer Ships Dreamer 4 Reproduction; Sakana AI Releases Fugu-Cyber Orchestration Model
- Poolside releases Laguna S 2.1 coding model; Runway launches model router for generative media
- Feyn Labs ships SQRL text-to-SQL family; Moonshot maxes GPU capacity on Kimi K3 in 48 hours
- Alibaba previews 2.4T-parameter Qwen3.8-Max; open weights, benchmarks and license still unpublished
- Google updates Gemma 4 with tool-calling fixes; xAI open-sources Grok-Build after data breach
- Thinking Machines releases Inkling open model; PrismML compresses 27B reasoning model to iPhone
- PrismML ships 1-bit and ternary Qwen3.6-27B builds; Mistral's Robostral Navigate runs on one RGB camera
- China's Orca world model rivals specialized robotics; Ant ships LingBot-VA 2.0
- GPT-5.6 Sol reported near Fable 5 at a third the cost; Kyutai ships open-weight MuScriptor
- OpenAI launches three-tier GPT-5.6 family; Meta enters AI coding with Muse Spark 1.1
- Z.ai's GLM-5.2 beats Claude Code on Semgrep security benchmarks; MCP beta SDKs drop
- Z.ai's GLM-5.2 leads open-weight models with 1M context under MIT; Majors says code economics flipped
- US export controls suspend Anthropic Fable 5 and Mythos 5 for foreign nationals
- Kimi K2.7-Code open-weights land; Claude Fable 5 goes public at $10/M input tokens