deepseek DeepSeek Unveils V3.2 Models With Record-Breaking Math Reasoning DeepSeek releases two new AI models that offer advanced reasoning and cost efficiency, alongside a math model that outperforms top human competitors.
deepseek DeepSeek Launches V3.1 Model with Hybrid Think and Non-Think Modes DeepSeek releases V3.1, introducing a hybrid inference system that lets users toggle between thinking and non-thinking modes. The update brings stronger agent capabilities, Anthropic API support, and new open-source weights.
meta Meta Unveils Llama 4 Herd With Native Multimodal Capabilities Meta launches Llama 4 Scout and Maverick as the first open-weight natively multimodal models using a mixture-of-experts architecture. The new models deliver industry-leading context windows and outperform major competitors across various benchmarks.
tencent Tencent Unveils Hunyuan-T1, the First Mamba-Powered Ultra-Large Reasoning Model Tencent releases the official version of its Hunyuan-T1, a cutting-edge reasoning model built on the world's first ultra-large Hybrid-Transformer-Mamba MoE architecture. The model leverages massive reinforcement learning to deliver top-tier performance with twice the decoding speed.
anthropic Anthropic Unveils Claude 3.7 Sonnet as First Hybrid Reasoning Model Anthropic releases Claude 3.7 Sonnet, a hybrid model that seamlessly switches between instant answers and visible step-by-step reasoning. The company also introduces Claude Code, an agentic terminal tool for developers.
deepseek DeepSeek Launches Open-Source R1 Model to Rival OpenAI-o1 DeepSeek releases DeepSeek-R1, a highly capable open-source reasoning model that matches OpenAI-o1 performance. The model and its API are available now under a permissive MIT license.
codeium AI Coding Platform Codeium Reaches $1.25 Billion Valuation in $150M Raise Codeium secures $150 million in a Series C funding round led by General Catalyst, pushing its valuation to $1.25 billion. The company plans to use the new capital to expand its workforce, accelerate product development, and build strategic partnerships.
google Google AI Overviews Evolve From SGE Into Conversational Search Experience Google's AI Overviews now use Gemini 3 to generate concise, formatted summaries at the top of search results. Recent updates integrate these overviews more closely with the interactive AI Mode interface.
aws Mistral Large 2 Arrives in Amazon Bedrock With Enhanced Reasoning Mistral AI's latest foundation model, Mistral Large 2, is now generally available in Amazon Bedrock, offering improved multilingual support, coding, and reduced hallucinations.
mistral ai Mistral Large 2 Delivers Frontier Performance With Reduced Hallucinations Mistral AI releases Mistral Large 2, a 123-billion-parameter model that rivals top competitors in coding and reasoning while significantly reducing hallucinations. The model features a massive 128k context window and supports dozens of human and programming languages.
openai OpenAI Cuts Off API Access for Developers in China and Hong Kong OpenAI is blocking mainland China and Hong Kong-based developers from accessing its AI services through APIs starting July 9. The restriction forces affected companies to abandon OpenAI's models and shift to domestic alternatives.
meta Meta Unveils Llama 3 as the Most Capable Open Source Language Model Meta releases Llama 3, an advanced open source large language model available in 8B and 70B parameter sizes. The new model achieves state-of-the-art performance and powers the upgraded Meta AI assistant.
xai xAI Releases Grok-1 Open-Weights Model for Public Download xAI officially releases the Grok-1 open-weights model, a massive 314-billion parameter language model, on GitHub. Developers need significant GPU memory to run the Mixture-of-Experts architecture locally.
xai xAI Releases Massive 314 Billion Parameter Grok-1 Model xAI makes its Grok-1 open-weights model available on Hugging Face, offering a massive 314 billion parameter language model under the Apache 2.0 license. Users need a multi-GPU setup to run the model locally.
anthropic Anthropic Launches Claude 3 Models to Challenge GPT-4 Dominance Anthropic introduces the Claude 3 model family, featuring three tiers of AI models that claim to outperform the original GPT-4 on various cognitive benchmarks. However, important caveats exist regarding these performance comparisons against newer GPT-4 Turbo versions.