google Google Enters Reasoning AI Race With Experimental Gemini Model Google introduces Gemini 2.0 Flash Thinking Experimental, a new AI model designed to work through complex problems by fact-checking itself. However, early testing shows the experimental system still struggles with basic tasks like counting letters.
ai agents Tech Industry Struggles to Define the Emerging AI Agent AI agents represent the next major shift in artificial intelligence, yet tech companies and experts lack a unified definition. Despite varying interpretations, these tools consistently aim to automate complex, multi-step tasks with minimal human intervention.
ai AI World Models Aim to Give Machines Human-Like Understanding Tech companies are pouring millions into "world models," AI systems that simulate physical reality to predict and generate more accurate video. These models attempt to mimic human subconscious reasoning to overcome the limitations of current generative AI.
ai AI Industry Embraces Reasoning Models Amid Rising Costs and Skepticism Following the launch of OpenAI's o1, rival AI labs are rapidly releasing their own reasoning algorithms as traditional scaling methods show diminishing returns. However, experts warn that these powerful models are incredibly expensive and heavily fueled by corporate marketing hype.
google Google DeepMind AI Model Beats Top Global Weather Forecasting System Google's DeepMind introduces GenCast, an AI weather prediction model that outperforms the leading global forecasting system 97.2% of the time. The tech giant plans to integrate the model into Search and Maps while releasing its data for public research.
meta Meta Releases Cost-Efficient Llama 3.3 Model to Rival Top AI Systems Meta unveils Llama 3.3 70B, a highly efficient AI model that matches the performance of its massive 405B model at a fraction of the cost. The release boosts Meta's push to dominate the AI assistant market alongside growing regulatory challenges.
openai OpenAI Unveils $200 ChatGPT Pro Tier With Advanced Reasoning Models OpenAI introduces a premium $200-per-month ChatGPT Pro subscription targeting power users with unlimited access to its latest reasoning models. The plan features an enhanced "o1 pro mode" that uses extra computing power to tackle complex math, coding, and visual tasks.
google Google Releases Open AI Models With Emotion Detection, Alarming Experts Google launches the PaliGemma 2 family of open AI models that claim to identify emotions in images. Critics warn the technology relies on flawed science and risks spreading harmful biases.
amazon Amazon Unveils Nova, a New Family of Multimodal AI Models AWS introduces Nova, a new suite of generative AI models that process text, images, and video across four tiers of capability. The models promise fast performance, low costs, and massive context windows for cloud customers.
alibaba Alibaba Launches Open Source Reasoning Model to Rival OpenAI's O1 Alibaba introduces QwQ-32B-Preview, an open-source reasoning AI that outperforms OpenAI's o1 on key math benchmarks. The 32.5-billion-parameter model self-fact-checks but exhibits Chinese political censorship.
uber Uber Launches Gig Worker Division to Label AI Training Data Uber creates a new division called Scaled Solutions that hires gig workers to label data for artificial intelligence models. The company serves both its internal units and external clients like Aurora Innovation and Niantic.
ai2 Ai2 Unveils OLMo 2, Truly Open AI Models Rivaling Meta's Llama Ai2 releases OLMo 2, a new family of fully open-source language models that meet strict open source definitions while competing with Meta's Llama 3.1 in performance.
openai OpenAI Backs Duke University Research to Predict Human Moral Judgments OpenAI provides a $1 million grant to Duke University researchers to develop algorithms that predict human moral judgments in complex scenarios. The project raises significant questions about whether artificial intelligence can accurately navigate the nuances of ethics.
ai2 Ai2 Releases Tulu 3 to Democratize AI Post-Training Process Ai2 introduces Tulu 3, an open-source post-training regimen that allows developers to shape raw language models into usable tools. This release helps bridge the secrecy gap between independent AI creators and large tech companies.
factiverse Norwegian Startup Factiverse Uses AI to Combat Rising Online Disinformation Factiverse develops a B2B tool that provides live fact-checking for text, video, and audio to combat the rapid spread of AI-generated disinformation. The startup relies on curated data rather than standard large language models to verify claims in real time.