google Google Unveils PaLM, a 540-Billion Parameter Language Model Google introduces PaLM, a massive 540-billion parameter language model trained using its new Pathways system. The model achieves breakthrough few-shot learning results and even outperforms average human performance on certain reasoning benchmarks.
google Google Unveils PaLM: 540-Billion Parameter Language Model Powered by Pathways Google introduces PaLM, a massive 540-billion parameter language model trained using its new Pathways system. The model achieves state-of-the-art few-shot performance across hundreds of tasks by efficiently scaling across thousands of TPU v4 chips.
google Google Unveils PaLM, a 540-Billion Parameter Language Model Google Research introduces PaLM, a massive 540-billion parameter language model trained using its new Pathways system. The model achieves state-of-the-art few-shot performance across hundreds of tasks.
microsoft Microsoft Expands Access to Cloud-Based GPT-3 AI Model for Developers Microsoft invites corporate developers to apply for its Azure OpenAI Service, moving away from a strict invitation-only policy. The cloud-based service utilizes the powerful GPT-3 language model to generate highly realistic human-like text.
sandboxaq SandboxAQ Raises $500 Million to Tackle Quantum Security Threats Alphabet spinoff SandboxAQ secures $500 million to build commercial products that combine artificial intelligence and quantum computing. The startup focuses heavily on helping enterprises and governments upgrade cryptographic defenses against future quantum attacks.
sandbox aq Alphabet AI and Quantum Spinoff Sandbox AQ Secures $500 Million Sandbox AQ, an independent company spun off from Google parent Alphabet, raises $500 million to advance its artificial intelligence and quantum computing technologies.
alphacode DeepMind's AlphaCode Achieves Median Human Level in Coding Competitions DeepMind introduces AlphaCode, an AI system that writes computer programs at a competitive level, ranking within the top 54% of participants on Codeforces. The system uses massive-scale code generation and smart filtering to solve novel programming problems.
nvidia Nvidia Unveils Hopper H100 Datacenter GPU With 80 Billion Transistors Nvidia introduces the Hopper H100 GPU for datacenters, packing 80 billion transistors on a custom TSMC 4N process to outperform the previous A100. The new chip focuses heavily on AI and supercomputing workloads with massive upgrades to memory and interconnect bandwidth.
ai Study Reveals Large Language Models Require Significantly More Training Data Researchers find that current large language models are severely undertrained, proving that model size and training data must scale equally for optimal performance. The resulting Chinchilla model outperforms massive rivals like GPT-3 using a fraction of the computing power.
nvidia Nvidia Unveils Hopper H100 Datacenter GPU With 80 Billion Transistors Nvidia introduces the Hopper H100 GPU at GTC 2022, packing 80 billion transistors on a custom TSMC 4N process to succeed the Ampere A100 in supercomputing and AI workloads.
nvidia Nvidia Unveils H100 Datacenter Chips and Grace CPU Superchip Nvidia CEO Jensen Huang reveals the powerful new H100 datacenter chips and Grace CPU Superchip at the 2022 GTC conference. The new hardware promises massive performance leaps for artificial intelligence and high-performance computing workloads.
midjourney Midjourney Emerges as Leading AI Image Generator from Text Prompts Midjourney is a generative artificial intelligence program that creates detailed images from natural language descriptions. Developed by an independent research lab in San Francisco, the service operates primarily through a popular Discord interface.
deepmind DeepMind AlphaCode Achieves Median Rank in Competitive Programming DeepMind's AlphaCode system writes computer programs at a competitive level, placing within the top 54% of participants on Codeforces. The AI system uses massive-scale code generation and smart filtering to solve novel problems.
deepmind DeepMind AI Achieves Breakthrough in Nuclear Fusion Plasma Control DeepMind and the Swiss Plasma Center successfully use deep reinforcement learning to control and sculpt superheated plasma inside a tokamak reactor. This breakthrough opens new avenues for advancing nuclear fusion research.
nvidia Nvidia Unveils Hopper H100 Datacenter GPU With 80 Billion Transistors Nvidia introduces the Hopper H100, a massive 80-billion-transistor datacenter GPU built on TSMC's 4N process. The new chip significantly outpaces the previous A100 with upgraded NVLink, PCIe 5.0, and HBM3 memory.