meta Meta Unveils Chameleon, a Native Multimodal AI Competitor Meta introduces Chameleon, a new family of AI models built from the ground up to natively process and generate mixed image and text sequences. The early-fusion architecture achieves state-of-the-art results on visual tasks while remaining competitive in text-only benchmarks.
google Google Unveils LaMDA 2 to Improve Conversational AI Accuracy Alphabet CEO Sundar Pichai introduces LaMDA 2 at the Google I/O conference, aiming to make computers more accessible through advanced natural language processing. The updated model undergoes extensive internal testing to significantly reduce inaccurate and offensive responses.
bert BERT Model Transforms Natural Language Processing With Bidirectional Training Researchers introduce BERT, a bidirectional language model that achieves state-of-the-art results across eleven NLP tasks by pre-training on unlabeled text.
nlp New Deep Contextualized Word Vectors Advance Natural Language Processing Researchers introduce deep contextualized word representations that capture complex syntax and semantics using a bidirectional language model. These new vectors significantly improve performance across six major NLP tasks.
large language models Researchers Introduce Massive Benchmark to Evaluate Language Model Capabilities A sweeping new study introduces the Beyond the Imitation Game benchmark to rigorously test the reasoning and knowledge of large language models. The project involves over 200 authors collaborating to push past traditional text generation metrics.
meta Meta Releases OPT-175B to Open Up Large Language Model Research Meta AI introduces OPT-175B, a 175-billion-parameter language model released alongside its training code to democratize access for researchers. The noncommercial release aims to boost reproducible studies on bias, toxicity, and model robustness.
opt Meta Releases Open-Source OPT Models to Rival GPT-3 Researchers introduce OPT, a suite of open pre-trained transformer language models that match GPT-3's performance while significantly reducing carbon emissions. The release includes full model weights and code to help the research community study large language models.
google Google Unveils PaLM, a 540-Billion Parameter Language Model Google Research introduces PaLM, a massive 540-billion parameter language model trained using its new Pathways system. The model achieves state-of-the-art few-shot learning performance across hundreds of tasks.
google Google Unveils PaLM, a 540-Billion Parameter Language Model Google Research introduces PaLM, a massive 540-billion parameter language model trained using its new Pathways system. The model achieves state-of-the-art few-shot performance across hundreds of tasks.
nlp New Survey Tackles Compression Techniques for Large Language Models A new paper accepted to AAAI 2023 reviews current methods for compressing and accelerating pretrained language models to reduce their massive energy costs and inference delays. The survey focuses specifically on the inference stage to help deploy these models on edge and mobile devices.
dynamic neural networks Dynamic Neural Networks Offer Efficient Alternative for Massive NLP Models A new survey highlights how dynamic neural networks effectively scale natural language processing models without linear increases in computing power. The research explores methods like skimming, mixture of experts, and early exit to handle trillion-parameter models.
openai Smaller AI Models Outperform GPT-3 Through Human Feedback Training OpenAI researchers introduce InstructGPT, a language model trained with human feedback to better follow user instructions. Despite having 100 times fewer parameters, this aligned model consistently outperforms the massive GPT-3 in human evaluations.
cohere Cohere for AI Unveils Aya Expanse Models to Close Global Language Gap Cohere for AI introduces Aya Expanse, an open-weights multilingual model family designed to bring state-of-the-art AI capabilities to underrepresented languages. The 32B model outperforms significantly larger competitors on multilingual understanding benchmarks.
artificial intelligence 2023 Marks a Pivotal Year for Artificial Intelligence Innovation The AI landscape experiences a transformative year as major tech companies introduce advanced language models and forge strategic partnerships.
baby agi Baby AGI Introduces Fully Autonomous Task Management Baby AGI is a new Python script that leverages OpenAI and Pinecone to autonomously create, prioritize, and execute business tasks. The system operates in an infinite loop to achieve predefined objectives without human intervention.