⚡ LIVE

Search AI News

Find articles from 100+ AI sources

🔍

Found 207 results for "LLaMA"

207 articles
Cameron Wolfe (AI) Models Jan 26, 2026 👁 17

Continual Learning with RL for LLMs

(from [1, 2, 3, 6, 11])Continual learning, which refers to the ability of an AI model to learn from new tasks…

ScienceDaily AI Research Jan 21, 2026 👁 15

The human brain may work more like AI than anyone expected

Science News from research organizations The human brain may work more like AI than anyone expected Date:…

VentureBeat Business Jan 19, 2026 👁 16

Claude Code costs up to $200 a month. Goose does the same thing for free.

The artificial intelligence coding revolution comes with a catch: it's expensive.Claude Code, Anthropic's…

Bounded Regret Research Jan 6, 2026 👁 16

Oversight Assistants: Turning Compute into Understanding

Currently, we primarily oversee AI with human supervision and human-run experiments, possibly augmented by…

AI Magazine (Raschka) Models Dec 30, 2025 👁 18

The State Of LLMs 2025: Progress, Problems, and Predictions

As 2025 comes to a close, I want to look back at some of the year’s most important developments in large…

Air Street Capital Business Nov 30, 2025 👁 14

State of AI: December 2025 newsletter

Dear readers, Welcome to the latest issue of the State of AI, an editorialized newsletter that covers the key…

AI Magazine (Raschka) Models Nov 4, 2025 👁 21

Beyond Standard LLMs

From DeepSeek R1 to MiniMax-M2, the largest and most capable open-weight LLMs today remain autoregressive…

DeepMind Research Oct 25, 2025 👁 18

Introducing Gemma 3n: The developer guide

The first Gemma model launched early last year and has since grown into a thriving Gemmaverse of over 160…

Air Street Capital Business Oct 9, 2025 👁 27

🪩 The State of AI Report 2025 🪩

Hi everyone!The day is finally here: I’m thrilled to share the State of AI Report 2025 with you!In short,…

AI Magazine (Raschka) Models Oct 5, 2025 👁 19

Understanding the 4 Main Approaches to LLM Evaluation (From Scratch)

How do we actually evaluate LLMs?It’s a simple question, but one that tends to open up a much bigger…

Cameron Wolfe (AI) Models Sep 29, 2025 👁 18

REINFORCE: Easy Online RL for LLMs

Reinforcement learning (RL) is playing an increasingly important role in research on large language models…

Cameron Wolfe (AI) Models Sep 8, 2025 👁 19

Online versus Offline RL for LLMs

(from [2, 5, 7, 9, 10])The alignment process teaches large language models (LLMs) how to generate completions…

SemiAnalysis Models Aug 20, 2025 👁 26

H100 vs GB200 NVL72 Training Benchmarks – Power, TCO, and Reliability Analysis, Software Improvement Over Time

Frontier model training has pushed GPUs and AI systems to their absolute limits, making cost, efficiency,…

AI Magazine (Raschka) Models Aug 9, 2025 👁 14

From GPT-2 to gpt-oss: Analyzing the Architectural Advances

OpenAI just released their new open-weight LLMs this week: gpt-oss-120b and gpt-oss-20b, their first…

Cameron Wolfe (AI) Models Jul 28, 2025 👁 16

Direct Preference Optimization (DPO)

(from [1, 2, 6, 9])Aligning large language models (LLMs) is a crucial post-training step that ensures models…

Hugging Face Open Source Jul 15, 2025 👁 15

Migrating the Hub from Git LFS to Xet

Back to Articles Migrating the Hub from Git LFS to Xet Published July 15, 2025 Update on GitHub Upvote 29 +23…

Air Street Capital Business Jul 13, 2025 👁 17

State of AI: July 2025 newsletter

Hi everyone!Welcome to the latest issue of the State of AI newsletter, an editorialized newsletter covering…

AI Tidbits Business Jul 13, 2025 👁 17

LinkedIn Highlights, June 2025 - AI Agents Edition

Welcome to LinkedIn Highlights!Each month, I'll share my five seven top-performing LinkedIn posts, bringing…

SemiAnalysis Models Jul 11, 2025 👁 15

Meta Superintelligence – Leadership Compute, Talent, and Data

Meta’s shocking purchase of 49% of Scale AI at a ~$30B valuation shows that money is of no concern for the…

AI Magazine (Raschka) Models Jul 1, 2025 👁 14

LLM Research Papers: The 2025 List (January to June)

As some of you know, I keep a running list of research papers I (want to) read and reference.About six months…

Cameron Wolfe AI Models Jun 30, 2025 👁 16

Reward Models

(from [1, 2, 4, 14])Reward models (RMs) are a cornerstone of large language model (LLM) research, enabling…

Hugging Face Open Source Jun 23, 2025 👁 13

Transformers backend integration in SGLang

Back to Articles Transformers backend integration in SGLang Published June 23, 2025 Update on GitHub Upvote…

AI Magazine (Raschka) Models Jun 17, 2025 👁 17

Understanding and Coding the KV Cache in LLMs from Scratch

KV caches are one of the most critical techniques for efficient inference in LLMs in production. KV caches…

Synced Review Research Jun 16, 2025 👁 17

MIT Researchers Unveil “SEAL”: A New Step Towards Self-Improving AI

The concept of AI self-improvement has been a hot topic in recent research circles, with a flurry of papers…