Search AI News
Find articles from 100+ AI sources
Found 207 results for "LLaMA"
Frontier post-training recipe review with Finbarr Timbers
As I’ve been recapping fundamentals of post-training to wrap up my RLHF / Post-training book I knew I…
Your AI bill is a tax on scale
Subscribe • Previous Issues The Hybrid AI Stack Is Coming for the Pricing Power of OpenAI and Anthropic…
DiffusionGemma: 4x faster text generation
Breadcrumb Innovation & AI Technology Developer tools DiffusionGemma: 4x faster text generation Jun 10, 2026…
NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI
Today, Google DeepMind released DiffusionGemma — an experimental open model built for exceptionally…
Introducing Gemma 4 12B: a unified, encoder-free multimodal model
Breadcrumb Innovation & AI Technology Developer tools Introducing Gemma 4 12B: a unified, encoder-free…
Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing
Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from…
Farewell Ai2
I’m departing the Allen Institute for AI (Ai2), where I got the great privilege to work on the Olmo models,…
Otari: Own Your AI Stack
Closed-source frontier providers offer what looks like a complete stack: tools, MCP server integrations,…
Reachy Mini goes fully local
Back to Articles Reachy Mini goes fully local Published May 27, 2026 Update on GitHub Upvote 64 +58 Amir…
AI Got Expensive. Now What?
Cloud AI got expensive in 2026. Now everyone's looking at local again, which would be great, except the local…
Some ideas for what comes next, May 2026
As the years of AI progress go by, it’s been accompanied by a slowly rising tide of consequence. Models are…
Build a Coding Assistant with Weaviate MCP: RAG over Code & Docs
Last week I asked Claude Code to implement something relatively trivial in my codebase. Three turns in, the…
Notes from inside China's AI labs
Staring out the window on a new, high-speed train from Hangzhou to Shanghai I’m gifted with views of…
Sovereign AI: Control, Choice, and Why It Goes Beyond Geopolitics
On Your Terms On Your Terms is a series of conversations with the builders at Mozilla.ai, going deep on the…
Announcing RAAIS 2026 headline speakers
The Research and Applied AI Summit (RAAIS) is a community for entrepreneurs and researchers who accelerate…
State of AI: April 2026 newsletter
Dear readers, Welcome to the latest issue of the State of AI, an editorialized newsletter that covers the key…
The inevitable need for an open model consortium
Recently, I was talking with Percy Liang, Stanford professor and lead of the Marin project (another…
Gemma 4: Byte for byte, the most capable open models
Breadcrumb Innovation & AI Technology Developer tools Gemma 4: Byte for byte, the most capable open models…
A Visual Guide to Attention Variants in Modern LLMs
I had originally planned to write about DeepSeek V4. Since it still hasn’t been released, I used the time…
The Memo - 28/Feb/2026
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…
A Dream of Spring for Open-Weight LLMs: 10 Architectures from Jan-Feb 2026
If you have struggled a bit to keep up with open-weight model releases this month, this article should catch…
How will OpenAI compete?
“Jakub and Mark set the research direction for the long run. Then after months of work, something…
Introducing Weaviate Agent Skills
It has never been easier to build software. With tools like Claude Code, Cursor, and GitHub Copilot, we’ve…
LWiAI Podcast #233 - Moltbot, Genie 3, Qwen3-Max-Thinking
Our 233rd episode with a summary and discussion of last week’s big AI news!Recorded on 01/30/2026Hosted by…