Search AI News
Find articles from 100+ AI sources
Found 1,190 results for "LLM"
Last Week to Register to the Build Production-Ready LLMs From Scratch Course!
This Saturday, we kick off the Build Production-Ready LLMs From Scratch course! This is the last week to…
A Guide for Debugging LLM Training Data
(from [2])Most discussions of LLM training focus heavily on models and algorithms. We enjoy experimenting…
Coding LLMs from the Ground Up: A Complete Course
I wrote a lot about reasoning models in recent months (4 articles in a row)! Next to everything "agentic,"…
PGN2FEN: A Benchmark for Evaluating LLM Chess Reasoning
Today, I’m releasing PGN2FEN — a new benchmark that tests the ability of language models to understand…
Introducing AutoRound: Intel’s Advanced Quantization for LLMs and VLMs
Running Agents Featured 219 Low-bit LLM Leaderboard 🏆 219 Track, rank and evaluate open LLMs and chatbots
All About The Modern Positional Encodings In LLMs
The Positional Encoding in LLMs may appear somewhat mysterious the first time we come across the concept, and…
Build Production-Ready LLMs From Scratch
Big news! I am now partnering with Maven as an instructor to teach the Build Production-Ready LLMs From…
The State of Reinforcement Learning for LLM Reasoning
A lot has happened this month, especially with the releases of new flagship models like GPT-4.5 and Llama 4.…
Prefill and Decode for Concurrent Requests - Optimizing LLM Performance
Back to Articles Prefill and Decode for Concurrent Requests - Optimizing LLM Performance Team Article…
The NLP Course is becoming the LLM Course
Back to Articles The NLP Course is becoming the LLM Course! Published April 3, 2025 Update on GitHub Upvote…
Efficient Request Queueing – Optimizing LLM Performance
Back to Articles Efficient Request Queueing – Optimizing LLM Performance Team Article Published April 2,…
Vision Large Language Models (vLLMs)
After the popularization of text-based large language models (LLMs), one of the most important questions…
🚀 Accelerating LLM Inference with TGI on Intel Gaudi
Back to Articles 🚀 Accelerating LLM Inference with TGI on Intel Gaudi Published March 28, 2025 Update on…
Welcome Gemma 3: Google's all new multimodal, multilingual, long context open LLM
Welcome Gemma 3: Google's all new multimodal, multilingual, long context open LLM. Hugging Face - artificial…
LLM Inference on Edge: A Fun and Easy Guide to run LLMs via React Native on your Phone!
HuggingFaceTB/SmolLM2-1.7B-Instruct Text Generation • 2B • Updated Apr 21, 2025 • 145k • 739
LLMs Turn Every Question Into an Answer
in Chain of ThoughtDALL-E/Every illustration.The world has changed considerably since our last ”think…
Fixing Open LLM Leaderboard with Math-Verify
Running on CPU Upgrade 14k Open LLM Leaderboard 🏆 14k Track, rank and evaluate open LLMs and chatbots
The Open Arabic LLM Leaderboard 2
The Open Arabic LLM Leaderboard 2. Hugging Face - artificial intelligence news.
Mastering Long Contexts in LLMs with KVPress
meta-llama/Llama-3.3-70B-Instruct Text Generation • 71B • Updated Dec 21, 2024 • 758k • 2.89k
Introducing multi-backends (TRT-LLM, vLLM) support for Text Generation Inference
Back to Articles Introducing multi-backends (TRT-LLM, vLLM) support for Text Generation Inference Published…
CO₂ Emissions and Models Performance: Insights from the Open LLM Leaderboard
Runtime error Agents 113 Open LLM Leaderboard Model Comparator 🏆 113 Compare Open LLM Leaderboard results
How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs
How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs. Hugging Face -…
Rethinking LLM Evaluation with 3C3H: AraGen Benchmark and Leaderboard
Running Agents 112 Judge Arena 💻 112 View and compare open‑source AI model rankings with ELO scores
Investing in Performance: Fine-tune small models with LLM insights - a CFM case study
EmergentMethods/gliner_medium_news-v2.1 Token Classification • 0.2B • Updated Jan 12 • 9.52k • 82