Search AI News
Find articles from 100+ AI sources
Found 781 results for "LLM"
Estimating worst case frontier risks of open weight LLMs
In this paper, we study the worst-case frontier risks of releasing gpt-oss. We introduce malicious…
📚 3LM: A Benchmark for Arabic LLMs in STEM and Code
📚 3LM: A Benchmark for Arabic LLMs in STEM and Code. Hugging Face - artificial intelligence news.
Accelerate a World of LLMs on Hugging Face with NVIDIA NIM
Accelerate a World of LLMs on Hugging Face with NVIDIA NIM. Hugging Face - artificial intelligence news.
Consilium: When Multiple LLMs Collaborate
Consilium: When Multiple LLMs Collaborate. Hugging Face - artificial intelligence news.
Last Week to Register to the Build Production-Ready LLMs From Scratch Course!
This Saturday, we kick off the latest cohort of the Build Production-Ready LLMs From Scratch course! This is…
Upskill your LLMs With Gradio MCP Servers
Upskill your LLMs With Gradio MCP Servers. Hugging Face - artificial intelligence news.
SmolLM3: smol, multilingual, long-context reasoner
SmolLM3: smol, multilingual, long-context reasoner. Hugging Face - artificial intelligence news.
LLM Research Papers: The 2025 List (January to June)
As some of you know, I keep a running list of research papers I (want to) read and reference.About six months…
Understanding and Coding the KV Cache in LLMs from Scratch
KV caches are one of the most critical techniques for efficient inference in LLMs in production. KV caches…
Build Production-Ready LLMs From Scratch Starting on July 12th!
Get ready! The latest iteration of the Build Production-Ready LLMs From Scratch live course is starting on…
How Long Prompts Block Other Requests - Optimizing LLM Performance
Back to Articles How Long Prompts Block Other Requests - Optimizing LLM Performance Team Article Published…
No GPU left behind: Unlocking Efficiency with Co-located vLLM in TRL
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models Paper • 2402.03300 •…
Last Week to Register to the Build Production-Ready LLMs From Scratch Course!
This Saturday, we kick off the Build Production-Ready LLMs From Scratch course! This is the last week to…
A Guide for Debugging LLM Training Data
(from [2])Most discussions of LLM training focus heavily on models and algorithms. We enjoy experimenting…
Coding LLMs from the Ground Up: A Complete Course
I wrote a lot about reasoning models in recent months (4 articles in a row)! Next to everything "agentic,"…
PGN2FEN: A Benchmark for Evaluating LLM Chess Reasoning
Today, I’m releasing PGN2FEN — a new benchmark that tests the ability of language models to understand…
Introducing AutoRound: Intel’s Advanced Quantization for LLMs and VLMs
Running Agents Featured 219 Low-bit LLM Leaderboard 🏆 219 Track, rank and evaluate open LLMs and chatbots
All About The Modern Positional Encodings In LLMs
The Positional Encoding in LLMs may appear somewhat mysterious the first time we come across the concept, and…
Build Production-Ready LLMs From Scratch
Big news! I am now partnering with Maven as an instructor to teach the Build Production-Ready LLMs From…
The State of Reinforcement Learning for LLM Reasoning
A lot has happened this month, especially with the releases of new flagship models like GPT-4.5 and Llama 4.…
Prefill and Decode for Concurrent Requests - Optimizing LLM Performance
Back to Articles Prefill and Decode for Concurrent Requests - Optimizing LLM Performance Team Article…
The NLP Course is becoming the LLM Course
Back to Articles The NLP Course is becoming the LLM Course! Published April 3, 2025 Update on GitHub Upvote…
Efficient Request Queueing – Optimizing LLM Performance
Back to Articles Efficient Request Queueing – Optimizing LLM Performance Team Article Published April 2,…
Vision Large Language Models (vLLMs)
After the popularization of text-based large language models (LLMs), one of the most important questions…