Search AI News
Find articles from 100+ AI sources
Found 1,191 results for "LLM"
Defining and evaluating political bias in LLMs
Learn how OpenAI evaluates political bias in ChatGPT through new real-world testing methods that improve…
Understanding the 4 Main Approaches to LLM Evaluation (From Scratch)
How do we actually evaluate LLMs?It’s a simple question, but one that tends to open up a much bigger…
AdalFlow: A PyTorch-Like Framework to Auto-Optimizing Prompt for your LLM agent
AI Agent frameworks are becoming just as important as model training itself! I am excited to introduce you to…
REINFORCE: Easy Online RL for LLMs
Reinforcement learning (RL) is playing an increasingly important role in research on large language models…
SyGra: The One-Stop Framework for Building Data for LLMs and SLMs
SyGra: The One-Stop Framework for Building Data for LLMs and SLMs. Hugging Face - artificial intelligence…
Fine-tune Any LLM from the Hugging Face Hub with Together AI
Fine-tune Any LLM from the Hugging Face Hub with Together AI. Hugging Face - artificial intelligence news.
Jupyter Agents: training LLMs to reason with notebooks
Jupyter Agents: training LLMs to reason with notebooks. Hugging Face - artificial intelligence news.
Online versus Offline RL for LLMs
(from [2, 5, 7, 9, 10])The alignment process teaches large language models (LLMs) how to generate completions…
Mixture-of-Experts: Early Sparse MoE Prototypes in LLMs
Mixture-of-Experts might be one of the most important improvements in the Transformer architecture! It allows…
Which Agent Causes Task Failures and When?Researchers from PSU and Duke explores automated failure attribution of LLM Multi-Agent Systems
Share My Research is Synced’s column that welcomes scholars to share their own research breakthroughs with…
🇵🇭 FilBench - Can LLMs Understand and Generate Filipino?
🇵🇭 FilBench - Can LLMs Understand and Generate Filipino?. Hugging Face - artificial intelligence news.
TextQuests: How Good are LLMs at Text-Based Video Games?
TextQuests: How Good are LLMs at Text-Based Video Games?. Hugging Face - artificial intelligence news.
Estimating worst case frontier risks of open weight LLMs
In this paper, we study the worst-case frontier risks of releasing gpt-oss. We introduce malicious…
📚 3LM: A Benchmark for Arabic LLMs in STEM and Code
📚 3LM: A Benchmark for Arabic LLMs in STEM and Code. Hugging Face - artificial intelligence news.
Accelerate a World of LLMs on Hugging Face with NVIDIA NIM
Accelerate a World of LLMs on Hugging Face with NVIDIA NIM. Hugging Face - artificial intelligence news.
Consilium: When Multiple LLMs Collaborate
Consilium: When Multiple LLMs Collaborate. Hugging Face - artificial intelligence news.
Last Week to Register to the Build Production-Ready LLMs From Scratch Course!
This Saturday, we kick off the latest cohort of the Build Production-Ready LLMs From Scratch course! This is…
Upskill your LLMs With Gradio MCP Servers
Upskill your LLMs With Gradio MCP Servers. Hugging Face - artificial intelligence news.
SmolLM3: smol, multilingual, long-context reasoner
SmolLM3: smol, multilingual, long-context reasoner. Hugging Face - artificial intelligence news.
LLM Research Papers: The 2025 List (January to June)
As some of you know, I keep a running list of research papers I (want to) read and reference.About six months…
Understanding and Coding the KV Cache in LLMs from Scratch
KV caches are one of the most critical techniques for efficient inference in LLMs in production. KV caches…
Build Production-Ready LLMs From Scratch Starting on July 12th!
Get ready! The latest iteration of the Build Production-Ready LLMs From Scratch live course is starting on…
How Long Prompts Block Other Requests - Optimizing LLM Performance
Back to Articles How Long Prompts Block Other Requests - Optimizing LLM Performance Team Article Published…
No GPU left behind: Unlocking Efficiency with Co-located vLLM in TRL
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models Paper • 2402.03300 •…