Search AI News
Find articles from 100+ AI sources
Found 618 results for "NVIDIA"
[AINews] The Field Guide to Fable
While we congratulate (friend of the show!) General Intuition on their new model and (friend of the show!)…
Bringing PyTorch Monarch to AMD GPUs: Single-Controller Distributed Training on ROCm
Training state-of-the-art large language models (LLMs) with billions of parameters requires distributed…
Your family’s $300 stake in OpenAI
This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in…
How Open Models Are Driving AI Research
Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers…
How Nations Are Deploying AI for Strategic Priorities
Nations have long invested in domestic infrastructure to advance their economies, protect and use their data,…
The Sequence Radar #889: Fable 5's Comeback, ZCode's Debut, Claude Science, and the $3.5B Deployment Land Grab
Next Week in The Sequence:We continue our series about model distillation. The AI of the Week, dives into…
[AINews] Sonnet 5 today, and Fable 5 tomorrow
In separate announcements, Sonnet 5 was released today, and Fable/Mythos 5 were approved to be released again…
Miles: A PyTorch-Native Stack for Large-Scale LLM RL Post-Training
TL;DR Miles is RadixArk’s open source framework for large-scale LLM RL post-training. It composes SGLang…
Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning
Editor’s note: This post is part of Into the Omniverse, a series focused on how developers, 3D…
Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era
Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from…
Latest open artifacts (#22): Zyphra, Cohere, and Poolside are expanding the breadth of the ecosystem
A trend we continue to see in open model releases is that the ecosystem is becoming more diverse, with an…
The Sequence Radar #885: Last Week in AI: Models, Games, and the Future of Evaluation
Next Week in The Sequence:We continue our series about distillation. In the AI of the week, we discuss…
The Memo - 28/Jun/2026
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…
Using Local Coding Agents
Many people reached out to me in the past asking about my local agent stack as well as how I set up my local…
Run a vLLM Server on HF Jobs in One Command
Back to Articles Run a vLLM Server on HF Jobs in One Command Published June 26, 2026 Update on GitHub Upvote…
TokenSpeed-Kernel: Portable APIs and High-Performance Kernels for Multi-Silicon LLM Inference
TL;DR The TokenSpeed-kernel is a standalone, open-source subsystem designed to solve backend complexity in…
The Memo - Special edition - Mid-2026 AI retrospective
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…
Serving DeepSeek-V4 on GB300 with SGLang: 5x Higher Throughput at the Same Interactivity Since Day-0
TL;DR: DeepSeek-V4 support was live in SGLang on Day-0, but the Day-0 stack was only the starting point.…
How Businesses Are Building Specialized AI They Can Trust
Editor’s note: This post is part of the Nemotron Labs blog series, which explores how the latest open…
GLM-5.2 is the step change for open agents
Housekeeping: Following my “State of the blog” post last week, noting a slight increase in paid features,…
AI Weekly Issue #505: 100 years from now : The Last War Between Countries
100 years from now : The Last War Between Countries June 19th 2026 Curated by Alexis This is 100 Years From…
Import & Vectorize Data with Weaviate at Scale
Most vector database pilots fail at ingest, not at search. You build a clever retrieval pipeline, you watch…
AI Weekly Issue #504: America blocked its best AI. China just raised $7.4 billion.
America blocked its best AI. China just raised $7.4 billion. June 17th 2026 Curated by Alexis Four days…
Coherent Breaks Ground on Expanded Texas Facility, Scaling AI’s Optical Backbone
AI runs at the speed of light. More and more, that light is made in Texas. Coherent broke ground today on an…