Search AI News
Find articles from 100+ AI sources
Found 323 results for "NVIDIA"
Miles: A PyTorch-Native Stack for Large-Scale LLM RL Post-Training
TL;DR Miles is RadixArk’s open source framework for large-scale LLM RL post-training. It composes SGLang…
Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning
Editor’s note: This post is part of Into the Omniverse, a series focused on how developers, 3D…
Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era
Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from…
Latest open artifacts (#22): Zyphra, Cohere, and Poolside are expanding the breadth of the ecosystem
A trend we continue to see in open model releases is that the ecosystem is becoming more diverse, with an…
The Sequence Radar #885: Last Week in AI: Models, Games, and the Future of Evaluation
Next Week in The Sequence:We continue our series about distillation. In the AI of the week, we discuss…
The Memo - 28/Jun/2026
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…
Using Local Coding Agents
Many people reached out to me in the past asking about my local agent stack as well as how I set up my local…
Run a vLLM Server on HF Jobs in One Command
Back to Articles Run a vLLM Server on HF Jobs in One Command Published June 26, 2026 Update on GitHub Upvote…
TokenSpeed-Kernel: Portable APIs and High-Performance Kernels for Multi-Silicon LLM Inference
TL;DR The TokenSpeed-kernel is a standalone, open-source subsystem designed to solve backend complexity in…
The Memo - Special edition - Mid-2026 AI retrospective
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…
Serving DeepSeek-V4 on GB300 with SGLang: 5x Higher Throughput at the Same Interactivity Since Day-0
TL;DR: DeepSeek-V4 support was live in SGLang on Day-0, but the Day-0 stack was only the starting point.…
How Businesses Are Building Specialized AI They Can Trust
Editor’s note: This post is part of the Nemotron Labs blog series, which explores how the latest open…
GLM-5.2 is the step change for open agents
Housekeeping: Following my “State of the blog” post last week, noting a slight increase in paid features,…
AI Weekly Issue #505: 100 years from now : The Last War Between Countries
100 years from now : The Last War Between Countries June 19th 2026 Curated by Alexis This is 100 Years From…
Import & Vectorize Data with Weaviate at Scale
Most vector database pilots fail at ingest, not at search. You build a clever retrieval pipeline, you watch…
AI Weekly Issue #504: America blocked its best AI. China just raised $7.4 billion.
America blocked its best AI. China just raised $7.4 billion. June 17th 2026 Curated by Alexis Four days…
Coherent Breaks Ground on Expanded Texas Facility, Scaling AI’s Optical Backbone
AI runs at the speed of light. More and more, that light is made in Texas. Coherent broke ground today on an…
Frontier post-training recipe review with Finbarr Timbers
As I’ve been recapping fundamentals of post-training to wrap up my RLHF / Post-training book I knew I…
The Memo - Special edition - Public access delays to intelligence & the Claude Fable 5 ban - 15/Jun/2026
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…
The Sequence Radar #877: Last Week in AI: Anthropic Ships, Apple Borrows, Musk Lists, Bezos Builds
Next Week in The Sequence:We continue our series about alternative to transformers. The AI of the week will…
The Memo - 13/Jun/2026
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…
DiffusionGemma: 4x faster text generation
Breadcrumb Innovation & AI Technology Developer tools DiffusionGemma: 4x faster text generation Jun 10, 2026…
Claude Fable 5 and new AI safety fables
Edit Jun. 11: Anthropic changed their silent model manipulation of AI research queries to also use a…
The Memo - Special edition - Claude Fable 5 - 9/Jun/2026
To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…