⚡ LIVE

Search AI News

Find articles from 100+ AI sources

🔍

Found 618 results for "NVIDIA"

618 articles
Latent Space Models Jul 7, 2026 👁 53

[AINews] The Field Guide to Fable

While we congratulate (friend of the show!) General Intuition on their new model and (friend of the show!)…

PyTorch Tools Jul 6, 2026 👁 42

Bringing PyTorch Monarch to AMD GPUs: Single-Controller Distributed Training on ROCm

Training state-of-the-art large language models (LLMs) with billions of parameters requires distributed…

MIT Tech Review AI Research Jul 6, 2026 👁 48

Your family’s $300 stake in OpenAI

This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in…

NVIDIA Generative AI Models Jul 6, 2026 👁 44

How Open Models Are Driving AI Research

Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers…

NVIDIA Generative AI Models Jul 6, 2026 👁 41

How Nations Are Deploying AI for Strategic Priorities

Nations have long invested in domestic infrastructure to advance their economies, protect and use their data,…

TheSequence Business Jul 5, 2026 👁 39

The Sequence Radar #889: Fable 5's Comeback, ZCode's Debut, Claude Science, and the $3.5B Deployment Land Grab

Next Week in The Sequence:We continue our series about model distillation. The AI of the Week, dives into…

Latent Space Models Jul 1, 2026 👁 49

[AINews] Sonnet 5 today, and Fable 5 tomorrow

In separate announcements, Sonnet 5 was released today, and Fable/Mythos 5 were approved to be released again…

PyTorch Tools Jun 30, 2026 👁 52

Miles: A PyTorch-Native Stack for Large-Scale LLM RL Post-Training

TL;DR Miles is RadixArk’s open source framework for large-scale LLM RL post-training. It composes SGLang…

NVIDIA Generative AI Models Jun 30, 2026 👁 43

Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning

Editor’s note: This post is part of Into the Omniverse, a series focused on how developers, 3D…

Import AI (Jack Clark) Research Jun 29, 2026 👁 39

Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era

Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from…

Interconnects AI Models Jun 28, 2026 👁 43

Latest open artifacts (#22): Zyphra, Cohere, and Poolside are expanding the breadth of the ecosystem

A trend we continue to see in open model releases is that the ecosystem is becoming more diverse, with an…

TheSequence Business Jun 28, 2026 👁 43

The Sequence Radar #885: Last Week in AI: Models, Games, and the Future of Evaluation

Next Week in The Sequence:We continue our series about distillation. In the AI of the week, we discuss…

Life Architect AI Models Jun 28, 2026 👁 51

The Memo - 28/Jun/2026

To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…

AI Magazine (Raschka) Models Jun 27, 2026 👁 52

Using Local Coding Agents

Many people reached out to me in the past asking about my local agent stack as well as how I set up my local…

Hugging Face Open Source Jun 26, 2026 👁 45

Run a vLLM Server on HF Jobs in One Command

Back to Articles Run a vLLM Server on HF Jobs in One Command Published June 26, 2026 Update on GitHub Upvote…

PyTorch Tools Jun 25, 2026 👁 43

TokenSpeed-Kernel: Portable APIs and High-Performance Kernels for Multi-Silicon LLM Inference

TL;DR The TokenSpeed-kernel is a standalone, open-source subsystem designed to solve backend complexity in…

Life Architect AI Models Jun 24, 2026 👁 46

The Memo - Special edition - Mid-2026 AI retrospective

To: US Govt, major govts, Microsoft, Apple, NVIDIA, Alphabet, Amazon, Meta, Tesla, Citi, Tencent, IBM, &…

PyTorch Tools Jun 23, 2026 👁 67

Serving DeepSeek-V4 on GB300 with SGLang: 5x Higher Throughput at the Same Interactivity Since Day-0

TL;DR: DeepSeek-V4 support was live in SGLang on Day-0, but the Day-0 stack was only the starting point.…

NVIDIA Generative AI Models Jun 23, 2026 👁 44

How Businesses Are Building Specialized AI They Can Trust

Editor’s note: This post is part of the Nemotron Labs blog series, which explores how the latest open…

Interconnects AI Models Jun 22, 2026 👁 44

GLM-5.2 is the step change for open agents

Housekeeping: Following my “State of the blog” post last week, noting a slight increase in paid features,…

AI Weekly Business Jun 19, 2026 👁 43

AI Weekly Issue #505: 100 years from now : The Last War Between Countries

100 years from now : The Last War Between Countries  June 19th 2026 Curated by Alexis This is 100 Years From…

Weaviate Tools Jun 18, 2026 👁 47

Import & Vectorize Data with Weaviate at Scale

Most vector database pilots fail at ingest, not at search. You build a clever retrieval pipeline, you watch…

AI Weekly Business Jun 17, 2026 👁 49

AI Weekly Issue #504: America blocked its best AI. China just raised $7.4 billion.

America blocked its best AI. China just raised $7.4 billion.  June 17th 2026 Curated by Alexis Four days…

NVIDIA Generative AI Models Jun 16, 2026 👁 44

Coherent Breaks Ground on Expanded Texas Facility, Scaling AI’s Optical Backbone

AI runs at the speed of light. More and more, that light is made in Texas. Coherent broke ground today on an…