⚡ LIVE

Search AI News

Find articles from 100+ AI sources

🔍

Found 168 results for "PyTorch"

168 articles
Latent Space Models Aug 8, 2026 👁 47

[AINews] Zawinski's Law of MultiAgents

We’ve discussed the HuggingFace-OpenAI security incident before, but OpenAI’s side of the story was the…

Latent Space Models Jul 29, 2026 👁 51

[AINews] AI is eating Finance; AIE NYC now open

We love writing a newsletter that cares more about being high signal than telling you there’s breaking news…

Gradient Flow Models Jul 29, 2026 👁 42

Specialized AI Is Getting Easier to Build

Last week I argued that open models will absorb most of the money and compute the world spends on AI. A week…

Berkeley AI Research Research Jul 29, 2026 👁 39

From CUDA to MLX: How K-Search Brings Decades of Kernel Expertise to Apple Silicon

Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into…

Gradient Flow Models Jul 28, 2026 👁 47

The Big AI Labs Are Suddenly Competing with Your Own Data

Subscribe • Previous Issues Specialized AI Is Getting Easier to Build Last week I argued that open models…

Latent Space Models Jul 28, 2026 👁 47

[AINews] Much ado about Open Weights

Everyone say hi to Richard MacManus, our new Head of Editorial!The current debate about Open Weights is the…

Linux Foundation AI Open Source Jul 27, 2026 👁 55

Open Models and Open Weights Are Foundational to Secure AI

Home Blog Open Models and Open Weights Are Foundational to Secure AI 6 MIN READ Open Models and Open Weights…

NVIDIA AI Models Jul 27, 2026 👁 48

Industry Leaders Unite in Open Secure AI Alliance for AI Safety and Security

Open source software is a critical pillar of the global economy. It underpins cloud computing, financial…

AWS Machine Learning Tools Jul 24, 2026 👁 52

Build an explainable next-best-product recommendation system for banking on AWS

Building a deep learning-based explainable next-best-product recommendation system helps banking institutions…

PyTorch Open Source Jul 23, 2026 👁 51

Helion on TPU: Towards Hardware Heterogeneous Kernel Authoring

TL;DR Helion is PyTorch’s high-level DSL for writing performance-portable ML kernels. Partnering with…

TechCrunch Business Jul 20, 2026 👁 38

OpenAI is scared of open-weight models. Should the US be?

The impressive capabilities of Chinese lab Moonshot’s Kimi K3, the biggest open-weight large language…

Hugging Face Models Jul 17, 2026 👁 44

Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers

Back to Articles Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers…

PyTorch Open Source Jul 15, 2026 👁 56

Triton Plugin Extensions: Enabling TLX and Custom Compiler Passes Out of the Box

TLDR The PyTorch-Triton 3.7 release introduces the Triton Plugin Extensions system, a framework for…

AWS ML Tools Jul 10, 2026 👁 40

Deploying quantized models on Amazon SageMaker AI with Unsloth

This post was co-written with Daniel Han and Michael Han from Unsloth. Deploying large foundation models…

PyTorch Tools Jul 10, 2026 👁 47

Towards Free Normalization: Fusing Normalization into GEMM and Attention Kernels

Code available at:…

Import AI (Jack Clark) Research Jul 6, 2026 👁 38

Import AI 464: Fables writes GPU kernels; AI automation; and analog computation

Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from…

Hugging Face Open Source Jul 6, 2026 👁 49

🤗 Kernels: Major Updates

Back to Articles 🤗 Kernels: Major Updates Published July 6, 2026 Update on GitHub Upvote 26 +20 Sayak Paul…

PyTorch Tools Jul 2, 2026 👁 38

Building the Future of On-Device AI at the ExecuTorch Hackathon

This past weekend in San Francisco, builders, researchers, mobile developers, and AI practitioners came…

PyTorch Tools Jun 25, 2026 👁 43

TokenSpeed-Kernel: Portable APIs and High-Performance Kernels for Multi-Silicon LLM Inference

TL;DR The TokenSpeed-kernel is a standalone, open-source subsystem designed to solve backend complexity in…

PyTorch Tools Jun 23, 2026 👁 67

Serving DeepSeek-V4 on GB300 with SGLang: 5x Higher Throughput at the Same Interactivity Since Day-0

TL;DR: DeepSeek-V4 support was live in SGLang on Day-0, but the Day-0 stack was only the starting point.…

Lambda Labs Models Jun 16, 2026 👁 58

MLPerf Training v6.0: Lambda delivers fastest LLM training on NVIDIA GB300 NVL72 and fastest MoE training on NVIDIA HGX B200

June 16, 2026 • 4 min read Lambda’s GB300 NVL72 Llama 3.1 8B MLPerf Training v6.0 submission improved…

Hugging Face Open Source Jun 8, 2026 👁 45

The Open Source Community is backing OpenEnv for Agentic RL

Back to Articles The Open Source Community is backing OpenEnv for Agentic RL Published June 8, 2026 Update on…

AI Magazine (Raschka) Models May 16, 2026 👁 45

Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention

After a short family break, I am excited to be back and catching up on a busy few weeks of open-weight LLM…

Hugging Face Open Source May 14, 2026 👁 44

Unlocking asynchronicity in continuous batching

Back to Articles Unlocking asynchronicity in continuous batching Published May 14, 2026 Update on GitHub…