⚡ LIVE

Search AI News

Find articles from 100+ AI sources

🔍

Found 323 results for "NVIDIA"

323 articles
Air Street Capital Business Apr 6, 2025 👁 24

State of AI: April 2025 newsletter

Hi everyone!Welcome to the latest issue of the State of AI newsletter, an editorialized newsletter covering…

Hugging Face Open Source Apr 2, 2025 👁 14

Efficient Request Queueing – Optimizing LLM Performance

Back to Articles Efficient Request Queueing – Optimizing LLM Performance Team Article Published April 2,…

Hugging Face Open Source Mar 28, 2025 👁 16

🚀 Accelerating LLM Inference with TGI on Intel Gaudi

Back to Articles 🚀 Accelerating LLM Inference with TGI on Intel Gaudi Published March 28, 2025 Update on…

The AI Edge Research Mar 26, 2025 👁 18

Reduce AI Model Operational Costs With Quantization Techniques

Model quantization is becoming a core strategy for training and deployment! I am excited to introduce you to…

The AI Edge Research Mar 21, 2025 👁 19

How To Construct Self-Attention Mechanisms For Arbitrary Long Sequences

With Gemini models having a 2M tokens context size and Claude having a 200K tokens context size while having…

Air Street Capital Business Mar 2, 2025 👁 19

State of AI: March 2025 newsletter

Hi everyone!Welcome to the latest issue of the State of AI newsletter, an editorialized newsletter covering…

Air Street Capital Business Feb 2, 2025 👁 17

State of AI: January 2025 newsletter

Dear readers,Welcome to the latest issue of the State of AI newsletter, an editorialized newsletter covering…

Hugging Face Models Jan 16, 2025 👁 14

Introducing multi-backends (TRT-LLM, vLLM) support for Text Generation Inference

Back to Articles Introducing multi-backends (TRT-LLM, vLLM) support for Text Generation Inference Published…

Hugging Face Models Oct 22, 2024 👁 14

Transformers.js v3: WebGPU Support, New Models & Tasks, and More…

Back to Articles Transformers.js v3: WebGPU Support, New Models & Tasks, and More… Published October 22,…

Hugging Face Models Sep 13, 2024 👁 13

Accelerate 1.0.0

Back to Articles Accelerate 1.0.0 Published September 13, 2024 Update on GitHub Upvote 54 +48 Zachary Mueller…

Replicate Open Source Sep 9, 2024 👁 15

Fine-tune FLUX.1 with an API

Replicate Blog Fine-tune FLUX.1 with an API Posted September 9, 2024 by zeke Info You can now fine-tune…

Replicate Open Source Aug 30, 2024 👁 13

Fine-tune FLUX.1 to create images of yourself

Replicate Blog Fine-tune FLUX.1 to create images of yourself Posted August 30, 2024 by zeke Info Update (May…

Hugging Face Models Aug 19, 2024 👁 13

Deploy Meta Llama 3.1 405B on Google Cloud Vertex AI

Back to Articles Deploy Meta Llama 3.1 405B on Google Cloud Vertex AI Published August 19, 2024 Update on…

Replicate Open Source Aug 15, 2024 👁 17

Fine-tune FLUX.1 with your own images

Replicate Blog Fine-tune FLUX.1 with your own images Posted August 15, 2024 by deepfates Info You can now…

Aidan Cooper (AI) Business Aug 12, 2024 👁 20

The Convergence of Proprietary and Open Source LLMs

A few months ago, I wrote about building LLM systems with self-hosted, open source models that beat…

Benedict Evans Business Jun 20, 2024 👁 15

Apple intelligence and AI maximalism

No-one outside Apple has really used any Apple Intelligence features yet. It won't launch until the autumn,…

Hugging Face Models Jun 4, 2024 👁 20

Faster assisted generation support for Intel Gaudi

Back to Articles Faster assisted generation support for Intel Gaudi Published June 4, 2024 Update on GitHub…

Hugging Face Models May 9, 2024 👁 19

Subscribe to Enterprise Hub with your AWS Account

Back to Articles Subscribe to Enterprise Hub with your AWS Account Published May 9, 2024 Update on GitHub…

Benedict Evans Business May 4, 2024 👁 17

Ways to think about AGI

The manuscript for ‘A Logic Named Joe’ In 1946, my grandfather, writing as ‘Murray Leinster’,…

Hugging Face Models Mar 20, 2024 👁 18

GaLore: Advancing Large Model Training on Consumer-grade Hardware

Back to Articles GaLore: Advancing Large Model Training on Consumer-grade Hardware Published March 20, 2024…

Hugging Face Models Mar 18, 2024 👁 20

Quanto: a PyTorch quantization backend for Optimum

Back to Articles Quanto: a PyTorch quantization backend for Optimum Published March 18, 2024 Update on GitHub…

Hugging Face Models Jan 25, 2024 👁 19

Hugging Face and Google partner for open AI collaboration

Back to Articles Hugging Face and Google partner for open AI collaboration Published January 25, 2024 Update…

Hugging Face Open Source Dec 5, 2023 👁 16

AMD + 🤗: Large Language Models Out-of-the-Box Acceleration with AMD GPU

Back to Articles AMD + 🤗: Large Language Models Out-of-the-Box Acceleration with AMD GPU Published…

AI Tuts Image AI Oct 26, 2023 👁 12

How to use AnimateDiff in ComfyUI (vid2vid)

AnimateDiff is an open source technology released in July 2023 [Github][research paper]; it's one of the best…