Search AI News
Find articles from 100+ AI sources
Found 129 results for "neural network"
Generating Human-level Text with Contrastive Search in Transformers 🤗
Back to Articles Generating Human-level Text with Contrastive Search in Transformers 🤗 Published November…
Deep Dive: Vision Transformers On Hugging Face Optimum Graphcore
Back to Articles Deep Dive: Vision Transformers On Hugging Face Optimum Graphcore Published August 18, 2022…
A Gentle Introduction to 8-bit Matrix Multiplication for transformers at scale using transformers, accelerate and bitsandbytes
Back to Articles A Gentle Introduction to 8-bit Matrix Multiplication for transformers at scale using Hugging…
Industry Perspective: Tree-Based Models vs Deep Learning for Tabular Data
For tabular data, gradient boosted trees (GBTs) perform better than neural networks (NNs). This is common…
Advantage Actor Critic (A2C)
Back to Articles Advantage Actor Critic (A2C) Published July 22, 2022 Update on GitHub Upvote 9 +3 Thomas…
Policy Gradient with PyTorch
Back to Articles Policy Gradient with PyTorch Published June 30, 2022 Update on GitHub Upvote - Thomas…
Learning to play Minecraft with Video PreTraining
We trained a neural network to play Minecraft by Video PreTraining (VPT) on a massive unlabeled video dataset…
Convert Transformers to ONNX with Hugging Face Optimum
Back to Articles Convert Transformers to ONNX with Hugging Face Optimum Published June 22, 2022 Update on…
Director of Machine Learning Insights [Part 3: Finance Edition]
Back to Articles Director of Machine Learning Insights [Part 3: Finance Edition] Published June 14, 2022…
The Annotated Diffusion Model
Back to Articles The Annotated Diffusion Model Published June 7, 2022 Update on GitHub Upvote 365 +359 Niels…
Deep Q-Learning with Space Invaders
Back to Articles Deep Q-Learning with Space Invaders Published June 7, 2022 Update on GitHub Upvote 2 Thomas…
How Sempre Health is leveraging the Expert Acceleration Program to accelerate their ML roadmap
Back to Articles How Sempre Health is leveraging the Expert Acceleration Program to accelerate their ML…
Machine Learning Experts - Sasha Luccioni
Back to Articles Machine Learning Experts - Sasha Luccioni Published May 17, 2022 Update on GitHub Upvote -…
Accelerate Large Model Training using PyTorch Fully Sharded Data Parallel
Back to Articles Accelerate Large Model Training using PyTorch Fully Sharded Data Parallel Published May 2,…
~Don't~ Repeat Yourself
Back to Articles Don't Repeat Yourself* Published April 5, 2022 Update on GitHub Upvote 55 +49 Patrick von…
Utility vs Understanding: the State of Machine Learning Entering 2022
The empirical utility of some fields of machine learning has rapidly outpaced our understanding of the…
Accelerating PyTorch distributed fine-tuning with Intel technologies
Back to Articles Accelerating PyTorch distributed fine-tuning with Intel technologies Published November 19,…
Scaling up BERT-like model Inference on modern CPU - Part 2
Back to Articles Scaling up BERT-like model Inference on modern CPU - Part 2 Published November 4, 2021…
Introducing Optimum: The Optimization Toolkit for Transformers at Scale
Back to Articles Introducing 🤗 Optimum: The Optimization Toolkit for Transformers at Scale Published…
Scaling-up BERT Inference on CPU (Part 1)
Back to Articles Scaling up BERT-like model Inference on modern CPU - Part 1 Published April 20, 2021 Update…
DALL·E: Creating images from text
We’ve trained a neural network called DALL·E that creates images from text captions for a wide range of…
CLIP: Connecting text and images
We’re introducing a neural network called CLIP which efficiently learns visual concepts from natural…
Block Sparse Matrices for Smaller and Faster Language Models
Back to Articles Block Sparse Matrices for Smaller and Faster Language Models Published September 10, 2020…
AI and efficiency
We’re releasing an analysis showing that since 2012 the amount of compute needed to train a neural net to…