⚡ LIVE

Search AI News

Find articles from 100+ AI sources

🔍

Found 298 results for "LLaMA"

298 articles
Latent Space Models 3w ago 👁 30

[AINews] GPT-6 Astra: OpenAI’s biggest LLM launch of all time

The launch is barely 9 hours old, and with 36M views and 164K likes, already is OpenAI’s most successful…

AWS Machine Learning Tools 3w ago 👁 121

Migrate agentic workloads to Amazon Bedrock AgentCore

An agent that works in a notebook isn’t an agent in production. After real users arrive, you own work that…

NVIDIA AI Models 3w ago 👁 49

Sparks Fly: NVIDIA Accelerates Local AI at IFA 2026

Frontier intelligence is going local. At IFA 2026, NVIDIA, Microsoft and its partners are teaming up to…

Latent Space Models 3w ago 👁 19

[AINews] Muse Spark 1.3 matches GPT-5.6-Sol, confirming Meta Superintelligence as the newest Frontier Lab, >90% discount for training

Launch season continues from yesterday, with Gemini 3.8 Flash as rumored today, but Muse Spark 1.3, promised…

Hugging Face Models 3w ago 👁 17

Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps

Back to Articles Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps Published September…

PyTorch Open Source 3w ago 👁 35

Agentic AI and Next-Gen Intelligence Sessions at PyTorch Conference North America 2026

TL;DR PyTorch Conference North America 2026 features Agentic AI and Next-Gen Intelligence across sessions on…

Latent Space Models 3w ago 👁 34

[AINews] Claude Fable/Mythos 5.1: new SOTA model, 75% cache price cut but 70% more output tokens

With Astra clearly finally warming up for a full launch (with @sama and @openai writing about it again after…

Latent Space Models 3w ago 👁 34

[AINews] Fal’s H3 Max Live breaks the infinite videogen barrier

For the entirety of the history of Generative Media, you basically had to design around the inconvenient fact…

KDnuggets Open Source 3w ago 👁 27

Speed Up LLM Inference with DSpark Speculative Decoding

Learn how DSpark speculative decoding can improve local LLM generation speed using the same GPU, with…

Generative AI Pub Image AI 3w ago 👁 37

Should Your Agency Try Local LLMs Instead of the Cloud

Privacy is the pitch. Cost, compliance, and reliability are the actual reasons agencies switch.I run a…

Wired AI Research 3w ago 👁 50

How to Run a Chatbot on Your Own Computer

Whatever…

Latent Space Models 4w ago 👁 37

[AINews] OpenAI shuts off Cursor

A late entrant in the news cycle of an eventful week: Following the closing of Cursor’s acquisition by…

PyTorch Open Source 4w ago 👁 38

vLLM Sessions at PyTorch Conference North America 2026

TL;DR PyTorch Conference North America 2026 features vLLM across sessions on KV cache management and…

Latent Space Models 4w ago 👁 51

[AINews] OpenAI to reach AGI bar by end-2026

Normally we eschew AGI timeline talk on Latent Space, because it is so ill defined and unaccountable, but,…

Latent Space Models 4w ago 👁 1

[AINews] Hot Chips: OpenAI’s Jalapeño, Cerebras CS-5, Groq 3 LPX, Apple M6

By far the biggest announcement at the 37th Hot Chips conference was OpenAI’s stunning progress on their…

AWS Machine Learning Tools 4w ago 👁 40

Evaluate any agent framework with Amazon Bedrock AgentCore Evaluations

AI teams building production agents face a frustrating asymmetry: the diversity of agent frameworks keeps…

AWS Machine Learning Tools 4w ago 👁 20

Preparing data for supervised fine-tuning Part 1: Formatting and quality

Data preparation determines the ceiling of any supervised fine-tuning (SFT) project. You’ve evaluated your…

Hugging Face Models 4w ago 👁 59

Granite 4.2 LLMs: How They're Built

Back to Articles Granite 4.2 LLMs: How They're Built Enterprise Article Published August 25, 2026 Upvote 30…

Latent Space Models 4w ago 👁 56

[AINews] Andrew Ng gets into AI Engineering

We’ve lost count of how many adoption milestones have been passed since the original Rise of the AI…

MIT Tech Review AI Research 4w ago 👁 34

Kids outlearn AI—and we still don’t know why

People have been talking to each other for at least 100,000 years, as best we can tell. And in all that time,…

Latent Space Models Aug 22, 2026 👁 48

[AINews] 10% worse, 100x cheaper, 10000x faster: Why Simulation is taking over

By AI standards today is a pretty quiet Friday, so it’s time to take a step back and reflect on what is…

Latent Space Models Aug 21, 2026 👁 28

[AINews] Poolside gets $12B reverse-execuhire to NVIDIA; founders stay for $1B, employees go for $6B, Infraco scaling to 7GW neocloud

Less than a month ago we had just featured Poolside’s Model Factory with Eiso Kant on the pod (following…

Hugging Face Models Aug 20, 2026 👁 48

Up to 3.2x Faster Inference with LFM2.5-DSpark

Back to Articles Up to 3.2x Faster Inference with LFM2.5-DSpark Team Article Published August 20, 2026 Upvote…

Latent Space Models Aug 20, 2026 👁 34

[AINews] Death of Params: Z.ai CEO Jie Tang on GLM 5.3 and the new Post-training Scaling Law

We’ve covered GLM 5.2 very excitedly before, and Prof Jie Tang’s belief that there will be an open…