|
How DeepSeek Is Running AI Coding Costs Into the Ground
|
|
1
|
5
|
3 August 2026
|
|
Running a 80B AI Model on a 8GB GPU, FAST (oLLM Guide)
|
|
1
|
0
|
2 August 2026
|
|
Stanford CS229 Machine Learning | Spring 2026 | Lecture 16: Basic Concept in RL, Policy Gradient
|
|
1
|
1
|
31 July 2026
|
|
Stanford CS229 Machine Learning | Spring 2026 | Lecture 14: Transformers, In-Context Learning
|
|
1
|
1
|
31 July 2026
|
|
Building the Automated AGI Lab: Core Automation's Jerry Tworek and Rohan Anil
|
|
1
|
0
|
29 July 2026
|
|
Speculative Decoding: How a Dumb Model Makes LLMs 3x Faster
|
|
1
|
0
|
29 July 2026
|
|
I Trained the TinyStories Model to Do Math
|
|
1
|
1
|
27 July 2026
|
|
Breaking Down Kimi K3's Architecture (Even For the Non-Technical)
|
|
1
|
1
|
23 July 2026
|
|
From Tokens to Cells: Foundation Models for Single-Cell Biology - Akram Baharlouei, Altos Labs
|
|
1
|
3
|
19 July 2026
|
|
DeepSeek's Deleted Vision Paper Is Nuts
|
|
1
|
1
|
8 July 2026
|
|
DeepSeek's New AI Speed Hack Is Amazing
|
|
1
|
1
|
7 July 2026
|
|
Deepseek drops another HUGE breakthrough
|
|
1
|
4
|
3 July 2026
|
|
LLM that loops instead of Doing Chain-of-Thought
|
|
1
|
3
|
1 July 2026
|
|
Research to Reality: Bringing Frontier ML Research to Production - Vaidas Razgaitis, Higharc
|
|
1
|
2
|
28 June 2026
|
|
Memory and Continual Learning: Engram's Dan Biderman and Jessy Lin
|
|
1
|
0
|
24 June 2026
|
|
The First Real LLM Breakthrough Is Here... SubQ (1000x Less Compute)
|
|
1
|
3
|
18 June 2026
|
|
GLM 5.2 - The Top NEW Open Weights Model
|
|
1
|
1
|
17 June 2026
|
|
World Models - The Next Big AI Revolution | AI Animated
|
|
1
|
1
|
15 June 2026
|
|
DeepMind Was Two Steps Ahead, AGAIN!
|
|
1
|
2
|
14 June 2026
|
|
LMS Are About to Hit a wall - The AI Scaling Law Might Be Breaking
|
|
1
|
0
|
11 June 2026
|
|
The Model Doesn't Unpack Its Memory
|
|
1
|
2
|
10 June 2026
|
|
Road to 5 Million Tokens: Breaking Barriers in Long Context Training — Max Ryabinin, Together AI
|
|
1
|
2
|
8 June 2026
|
|
DeepMind’s New AI Found A Strange New Way To Think
|
|
1
|
13
|
5 June 2026
|
|
Text Diffusion — Brendon Dillon, Google DeepMind
|
|
1
|
2
|
4 June 2026
|
|
I was wrong about ai
|
|
1
|
1
|
31 May 2026
|
|
How to 2x Speed LOCAL AI for only 265MB RAM 🤯 | MTP + Qwen Guide
|
|
1
|
2
|
23 May 2026
|
|
Personalization in the Era of LLMs - Shivam Verma, Spotify
|
|
1
|
5
|
19 May 2026
|
|
KV Cache as the New AI Memory Abstraction
|
|
1
|
1
|
14 May 2026
|
|
Multi-Token Prediction (MTP): Accelerating Local Models with no Quality Loss
|
|
1
|
3
|
13 May 2026
|
|
Predictive vs Generative AI: How They Work and When to Use Each
|
|
1
|
3
|
11 May 2026
|