9 articles · page 1 of 1
Hot Chips 2026: CUDA Targets RISC-V – By Chester Lam
2026-08-24
Auto-research with codex: How I achieved a 232x Faster Kernel
2026-08-15
Self-hosting Kimi K3: 20% more hardware cost, 20% better task resolution
2026-07-29
RTX 5080 and RTX 3090 Setup: 80 Tok/s on Qwen 3.6 27B Q8
2026-06-13
Zero-Copy GPU Inference from WebAssembly on Apple Silicon
2026-04-18
Taking on CUDA with ROCm: 'One Step After Another'
2026-04-12
MegaTrain: Full Precision Training of 100B+ Parameter LLMs on a Single GPU
2026-04-08
Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
2026-03-19
A CPU that runs entirely on GPU
2026-03-04