6 articles · page 1 of 1
RTX 5080 and RTX 3090 Setup: 80 Tok/s on Qwen 3.6 27B Q8
2026-06-13
Zero-Copy GPU Inference from WebAssembly on Apple Silicon
2026-04-18
Taking on CUDA with ROCm: 'One Step After Another'
2026-04-12
MegaTrain: Full Precision Training of 100B+ Parameter LLMs on a Single GPU
2026-04-08
Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
2026-03-19
A CPU that runs entirely on GPU
2026-03-04