10 articles · page 1 of 1
Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
2026-08-25
AirLLM 70B inference with single 4GB GPU
2026-08-03
A $500 RL fine-tune of a 9B open model beat frontier models on catalog review
2026-07-28
Kimi K3 Now Available via Telnyx Inference API
2026-07-27
Ornith-1.0: self-improving open-source models for agentic coding
2026-06-29
AI Coding at Home Without Going Broke
2026-06-13
Granite 4.1: IBM's 8B Model Matching 32B MoE
2026-04-30
Kimi vendor verifier – verify accuracy of inference providers
2026-04-20
LLM Neuroanatomy II: Modern LLM Hacking and Hints of a Universal Language?
2026-03-24
Sarvam 105B, the first competitive Indian open source LLM
2026-03-07