741 articles · page 3 of 15
Does code cleanliness affect coding agents? A controlled minimal-pair study
2026-07-05
Dungeon Proof Crawler: learn how to write proofs with RPG
2026-07-05
Mark Zuckerberg tells staff that AI agents haven't progressed enough
2026-07-05
Knowledge Should Not Be Gated
2026-07-05
The Log is the Agent
2026-07-05
Potential session/cache leakage between workspace instances or consumer accounts
2026-07-04
Agentic coding notes from Galapagos Island
2026-07-04
Leanstral 1.5: Proof abundance for all
2026-07-03
Claude, please stop trying to memorize random crap
2026-07-03
Costco is the anti-Amazon
2026-07-03
It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)
2026-07-03
Give Smart People the Tools to Do Smart Things
2026-07-03
Zuckerberg 'Admits' Meta's Layoffs Were Ineffective
2026-07-03
The short leash AI coding method for beating Fable
2026-07-02
Claude's AskUserQuestion: "No response after 60s – continued without an answer"
2026-07-02
CursorBench 3.1
2026-07-02
Senior SWE-Bench: open-source benchmark that assesses agents as senior engineers
2026-07-02
Fable 5 Is Back
2026-07-01
Claude Fable 5 Promotional Access
2026-07-01
ZCode: Claude Code from the Makers of GLM
2026-07-01
Weave Robotics launches Isaac 1, a $7,999 home robot with Fall 2026 deliveries
2026-07-01
Claude Fable 5 export control lifted
2026-07-01
Claude Code Just Got 5x More Expensive
2026-06-30
Claude Sonnet 5
2026-06-30
Claude Science
2026-06-30
Claude Code Is Steganographically Marking Requests
2026-06-30
Micro-Agent: Beat Frontier Models with Collaboration Inside Model API
2026-06-29
Ornith-1.0: self-improving open-source models for agentic coding
2026-06-29
Mag 7 starting to underperform [pdf]
2026-06-29
Herdr: Agent multiplexer that lives in your terminal
2026-06-29
Xonaly – Canada's Independent Search Engine
2026-06-28
Do LLMs pass the mirror test?
2026-06-28
GLM 5.2 beats Claude in our benchmarks
2026-06-28
I used Claude Code to get a second opinion on my MRI
2026-06-28
Tokenmaxxing is dead, long live tokenmaxxing
2026-06-28
A way to exclude sensitive files issue still open for OpenAI Codex
2026-06-28
Wayfinder Router: deterministic routing of queries between local and hosted LLM
2026-06-28
Anonymous GitHub account mass-dropping undisclosed 0-days
2026-06-27
Post-Mythos Cybersecurity: Keep calm and carry on
2026-06-27
Anatomy of a Failed (Nation-State?) Attack
2026-06-27
The Exhaustion of Talking to a Tool
2026-06-26
Captcha proves you're human. HATCHA proves you're not
2026-06-26
OpenAI to Stagger Release of GPT 5.6 at Request of U.S. Government
2026-06-25
Bible as RAG Database
2026-06-25
I rewrote PostHog's SQL parser, 70x faster, while barely looking at the code
2026-06-24
Computer use in Gemini 3.5 Flash
2026-06-24
The CAPTCHA arms race: from distorted text to browser identity
2026-06-24
RubyLLM: A Ruby framework for all major AI providers
2026-06-24
Haystack: Open-Source AI Framework for Production Ready Agents, RAG
2026-06-24
Qwen-AgentWorld: Language World Models for General Agents
2026-06-24