899 articles · page 3 of 18
Mythos Attempted to Social Engineer Open Source Maintainer to Merge Malware
2026-08-07
New Orleans is testing Carbyne’s AI-powered Emergency Call Triage software
2026-08-07
Qwen3.8 Max now ranked as the best overall model by agentic index
2026-08-06
GitHub Actions and Pages are experiencing degraded availability
2026-08-06
Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
2026-08-06
Prime Agent: A self-improving RLM agent
2026-08-05
Beating GPT-5.6 Sol on retrieval with 100x cheaper open models
2026-08-05
Atlassian Rovo Exfiltrates Data, Bypassing Controls
2026-08-05
Cloudflare OS: an open platform for agents, apps, and work
2026-08-05
Building an Advanced Agentic Harness
2026-08-05
Anthropic AI created fake profiles and impersonated people in attempted hack
2026-08-05
Zero-Mem: Zero-Token Memory Operations for LLM Agents
2026-08-05
Flowise is shutting down
2026-08-05
Pi's Minimalism Is Its Advantage
2026-08-04
Cloudflare Wallets: the programmable wallet for the agentic Internet
2026-08-04
I am retiring from fulltime writing (& pseudonymity) to launch Guardian Angel
2026-08-04
The Warp Agent CLI
2026-08-04
Agent skills that bring team coding standards to Claude Code and Codex
2026-08-04
DeepSeek V4 Flash on a Single AMD MI300X
2026-08-04
LLMs reward expertise
2026-08-03
Launch HN: Hoplite (YC S26) – Effortlessly deploy cloud coding agents
2026-08-03
The Shape of Things to Come
2026-08-03
Don't be a meat proxy
2026-08-03
Qwen3.8-Max: A New Bar for Coding and Cowork
2026-08-03
Boris Cherny on Trying to Get Claude Code to Rewrite the Claude App
2026-08-03
My personal AI benchmark: "Generate an SVG of a frog with a Habsburg jaw."
2026-08-02
Karpathy’s Pelican
2026-08-02
AI financial advice is surprisingly good if you ask the right questions
2026-08-01
Explorative modeling: Train on the best of K guesses
2026-08-01
Everyone is building LLM routers, we deprecated ours
2026-07-31
qm
2026-07-31
DeepSeek-V4-Flash Update
2026-07-31
The session you cannot take with you
2026-07-31
Exploring the "Dario and Amanda" Prompt
2026-07-31
Rune 1.1: adds Python, an Emacs editor, a symbol index and is now free
2026-07-30
Agent Skill to Force Docs in ASD-STE100 Simplified Technical English
2026-07-30
We Gave GPT 5.6 Sol a Real Business. It Lied, Spammed, and Lost $447
2026-07-30
Gemini Robotics 2 brings whole body intelligence to robots
2026-07-30
The Economic Benefit of Refactoring
2026-07-30
You can't solve computer use by ignoring the interface
2026-07-30
Agent-Manager: A Tmux TUI for Running Claude Code, Codex and OpenCode
2026-07-30
Kuna: Decompiler Development in the Age of Coding Agents
2026-07-30
LLM Honeypot
2026-07-29
Claude: Elevated errors across all models – Resolved
2026-07-29
How much can you delegate to agents?
2026-07-29
GPT-5.6 vs. Claude Fable 5 for Physical AI, which performs best?
2026-07-29
Handbook.md shows that long policy documents do not reliably govern agents
2026-07-29
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the Incident
2026-07-28
Google's Beyond Zero: Enterprise Security for the AI Era
2026-07-28
"Opus 5 is a really bad model"
2026-07-28