GPT-5.2 derives a new result in theoretical physics
Also: GPT 4o leaving • Visualizing the Difference Between GPT-4o and GPT-5.2 on a Spatial Reasoning Benchmark (MineBench)
Show HN: Skill that lets Claude Code/Codex spin up VMs and GPUs
Also: I built an macOS app to monitor running Claude Code sessions • [Show & Tell] Herald — How I used Claude Chat to orchestrate Claude Code via MCP
Expensively Quadratic: The LLM Agent Cost Curve
Safe YOLO Mode: Running LLM agents in vms with Libvirt and Virsh
Anthropic Released 32 Page Detailed Guide on Building Claude Skills
Also: Pentagon Used Anthropic’s Claude in Maduro Venezuela Raid • Why can't Anthropic increase the context a little for Claude Code users?
The "AI agent hit piece" situation clarifies how dumb we are acting
CBP Signs Clearview AI Deal to Use Face Recognition for 'Tactical Targeting'
OpenAI has deleted the word 'safely' from its mission
IBM Triples Entry Level Job Openings. Finds Limits to AI
Stop Shipping AI Slop. Design with Weavy AI, Claude etc.
SWE-rebench Jan 2026: GLM-5, MiniMax M2.5, Qwen3-Coder-Next, Opus 4.6, Codex Performance
Also: Claude Code (Opus 4.6 High) for Planning & Implementation, Codex CLI (5.3) for Review & QA — still took 8 phases for a 5-phase plan • The gap between open-weight and proprietary model intelligence is as small as it has ever been, with Claude Opus 4.6 and GLM-5'