GPT-5.2 derives a new result in theoretical physics
Also: GPT-5.2 just contributed to a new theoretical physics result. Are we now in Stage 4 “AI as innovator” territory? • Difference Between Opus 4.6 and GPT-5.2 Pro on a Spatial Reasoning Benchmark (MineBench)
Show HN: Skill that lets Claude Code/Codex spin up VMs and GPUs
Also: [Show & Tell] Herald — How I used Claude Chat to orchestrate Claude Code via MCP • I built a Claude Code Skill that gives agents persistent memory — using just files
Expensively Quadratic: The LLM Agent Cost Curve
Safe YOLO Mode: Running LLM agents in vms with Libvirt and Virsh
Anthropic Released 32 Page Detailed Guide on Building Claude Skills
Also: WSJ: Pentagon Used Anthropic’s Claude in Maduro Venezuela Raid • Anthropic Released 32 Page Detailed Guide on Building Claude Skills
The "AI agent hit piece" situation clarifies how dumb we are acting
CBP Signs Clearview AI Deal to Use Face Recognition for 'Tactical Targeting'
OpenAI has deleted the word 'safely' from its mission
IBM Triples Entry Level Job Openings. Finds Limits to AI
Stop Shipping AI Slop. Design with Weavy AI, Claude etc.
SWE-rebench Jan 2026: GLM-5, MiniMax M2.5, Qwen3-Coder-Next, Opus 4.6, Codex Performance
Also: Claude Code (Opus 4.6 High) for Planning & Implementation, Codex CLI (5.3) for Review & QA — still took 8 phases for a 5-phase plan • The gap between open-weight and proprietary model intelligence is as small as it has ever been, with Claude Opus 4.6 and GLM-5'