Pi 1.0 sat at #1 on Hacker News today with 1,200+ points. Pi Durable, its sibling release, is also in the top 10.
The part that got people typing: the agent that spent over a year dunking on MCP now supports MCP. Here's what actually shipped, and why I think the design matters more than the drama.
Hey everyone! 👋
When expanding JED AI (https://jedcorp.ink) beyond automated vulnerability scans and terminal orchestration, we wanted to tackle a different kind of challenge: real-time global telemetry and visual intelligence.
Enter J3D Eye: our global camera intelligence and OSINT d...
My codebase Q&A agent ran 1,200 times in its first week. 388 of those runs died with the same HTTP 400. The other 812 worked fine, which is the worst possible ratio: too low to look like a broken deploy, too high to ignore.
The cause was one line I'd copied from my own earlier prototype: next(b for b in resp.content if b.type == "tool_use"). It grabs the first tool call...
TL;DR: “It passed yesterday” is not evidence about the code in front of you. A command result must be bound to the exact commit it observed. The accountability apparatus needs a record tied to the artifact, not a comforting memory.
You have seen the familiar green check on a branc...
LLM evaluation only becomes useful when every model faces the same prompts, the same fixed judge, per-axis rubrics, and a public verbatim trail. The LFORLA Reverse Engineering benchmark does exactly that: it restores C source from stripped binaries, scores with deterministic token similarity against server-only references, and publishes what each model actually answere...
A weird thing is happening in software development.
Building software has become ridiculously fast.
You can describe an application to an AI coding tool and have a working interface within minutes.
Need an API?
Generated.
Database schema?
Generated.
Authentication flow?
Probably generated too.
A solo developer can now build...