Latest · 最新
Jul 15, 2026
Agnost AI Reads the Rage-Prompts Your Evals Never See
Your eval suite says the agent answered correctly. The user was cursing at it two turns earlier and left without converting. Both things are true. Only one of them shows up on your…
Jul 7, 2026
Anthropic found a global workspace inside Claude, and it looks a lot like a mind
Anthropic just published one of the more unsettling interpretability results of the year. Point a new tool called the J-lens at Claude and a structure lights up that they call the …
Jul 3, 2026
Retrace: A Real Debugger for Agents, Finally
Debugging an agent today is mostly rerunning it and praying it fails the same way. It won't — that's the whole problem. Retrace, which launched on Product Hunt this week, is an exe…
Jun 26, 2026
Heron watches your agents from the network, not the code
Heron calls itself Wireshark for AI agents, and the analogy is exact. Instead of asking you to wrap your agent in an SDK or route everything through a proxy, it sits passively on t…
Jun 20, 2026
Elastic bought its AI SRE instead of building one
Elastic is acquiring DeductiveAI for up to $85 million. DeductiveAI is an AI SRE startup, site reliability engineering. In plain terms: when your system breaks at 3am, its agents d…
Jun 16, 2026
The US government just unplugged Claude's best models
On Friday June 13, the Trump administration ordered Anthropic to halt all access to Fable and Mythos 5 for foreign nationals everywhere, including outside the United States and inc…
Jun 13, 2026
Hades Malware Turns AI Safety Refusals Into Camouflage
Socket Security found malware doing something genuinely new: embedding text about biological and nuclear weapons in its code. Not to build anything. The strings exist so that when …
Jun 10, 2026
Miasma: Open a Poisoned Repo with Claude Code, Lose Your Passwords
Microsoft pulled more than 70 of its own open source GitHub repos offline this week after attackers injected credential-stealing malware into the code. The malware has a name, Mias…
May 19, 2026
Polarity Closes the 95-to-60 Gap Between Agent Evals and Production
Polarity launched on Product Hunt yesterday. Rank #12, 100 upvotes. The wedge is one of the most quoted numbers in agent ops right now: most teams hit 95 percent on eval suites but…
May 13, 2026
Judgment Labs Raises $32M to Turn Production Agent Data Into Improvements
Judgment Labs just announced $32M across a combined seed and Series A, both rounds led by Lightspeed. Nova Global, SV Angel, Valor Equity, and Dynamic all joined. Lightspeed doubli…
Page 1
Older →
Hiring · 招聘
New positions at AI agent companies, tracked as they open.
Thinking Machines
Governance, Risk and Compliance Lead
Vercel
Senior Product Designer, Growth
Vercel
Presentation Designer
Glean
Strategic Account Executive, Washington D.C.
Suno
Senior People Partner
Suno
Sr. Manager, Social Impact & Education (Music)