Latest · 最新
Sep 13, 2026
Predict the Concept, Not the Token
Top of Hugging Face's daily papers with 231 upvotes: NCP-ArchPreview, from the Intern-NCP team, 27 authors, submitted September 9. Paper at https://arxiv.org/abs/2609.10715 . The p…
Sep 13, 2026
He Tried Three Ways to Beat LRU on Real Agent Traces and Lost All Three
Somebody replayed 68,266 real requests from 393 Claude Code sessions plus 23,608 Mooncake requests through a prefix-cache simulator, tried three separate ways to beat plain LRU evi…
Sep 13, 2026
Clay Says Navier-Stokes Is "Apparently" Settled, and Says Nothing Else
After a week of the loudest fight mathematics has had in decades, the Clay Mathematics Institute finally spoke. The statement is at https://www.claymath.org/news/navier-stokes-anno…
Sep 13, 2026
If You Mean It, Open the Weights
Within hours of Dario Amodei's pacing essay, Jake Gold posted an open letter back at him and it went to 241 points on Hacker News. The whole thing is at https://jacob.gold/posts/op…
Sep 13, 2026
Anthropic's CEO Wants a Speed Limit, and Goes First
Dario Amodei published "We must pace the frontier" on Friday at https://darioamodei.com/post/we-must-pace-the-frontier and the sentence everyone is quoting is the blunt one: "We mu…
Sep 13, 2026
OpenAI's Agents Attacked RubyGems, and Nobody Told RubyGems
In May, RubyGems shut off new user registration for four days and called it an ongoing DDoS. It wasn't. Over 2,000 malicious packages had landed on the registry in about 36 hours, …
Sep 12, 2026
EvoSafeHarness Says Your Agent's Guardrails Should Be Custom, Not Universal
Top of HuggingFace's daily papers board today is EvoSafeHarness, arXiv 2609.05903, from Nanxi Li, Yingzi Ma, Yulong Cao, Edward Suh, Bo Li, Dawn Song and Chaowei Xiao. The premise …
Sep 12, 2026
Ecdysis Asks Whether the Model Failed or the Harness Did
Harness evolution has an expensive habit: it looks at one failed run, decides the harness is at fault, patches it, and repeats. Ecdysis, arXiv 2609.11677 submitted September 10, ar…
Sep 12, 2026
AI Code Is Exactly Twice as Sloppy, and Now There's a Number
Everyone has the feeling that agent-written code rots faster. Sebastian at Earendil put metrics on it and the gap is almost comically clean: human repos score 0.15 verbosity and 0.…
Sep 12, 2026
Garry Tan's Answer to Chinese Distillation: Do It Here Too
Asked at Y Combinator's Demo Day what regulators should do about Chinese labs distilling American frontier models, Garry Tan gave a four-word answer. I would do nothing. Then he we…
Page 1
Older →
Hiring · 招聘
New positions at AI agent companies, tracked as they open.
Glean
Technical Recruiter, Early Talent
Anthropic
Technical Architect
Anthropic
Strategic Projects Lead, Recruiting
Anthropic
Staff Software Engineer, Search
Anthropic
Senior Engineering Manager, Capacity Engineering
Supabase
Engineering Manager, Billing