September 7, 2026super-user

Super User Daily: 2026-09-07

The most striking thing about this batch: the money stories are getting specific. A Japanese user had Claude Code build a one-person company that keeps selling 500-yen notes, an OpenClaw agent negotiated an insurance renewal and saved $973.70, and a two-person startup crossed $1M ARR with a customer-intelligence dashboard the non-technical cofounder built himself. Meanwhile the power users have moved past "one agent, one terminal" entirely - Mac Mini farms running dozens of agents in Slack, cross-harness inboxes where Codex workers ask Claude workers for function signatures, and a guy rotating 20 subscriptions across two laptops running 24/7. The dark cloud: Astra's launch turned the pricing-and-limits grumble into an open revolt, and Claude Code's own Rules system was caught silently not firing.
@buttanoteragoya [Claude Code]
Claude Code#1
https://x.com/buttanoteragoya/status/2096342607932879254
Asked Claude Code to "build a company where 500-yen notes keep selling" - and got a working one-person company. The setup ships with 30 evergreen topic datasets, a prompt that generates a 4,000-character note in one shot, an automatic detector for dated expressions, a template that turns one note into 10 promo posts, and a magazine structure to raise prices. The daily job is literally replying "OK" to proposals for 15 minutes. Past notes keep selling even in months when nothing new is written.
@undefinedKi [Claude Code]
Claude Code#2
https://x.com/undefinedKi/status/2096224107121516997
A two-person AI startup hit $1M annualized revenue two months after launch, and the layer behind it is the real story. Every customer call runs with an AI notetaker into Notion, then the non-technical cofounder pointed Claude Code at that database via MCP and built a "customer intelligence" dashboard in plain English. It scores every paying customer on product love and indexes the roadmap on that score instead of feature requests. The scoring surfaced their key insight: curiosity signups churn, "I need marketing now" signups retain.
@Ryrenz [Claude Code]
Claude Code#3
https://x.com/Ryrenz/status/2096162673821921424
A laid-off geophysicist open-sourced his entire job-search pipeline as ai-job-search, a skill pack built on Claude Code that now sits at 40k GitHub stars. One "apply" run rewrites the resume for the specific job description, compiles the PDF, checks that ATS systems can parse it, and only then hands it to the human for final approval; results sync to Notion and Gmail. He used it to send 69 tailored applications, got 20 first-round interviews, and landed an AI engineering job. Thirteen thousand forks say the "fork it and fill in your own profile" model works.
@nakamura [Claude Code]
#4
https://x.com/nakamura/status/2096178472108642553
A professor built everything his 35-person research lab needs to operate: 105 features and 11 services in 3 months, including shared authentication and even a lab payment service. His conclusion is that the era of "just build every service your lab needs" has arrived, and he is telling other PIs to do the same. Running a lab is a small-business workload, and it just got automatable by one person.
@vishnuunhsiv9 [OpenClaw]
OpenClaw#5
https://x.com/vishnuunhsiv9/status/2096308405367456142
Saved $973.70 on a home/auto umbrella insurance renewal by handing OpenClaw last year's declarations and asking it to find the best local agents. The agent emailed brokers, negotiated, and came back with better declarations at a better price. Insurance shopping is exactly the kind of chore humans skip because the payoff-per-hassle feels low - an agent flips that equation.
@fjpedrosa86 [OpenClaw]
OpenClaw#6
https://x.com/fjpedrosa86/status/2096292811410751691
Running a genuinely fun autonomy experiment: an OpenClaw agent with its own Revolut card and a 100-euro budget must order his weekday lunches, delivered between 14:00 and 14:15. The agent has to optimize price, restaurant reputation, his food preferences, and macro/calorie targets simultaneously. Giving an agent a card and a recurring physical-world job is the next trust level up from letting it edit files.
@jcfmunoz [OpenClaw]
OpenClaw#7
https://x.com/jcfmunoz/status/2096156313986195656
Used his OpenClaw agent (named Beru) with the App Store Connect API to upload 306 marketing images and create 16 localized language versions for a new release of his app - in 45 minutes, with contextual supervision. Then one more hour for a full ASO study that set the best keywords per language. App store operations, one of the most tedious parts of indie iOS development, compressed from days to a lunch break.
@voltsandvision [OpenClaw]
OpenClaw#8
https://x.com/voltsandvision/status/2096231745410498564
A contractor runs OpenClaw as "Jarvis" for work: he uploads screenshots to Telegram and the agent keeps his spreadsheets updated. Now he is planning a crew of specialized agents - estimator, project manager, logistics, trip planner - taught to quote his jobs. This is the trades-business version of the agent-org pattern, arriving from a self-described normie.
@thewebbie [OpenClaw]
OpenClaw#9
https://x.com/thewebbie/status/2096279017657643264
A look at what a non-toy OpenClaw install actually looks like: LaunchAgents keeping local services alive, Telegram delivery, Gateway/exec access, Tailscale routes, story-media workflows, backups, and personal automation glued around it. The boring reliability layer is the difference between a demo and a daily driver. As another user put it in the same thread - plenty of automations still run on OpenClaw, you just don't hear about them because they're boring and reliable.
@cyburke [OpenClaw]
OpenClaw#10
https://x.com/cyburke/status/2096303369984028881
The counterweight: a fresh OpenClaw 2.0 install that went sideways. He wanted only Meta's Muse Spark model, but Meta wasn't in the dropdown, there was no custom endpoint, setup kept grabbing Codex as fallback, and it imported his Claude and Codex sessions unasked. He had to use another agent to debug the installer. His verdict is fair: a 2.0 first-run experience should not require this kind of orchestration.
@rlaope [Claude Code]
Claude Code#11
https://x.com/rlaope/status/2096111328528318713
Enterprise-scale agent operations at Sionic AI: the CEO bought a fleet of Mac Minis and built an agent production factory, with dozens of agents covering service development, DevOps, AI research, and BD. Engineers stopped opening coding sessions - they message agents through Slack, which he strongly recommends over desktop/TUI because company context already lives there. The post includes hard-won config tips (disable tool-progress noise for Slack) and a candid admission: he barely opens Claude Code anymore, and sometimes has his agents trigger Claude Code just to burn the subscription quota he already paid for.
@hraness [Claude Code]
Claude Code#12
https://x.com/hraness/status/2096082136285528496
The most extreme manual setup we've seen: two laptops running 24/7 with the Codex app and Claude Code, no orchestrator tools at all. He tracks 20 subscriptions and their exact reset times in Apple Notes, rotates them by hand, and fires "monster prompts that will run for hours or days." He describes pacing weekly limits as a weird new skill - he can burn a Fable limit in under an hour if not careful. Rawdogging the swarm is a real workflow now.
@coreyganim [Claude Code]
Claude Code#13
https://x.com/coreyganim/status/2096183307822420031
A complete business AI stack, laid out plainly: Claude Code/Codex as the workbench, GrokBot as the workforce, Slack as the communication layer, a "Second Brain" living in GitHub, Composio for tool access, GHL for CRM. Deploying a new specialized agent takes 5 minutes, and every agent runs a weekly ingest skill that pushes learnings back to the Second Brain - so the whole system gets smarter as it works. That last loop is the difference between hiring agents and growing an organization.
@JohnRTyndall [Claude Code]
Claude Code#14
https://x.com/JohnRTyndall/status/2096348757411672447
Runs Claude Code and Codex agents simultaneously under a single agent lead, with a shared inbox built from hooks and MCP so workers can message each other across harnesses. He watched a Codex worker ask a Claude worker for the signature of a function still being written, get the answer, and build tests against it; two other workers negotiated checksum rules and cleaned up their scratch databases before signing off. Cross-vendor agent teams coordinating over an inbox - this is what the multi-agent future actually looks like in practice.
@0xrux [Claude Code]
Claude Code#15
https://x.com/0xrux/status/2096129729178849480
Breakdown of a setup where the human only ever talks to one agent: Pluto, a Chief of Staff GrokBot that routes everything to Ledger (accounting), Kin (referrals), and other specialists. The clever part: all bots share one VM whose terminal is logged into Claude Code and Codex, running Karpathy's LLM-council idea - Claude and Codex review each other's work automatically. And when work routes to Claude Code or Codex, it spends those subscriptions' tokens, not the GrokBot quota. Cost arbitrage between agent platforms has arrived.
@Austen [Claude Code]
Claude Code#16
https://x.com/Austen/status/2096238185760350233
Turned the model-council idea into an installable skill: it runs Grok, Claude Code, and Codex inside a Grok Bot and has them argue until they reach consensus. Adversarial multi-vendor review as a one-liner. The quiet assumption here - that no single lab's model should be trusted alone - is becoming a default working style.
@k8adev [Claude Code]
Claude Code#17
https://x.com/k8adev/status/2096294240934400258
His new workflow: discuss the implementation with ChatGPT (Astra), have it write the detailed technical spec, then let Astra itself coordinate and monitor the implementation in Claude Code through tmux. All driven by voice. One model as architect and supervisor, another as the hands - by talking, not typing.
@sanchitmonga22 [Claude Code]
Claude Code#18
https://x.com/sanchitmonga22/status/2096372075439235580
OpenAI shipped a plugin putting Codex inside Claude Code, so he built the reverse: Claude Code inside Codex. Astra writes the code, Claude reviews it, one prompt. His honest day-one take: Astra is a monster on benchmarks and half the cost per task, but Claude still caught things Astra wrote and walked right past. Different lab, different blind spots - which is exactly why the cross-review pattern works.
@claudecode84 [Claude Code]
#19
https://x.com/claudecode84/status/2096083785938854376
Published the actual folder structure that runs his night shift: CONTRACT.md with the shift rules, a harness folder with spend caps and timeouts, schedule.yml for when shifts fire, rubrics for code/writing/safety that catch what he misses, plus receipts (5,382 kept), a trace log, and a checkpoint file so the next shift resumes where the last stopped. He used to wake up and check what broke overnight; now he wakes up to dated, graded receipts awaiting review. Agent ops is becoming a folder you can copy.
@lidangzzz [Claude Code]
#20
https://x.com/lidangzzz/status/2096078265534292180
A sharp multi-agent philosophy: in any agent's eyes, every other agent's output is wrong until verified. He argues effective multi-agent systems are not 1000 agents having a party but agents that challenge each other - default-distrust, review-first, test-driven, refuse-to-proceed-without-verification. His trade: burn 3-10x tokens for 50% more reliability, and get either a verified deliverable or an honest "I failed." The anxiety of staring at a pile of unverified agent output is the real enemy this solves.
@ByteMohit [Claude Code]
#21
https://x.com/ByteMohit/status/2096269603592982835
Best articulation this week of why Plan Mode matters: a prompt that says "don't write files yet" is not a plan - if write tools are still in the schema, the model can still call them. Plan mode is an action space: read, search, ask, no mutation, enforced by the harness rather than begged for in prose. He wants a plan the human can reject, write tools removed until approval, and a bound on the planning itself. Skip the plan and you're not moving fast, you're paying to undo.
@martinvars [Claude Code]
Claude Code#22
https://x.com/martinvars/status/2096329310546141397
Spotify published how one engineer cut Claude Code token use by 90 percent: stop letting the frontier model do grunt work. Most of what an agent does is I/O - reading five files to answer one question, writing the twentieth lookalike test - so they routed bulk reads and boilerplate to a cheap model and kept Claude for hard problems. With engineering leaders already spending $200-500 per developer per month on tokens, his conclusion lands: route the workload or the token invoice owns you.
@socialwithaayan [Claude Code]
#23
https://x.com/socialwithaayan/status/2096190823012684265
A dense checklist of 12 cost levers for Anthropic bills that don't touch output quality: cache-first ($0.25/M cache reads), effort-per-route (most people set it one rung too high), task budgets to end runaway loops, the /claude-api cost-optimize and prompt-audit commands, plus community tools RTK, CodeGraph, Ponytail, and Caveman. His own audit found 14 prompt lines still aimed at Claude 3 in a skill file named final-final-v3. Cost literacy is now a discipline with its own tooling.
@SKatalystAI [Claude Code]
Claude Code#24
https://x.com/SKatalystAI/status/2096028799489933616
Zeroes in on the update that matters more than benchmarks: Fable 5.1 on medium effort roughly equals Fable 5 on high, and medium is now the default inside Claude Code per Anthropic's own Lydia. His frustration was never raw capability - it was constantly thinking about effort levels and whether a task was "worth" spending Fable on. If medium delivers most of the capability without chewing the session, the day-to-day experience changes more than any leaderboard.
@dotey [Claude Code]
#25
https://x.com/dotey/status/2096120681104622049
A heavy user's honest ledger after intensive Astra use: Fable is strong everywhere but so expensive that even on the $200 plan he never dared use it freely, and directing cheaper models with Fable degrades UI details. Astra now handles prototypes and iterates them with computer use that finds and fixes small bugs by itself - the bugs he used to find manually and then begrudge spending Fable tokens to fix. He's canceling one of his two Claude Max subscriptions and is glad Anthropic finally has real competition at the Fable tier.
@cwmasaki [Claude Code]
Claude Code#26
https://x.com/cwmasaki/status/2096080170520306123
The most thorough Codex-vs-Claude-Code comparison in the batch, from someone using both: Codex wins on no 5-hour cap on upper plans, honest 20x quota math, a better desktop app, computer use, smarter compaction, /goal ergonomics, image generation, and Luna's cost-performance. Claude Code still wins on hooks and managed services like Routines. His read on the cause is structural: Anthropic looks compute-constrained - renting GPUs at high prices, capping Fable at 50%, fewer resets - and can recover when that eases.
@kawasin73 [Claude Code]
Claude Code#27
https://x.com/kawasin73/status/2096030851217887431
Wrote an emergency blog post titled "Claude Code's Rules are already dead" that ricocheted around Japanese dev Twitter. The finding: path-scoped Rules can silently stop firing. When your guardrails depend on the harness's file-access patterns, a harness optimization can quietly delete your safety net.
@connect24h [Claude Code]
Claude Code#28
https://x.com/connect24h/status/2096150936876147036
The mechanism behind the "Rules are dead" alarm: Claude Code started using cat/sed/grep via Bash instead of its dedicated Read tool, which bypasses the triggers Rules and Read-based hooks rely on - path rules don't fire, subdirectory CLAUDE.md files go unread. The workaround people dug up is disabling CLAUDE_CODE_THRIFTY_SONIC. His framing is the keeper: the scary failure mode in AI coding isn't an error, it's a constraint you believed was active quietly vanishing. Harness engineering now needs regression tests of its own.
@rohanpaul_ai [Claude Code]
Claude Code#29
https://x.com/rohanpaul_ai/status/2096117661172384137
Claude Code creator Boris Cherny's advice at YC Startup School: every 6 months, delete your claude.md, delete your skills, delete your hooks, and see what the model does - it might surprise you. For Opus 5 he strongly recommends deleting all of it, because the model may no longer need instructions written for weaker predecessors. Your carefully tuned harness has a shelf life, and the expiry date is every model release.
@norvex1029 [Claude Code]
Claude Code#30
https://x.com/norvex1029/status/2096136031221207380
Inside the Claude Code team, some engineers say 70-80% of daily work now happens through an agent inside Slack - not the terminal, not the IDE. They give Claude a goal, it pulls context from team Slack conversations, runs loops, parallel research, code reviews and verification, and tightens until output is review-ready. And they keep deleting parts of their own harness as models improve. The lesson he draws: the advantage isn't building the most complicated workflow, it's knowing what you can stop doing.
@coinbureau [Claude Code]
Claude Code#31
https://x.com/coinbureau/status/2096207830571479363
A developer (credit: @mertcobanov) pointed Claude Code at his 4-year-old Android TV via Developer Options and had it debloat the thing - removing RAM-hogging apps from the pipeline, installing FLauncher, and tightening animation timings. Nothing uninstalled, every change reversible, no root, and he published a 4-step guide with the exact prompt. Old hardware revival may be the most relatable agent use case yet.
@om_patel5 [Claude Code]
#32
https://x.com/om_patel5/status/2096076206026195203
Wanted a spy-gadget watch as a kid, so he vibe-coded one for his Garmin: radar sweeps for nearby Bluetooth devices plotted by distance, a tailscan that flags signals still following you after you move, a seismic tool reading the accelerometer, GPS dead drops, and a fingerprint gate on the watch face. When prompt-designing the dial went nowhere, he vibe-coded his own browser-based watch editor with pixel precision - then used it as the renderer and shipped to 43 Garmin models in a single night. Build the tool that builds the thing.
@xevrion_the1 [Claude Code]
Claude Code#33
https://x.com/xevrion_the1/status/2096345842001367496
Hooked a cheap Bluetooth LED strip to his desktop wallpaper so the room recolors when the desktop does, 0.12 seconds behind. Claude Code cracked the strip's Bluetooth protocol - the entire "phone app" turned out to be 7 unencrypted bytes. The strip only allows one connection, so the vendor app is now permanently locked out. Reverse-engineering consumer hardware has become a weekend prompt.
@yuris [Claude Code]
#34
https://x.com/yuris/status/2096062416841028017
In three weeks: a fully featured multi-user CRM with email parsing and deep internal-data integrations, a robust Chrome extension, and a reverse-engineered plugin adding Affinity support natively inside Superhuman. All worked essentially bug-free on the first try, and he read zero lines of code - he even stopped using the terminal and switched to the app. His claim: for ~90% of the software you want for yourself, you no longer need to know how to program.
@bokuwalily [Claude Code]
Claude Code#35
https://x.com/bokuwalily/status/2096040820428419266
A counter-take on tool-hopping: he has shipped 13 iOS apps almost entirely with Claude Code, as a non-coder. His reasoning - learning one tool's quirks completely beats switching between tools, when you can't read the code yourself. Consistency is a feature when the agent is your only engineer.
@Minimalist_Quin [Claude Code]
Claude Code#36
https://x.com/Minimalist_Quin/status/2096231518720971228
Started building a portfolio two months ago with absolutely no idea what she was doing - just curiosity, Claude Code, trial and error, and a lot of "why isn't this working?" moments. Now proudly shipping it. The on-ramp for non-developers keeps getting shorter, and the emotional arc (confusion to pride in 8 weeks) is the actual product pitch.
@PoShenLoh [Claude Code]
Claude Code#37
https://x.com/PoShenLoh/status/2096060001026527689
The CMU mathematician spent a weekend and $100 creating an entire video ad campaign with a distinct identity - his prompt imagined an Apple/Buc-ee's/Michael Jackson collab - and published the full chain-of-thought workflow. He now runs 4 computers with Claude Code or Codex from a 5-screen workstation and says his productivity across the company has skyrocketed. His jobs take: GPT-6-class tools will shatter work into fragments where small teams do the work of thousands - and 50+ students showed up to his 18-person Friday class, apparently agreeing.
@godai_ceo [Claude Code]
Claude Code#38
https://x.com/godai_ceo/status/2096102951710548288
Edited a 40-minute video in 3 days entirely with AI, projecting ¥300k/month in cost savings: Claude Code (Fable 5.1) analyzed reference videos, Codex ran the main editing, Astra joined mid-project, HeyGen handled lip sync, Genspark and Seedance generated assets. A video-editing beginner shipping broadcast-length content by orchestrating five tools is the multi-agent pipeline story in miniature.
@irabukht [Claude Code]
Claude Code#39
https://x.com/irabukht/status/2096373848262275569
Got 200+ new customers from 4 product launches in 2 weeks, with each launch video made in Claude Code in under an hour - 300k+ views across the four. His honest framing: not influencer numbers, but absolutely worth it given how easy the videos are to make. Launch content as a near-free byproduct of the dev tool.
@every [Claude Code]
Claude Code#40
https://x.com/every/status/2096267894401409172
Every packaged "Compound Writing" as a free open-source plugin for Claude Code and Codex: before drafting, the model interviews you with 15-30 questions to draw out your thinking, then outlines, drafts, and reviews one section at a time, with reviewer personas like /hitchcock (does the reader lean in?) and /sorkin (does it move?). Sessions that teach it something get saved to voice and style files, so every piece starts ahead of the last. The anti-vending-machine model of AI writing, packaged.
@cyrilXBT [Claude Code]
Claude Code#41
https://x.com/cyrilXBT/status/2096179280661057704
The Karpathy second-brain recipe keeps spreading: point Claude Code at an Obsidian vault, drop any source into a raw folder, say "ingest this," and it links and files everything into a living wiki governed by a CLAUDE.md. Five minutes to set up, compounds like interest, and you never start from a blank chat again. Multiple variants of this workflow surfaced today, which tells you it's crossing from trick to standard practice.
@asahi_ai_x [Claude Code]
Claude Code#42
https://x.com/asahi_ai_x/status/2096140897683681390
A dead-simple anti-hallucination ritual: after Claude Code finishes research, he runs it again with one instruction - open every URL you cited and verify the text actually matches your claims. On a recent 12-URL check it found 6 discrepancies. Don't trust the AI; make the AI distrust itself. Cheap, mechanical, and it catches exactly the failure mode that burns publishers.
@dr_cintas [Claude Code]
#43
https://x.com/dr_cintas/status/2096297455092457532
The new /skill-doctor command is the audit tool skills needed: it lists every skill loaded in your session, which ones you've never used, and how many context tokens each is costing. His prompt: "grade which skills are actually earning their token cost and which are dead weight." Skills went from collect-them-all to measured payroll in one release.
@masahirochaen [Claude Code]
Claude Code#44
https://x.com/masahirochaen/status/2096069173134959060
His ritual for every major model release: run a full PC diagnosis - folders, skills, MCP servers, instruction files - plus vulnerability checks on existing projects, using a carefully scoped read-only prompt that works in both Claude Code and Codex. The insight is that your agent environment accumulates cruft exactly like a codebase, and a new model is the right moment to re-audit what's still earning its place. He published the full prompt, complete with strict no-write guardrails.
@ikinokore_3k [Claude Code]
Claude Code#45
https://x.com/ikinokore_3k/status/2096180414595715168
Deep write-up of text-to-cad v0.5: an open-source skill set that turns "I need this part" into actual STEP/STL/DXF manufacturing files by having the agent write build123d Python code, running on the same OCCT kernel as FreeCAD. Works with Claude Code, Codex, and Grok Build; demos include a W16 engine, a watch movement, and a 16-DOF robot hand built in about an hour and 300k tokens. His sober caveat: demos are strong, assembly-grade precision isn't there yet - but software people are already ordering sheet-metal parts on weekends.
@teach_fireworks [Claude Code]
Claude Code#46
https://x.com/teach_fireworks/status/2096252438277915035
A window into what open-source maintenance looks like now: he ran a full upgrade round on his fireworks-tech-graph diagram skill (used from Codex and Claude Code), catching a subtle bug where text width was computed correctly but a fixed CSS font size silently overrode the layout - the kind of thing that passes if you only read code and reports. He shipped Chinese line-wrapping, truncation checks, arrow collision fixes, and ran all 141 tests plus real renders across 12 styles. Agents wrote the code; the human's job was refusing to trust "looks done."
@josesilesdata [Claude Code]
#47
https://x.com/josesilesdata/status/2096146435733393793
reverse-skill turns your agent into a reverse engineer that knows which tool to use: drop an APK, a binary, obfuscated JS, or a CTF challenge, and it routes to the right methodology through 41 decision rules, checks which tools you actually have installed, stands up the missing toolchain on demand, and runs a repeatable flow with a timeline and evidence chain. It ships with a scope gate that blocks touching a target until you confirm authorization, and 163 regression cases run in CI on Windows and Ubuntu. Security tooling with governance built in, not bolted on.
@bkdgiffug [Claude Code]
Claude Code#48
https://x.com/bkdgiffug/status/2096253521020518767
ClaudeBrain gives Claude Code a security-research brain: a 500+ page offensive/defensive knowledge base wired to vulnerability-hunting skills covering XSS, SQLi, SSRF, RCE, API, cloud, and LLM security. It searches the knowledge base semantically before acting, tracks testing state to avoid repeated attempts, isolates client data, and checks for leak risks before submission. Built for authorized pentesting, bug bounty, and CTF work - the experience library finally becomes something the agent can actually call.
@NFTCPS [Claude Code]
Claude Code#49
https://x.com/NFTCPS/status/2096056832305594383
shuohao-skills runs a whole short-drama production pipeline inside Claude Code or Codex: feed it a novel and it extracts characters, builds the outline, generates scene and prop settings, writes the script, and cuts storyboards - producing material ready for AI video generation, using only your session quota with zero API keys or dependencies. While some people hand-type scripts line by line, others have automated the entire assembly line.
@AmiOtsuka_SE [Claude Code]
Claude Code#50
https://x.com/AmiOtsuka_SE/status/2096079781649695064
Published a book called "Claude Code Without Writing Code" - 16 chapters on handing an entire one-person company to AI over a year, typing nothing but Japanese sentences. Each chapter ends with one line to append to your next AI request, designed to be tried immediately. The non-programmer Claude Code literature is becoming its own genre, and it's shipping in paperback.
@tsutomu_kumagai [Claude Code]
Claude Code#51
https://x.com/tsutomu_kumagai/status/2096176593777623442
Ran the same prompt through Codex (Astra) and Claude Code (Fable 5.1): make a show-off video for social media. Astra storyboarded with generated images, composed music, added a whoosh sound effect synced to motion, and ended the music on the final cut - 16.5 minutes, about 30% of a 5-hour window. Fable's result he compares, brutally, to Mario Paint. His company runs Claude Code as its main tool, and he says this test alone forced a rethink.
@phreakv6 [Claude Code]
Claude Code#52
https://x.com/phreakv6/status/2096214499107938421
A year after switching to Claude, he's back on ChatGPT post-Astra - not for investing research (Claude suffices) but for computer use: he loves how it drives Blender, Unreal, and Fusion and wants his son using them. His observations: visual-spatial intelligence is off the charts, people are building game maps and building plans with it, and it's a massive driver of local compute - game-dev and CAD folks are buying powerful secondary machines just to run Astra. He calls it as much an inflection point as Claude Code was last year.
@kutaro_ai [Claude Code]
Claude Code#53
https://x.com/kutaro_ai/status/2096387388771844192
A self-described total AI beginner reached full YouTube Shorts automation with Claude Code after sleepless nights, and is now writing up the method for other beginners before tackling TikTok. Peak grind-story energy, but the underlying signal matters: the automation workflows that were power-user territory six months ago are now being reached by people who can barely code, one all-nighter at a time.
@brehm_shaun [Claude Code]
Claude Code#54
https://x.com/brehm_shaun/status/2096230759103180857
Devlog numbers from an idle reactor game built with Claude Code: ~60 features and fixes landed in 3 days, gated by ~250 automated probes on every change, hitting 115 fps after GPU work, headed to Steam Next Fest. One human at the controls. The probe-gate detail is the part worth stealing - velocity without a verification gate is just noise.
@nicochristie [Claude Code]
Claude Code#55
https://x.com/nicochristie/status/2096258046154743830
Ran sentiment analysis across 13 models on X: Astra has the highest day-one sentiment (+64%) and top intelligence score, Opus 5 is the only model net-negative, and people prefer Codex to Claude Code 121:60 - with 29 observed switches to Codex versus 9 the other way. The single biggest driver: limits and quota, where Claude Code's sentiment sits at -64. Fable 5.1 is admired for intelligence and dragged down by token costs. The market data version of everything in the replies.
🗣 User Voice
User Voice

1. The limits revolt is now the loudest signal in the dataset. @theo says he'd probably pay for a $10,000 Claude Code or Codex plan if it existed - while @ashen_one claims the $200 Claude plan delivers "barely 2x" real usage versus Codex's honest 20x, @JamesArslanSwe cancelled after filling his Claude Code limit in under an hour while Astra ran all night, and @xiaomabosn sarcastically thanks Anthropic for the Max 5x/20x "design." Users aren't asking for cheaper - they're asking for honest meters and bigger buckets.

2. Trust in the harness itself cracked this week. @kawasin73's "Rules are already dead" post and @connect24h's breakdown of silent rule-bypass hit a nerve because the failure was invisible - as @izutorishima puts it, Claude Code's harness design is starting to feel over-built and outdated versus Codex's cleaner loop. Users want regression tests and observability for their own guardrails, not just for their code.

3. The human is now the bottleneck. @ds_nakajima runs 5 parallel sessions in each of Codex and Claude Code and finds his own management capacity is the limiting factor; @hamko_intel notes the freed-up time just fills with more parallel tasks. Expect demand for orchestration layers, not more raw agent capability.

4. Cross-agent context portability is a rising ask. @ClaudeCode_UT highlights the pain of re-explaining project background every time you switch agents - a tax that grows with every new tool you adopt. Tools like handoff skills, memex, and claude-mem exist precisely because the platforms don't solve this natively.

5. Support and safety edges are hurting real users: @John_Bailey had Claude Code flatly refuse to build a clinical-trial search tool over imagined legal violations, @broadvideo_us lost a month of Pro access to an unresolved billing bug, and @Isseyv4z reports a ban with no stated reason. As agents become load-bearing infrastructure, support SLAs matter as much as model quality.
📡 Eco Products Radar
Eco Products Radar

Products and tools mentioned 3+ times in today's dataset:

Codex / GPT-6 Astra - the launch dominated every thread; the new default comparison point for Claude Code
Hermes Agent - session pickup from Claude Code/Codex (hermes --resume), positioned as the OpenClaw successor
Grok Bot - the "agent team" platform users route Claude Code/Codex work through
OpenClaw - v2026.9.2 shipped restart-resume; community split between production users and eulogies
Obsidian - the default local knowledge base for Claude Code second-brain workflows
Remotion - the code-driven video layer behind multiple Claude Code motion-design workflows
RTK - terminal-output compressor, cited repeatedly as a ~70% input-token saver
Ponytail - the "lazy senior dev" skill claiming 54% less code on measured runs
Humanizer - 42k-star skill for stripping AI tells from writing, now with a Chinese fork
Video Use - timeline-free video editing skill (24k stars) that reviews its own cuts
skill-doctor - the new session-skill auditor everyone ran this week
M3E Canvas - drag-and-drop Material 3 UI builder that exports prompts for coding agents
Magnitude - local-model advisor/runtime that benchmarks your machine and configures your agent
claude-mem / memex / Mnemosyne - the cross-agent memory layer race, all three trending
Chrome DevTools MCP - Google's official browser-debugging MCP, 50k+ stars
OfficeCLI - Word/Excel/PPT manipulation binary built for agent pipelines
free-claude-code / cc-switch / router setups - proxies pointing the Claude Code harness at cheap or free models
Composio - the tool-access layer in multiple business-stack posts
← Previous
Your LLM Judge Drifts, and Its Reasons Are Confabulated
Next →
Loop Daily: 2026-09-07
← Back to all articles

Comments

Loading...
>_