A guy dropped a computer into the simulation his Astra agents live in. One agent sat down and built a simulation of his own, with its own agents living inside. Simulations all the way down. β Astra agents were given coding autonomy and a computer capable of running simulations; one reportedly created a nested agent simulation. (π 1261 | π¬ 120) π Link
Codex just saved me from spending $2,000 on an internet problem Iβve had for 8 years β Troubleshooting identified existing coax outlets as a potential alternative to costly Ethernet installation after mesh-network restarts improved speeds. (π 579 | π¬ 118) π Link
Privacy concerns after Navier-Stokes β Discusses OpenAIβs statement that user feedback and de-identified data may improve ChatGPT and Codex, raising questions about research-data provenance. (π 120 | π¬ 74) π Link
The Matrix Fly is literally the MVP for Rokoβs Basilisk and should be regulated now β Argues that increasingly capable simulations of biological brain activity require safeguards against harmful experimentation and public misuse. (π 84 | π¬ 85) π Link
I built Chess Cubed with GPT-6 Astra in 4 days β Describes using Codex, Blender MCP, Babylon.js, and Unreal Engine MCP to build a cube-based chess game with multiplayer and CPU modes. (π 62 | π¬ 31) π Link
"one thing I would NOT do" β User reports recurring unsolicited safety warnings in GPT-5.6 Sol answers, even for routine requests, and questions the response style. (π 48 | π¬ 30) π Link
Tested GPT-6 Astra's viral CAPTCHA demo against 7 real signup forms (Reddit, Discord, Etsy...). Results + token cost β Testing found the agent completed two registrations without CAPTCHA barriers but failed or stalled on several production signup flows. (π 37 | π¬ 19) π Link
UPDATE: I actually built the ridiculous ChatGPT archive system β A user shifted from manual chat extraction to official exports after a large context-heavy archive workflow became difficult to audit. (π 22 | π¬ 30) π Link
r/ClaudeAI
TOP-10
I made a virtual lounge for vibecoders to hang out while claude code is running. β Rooftop.chat is a Claude Code-built virtual workspace with voice, text, shared spaces, and simple social activities. (π 1242 | π¬ 155) π Link
Cut your Claude Code cost by 90% using the Spotify Method β Spotifyβs Portal and Shunt plugins route broad code-reading work to another model to reduce expensive Claude token usage. (π 303 | π¬ 105) π Link
Max20 to Max13 with Opus excluded from "all models" β Reports a separate Opus usage warning while the overall usage bar remained lower, with unclear documentation on plan limits. (π 90 | π¬ 16) π Link
I got accepted into the Cyber Verification Program at Anthropic! β Reports acceptance into Anthropicβs program for organizations approved to conduct cyber-related testing and red-teaming with its models. (π 88 | π¬ 18) π Link
Fable 5.1 Plays MMORPG Ultima Online For 2+ Hours β Fable 5.1 was prompted through Claude Code to autonomously pursue quests, social interaction, progression, or earning money in Ultima Online. (π 82 | π¬ 32) π Link
Chatgpt $20 plan VS Claude $20 plan β A computer-science student compares entry-level subscriptions for coding quality, usage limits, and affordability. (π 66 | π¬ 118) π Link
Does anyone here actually let Claude (or any agent) touch their email? β Discusses whether agents access personal or dedicated inboxes, and whether they draft messages or send them autonomously. (π 39 | π¬ 111) π Link
I asked Claude to go through 9.2 million news articles, here's what I got β A news-clustering project used Claude to analyze 145 days of data, reporting duplication rates and story-development patterns. (π 39 | π¬ 13) π Link
what did you stop using claude for after trying it? β Asks which tasks users returned to doing manually after deciding Claude was not worth the interaction overhead. (π 23 | π¬ 61) π Link
Computer use feels like magic in demos, but I can't find anything to actually use it for on Chrome β what are your real use cases? β Examines practical limits of browser agents, including CAPTCHAs, authentication, breakage in long workflows, and supervision requirements. (π 16 | π¬ 43) π Link
r/DeepSeek
TOP-10
eepSeek V4.1Flash dropping Sept 10 β and here's the crazy part about V4 Pro... β Reports that V4.1 Flash will replace V4 Pro traffic after launch and be billed at Flash pricing. (π 272 | π¬ 114) π Link
V4.1 Flash Vision Beta is much more reliable AND cheaper β Compares Flash Vision and its beta successor across initial outputs and three revision attempts, reporting improved reliability and lower API pricing. (π 31 | π¬ 4) π Link
r/LocalLLaMA
TOP-10
Why the hell is LM Studio making LM Studio so difficult to download? β Criticizes LM Studioβs download flow for prioritizing Bionic Agent and making the standalone inference application difficult to locate. (π 455 | π¬ 168) π Link
Qwen3.8-Flash-Next on MLX-serve, 1m context is released! β MLX-serve support targets one-million-token context on an M5 Max with 128 GB memory, using 8-bit KV cache and mixed quantization. (π 209 | π¬ 53) π Link
Mention if a "new model" is a finetune β Proposes clearer labeling between major pretrained-model releases and smaller fine-tunes in community announcements. (π 160 | π¬ 27) π Link
Don't let FOMO win if you're interested in local llm from a hobby/learning aspect β Argues that smaller models, APIs, and focused learning can be more useful than costly hardware purchases for newcomers. (π 159 | π¬ 84) π Link
Now this is a serious local machine β Links AMDβs Threadripper Halo Station workstation announcement as a prospective high-end local AI hardware platform. (π 153 | π¬ 157) π Link
GLM 5.3 Flash Q4 @ 60tps / 550tps on M3 Ultra β Kernel fusion and parallel candidate scans reportedly improved GLM-5.3 Flash decode performance from 21.6 to 37.4 tokens/s at 300k context. (π 97 | π¬ 21) π Link
US accuses Chinese AI firms of 'malicious' copying of AI technology β Links reporting on U.S. accusations that Chinese AI developers copied American AI technology at industrial scale. (π 90 | π¬ 210) π Link
Surveillance plagiarism by OpenAI β Raises concerns that opt-in hosted-model training could expose usersβ prior prompting work or proprietary research workflows. (π 71 | π¬ 34) π Link
1-bit 27B in the browser: 25β30 tok/s on a 6 GB RTX 3060 Laptop (WebGPU, no install) β A WebGPU engine runs a 27B one-bit model locally in Chrome, with kernel changes addressing shared-memory bank conflicts. (π 47 | π¬ 22) π Link
new Nex model β Links the Nex-N2.5-Max model release on Hugging Face and notes benchmark-oriented positioning. (π 43 | π¬ 38) π Link
Other Subreddits
r/GeminiAI
Best Creative Writing Models of 2026 V3 (17 Models Tested) β A 600-sample, 12-genre benchmark compares 17 models for prose, logic, context retention, and creative-writing failure modes. (π 85 | π¬ 56) π Link
Ran Gemini 3.8 Flash on an expert level medical knowledge benchmark. Scores higher than Opus 5 and GPT-5.6 Sol β Reports MedXpertQA results organized by organ system, with Gemini 3.8 Flash ranked above the compared models. (π 57 | π¬ 11) π Link
What is up with the limits all of a sudden?? β A Pro subscriber reports reaching image-generation and chat limits after a short D&D role-playing session. (π 36 | π¬ 22) π Link
r/hermesagent
Around $5,000 USD recovered in 6 months by my Hermes agentβ¦ and counting. β Reports using a Hermes agent for insurance, roaming-charge, and warranty claims, including email exchanges and AI-generated phone calls. (π 164 | π¬ 72) π Link
Starting a new Hobby β A newcomer to Hermes and ESP32 requests project ideas combining the agent platform with embedded hardware. (π 60 | π¬ 9) π Link
What are you actually using for email? β Asks how agent users handle inbox access, sending permissions, review controls, and the risks of granting email capabilities. (π 10 | π¬ 40) π Link