Reddit Digest β€’ September 25, 2026

642
posts processed
106
posts filtered
8
subreddits

r/ChatGPT

TOP-10

  1. Made an AR Yu-Gi-Oh prototype with ChatGPT, Astra, CLAD, and Lens Studio β€” The prototype turns physical cards into 3D monsters and adds 3D visuals for Spell Cards. (πŸ‘ 1574 | πŸ’¬ 77) πŸ”— Link
  2. OpenAI prepares new $500 per month Pro Max plan for ChatGPT β€” The post reports that OpenAI is preparing a $500 monthly ChatGPT tier. (πŸ‘ 258 | πŸ’¬ 132) πŸ”— Link
  3. Give Gpt more time ... β€” The author compares four models recreating the same ink-style image as SVGs, with Astra and Opus rated ahead of the two Sol versions. (πŸ‘ 246 | πŸ’¬ 40) πŸ”— Link
  4. ChatGPT Voice can use connected apps now, and it’s kind of a game changer β€” The author tested Voice using connected Gmail and Google Drive apps, including summarizing email and writing to a document. (πŸ‘ 206 | πŸ’¬ 54) πŸ”— Link
  5. I gave my AI its own Google Doc diary. β€” A scheduled workflow reviews daily conversations and writes selected moments into a Google Doc. (πŸ‘ 118 | πŸ’¬ 29) πŸ”— Link
  6. New ChatGPT Update Sucks β€” The author lists reported UI changes and issues, including hidden controls, long-chat freezes, and difficulty saving temporary chats. (πŸ‘ 80 | πŸ’¬ 23) πŸ”— Link
  7. The UI update to the chatgpt website is absolutely awful and breaks a lot of basic functionlity. β€” The author says the redesigned navigation ladder and quote links no longer support their usual workflow. (πŸ‘ 75 | πŸ’¬ 40) πŸ”— Link
  8. ChatGPT Plus free trial made me a TV remote (first try!) β€” The author says Codex built a remote-control app for their TV that works from a PC or Android phone. (πŸ‘ 74 | πŸ’¬ 25) πŸ”— Link
  9. Why did they make ChatGPT nitpick and 'correct' everything instead of strongmanning what I said and continuing to build on it with momentum β€” it did well for the last few months until these two weeks. β€” The author reports more semantic corrections, assumptions, and changes in how ChatGPT interprets prompts. (πŸ‘ 46 | πŸ’¬ 31) πŸ”— Link
  10. Chatgpt Breathing/coughing during voice reading? β€” The author heard breathing and coughing sounds during voice playback and says ChatGPT did not directly address the question afterward. (πŸ‘ 35 | πŸ’¬ 56) πŸ”— Link

r/ClaudeAI

TOP-10

  1. Testing Claude for 3D creation. Max took a whole hour, but just look at the result πŸ‘€ β€” The prompt asks Claude to create a looping low-poly horse animation in a self-contained Three.js HTML file. (πŸ‘ 1411 | πŸ’¬ 240) πŸ”— Link
  2. Real talk: If Anthropic never nerfs Opus 5.5, I will keep my Max subscription for years... β€” The author reports using Opus 5.5 to process 30,000–50,000-word documents and says it retains their contents well. (πŸ‘ 741 | πŸ’¬ 103) πŸ”— Link
  3. Anthropic signs $11.6B cloud deal with Akamai β€” The post reports an $11.6 billion cloud deal between Anthropic and Akamai. (πŸ‘ 458 | πŸ’¬ 31) πŸ”— Link
  4. Your AI games suck, and it's not the AI's fault β€” The author argues that AI-generated games still need hands-on playtesting and iteration to address problems with gameplay. (πŸ‘ 368 | πŸ’¬ 132) πŸ”— Link
  5. Claude Code Wrap Up Allowance! β€” The feature allows Claude Code to use capped extra capacity to reach a stopping point when a plan limit is hit mid-response. (πŸ‘ 278 | πŸ’¬ 16) πŸ”— Link
  6. claude starting planning a pizza party for me? β€” The author says Claude inserted an unsolicited pizza-party plan and a system warning into a response about diet and nutrition. (πŸ‘ 266 | πŸ’¬ 74) πŸ”— Link
  7. Week 3 Update: Building a cozy game with no game dev experience. He rolls now!! β€” The author describes adding a rolling animation, a bouncier jump, and a sit animation to a game project. (πŸ‘ 259 | πŸ’¬ 44) πŸ”— Link
  8. A "made entirely with Opus 5.5" explainer video with $0 additional costs. Insane results. β€” The author describes generating an animated explainer with local Kokoro speech, JavaScript music, synthesized effects, and Whisper transcription. (πŸ‘ 138 | πŸ’¬ 37) πŸ”— Link
  9. If opus 5.5 is basically fable level, what are you still using fable for? β€” The post asks users to compare Opus 5.5 and Fable 5.1, citing Anthropic’s performance and pricing claims. (πŸ‘ 131 | πŸ’¬ 81) πŸ”— Link
  10. What exactly is an "agent" β€” The author asks whether β€œagentic” use refers to a single LLM instance, such as an API query or web session. (πŸ‘ 67 | πŸ’¬ 41) πŸ”— Link

r/codex

TOP-10

  1. GPT-6 feels like a downgrade for Codex subscribers, and β€œbut it's cheaper” doesn't really excuse it β€” The author cites Bug Hunt Bench results and compares API price reductions with changes to subscription message allowances. (πŸ‘ 293 | πŸ’¬ 62) πŸ”— Link
  2. The downhill begins β€” The author argues that a proposed higher-priced plan could coincide with lower limits or fewer resets for existing subscribers. (πŸ‘ 250 | πŸ’¬ 190) πŸ”— Link
  3. GPT-6 Astra seems unusable due to Token burn, GPT-6 Sol is too bad to be trusted - Any reason not to switch? β€” The author reports Sol ignoring project instructions and taking incorrect implementation paths, then switching back to Astra. (πŸ‘ 193 | πŸ’¬ 112) πŸ”— Link
  4. Opus wipes the floor with Sol β€” The author compares token usage and subscription limits, arguing that Opus provides better value for their workload. (πŸ‘ 160 | πŸ’¬ 49) πŸ”— Link
  5. Absolutely 0 doubt in my mind Astra has been lobotomized and compared to Opus 5.5 it's absolutely not even close. β€” The author reports that Astra became slower and less efficient, and says Opus 5.5 performed better in their recent use. (πŸ‘ 98 | πŸ’¬ 49) πŸ”— Link
  6. I did it, you got me ! β€” The author describes switching from Codex to Claude and says a reset came with their new subscription. (πŸ‘ 84 | πŸ’¬ 50) πŸ”— Link
  7. Astra vs Opus 5.5, my impressions on hard project β€” The author compares both models on a complex async-runtime design task, finding Astra identified additional edge cases while Opus used fewer limits. (πŸ‘ 83 | πŸ’¬ 87) πŸ”— Link
  8. How is Anthropic beating OpenAI on compute all of a sudden? β€” The author questions how recent reports of OpenAI constraints align with Anthropic’s cloud-compute position. (πŸ‘ 81 | πŸ’¬ 72) πŸ”— Link
  9. Web usage got merged with Codex/Work ? β€” The author reports seeing a new window and left sidebar during a web chat session, despite not using Work. (πŸ‘ 65 | πŸ’¬ 54) πŸ”— Link
  10. Did they reduce GPT 6 Sol usage? β€” The author reports that GPT-6 Sol now uses their five-hour allowance more quickly than it did two days earlier. (πŸ‘ 57 | πŸ’¬ 33) πŸ”— Link

r/GeminiAI

TOP-10

  1. Gemini's speed is actually breaking my brain. How is this even real? β€” The author says Gemini returned a response to a complex prompt much faster than ChatGPT, which took several minutes. (πŸ‘ 159 | πŸ’¬ 69) πŸ”— Link
  2. Google's Next Image Model (Spicy Mayo) β€” The post summarizes reports of a Google image model in Arena, mentioning speed, text rendering, editing features, and possible 4K output. (πŸ‘ 74 | πŸ’¬ 17) πŸ”— Link
  3. Google has become the Nintendo of the AI Labs β€” The author argues that Google’s products serve broad audiences but are less focused on competing at the AI frontier. (πŸ‘ 66 | πŸ’¬ 27) πŸ”— Link
  4. Are we underestimating Gemini Flash 3.8 (high) on Antigravity? β€” The author reports that Flash 3.8 reconciled a lengthy technical draft while using about 10% of their plan’s tokens. (πŸ‘ 23 | πŸ’¬ 31) πŸ”— Link

r/hermesagent

TOP-10

  1. Hermes agent was used to attack 27 companies. Attacker Stole 600,000 payment cards!! β€” The post summarizes a reported campaign using Hermes, Strix, and Cairn for scanning and exploitation, with more than 600,000 payment-card records exposed. (πŸ‘ 163 | πŸ’¬ 59) πŸ”— Link
  2. Can someone explain the actual advantage of Bot Mode over Profiles + long-running chats? β€” The author asks how Bot Mode’s persistent chats, routines, and bot-to-bot features differ from Profiles with dedicated conversations. (πŸ‘ 47 | πŸ’¬ 10) πŸ”— Link
  3. Making Hermes more proactive β€” The author asks how to make Hermes respond to events, such as an upcoming meeting, rather than relying only on scheduled briefings. (πŸ‘ 45 | πŸ’¬ 19) πŸ”— Link
  4. Hermes Redde β€” This native iOS and iPadOS interface supports voice, live tool approvals, subagents, sessions, scheduled jobs, skills, and local-server connections. (πŸ‘ 35 | πŸ’¬ 46) πŸ”— Link
  5. Using Jev for Magic the Gathering Card Classification β€” The author classified 2,527 unique cards across 61 archetypes for about $0.22 and describes plans to refine the categories. (πŸ‘ 30 | πŸ’¬ 9) πŸ”— Link

r/LocalLLaMA

TOP-10

  1. Qwengram-0.8B: I transferred Qwen3.8 Flash-Next’s n-gram memory into Qwen3.5-0.8B β€” 5.05% lower validation perplexity β€” The author reports a 5.05% perplexity reduction using frozen PLE memory and a trained reader, with an inference implementation in llama.cpp. (πŸ‘ 305 | πŸ’¬ 86) πŸ”— Link
  2. Former Intel CEO: "HBM is lousy". High Bandwidth Flash Is Coming β€” The post discusses criticism of HBM and claims that high-bandwidth flash may offer another approach to the memory wall. (πŸ‘ 183 | πŸ’¬ 75) πŸ”— Link
  3. M5 Ultra 80Core GLM-5.3-Flash on DwarfStar Speeds β€” The author shares agentic-inference tests on an 80-core, 256GB M5 Ultra and questions whether its GPU can use additional RAM effectively. (πŸ‘ 176 | πŸ’¬ 108) πŸ”— Link
  4. Swift1.5-Qwen3.8-Flash-Next is phenomenal vs. base 3.8-Flash! β€” In an Aider coding test, Swift reduced tokens per case from 17,646 to 6,991 while producing similar pass rates. (πŸ‘ 154 | πŸ’¬ 74) πŸ”— Link
  5. I ran the actual break-even math on buying vs renting an H200 box, and it is not where I expected β€” The author estimates ownership breaks even after roughly 24 months at 60% utilization, before power, cooling, depreciation, and other costs. (πŸ‘ 116 | πŸ’¬ 89) πŸ”— Link
  6. Jev vs. Kev: open-source Jev alternative tested side by side β€” On 362 recent items, the author reports similar accuracy, stronger Jev calibration, and a higher relative cost for short requests. (πŸ‘ 81 | πŸ’¬ 26) πŸ”— Link
  7. Is Qwen Flash Next at like Q2 better than 27B at Q4? β€” The post asks users to compare Q2 quantization of Flash Next with Q4 quantization of a 27B model. (πŸ‘ 64 | πŸ’¬ 140) πŸ”— Link
  8. How long can I expect to wait until the local ~30B A3B frontier catches up to GLM 5.3 Flash quality? β€” The author asks whether local models will reach that quality on a system with 16GB RAM and 8GB VRAM. (πŸ‘ 42 | πŸ’¬ 87) πŸ”— Link
  9. 4-5 days replacing Claude w Qwen 3.8 Next β€” The author reports using Qwen locally on an M1 Ultra for research and business tasks, while noting differences in speed and accumulated context. (πŸ‘ 31 | πŸ’¬ 36) πŸ”— Link
  10. Make Volta Fast Again β€” The post describes testing a vLLM fork optimized for V100 cards against a llama.cpp setup on Strix Halo. (πŸ‘ 23 | πŸ’¬ 34) πŸ”— Link

r/vibecoding

TOP-10

  1. Fable + Three.js + 3D AI β€” It Works and Saves a Ton of Time β€” The author describes a dice-based roguelike using a Three.js UI, procedurally generated dice, and 3D assets from Tripo P2. (πŸ‘ 102 | πŸ’¬ 5) πŸ”— Link
  2. GPT-6 Astra made me an ecom brand, built the ad campaigns, and got its first order β€” The author describes an AI-built storefront, product catalog, and ad campaign that generated a first $64 order after human quality checks. (πŸ‘ 97 | πŸ’¬ 34) πŸ”— Link
  3. Swift Qwen 3.8 27b is insane β€” The author reports that Swift Qwen roughly doubled inference speed on an M1 Max, from 7–8 to 15–16 tokens per second. (πŸ‘ 39 | πŸ’¬ 16) πŸ”— Link
  4. Why aren’t more vibe coders building their own self-hosted platform instead of depending on Base44/Lovable/etc? β€” The author argues that self-hosting can provide more control over infrastructure, updates, and app data than hosted builders. (πŸ‘ 14 | πŸ’¬ 70) πŸ”— Link
  5. Opust 5.5 js One SHOTTED this (Prompt Below) β€” The author says Opus 5.5 generated motion design, audio, and video using access to a project repository. (πŸ‘ 4 | πŸ’¬ 41) πŸ”— Link

Other Subreddits

r/DeepSeek

  1. I built OpenGhost - a fully open-source AI agent for DeepSeek focused on maximum visualization and a local browser the agent controls itself β€” OpenGhost combines a local browser, automatic context compression, and a custom rendering engine; the project is open source. (πŸ‘ 114 | πŸ’¬ 53) πŸ”— Link
  2. God I love 4.1 β€” The author reports spending under $4 on half a billion tokens with 4.1 Flash and says it runs slower but performs thorough work. (πŸ‘ 57 | πŸ’¬ 20) πŸ”— Link