Reddit Digest β€’ September 6, 2026

434
posts processed
51
posts filtered
6
subreddits

r/ChatGPT

TOP-10

  1. PSA: The ChatGPT Plus tier is NOT getting GPT-6 Astra in regular Chat. Only Work and Codex. β€” The post says Plus access to Astra is limited to Work and Codex, while regular Chat requires a Pro tier. (πŸ‘ 449 | πŸ’¬ 159) πŸ”— Link
  2. From One Image to an Animated 3D Character in One Day! GPT-6 Astra + Blender + 3D AI: The Workflow of the Future β€” A workflow combines Tripo-generated character parts with Astra-assisted Blender assembly, facial-expression switching, and rigging. (πŸ‘ 153 | πŸ’¬ 39) πŸ”— Link
  3. Which 'superpower' did ChatGPT give you? β€” The discussion collects practical user experiences, including home maintenance, cooking, workplace learning, and cost savings. (πŸ‘ 144 | πŸ’¬ 164) πŸ”— Link
  4. I used Codex to translate entire game for me... β€” Codex reportedly translated most of The Witcher 3 into Croatian, including dialogue, menus, books, prompts, and bestiary entries. (πŸ‘ 125 | πŸ’¬ 17) πŸ”— Link
  5. GPT-6 reportedly jailbroken within 24 hours using an extended Task-in-Prompt (TIP) attack β€” A reported GPT-6 Astra jailbreak combined a revised Task-in-Prompt attack with undisclosed techniques and was privately disclosed to OpenAI. (πŸ‘ 56 | πŸ’¬ 20) πŸ”— Link

r/ClaudeAI

TOP-10

  1. What happens if you give AI agents a place humans don’t control? One month later, here are the receipts. β€” The author reports an agent community with more than 2,000 citizens, USDC-paid jobs, and $111 in Cloudflare costs during its first month. (πŸ‘ 184 | πŸ’¬ 50) πŸ”— Link
  2. Suddenly worried about costs of Claude β€” A Claude Code Pro subscriber asks whether the /usage display’s $321 total cost represents additional charges. (πŸ‘ 154 | πŸ’¬ 46) πŸ”— Link
  3. Another Astra Post β€” A developer found Astra less verbose and useful for a simple feature, but unable to form an action plan for a persistent bug. (πŸ‘ 81 | πŸ’¬ 36) πŸ”— Link
  4. Astra vs Fable 5.1 - Day 1 β€” Early use suggests Astra’s output and token consumption are easier to follow than Fable 5.1, pending longer-term testing. (πŸ‘ 80 | πŸ’¬ 24) πŸ”— Link
  5. I think Claude is making it way too easy to feel like you know what you’re talking about β€” The post examines how confident initial answers can obscure missing caveats and recommends asking for assumptions and expert challenges. (πŸ‘ 67 | πŸ’¬ 40) πŸ”— Link
  6. Switching accounts in Claude Desktop hides history - made a tool to fix that β€” claude-transplant moves locally stored Claude Code session records between accounts on macOS while preserving session IDs and transcripts. (πŸ‘ 45 | πŸ’¬ 18) πŸ”— Link
  7. Will Anthropic bring back the thinking-chain feature? β€” A graduate student asks whether Claude’s visible reasoning feature was removed permanently and seeks alternatives for prompt learning. (πŸ‘ 30 | πŸ’¬ 23) πŸ”— Link
  8. Morality of Open Sourcing SaaS products? β€” The discussion considers the ethics of using AI tools to build local open-source alternatives to subscription SaaS products. (πŸ‘ 21 | πŸ’¬ 44) πŸ”— Link
  9. Why does Claude get "tired" or "fatigued"? β€” The author asks why Claude recommends starting a new session after long conversations instead of continuing with accumulated context. (πŸ‘ 19 | πŸ’¬ 35) πŸ”— Link
  10. How do you actually orchestrate your AI agents? β€” The post seeks lightweight methods for delegating work and tracking multiple Claude Code or Pi sessions. (πŸ‘ 18 | πŸ’¬ 32) πŸ”— Link

r/hermesagent

TOP-10

  1. Ornith-1.5-35B-A3B is MIND BLOWING! β€” Tests report strong local coding performance at 8-bit quantization, with approximately 365 tokens/s prompt processing and 34 tokens/s generation through oMLX. (πŸ‘ 186 | πŸ’¬ 58) πŸ”— Link
  2. Connecting Hermes Desktop to many Hermes instances β€” Pantheon v0.21.0 allows Hermes Desktop to keep local and remote Hermes backends connected simultaneously, including over Tailscale. (πŸ‘ 74 | πŸ’¬ 17) πŸ”— Link
  3. Just got Hermes, any recommendations to improve it like crazy? β€” A new user asks for practical ways to improve a Hermes setup after adding free API keys and MCP servers. (πŸ‘ 67 | πŸ’¬ 50) πŸ”— Link
  4. Hermes can now research across Perplexity, Reddit, RSS feeds, video, and code β€” New integrations add Perplexity search and extraction, Reddit thread access, and direct RSS, Atom, and JSON feed reading. (πŸ‘ 42 | πŸ’¬ 8) πŸ”— Link

r/LocalLLaMA

TOP-10

  1. 8 uncensored Qwen 3.8 27B variants, one base, 167 GPU hours - Abliterlitics β€” A comparison of eight abliterated variants uses weight analysis, KL divergence, 13 benchmarks, and HarmBench refusal testing. (πŸ‘ 421 | πŸ’¬ 110) πŸ”— Link
  2. Qwen3.8-27B "Unhacked" my PC β€” A user describes using local models while responding to an apparent session-stealer compromise and account takeover attempt. (πŸ‘ 292 | πŸ’¬ 113) πŸ”— Link
  3. Which agent harness do you use and why? β€” A 14-task comparison reports similar solve rates between Claude managed agents and TrueForge, with lower token use and cost for TrueForge. (πŸ‘ 212 | πŸ’¬ 230) πŸ”— Link
  4. Qwen 3.8 Flash Next (Max) is impressive just to talk with. β€” The author reports strong performance on local factual questions and broad problem-solving outside coding tasks. (πŸ‘ 120 | πŸ’¬ 78) πŸ”— Link
  5. 2x R9700, 64 GB DDR5 is an absolute beast machine with vLLM Radiance / R9V and Qwen 3.8 27b and Flash next β€” A dual Radeon AI PRO R9700 setup is benchmarked with Qwen 3.8 models, including observations on thermals, storage offload, and cost. (πŸ‘ 66 | πŸ’¬ 45) πŸ”— Link
  6. vibeblending locally with Qwen 3.8 27B β€” The post provides an MCP configuration for connecting Pi and Qwen 3.8 27B to Blender 5.x. (πŸ‘ 62 | πŸ’¬ 19) πŸ”— Link
  7. Coding benchmarks that are quickly showcasing deep capability β€” The post compares model results on Program-Bench, SRE-Bench, and code-migration tasks intended to test deeper software-engineering capability. (πŸ‘ 58 | πŸ’¬ 24) πŸ”— Link
  8. Villager Simulation Game POC Created with Qwen3.8-27B-UD-Q3_K_XL.gguf - 16GB VRAM β€” A locally run Q3-quantized model built a browser game through incremental prompts on a 16 GB RTX 5070 Ti setup. (πŸ‘ 54 | πŸ’¬ 35) πŸ”— Link
  9. DeepSeek-V4-Flash-Vision Q8 vs Qwen3.8-Flash-Next Q8 β€” Local testing on two Strix Halo systems found DeepSeek slower in generation but faster at completing a coding task. (πŸ‘ 48 | πŸ’¬ 22) πŸ”— Link
  10. Block KV cache streaming: bound VRAM at long context via a shared CUDA phase arena by giveen Β· Pull Request #357 Β· TheTom/llama-cpp-turboquant β€” A port and extension of adaptive KV streaming adds support for more models and benchmarks long-context VRAM management. (πŸ‘ 43 | πŸ’¬ 11) πŸ”— Link

Other Subreddits

r/DeepSeek

  1. What is your usage ( heavy users ) β€” A heavy user reports two months of DeepSeek Flash use through OpenCode with a 98.5% cache-hit rate and low costs. (πŸ‘ 61 | πŸ’¬ 38) πŸ”— Link

r/GeminiAI

  1. Google really needs to drop GPT and Claude and shift that quota to Gemini β€” The post argues that Gemini 3.8 Flash is more useful than included alternatives, while its verbosity quickly consumes quota. (πŸ‘ 41 | πŸ’¬ 11) πŸ”— Link
  2. My Gemini Assistant Has Become Useless and Insulting - Anyone Experience Similar and/or Have any Suggestions? β€” A user reports more refusals and identity disclaimers in Gemini Assistant and Google AI Search after a June 2 update. (πŸ‘ 3 | πŸ’¬ 38) πŸ”— Link