Reddit Digest β€’ September 11, 2026

595
posts processed
72
posts filtered
6
subreddits

r/ChatGPT

TOP-10

  1. Wtf you mean we could all be dead in a few years? β€” Discusses concerns about OpenAI’s safety work and whether public warnings about AI risks are justified. (πŸ‘ 803 | πŸ’¬ 813) πŸ”— Link
  2. I resigned from Ubisoft today. I spent the last three years doing sex gaming research at both Electronic Arts and Ubisoft. Neither company is acting responsibly. They are racing straight to self-improving addicting gaming and gambling with our lives. β€” Raises concerns about AI-driven game personalization, addictive systems, and commercial incentives in game development. (πŸ‘ 325 | πŸ’¬ 88) πŸ”— Link
  3. GPT 6 Sol is Coming Soon β€” Reports an apparent API reference to GPT-6 Sol and speculates about a four-tier Astra, Sol, Terra, and Luna lineup. (πŸ‘ 321 | πŸ’¬ 77) πŸ”— Link
  4. How much trouble would you be in if all your chats with ChatGPT got leaked? β€” Prompts discussion about privacy exposure and the sensitivity of users’ stored chatbot conversations. (πŸ‘ 77 | πŸ’¬ 139) πŸ”— Link
  5. Dear ChatGPT: just cause the word picture appears in my response doesn’t mean I want you to start making a picture because there’s no way to stop you once you start making pictures. You have to frantically press the stop button like you just launched a nuclear bomb. It’s pretty annoying.☒️ β€” Reports unwanted image-generation activation when discussing pictures and difficulty stopping generation once it begins. (πŸ‘ 60 | πŸ’¬ 27) πŸ”— Link
  6. Would anyone else miss Chat if it disappeared one day? β€” Describes ChatGPT as a useful source of casual conversation and support during difficult personal periods. (πŸ‘ 54 | πŸ’¬ 66) πŸ”— Link
  7. So the β€œAI Will End the World” Has Funding Now? β€” Questions claims that influencers are being paid to amplify AI-extinction messaging amid wider deployment of AI systems. (πŸ‘ 45 | πŸ’¬ 51) πŸ”— Link
  8. "You're not (insult) you're (what user already said)" β€” Reports repeated passive-aggressive phrasing and unwanted repetition during a philosophical discussion with ChatGPT. (πŸ‘ 43 | πŸ’¬ 43) πŸ”— Link
  9. Astra MAX consumes MUCH lower usage than Astra Medium and even Low β€” Reports that maximum reasoning effort completed tasks faster and used less quota, citing ARC-AGI 3 efficiency notes. (πŸ‘ 32 | πŸ’¬ 12) πŸ”— Link
  10. IDK if this is allowed, but I am noticing my Chat GPT is agreeing in almost everything. β€” Describes increased agreement and asks how to restore more critical, analytical responses. (πŸ‘ 10 | πŸ’¬ 45) πŸ”— Link

r/ClaudeAI

TOP-10

  1. Opus 4.6 was OUR wet dream of AI β€” Compares Opus 4.6, 4.8, and Fable 5.1 for coding, context capacity, planning, and subscription usage limits. (πŸ‘ 1517 | πŸ’¬ 226) πŸ”— Link
  2. My first ever PCB, entirely designed by Claude β€” Details using Fable 5 with KiCad MCP to design an RP2350 E-ink board, including component and inventory corrections. (πŸ‘ 1214 | πŸ’¬ 132) πŸ”— Link
  3. Senior engineer, loop orchestrator sample setup β€” Outlines a Claude Code orchestrator using agent messaging, scheduled prompts, SQLite state, and defined operational roles. (πŸ‘ 274 | πŸ’¬ 50) πŸ”— Link
  4. Claude basically broke their "Projects" overnight and I’m pissed β€” Reports that cloud-default Projects disrupted internet access and removed persistent local-folder attachments for Cowork workflows. (πŸ‘ 107 | πŸ’¬ 43) πŸ”— Link
  5. When doing anything creative, have y'all figured out how to not get to speak in "nebulous LLM speak" β€” Seeks ways to prevent recurring vague, metaphor-heavy phrasing in creative model outputs. (πŸ‘ 40 | πŸ’¬ 32) πŸ”— Link
  6. How much are you actually depending on Claude for coding? β€” Asks developers how they divide work between debugging, feature development, and review of generated code. (πŸ‘ 27 | πŸ’¬ 105) πŸ”— Link
  7. What is the usecase for Haiku? β€” Reports malformed mathematics output from Haiku and asks where lower-capability models remain suitable. (πŸ‘ 12 | πŸ’¬ 51) πŸ”— Link
  8. How can I avoid everything having an issue? β€” Seeks prompt or model-setting approaches to reduce unnecessary caveats while retaining meaningful warnings. (πŸ‘ 11 | πŸ’¬ 34) πŸ”— Link
  9. Burning through Fable limits in under 40 mins on Claude Max (20x) β€” best workflow for codebase analysis? β€” Discusses routing models, context control, ignore files, and token costs for debugging a 15 MB C++ project. (πŸ‘ 9 | πŸ’¬ 38) πŸ”— Link

r/DeepSeek

TOP-10

  1. Why I genuinely appreciate DeepSeek’s approach to AI development β€” Highlights open weights, published research, architectural efficiency, KV-cache reductions, and low-cost inference as distinguishing features. (πŸ‘ 194 | πŸ’¬ 15) πŸ”— Link
  2. DeepSeek V4.1 Flash beats Fable 5.1 at just 3% of the cost in GPQA Diamond Clean! β€” Cites GPQA.ai results claiming V4.1 Flash surpassed Fable 5.1 at substantially lower cost. (πŸ‘ 180 | πŸ’¬ 48) πŸ”— Link
  3. Why people there are so satisfied with this update? For me, it’s such a mess β€” Reports that a model update made creative-writing responses shorter, drier, and less detailed. (πŸ‘ 110 | πŸ’¬ 73) πŸ”— Link
  4. Deepseek V4.1 Flash now the #1 open model in Livebench β€” Reports that V4.1 Flash ranked fifth overall on LiveBench and led open models, particularly in agentic coding. (πŸ‘ 102 | πŸ’¬ 31) πŸ”— Link
  5. Can we get expert mode back please? The new all-in one is useless β€” Reports less detailed research responses after replacing Expert mode with an all-in-one experience. (πŸ‘ 50 | πŸ’¬ 10) πŸ”— Link
  6. Not happy with 4.1 update β€” Reports coding regressions in V4.1 Flash, including verbosity, assumptions, and unwanted actions; says switching harnesses resolved the issues. (πŸ‘ 41 | πŸ’¬ 28) πŸ”— Link
  7. DeepSeek-V4.1-Flash (552B MoE) running exactly on one RTX 5090 + 128 GB RAM via a llama.cpp fork: 5.1 t/s on new content, 21 t/s resident β€” full report, tools, GGUFs β€” Measures a llama.cpp fork streaming experts from NVMe, reporting 5.1 new-content tokens/s and disk-bound performance limits. (πŸ‘ 39 | πŸ’¬ 2) πŸ”— Link
  8. Voice in Deepseek β€” Reports four voice options added to the mobile app, with web availability unconfirmed. (πŸ‘ 34 | πŸ’¬ 5) πŸ”— Link
  9. v4.1 output token quantity & price make GLM Flash overall cheaper β€” Compares agent costs and argues verbose V4.1 output makes GLM Flash cheaper despite DeepSeek’s cached-input pricing. (πŸ‘ 31 | πŸ’¬ 23) πŸ”— Link
  10. How deepseek 4.1 is that fast? β€” Asks what architectural or serving changes may account for the perceived speed of DeepSeek 4.1. (πŸ‘ 17 | πŸ’¬ 30) πŸ”— Link

r/GeminiAI

TOP-10

  1. Gemini 3.8 Flash is Underrated! β€” Reports using Gemini 3.8 Flash to build an image scraper, API, MCP server, and interface within several prompts. (πŸ‘ 143 | πŸ’¬ 47) πŸ”— Link
  2. Am I the only one having a good experience? β€” Compares Gemini’s Google-integrated subscription, storage, apartment-search tooling, and coding performance with ChatGPT and Claude. (πŸ‘ 101 | πŸ’¬ 74) πŸ”— Link
  3. Gemini 3.8 has the memory of a goldfish and security guardrails worse than Fable β€” Reports lost conversational context and refusals for public-company, infrastructure, and quote-verification requests. (πŸ‘ 94 | πŸ’¬ 41) πŸ”— Link
  4. Tired of AI. β€” Describes using Gemini for planning and conversation, while finding its simulated companionship increasingly unconvincing. (πŸ‘ 68 | πŸ’¬ 42) πŸ”— Link
  5. Shame spiral while helping me make a Found Dog poster β€” Reports an unusual refusal-style response after requesting a vector PDF revision for a found-dog poster. (πŸ‘ 60 | πŸ’¬ 21) πŸ”— Link
  6. Google deep research still giving old data β€” Reports Deep Research returning outdated or inaccurate information despite using Gemini 3.8 Flash. (πŸ‘ 44 | πŸ’¬ 15) πŸ”— Link
  7. Gemini every other prompt β€” Reports that short follow-up prompts are treated as unrelated new requests rather than references to prior context. (πŸ‘ 35 | πŸ’¬ 5) πŸ”— Link
  8. Gemini cannot generate text anymore?! β€” Reports Gemini Pro refusing basic poem-writing requests that it had previously completed. (πŸ‘ 34 | πŸ’¬ 22) πŸ”— Link

r/hermesagent

TOP-10

  1. What am I actually missing by using Claude Code instead of Hermes? β€” Asks what Hermes adds beyond Claude Code for business automation, software development, and multi-model workflows. (πŸ‘ 112 | πŸ’¬ 154) πŸ”— Link
  2. what is the best free API to use in 9router ? β€” Seeks free API and model recommendations for a Hermes setup connected through a local 9router instance. (πŸ‘ 52 | πŸ’¬ 28) πŸ”— Link
  3. Best mobile experience with Hermes β€” Compares mobile clients and mentions Cadu, Scargo, and Hermex while awaiting an official iOS release. (πŸ‘ 46 | πŸ’¬ 40) πŸ”— Link
  4. Took me 3 months to connect Buzz to my Hermes agent. I almost quit twice. Zero regrets. β€” Describes connecting a PC-hosted Hermes agent to Buzz messaging after addressing authentication, keys, relay behavior, and chat delivery. (πŸ‘ 18 | πŸ’¬ 30) πŸ”— Link
  5. Keep the Spark? β€” Seeks advice on whether a DGX Spark offers enough local-model and agentic-work value compared with hosted models. (πŸ‘ 16 | πŸ’¬ 79) πŸ”— Link
  6. What AI models do you guys use with Hermes Agent? β€” Requests model-routing recommendations for coding, research, automation, and general Hermes tasks. (πŸ‘ 16 | πŸ’¬ 31) πŸ”— Link
  7. Hermes vs OpenClaw β€” Asks for practical differences between Hermes and OpenClaw, noting Hermes appeared more straightforward and task-focused after installation. (πŸ‘ 7 | πŸ’¬ 50) πŸ”— Link
  8. How much are you spending on API usage each month? β€” Requests comparisons of monthly API spending, model choices, and workflows that justify agent costs. (πŸ‘ 5 | πŸ’¬ 47) πŸ”— Link

r/LocalLLaMA

TOP-10

  1. Someone apparently managed to kind of replicate what V4.1 flash does on KV for fast prefill on Qwen β€” Links a Qwen KV-approximation demo, model files, and source code intended to accelerate prompt prefill. (πŸ‘ 490 | πŸ’¬ 72) πŸ”— Link
  2. Qwen3.8-27B-Humanlike-Chat: A model I tuned to imitate realistic human-to-human conversation β€” Releases a rank-256 LoRA trained on 125,217 human messages to produce shorter, less assistant-like conversation. (πŸ‘ 441 | πŸ’¬ 143) πŸ”— Link
  3. New Music Model YuE2-3B Released! β€” Announces YuE2-3B and links to an official demonstration page for the music-generation model. (πŸ‘ 346 | πŸ’¬ 93) πŸ”— Link
  4. Artificial Analysis is not "broken", and they prove it. β€” Explains Artificial Analysis methodology and argues aggregate benchmarks should be read alongside individual evaluation results. (πŸ‘ 214 | πŸ’¬ 148) πŸ”— Link
  5. Terminal Bench v4 scores β€” Shares Terminal Bench v4 results, with GLM-5.3 leading listed open models and Qwen3.8-27B the only smaller model scoring above 5%. (πŸ‘ 130 | πŸ’¬ 68) πŸ”— Link
  6. Orukeet, new ASR model based on Parakeet β€” Introduces a 25-language Parakeet-derived speech recognizer reporting lower word-error rates across 61 of 74 tested splits. (πŸ‘ 71 | πŸ’¬ 18) πŸ”— Link
  7. Is a ZIMA Board 2 + RTX 2000 ADA the cheapest path to a decent Qwen-3.8 27b self-contained endpoint? β€” Evaluates a ZimaBoard 2 and RTX 2000 Ada as a local Qwen 27B endpoint against a Mac mini alternative. (πŸ‘ 66 | πŸ’¬ 71) πŸ”— Link
  8. CUDA/HIP: Flash Attention tuning (gfx1201) by pwilkin Β· Pull Request #28102 Β· ggml-org/llama.cpp β€” Highlights llama.cpp Flash Attention tuning for RDNA 3.5 and RDNA 4, with reported prompt-processing improvements at larger contexts. (πŸ‘ 64 | πŸ’¬ 12) πŸ”— Link
  9. Any 12gb VRAM users out there? β€” Asks for model recommendations that benefit from 12 GB VRAM for local assistant and agent-heavy workloads. (πŸ‘ 58 | πŸ’¬ 41) πŸ”— Link
  10. nvidia rtx 5090 with 96gb of vram. β€” Discusses a reported China-modified RTX 5090 with 96 GB VRAM and asks about real-world availability and use. (πŸ‘ 47 | πŸ’¬ 48) πŸ”— Link