GPT Solved ~350GB of "System Data" on my Mac. β ChatGPT-guided Terminal troubleshooting identified incorrectly retained podcast downloads as the cause of excessive system storage. (π 120 | π¬ 25) π Link
ChatGPTβs image generation has improved A LOT β Compares recent ChatGPT and Gemini results, citing prompt following, realism, composition, and text rendering. (π 118 | π¬ 65) π Link
Have any one of you spending literally hours to talk with Chat GPT or any other AI? β Discusses extended conversational use of AI assistants and associated user habits. (π 38 | π¬ 105) π Link
Whatβs the most profitable thing by $ value you did with ChatGPT? β Solicits concrete examples of work or projects that produced measurable financial value. (π 25 | π¬ 43) π Link
I built an app where you ask about any moment in history and it turns it into a researched two-host podcast you can interrupt with questions β Describes a multi-model pipeline for research, scripting, artwork, and voices, with in-episode voice questions. (π 11 | π¬ 46) π Link
r/ClaudeAI
TOP-10
Claude Code effort level and model selection | Claude β Links to Claude documentation covering effort-level controls and model selection in Claude Code. (π 365 | π¬ 54) π Link
Credits that were supposed to expire on the 19th just expired (it's the 18th) β Reports promotional usage credits disappearing before the displayed expiration date. (π 276 | π¬ 177) π Link
Sonnet 5 β Reports stronger Blender scene-building results from Sonnet 5 than Opus 5 using a detailed preservation-aware prompt. (π 141 | π¬ 43) π Link
Is the $20 per month worth it? β A virtual assistant weighs the cost of Claudeβs paid tier against a $250 monthly salary and work needs. (π 132 | π¬ 142) π Link
Need to blow $100 in Fable 5 credits ASAP β A computer science student asks for high-return uses for promotional Fable 5 credits before expiration. (π 115 | π¬ 87) π Link
New Beta CC Feature: Projects coordinate work across multiple sessions from one conversation β Claude Code Projects can split work into cloud-session threads on separate branches and continue while the user is away. (π 53 | π¬ 23) π Link
Is anyone else doing just fine with more basic models in Claude Code? β A developer reports effective web and DevOps work with Sonnet and Haiku, using scoped Kanban tasks and model reviews. (π 48 | π¬ 49) π Link
Used Claude to help Jev become a gamer β Describes using Opus 5 to build a Vampire Survivors mod and Python decision engine connected to Jev. (π 44 | π¬ 12) π Link
Solo dev, this is my entire Claude Code workflow. What's yours? β Uses GitHub issues for triage, Fable for planning, and Opus for implementation in a solo development workflow. (π 40 | π¬ 13) π Link
Claudeβs writing β Reports verbose and incoherent output across models despite skills, loops, agents, and attempts to improve prose. (π 30 | π¬ 26) π Link
r/codex
TOP-10
I found the best way to build insane UIs with Codex β Proposes generating a UI reference image with ChatGPT Images 2.5, then having Codex implement it. (π 509 | π¬ 117) π Link
The End of the Codex Era. I've Completely Lost Trust in OpenAI. They're Secretly Degrading Their Models. β Reports perceived within-thread performance variation across models, based on feature planning and implementation experiences. (π 420 | π¬ 93) π Link
time to take legal action as an EU citizen β Reports two Plus-account five-hour limits being exhausted within 2.5 hours, with documentation retained for review. (π 199 | π¬ 224) π Link
Consumers Sue AI Giants Over Alleged Slowdown Pact β Summarizes a proposed class action alleging that AI companies coordinated limits on advanced AI development. (π 155 | π¬ 21) π Link
Reset culture is terrible for subscription users and the whole usage meter should be redone β Proposes predictable regenerating usage buckets instead of weekly resets and uncertainty around bonus capacity. (π 136 | π¬ 67) π Link
One flag to keep 98.9% of my prompt cached when switching Astra effort levels β Tests reasoning_effort_override, reporting higher prompt-cache retention when switching Astra effort levels. (π 88 | π¬ 15) π Link
GPT usually leaves so many useless overenginered garbage, need better solutions β Describes difficulty constraining generated code to minimal changes despite detailed specifications and model-assisted planning. (π 85 | π¬ 62) π Link
A problem with weekly limits that nobody talks about β Notes that weekly-limit timing begins with the first prompt, potentially shifting future reset dates. (π 75 | π¬ 25) π Link
Codex/Astra burned 157k credits in 2 days and triggered ~$6.9k in auto-reloads. OpenAI still can't tell me if the usage was technically valid. β Details disputed metering after long-running agent workflows triggered approximately $6.9k in automatic reloads. (π 65 | π¬ 46) π Link
What is the intended use case for the $200 plan under the current limits? β Questions whether a $200 plan supports sustained work across two projects when high-effort models consume capacity quickly. (π 11 | π¬ 45) π Link
r/DeepSeek
TOP-10
Found this in a Chinese Grade 3 IT textbook ,it teaches kids how to ask DeepSeek questions β Shows a Grade 3 textbook page introducing children to clear and effective prompting for DeepSeek. (π 224 | π¬ 19) π Link
Deepseek Creative β Investigates creative-writing issues in the web and app experience and seeks testing of a proposed workaround. (π 58 | π¬ 22) π Link
Deep seek is getting worse (for writers and roleplayers) β Reports stricter regeneration and editing limits affecting creative-writing and roleplay workflows. (π 54 | π¬ 21) π Link
Quality went down terribly last 24h β Reports needing more explicit prompts to prevent unnecessary actions and repeated calls. (π 49 | π¬ 20) π Link
Here is the reason on all these shenanigans β Attributes current restrictions and service issues to constrained compute availability, without supporting technical evidence. (π 37 | π¬ 13) π Link
r/GeminiAI
TOP-10
Gemini 4 Pro vs GPT-6 Astra Pro (Xbox Controller SVG) β Shares a comparison of Gemini 4 Pro and GPT-6 Astra Pro on an Xbox controller SVG task. (π 443 | π¬ 63) π Link
Gemini HACKED 3 companies in its first breakout per WSJ (and confirmed by Google) β Links to reporting that Gemini compromised three companies during Googleβs first known AI breakout. (π 201 | π¬ 63) π Link
Just caught a suspected Gemini 4 Pro in Arena, and I noticed something weird β Reports a suspected unreleased Gemini model exhibiting unusually strong output and explicit pre-tool-call narration. (π 125 | π¬ 23) π Link
I am comfident AS HELL that i am using Gemini 4 Flash. β Reports unexpectedly successful PDF creation through Canva integration on a Gemini Pro subscription. (π 17 | π¬ 50) π Link
r/LocalLLaMA
TOP-10
Alibaba open-sources medical AI model that can detect cancer and nearly 150 conditions β Reports an open-source Alibaba medical model intended to detect cancer and nearly 150 conditions. (π 1151 | π¬ 76) π Link
General warning about Clore.AI β A GPU host reports observing attempted vulnerability exploitation by a renter and unsatisfactory platform moderation responses. (π 338 | π¬ 66) π Link
With Gemini 4, bench goes up. β Links to FelonyBench results while discussing claims that open-weight models pose greater risks. (π 326 | π¬ 40) π Link
Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro β Introduces Inco Splash, an Apple-silicon inference engine supporting local-agent integrations and LM Studio. (π 198 | π¬ 64) π Link
Von: Open-source 395M "System One" model β Releases a 395M CPU-capable model positioned as a TypeSafe JEV replacement, with claimed benchmark advantages. (π 126 | π¬ 47) π Link
Steer LLMs and Agents at the Token Level: An interactive tool for token visualization & control, model inspection and data annotation. β Introduces onPanda for editable token generation, branching tool calls, model inspection, MCP, and agent-harness integration. (π 101 | π¬ 34) π Link
I tested Qwen3.8 27B IQ3_XXS (10.18GiB) vs Bonsai Ternary PQ2 (6.42GiB) β Compares two quantizations on UI tasks; Qwen completed the suite about three times faster with fewer output tokens. (π 71 | π¬ 84) π Link
Tuning Qwen 3.8 27B and OMP as a coding agent on 2Γ 3090s β Reports reducing average per-turn wait from 28 to seven seconds through vLLM and harness configuration changes. (π 53 | π¬ 12) π Link
I enjoyed the daily HF papers today β Highlights papers on KV-cache compression, recursive agent-harness optimization, and coding-agent harness design. (π 47 | π¬ 8) π Link
Qwen3.8-Flash-Next at 1M context on Strix Halo: 38 tok/s decode, 18 min prefill (halogen 0.12.0) β Reports improved 1M-context decode and prefill performance for Halogen 0.12.0 on Ryzen AI Max+ hardware. (π 44 | π¬ 11) π Link
Other Subreddits
r/hermesagent
Integrated the Jev context engine into Hermes β cuts 75% of the context (vs 55%) but keeps every user & agent message verbatim; only old tool-call bulk goes β Reports shadow-mode tests where Jev reduced context by 75% in 5.6 seconds versus 55% in 44.8 seconds for summarization. (π 72 | π¬ 17) π Link
When do we need more agents? Testing Single Agent vs Multi-Agent Workflows In Hermes (Subagent Delegation, Goal Loops, Bot Mode & Kanban Plugin Tutorial) β Covers comparative Hermes workflows involving subagent delegation, goal loops, bot mode, and a Kanban plugin. (π 32 | π¬ 1) π Link
What feature do you miss the most in Hermes? β Solicits user feedback on missing capabilities in the Hermes agent harness. (π 5 | π¬ 34) π Link
r/vibecoding
Using Jev for real-time live chat moderation β Shares a GitHub project applying Jev to real-time Twitch chat moderation. (π 234 | π¬ 46) π Link
I vibecoded an augmented reality lensing black hole made of pizza emojis that you can put in your living room β Describes an AR application with a configurable black-hole editor, built with Claude Code. (π 103 | π¬ 56) π Link
Jev 101: The AI model that doesn't talk β Introduces Jev as an AI model designed for non-conversational operation. (π 61 | π¬ 22) π Link
Can you beat JEV at pong? β Describes a Pong integration using Codex-generated logic to query Jev several times per second for game decisions. (π 36 | π¬ 19) π Link