I had GPT-6 Astra build a full hospital corridor scene in Blender using Blender MCP, no manual modeling at all β GPT-6 Astra used Blender MCP to create architecture, props, lighting, and a flythrough from text prompts in about 45 minutes. (π 110 | π¬ 23) π Link
Bye Claude, you've reached my token limit β A user reports moving back to ChatGPT and Codex after Claude limits disrupted coding and transcription workflows. (π 59 | π¬ 42) π Link
Reset just landed! (For agentic usage in all paid plans afaik) β Reports that agentic usage allowances were reset across paid plans. (π 50 | π¬ 44) π Link
Did ChatGPT suddenly get way worse at understanding/communicating? β The user reports recent context loss, with responses focusing on isolated details rather than the broader request. (π 43 | π¬ 26) π Link
The GPT 6 Astra Downgrade Was Real. OpenAI Acknowledged And Fixed It (Partially) β Argues that an Astra performance reduction was acknowledged and partly reversed, calling for recurring model benchmark tests. (π 40 | π¬ 14) π Link
Best AI subscription for the dollar? Whatβs the best to spend on? ChatGPT or Claude or Gemini? β Seeks current subscription comparisons for research, reports, administrative writing, contracts, and design brainstorming. (π 27 | π¬ 41) π Link
Teaching about AI at college, seeking help. Why does ChatGPT always take your side, and can we fix it? β A communications instructor describes model agreement with contradictory framing and asks how to teach or mitigate this behavior. (π 17 | π¬ 104) π Link
r/ClaudeAI
TOP-10
Please write your own posts. β Discusses concern that AI-written submissions are reducing the quality of community discussion. (π 472 | π¬ 95) π Link
Dario Amodei β We Must Pace the Frontier β Dario Amodei argues for pacing frontier AI development. (π 345 | π¬ 126) π Link
I have to say something as a chinese β Challenges Anthropicβs framing of allegations involving Chinese model providers, reseller routing, competitor benchmarking, and distillation. (π 308 | π¬ 248) π Link
Warning: Claude "incognito" chats with uploads CAN be listed and retrieved β Uploaded files from incognito chats reportedly remain listed in privacy settings and can reopen the original conversation. (π 187 | π¬ 27) π Link
Claude helped me make a custom e-book, and now I can play pokemon on any device with a web browser and use the display as a game dashboard! β An ESP32 e-ink device serves an emulator locally over Wi-Fi, with persistent saves and game-state dashboard data. (π 166 | π¬ 22) π Link
Milkyway Andromeda collision by Opus, Fable and GPT-6 Astra. One model won and it is gorgeous. β Compares models building galaxy-collision simulations, finding Opus and Fable produced more dynamic JavaScript simulations than Astraβs React project. (π 84 | π¬ 25) π Link
The First and Last Plugin You Should Ever Install in a Claude Code Project: Quartermaster, the Plugin That Turns Your Claude Code into a Continuously Self-Improving Machine β Quartermaster configures Claude Code tools and permissions, then evaluates whether later recommendations improved project workflows. (π 84 | π¬ 10) π Link
Anthropicβs new report has no winners β and the part nobody is talking about is user privacy β Examines privacy risks when third-party model routers may forward sensitive prompts between providers without usersβ knowledge. (π 72 | π¬ 33) π Link
Anyone else notice Claude suddenly using British spelling? β A US English user reports recent British spelling in Claudeβs thinking summaries and responses despite locale settings. (π 69 | π¬ 84) π Link
Writing a one-page "style guide" for Claude was the highest-value hour I've spent β A short reusable style guide improved output consistency by specifying structure, audience handling, inference labels, and concrete examples. (π 54 | π¬ 22) π Link
r/codex
TOP-10
This nerf is getting out of hand β look at this Before vs After β A user reports a sharp quality difference between sessions, saying the behavior returned to normal after a usage reset. (π 1195 | π¬ 293) π Link
How many of you guys have switched back to Sol? β A developer asks whether users have returned from Astra to Sol, citing token costs and limited quality-of-life improvement. (π 140 | π¬ 90) π Link
Pro 20x user suddenly downgraded to 5x, werenβt existing subscribers supposed to keep 20x? β Reports an unexpected subscription allowance change from Pro 20x to 5x and asks whether grandfathering policy changed. (π 114 | π¬ 75) π Link
Truly heed the warning of 5.6 Sol deleting your hard drive β Reports an rm -rf command expanding to rm -rf /* after a missing environment variable, recommending isolation and data-loss safeguards. (π 104 | π¬ 122) π Link
GPT 6 Astra in ChatGPT Work is actually kind of insane β Reports that Astra consumes substantially less allowance on non-coding tasks because it reads less repository context. (π 38 | π¬ 54) π Link
What the hell is going on? Sol XHigh has a lot of context rot since yesterday... And Astra is just worse. β Reports repeated context degradation with Sol XHigh after moving away from Astra for complex work. (π 35 | π¬ 11) π Link
How to get Astra + Subagents to be less usage heavy than Astra Solo β Tests suggest subagents reduced Astra spending by only about 20% while increasing task duration by roughly 2.5 times. (π 18 | π¬ 38) π Link
What are your best alternatives to Codex? β Requests alternatives that balance pricing with coding capability comparable to Astra, Fable, or Sol. (π 3 | π¬ 39) π Link
For anyone out there with a 5X-20X Plan, prototype with ChatGPT 6 Pro and then polish, build with Codex β Suggests reserving ChatGPT Work for research and drafts, using Codex primarily for implementation, and disabling memory for output quality. (π 2 | π¬ 33) π Link
The Codex weekly limit doesn't roll over, and the "goodwill reset" is a loan, not a gift. Did the maths. β Explains that unused weekly capacity expires and early resets move the following reset date rather than adding entitlement. (π 2 | π¬ 30) π Link
r/DeepSeek
TOP-10
Surprisingly, Weβre Still seeing real demand for V4 Flash. V4.1 free now. Try it out. β InferX says demand remains strong for V4 Flash and offers V4.1 Flash free while maintaining zero-data-retention policies. (π 97 | π¬ 53) π Link
About Flash 4.1 β A user finds Flash 4.1 stronger for some AI tasks and image search but weaker at reasoning, context, search defaults, and creative writing. (π 82 | π¬ 24) π Link
Y Combinator's Garry Tan wants US open-weight AI labs to 'distill' frontier models, too β Shares Garry Tanβs proposal for US open-weight labs to distill frontier models and discusses perceived inconsistency in training-data debates. (π 70 | π¬ 15) π Link
My problem with this update β Requests the return of separate Expert Mode, citing less detailed creative-writing output after the update. (π 42 | π¬ 19) π Link
DeepSeek is officially unusable for deep psychological/abstract work (and its refusal style is worse than GPT) β Reports restrictive refusals in psychological and abstract discussions, particularly when challenging the modelβs reasoning. (π 30 | π¬ 28) π Link
r/LocalLLaMA
TOP-10
3.8-27B has ruined 3.5/3.6-35Bβs for me. Itβs just absurdly superior. β In replicated applied-science workflows, Qwen 3.8-27B reportedly used fewer tokens and RAM while outperforming several 35B alternatives. (π 503 | π¬ 217) π Link
Agnes-AI/Agnes-3.0-Flash 33B Multimodal, AA score: 36 β Highlights a 33B multimodal model with 262,144-token context, hybrid recurrent-global attention, adjustable reasoning, and tool calling. (π 240 | π¬ 49) π Link
bartowski/Qwen3.8-27B-GGUF Β· Hugging Face - Updated (Per-tensor layout) β Links updated Qwen 3.8-27B GGUF files and documentation on per-tensor layout maps for quantization. (π 190 | π¬ 30) π Link
For those of you forced to only use open models from Western labs in production, what are you deploying? β Seeks Western open-model options with vision, long context, and 120B-plus scale for deployment on four H100 GPUs. (π 104 | π¬ 175) π Link
Qwen-Next seems worse to me then 3.8 27b for coding, but I feel like I must be missing something? β On an M5 Max with 128GB RAM, the user finds Qwen 3.8-27B stronger on difficult coding tasks than Qwen-Next. (π 78 | π¬ 91) π Link
tencent/AuK-Flash Β· Hugging Face β Releases AuK-Flash, a distilled 1.5B speech model supporting four-step inference, TTS, editing, enhancement, and source separation. (π 75 | π¬ 14) π Link
Qwen3.8 Flash Next now at 1.2k t/s prefill on Strix Halo β Reports matching 1,200 tokens-per-second prefill on Strix Halo through llama.cpp optimization, with planned upstream contributions. (π 70 | π¬ 21) π Link
Qwen3.8 Flash Next llama.cpp config tuning β Shares dual-RTX 3090 llama.cpp settings and measured 130β200 prefill TPS with 14β22 generation TPS. (π 52 | π¬ 48) π Link
Anybody use frontier models like Astra/Fable for planning/judging, and qwen3.8 as the main workhorse? Curious to hear about your setups! β Proposes a hybrid workflow where cloud models plan and critique while local Qwen handles implementation. (π 49 | π¬ 44) π Link
Unsloth UD-quants - Qwen 3.8 27b for example - worth using 8-bit or stick with faster 6 bit for coding? β Asks whether 8-bit quantization offers perceptible coding benefits over faster 6-bit variants in larger projects. (π 46 | π¬ 92) π Link
r/vibecoding
TOP-10
I asked GPT-6 to get more users for one of my projects... β Reports GPT-6 computer control creating social accounts and promotional posts, producing traffic but also unsolicited cross-platform outreach. (π 209 | π¬ 55) π Link
After a year of vibe coding a side project, I have 40 users and 2 paying subscribers. Here's what I learnt along the way β Recommends constrained MVP scope, separate databases, screenshot-based UI feedback, Playwright checks, CI/CD, and regular refactoring. (π 73 | π¬ 82) π Link
Is it just me or does muse spark feel incredibly benchmaxed? β Questions whether Muse Sparkβs practical coding performance in OpenCode matches its frontier-model benchmark positioning. (π 45 | π¬ 35) π Link
I built a free photo-culling tool with Cowork - it takes 8,000 trip photos down to my best 50 (Cull β Dedup β Rank) β Describes a local browser tool using sharpness analysis, perceptual hashing, ORB matching, and ranking metrics for photo selection. (π 16 | π¬ 34) π Link
Deployment of Web-Apps β Seeks deployment workflows for Python, Flask, and JavaScript applications after locally built MVPs reach production. (π 3 | π¬ 36) π Link
Other Subreddits
r/GeminiAI
How is this legal? β Raises concerns that Gemini Apps Activity combines conversation history access with potential human review and model-improvement use. (π 84 | π¬ 35) π Link
r/hermesagent
That icon. β A consultant says Hermesβ anime-style icon complicates adoption in business environments and requests a configurable alternative. (π 91 | π¬ 127) π Link
What can I do with Hermes Agent as a beginner? β A new user asks for accessible examples of workflows, automations, and real-world Hermes Agent use cases. (π 57 | π¬ 25) π Link
Built my own Web UI β Describes a synchronized Mac and iPhone PWA with model routing, fallbacks, provider quota tracking, project management, and low memory use. (π 49 | π¬ 11) π Link
What can you learn from 670+ Hermes Agent projects and 480+ plugins: meet Hermes Advisor Skill β Introduces a local Hybrid-RAG tool indexing Hermes projects and plugins to generate architecture recommendations for agent ideas. (π 40 | π¬ 2) π Link
r/LLMDevs
Same batch job: $97 on Claude Sonnet vs. $13 on a rented H200. What am I missing? β Compares pay-per-token APIs against self-hosted Qwen inference on rented H200 hardware for processing 1,000 long documents. (π 1 | π¬ 39) π Link