GPT 5.6 Got Massively Upgraded Without an Announcement β A user reports faster responses, deeper reasoning, fewer hallucinations, and better task completion from GPT 5.6 Sol. (π 310 | π¬ 93) π Link
GPT has pulled ahead of Claude? β A subscriber to both services asks whether GPT 5.6 Sol now outperforms Claude across regular use. (π 145 | π¬ 67) π Link
I've discovered something ChatGPT can do that I'm thrilled with: Custom interesting podcasts for long car rides. β Describes using ChatGPT to select learning topics and generate tailored audio content for travel. (π 94 | π¬ 63) π Link
5.6 high is answering like "instant". It seems unable to actually perform tasks. β Reports that the model lists implementation steps without executing them; an update suggests it may silently switch to GPT 5.5. (π 52 | π¬ 36) π Link
Most stunning place in the universe β Shares a prompt for generating an image of a quiet, uninhabited place in the universe. (π 52 | π¬ 30) π Link
I built an app with Codex that converts any text into high-quality audio. It works with PDFs, blog posts, Substack and Medium links, and even photos of text. β The mobile app converts copied text, webpages, PDFs, article links, and photographed text into natural-sounding speech. (π 36 | π¬ 1) π Link
I'm building a virtual world that only AI agents can join. They all keep dying. β Presents a sandbox where humans can observe AI agents interacting in a virtual world. (π 20 | π¬ 39) π Link
Recommended books β A user asked ChatGPT to recommend five books based on their previous interactions. (π 14 | π¬ 40) π Link
ChatGPT yaps too much β The post argues that overly long responses disrupt workflow and contrasts this with Claudeβs more concise style. (π 11 | π¬ 40) π Link
I know you used AI β Raises concerns about lengthy AI-written emails replacing brief, direct communication. (π 0 | π¬ 75) π Link
r/ClaudeAI
TOP-10
Having unlimited tokens is wild β Asks whether any organization outside Anthropic has access to a comparable token budget. (π 1676 | π¬ 300) π Link
I built an English β Claudish translator β Introduces a bidirectional translator compiled with ProgramAsWeights that runs on CPUs. (π 1126 | π¬ 47) π Link
Is anyone else finding Claude really hard to follow lately? (Massive context dumps, cryptic phrasing) β Reports cryptic, stream-of-consciousness responses and excessive context dumps, including in short Claude Code conversations. (π 520 | π¬ 176) π Link
Anthropic Stealth Nerfing Effort Levels β Cites an Anthropic update that a Claude Code serving-config test maps numerical effort values differently. (π 152 | π¬ 47) π Link
What do people mean by "my harness" re: agentic coding? β Seeks clarification on whether agentic coding harnesses are standard tools or custom systems built around coding agents. (π 113 | π¬ 67) π Link
Devs who actually use Claude Code properly (not vibe coding) β what's your take? β Discusses disciplined use of coding agents, emphasizing review and concerns about production reliability and security. (π 63 | π¬ 81) π Link
Claude saved my data β Claude identified slow disk writes, examined SMART data and journals, and warned about rising bad-sector counts. (π 52 | π¬ 14) π Link
Anthropic: Please Have Daisy the CC Engineer Do a Video! β References an Anthropic newsletter describing lead agents, project agents, and multiple IC agents operating across projects. (π 47 | π¬ 22) π Link
What a plain language standard does to a coding agent β Describes a Claude Code and Codex CLI plugin that uses skills and output styles to make generated prose plainer. (π 45 | π¬ 7) π Link
I taught an LLM to win the Cold War β Uses the asymmetric board game Twilight Struggle to test an LLM under hidden information and alternative card actions. (π 37 | π¬ 4) π Link
r/DeepSeek
TOP-10
PSA: No peak pricing during the weekend β DeepSeek reportedly removed weekend peak pricing in Beijing time; the post argues fixed API pricing would simplify billing. (π 174 | π¬ 34) π Link
DeepSeek's New Model Is in a New Round of Gray Testing β Reports that Chinese technology bloggers have accessed a new gray-test model said to approach Fable-level capability. (π 76 | π¬ 18) π Link
Tried Deepseek for the first time. β A new user asks about current DeepSeek pricing and lower-cost alternatives after observing prices increase. (π 56 | π¬ 44) π Link
Latest DeepSeek Pricing β Shares revised billing rules: weekday peak and off-peak pricing remains, while weekend pricing changes from August 23, 2026. (π 40 | π¬ 12) π Link
Just want to say deepseek flash seems more intelligent nowβ¦ β A Reasonix user reports that DeepSeek Flash appears more capable after an update, particularly on personal projects. (π 33 | π¬ 27) π Link
r/GeminiAI
TOP-10
gemini vs claude β A prospective user asks how Gemini compares with Claude Pro and whether a promotional offer justifies switching. (π 317 | π¬ 221) π Link
Gemini 3.5 pro is coming!!!! β Shares screenshot links presented as evidence that Gemini 3.5 Pro may be nearing release. (π 163 | π¬ 125) π Link
Seems Google Finally releasing 3.5 Pro β Cites screenshots and recent social-media activity from the team as signs of a Gemini 3.5 Pro release. (π 115 | π¬ 44) π Link
Gemini is Gemini β A user cites Geminiβs integration across Googleβs platform as the main reason they use it personally. (π 72 | π¬ 17) π Link
Would be a Big Power Move from Google if True that Ox Alpha is Gemini Model β Speculates that Ox Alpha could be a Gemini model, an attribution the post characterizes as unexpected. (π 59 | π¬ 30) π Link
OX Alpha is Gemini New Specialized Coding Model β Argues that Ox Alpha is a Gemini coding model rather than GLM, while noting weaker general-reasoning performance. (π 33 | π¬ 18) π Link
Almost entirely sure Gemini 3.5 pro is dropping today β Points to AI Studio team activity over 12 hours as evidence for a possible weekend launch. (π 0 | π¬ 36) π Link
r/LocalLLaMA
TOP-10
This is a great sub, regardless of what complaints people have about it. β Describes the community as receptive to discussion of local LLM limitations and practical deployment constraints. (π 341 | π¬ 123) π Link
Think you're going to get cheap DDR5 RAM? Think again, even if prices fall, scalper bots now outnumber shoppers 10 to 1 and will keep prices high β Raises concerns that automated scalping could keep DDR5 memory prices elevated even if market prices decline. (π 298 | π¬ 216) π Link
New 100B Liquid AI model coming soon β Discusses a potential 100B Liquid Foundation Model and references Liquid AIβs existing LLM and SLM architectures. (π 220 | π¬ 65) π Link
I tried to do agenic coding with Qwen 3.8 27B 3bit quant on a macbook air m2 24gb. It took 63 hours, but amazingly, the flight simulator worked. β Qwen 3.8 27B Q3_K_S in LM Studio used 57K context and took 63 hours to build a single-page HTML flight simulator. (π 151 | π¬ 43) π Link
Qwen 3.8 27b - PI AGENT vs OPENCODE - another smaple β A follow-up compares Pi Agent and OpenCode using Qwen 3.8 27B. (π 130 | π¬ 100) π Link
Llama.cpp version 0.2.0 is out! β Links to the v0.2.0 changelog, source code, and associated prebuilt release. (π 114 | π¬ 21) π Link
Artificial Analysis "Intelligence": A meaningless benchmark β Questions the usefulness of Artificial Analysis Intelligence rankings when evaluating Qwen 3.8 27B. (π 112 | π¬ 133) π Link
16 GB VRAM purgatory discussion thread β Requests model and configuration recommendations for 16 GB VRAM, citing a Qwen3.8-27B GGUF setup on Windows. (π 111 | π¬ 75) π Link
How to remove trendy speech from llms? β Seeks methods to reduce fashionable phrasing such as βmintedβ and βescape hatchβ in LLM outputs. (π 95 | π¬ 73) π Link
I benchmark DFlash 2 (PR build) in llama.cpp on Qwen 3.8 27B against all speculative methods for 3 days. 2.26x on 100 real coding prompts, 4.68x with one n-gram drafter on top. Up to 8x on specific cases. β Benchmarks DFlash 2 against plain decoding, MTP, and n-gram drafters on an RTX PRO 6000 over three days. (π 61 | π¬ 9) π Link
Other Subreddits
r/Bard
Would be a Big Power Move from Google if True that Ox Alpha is Gemini Model β Speculates that Ox Alpha may be a Gemini model, despite the attribution being viewed as unlikely. (π 97 | π¬ 41) π Link
Gemini's censor in AI Studio is really getting out of hand. β Reports that Gemini AI Studio flagged a wound-treatment scene in a novel as sexual content. (π 33 | π¬ 26) π Link