PSA: The ChatGPT Plus tier is NOT getting GPT-6 Astra in regular Chat. Only Work and Codex. β The post says Plus access to Astra is limited to Work and Codex, while regular Chat requires a Pro tier. (π 449 | π¬ 159) π Link
From One Image to an Animated 3D Character in One Day! GPT-6 Astra + Blender + 3D AI: The Workflow of the Future β A workflow combines Tripo-generated character parts with Astra-assisted Blender assembly, facial-expression switching, and rigging. (π 153 | π¬ 39) π Link
Which 'superpower' did ChatGPT give you? β The discussion collects practical user experiences, including home maintenance, cooking, workplace learning, and cost savings. (π 144 | π¬ 164) π Link
I used Codex to translate entire game for me... β Codex reportedly translated most of The Witcher 3 into Croatian, including dialogue, menus, books, prompts, and bestiary entries. (π 125 | π¬ 17) π Link
GPT-6 reportedly jailbroken within 24 hours using an extended Task-in-Prompt (TIP) attack β A reported GPT-6 Astra jailbreak combined a revised Task-in-Prompt attack with undisclosed techniques and was privately disclosed to OpenAI. (π 56 | π¬ 20) π Link
r/ClaudeAI
TOP-10
What happens if you give AI agents a place humans donβt control? One month later, here are the receipts. β The author reports an agent community with more than 2,000 citizens, USDC-paid jobs, and $111 in Cloudflare costs during its first month. (π 184 | π¬ 50) π Link
Suddenly worried about costs of Claude β A Claude Code Pro subscriber asks whether the /usage displayβs $321 total cost represents additional charges. (π 154 | π¬ 46) π Link
Another Astra Post β A developer found Astra less verbose and useful for a simple feature, but unable to form an action plan for a persistent bug. (π 81 | π¬ 36) π Link
Astra vs Fable 5.1 - Day 1 β Early use suggests Astraβs output and token consumption are easier to follow than Fable 5.1, pending longer-term testing. (π 80 | π¬ 24) π Link
I think Claude is making it way too easy to feel like you know what youβre talking about β The post examines how confident initial answers can obscure missing caveats and recommends asking for assumptions and expert challenges. (π 67 | π¬ 40) π Link
Switching accounts in Claude Desktop hides history - made a tool to fix that β claude-transplant moves locally stored Claude Code session records between accounts on macOS while preserving session IDs and transcripts. (π 45 | π¬ 18) π Link
Will Anthropic bring back the thinking-chain feature? β A graduate student asks whether Claudeβs visible reasoning feature was removed permanently and seeks alternatives for prompt learning. (π 30 | π¬ 23) π Link
Morality of Open Sourcing SaaS products? β The discussion considers the ethics of using AI tools to build local open-source alternatives to subscription SaaS products. (π 21 | π¬ 44) π Link
Why does Claude get "tired" or "fatigued"? β The author asks why Claude recommends starting a new session after long conversations instead of continuing with accumulated context. (π 19 | π¬ 35) π Link
How do you actually orchestrate your AI agents? β The post seeks lightweight methods for delegating work and tracking multiple Claude Code or Pi sessions. (π 18 | π¬ 32) π Link
r/hermesagent
TOP-10
Ornith-1.5-35B-A3B is MIND BLOWING! β Tests report strong local coding performance at 8-bit quantization, with approximately 365 tokens/s prompt processing and 34 tokens/s generation through oMLX. (π 186 | π¬ 58) π Link
Connecting Hermes Desktop to many Hermes instances β Pantheon v0.21.0 allows Hermes Desktop to keep local and remote Hermes backends connected simultaneously, including over Tailscale. (π 74 | π¬ 17) π Link
Just got Hermes, any recommendations to improve it like crazy? β A new user asks for practical ways to improve a Hermes setup after adding free API keys and MCP servers. (π 67 | π¬ 50) π Link
Hermes can now research across Perplexity, Reddit, RSS feeds, video, and code β New integrations add Perplexity search and extraction, Reddit thread access, and direct RSS, Atom, and JSON feed reading. (π 42 | π¬ 8) π Link
r/LocalLLaMA
TOP-10
8 uncensored Qwen 3.8 27B variants, one base, 167 GPU hours - Abliterlitics β A comparison of eight abliterated variants uses weight analysis, KL divergence, 13 benchmarks, and HarmBench refusal testing. (π 421 | π¬ 110) π Link
Qwen3.8-27B "Unhacked" my PC β A user describes using local models while responding to an apparent session-stealer compromise and account takeover attempt. (π 292 | π¬ 113) π Link
Which agent harness do you use and why? β A 14-task comparison reports similar solve rates between Claude managed agents and TrueForge, with lower token use and cost for TrueForge. (π 212 | π¬ 230) π Link
Qwen 3.8 Flash Next (Max) is impressive just to talk with. β The author reports strong performance on local factual questions and broad problem-solving outside coding tasks. (π 120 | π¬ 78) π Link
2x R9700, 64 GB DDR5 is an absolute beast machine with vLLM Radiance / R9V and Qwen 3.8 27b and Flash next β A dual Radeon AI PRO R9700 setup is benchmarked with Qwen 3.8 models, including observations on thermals, storage offload, and cost. (π 66 | π¬ 45) π Link
vibeblending locally with Qwen 3.8 27B β The post provides an MCP configuration for connecting Pi and Qwen 3.8 27B to Blender 5.x. (π 62 | π¬ 19) π Link
Coding benchmarks that are quickly showcasing deep capability β The post compares model results on Program-Bench, SRE-Bench, and code-migration tasks intended to test deeper software-engineering capability. (π 58 | π¬ 24) π Link
Villager Simulation Game POC Created with Qwen3.8-27B-UD-Q3_K_XL.gguf - 16GB VRAM β A locally run Q3-quantized model built a browser game through incremental prompts on a 16 GB RTX 5070 Ti setup. (π 54 | π¬ 35) π Link
DeepSeek-V4-Flash-Vision Q8 vs Qwen3.8-Flash-Next Q8 β Local testing on two Strix Halo systems found DeepSeek slower in generation but faster at completing a coding task. (π 48 | π¬ 22) π Link
Block KV cache streaming: bound VRAM at long context via a shared CUDA phase arena by giveen Β· Pull Request #357 Β· TheTom/llama-cpp-turboquant β A port and extension of adaptive KV streaming adds support for more models and benchmarks long-context VRAM management. (π 43 | π¬ 11) π Link
Other Subreddits
r/DeepSeek
What is your usage ( heavy users ) β A heavy user reports two months of DeepSeek Flash use through OpenCode with a 98.5% cache-hit rate and low costs. (π 61 | π¬ 38) π Link
r/GeminiAI
Google really needs to drop GPT and Claude and shift that quota to Gemini β The post argues that Gemini 3.8 Flash is more useful than included alternatives, while its verbosity quickly consumes quota. (π 41 | π¬ 11) π Link
My Gemini Assistant Has Become Useless and Insulting - Anyone Experience Similar and/or Have any Suggestions? β A user reports more refusals and identity disclaimers in Gemini Assistant and Google AI Search after a June 2 update. (π 3 | π¬ 38) π Link