🤖 AI News Summary
2026-07-23 13:20 GMT+8 · summary_2026-07-23_13-20.md

🤖 AI News Summary - 2026-07-23 13:20 GMT+8

Focused AI/dev subreddit roundup.

Full site: https://ai-news-summary.pages.dev/

What changed since last run


r/openai

#PostSummaryTimeScoreAuthorCommunity reaction
1Gemini 3.6 Flash: twice as fast, 18% cheaper, and precisely 0% smarter🥲[Image: Gemini 3.6 Flash: twice as fast, 18% cheaper, and precisely 0% smarter🥲] Google released Gemini 3.6 Flash and independent testing found exactly zero intelligence improvement over 3.5 Flash. It is basically 3.5 Flash after an inference-cost consultant optimized the serving stack.2026-07-22 21:28 GMT+8/u/etherd0tCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly treat Gemini 3.6 Flash as an economics and serving-efficiency story rather than a capability jump: several argue that big AI labs are losing money, subsidizing enterprise usage, and therefore have to optimize inference, raise prices later, or sell compute rather than keep pushing frontier models for their own sake. The main disagreement is whether Google has genuinely given up on SOTA or is simply prioritizing a huge consumer workload like Search while still funding frontier work; the practical operator takeaway is that token costs, utilization, and deployment efficiency are now as important as benchmark gains, especially when consumer-scale products and enterprise subsidies shape the serving stack. Overall sentiment — post: mixed; author: neutral. Reply threads: 2026-07-22 21:34 GMT+8: post=skeptical, author=neutral — They argue that Google has effectively given up chasing SOTA and frontier models. | 2026-07-22 23:29 GMT+8: post=positive, author=neutral — They say big AI labs are operating at a loss, subsidizing enterprise customers, and will eventually need… | 2026-07-23 02:43 GMT+8: post=positive, author=neutral — They note that lab token costs have dropped by about 100x over the last couple of years, that consumer usage…
2People who got their OpenAI made $230 mini keyboard, what are your reviews?[Image: People who got their OpenAI made $230 mini keyboard, what are your reviews?] Is it worth it?2026-07-23 11:45 GMT+8/u/ImaginaryRea1ityCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly dismiss the $230 mini keyboard as an overpriced macropad, saying the same workflow can be replicated for under half the price with better keycaps/switches or with an Elgato Stream Deck plus knobs, especially if you can use software/firmware like OpenDeck. One commenter argues it is “not even close to the Codex Micro” and asks whether critics have seen it work, but the thread is otherwise dominated by comparisons to cheaper alternatives and outright ridicule, with no practical endorsement of the price/performance tradeoff. Overall sentiment — post: critical; author: critical. Reply threads: 2026-07-23 12:00 GMT+8: post=critical, author=neutral — They say macropods have existed for a long time and that the same functionality can be built for less than… | 2026-07-23 12:35 GMT+8: post=neutral, author=neutral — They prefer an Elgato Stream Deck with knobs and note that it cost about $120. | 2026-07-23 12:45 GMT+8: post=positive, author=neutral — They argue it is not comparable to the Codex Micro and challenge critics to look at how it actually works.

r/LocalLLaMA

#PostSummaryTimeScoreAuthorCommunity reaction
1Laguna S 2.1 looping fix incoming[Image: Laguna S 2.1 looping fix incoming] EDIT - Poolside have updated the INT4, NVPF4, and FP8 versions with a fix for the looping issue many of us have been seeing. Full precision and GGUFs remain unchanged.2026-07-23 04:32 GMT+8/u/rmhubbertCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly agree the model is strong when it avoids the looping bug, with one calling Laguna S 2.1 the best ~120B-class model they have used and another saying their only issue so far has been the thought-looping. The main caveat is deployment timing and re-quantization: people are waiting to see whether the INT4, NVPF4, and FP8 fixes actually hold up, while GGUF and full-precision builds are still unchanged and Unsloth 6-bit users were told they would need new quants based on the updated version. Practical operator takeaways were to hold off for a couple weeks if you want to avoid chasing regressions, and to expect some tuning work on sampling/settings because users are still asking what works best, while q3 testing suggests it can catch Claude-written bugs but may overthink heavily. Overall sentiment — post: mixed; author: neutral. Reply threads: 2026-07-23 05:03 GMT+8: post=positive, author=neutral — They say Laguna S 2.1 is the best ~120B-class model they have used as long as it does not fall into… | 2026-07-23 05:00 GMT+8: post=skeptical, author=neutral — They plan to wait a couple more weeks before trying Laguna S 2.1 and only revisit it if people stop reporting… | 2026-07-23 07:09 GMT+8: post=positive, author=neutral — They report that q3 still catches Claude-written bugs and that Claude validated the analysis, but also note…
2🇦🇹 Austria is rolling out a government AI-platform using Mistral models and Open WebUI[Image: 🇦🇹 Austria is rolling out a government AI-platform using Mistral models and Open WebUI] This is a surprisingly large real-world deployment: “GovGPT” is part of Austria’s Public AI initiative, running on sovereign infrastructure (in their BRZ - federal datacenter) with Mistral open-weight models. Trending…2026-07-22 22:28 GMT+8/u/ClassicMainCommunity reaction (frontier/gpt-5.4-mini): Most commenters see the platform as a useful sovereign-AI proof of concept for government workflows, especially if it can ingest internal documents and reduce mundane admin work, and one Austrian commenter explicitly frames it as a productivity win for trivial tasks. The main split is over trust and capability: some want open-source/open-weights as a transparency signal while others say it does not change much when the government controls the web UI and inference stack, and a practical caveat is that Mistral 14B was called too weak for tasks like editing Excel files in Open Terminal, with Mistral Small 3.5 suggested instead. Overall sentiment — post: mixed; author: neutral. Reply threads: 2026-07-22 22:37 GMT+8: post=positive, author=neutral — They argue that AI is especially useful when it can use government document context, making it a practical… | 2026-07-22 23:16 GMT+8: post=concerned, author=neutral — They support the idea only if it uses an open-source model and treat missing clarification as a red flag that… | 2026-07-23 01:53 GMT+8: post=skeptical, author=neutral — They say model openness matters little because the government controls both the web UI platform and the…

r/llmdevs

#PostSummaryTimeScoreAuthorCommunity reaction
1One-shot HTML benchmarks should probably show cost next to quality[Image: One-shot HTML benchmarks should probably show cost next to quality] The thing that stood out to me wasn’t just which model looked best. It was how different the cost/quality tradeoff looked once price was shown.2026-07-22 19:21 GMT+8/u/TarandjpopCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly agreed that the post’s useful angle is to show cost next to quality, with multiple people arguing that Gemini Flash/3.6 Flash can be compelling on results per dollar even if it is not absolute SOTA. A recurring operator takeaway was that Claude feels smoother and needs less steering/context than GPT in long iteration loops, while GPT can go rogue and force more backtracking; the main caveat was that a narrow one-shot HTML benchmark may not reflect frontier-task behavior, where cheaper 27B-class models or Kimi can still be sufficient for simpler workloads. Overall sentiment — post: positive; author: neutral. Reply threads: 2026-07-22 22:08 GMT+8: post=positive, author=neutral — They said the benchmarks are a good baseline and added that, in long iterative work, Claude has required less… | 2026-07-23 00:02 GMT+8: post=critical, author=critical — They complained that 3.6 Flash is not SOTA and questioned what the author was thinking. | 2026-07-23 00:30 GMT+8: post=positive, author=neutral — They defended the post by saying its point is pricing and that 3.6 Flash is likely near SOTA for results per…
2Using Claude Opus, GPT-5.5, or GLM-5.2 for every agent turn is surprisingly wasteful[Image: Using Claude Opus, GPT-5.5, or GLM-5.2 for every agent turn is surprisingly wasteful] We noticed Claude Opus, GPT-5.5 and GLM-5.2 were spending most of their time doing routine work like searching files, rerunning tests and updating code, instead of actual hard reasoning. So we built a router that picks the…2026-07-23 12:58 GMT+8/u/entelligenceai17Community reaction (frontier/gpt-5.4-mini): Commenters focus on cache behavior and model-routing economics: one argues it is better to use a cheap subagent because switching models kills the cache, while another says the router design avoids that by keeping most turns on the cheap model so the cache stays warm and only escalates on misses. The practical takeaway is that the post’s idea is only attractive if the router meaningfully reduces expensive-model turns without paying too much cache churn, and the main caveat raised is the cache penalty of frequent model switching. Overall sentiment — post: mixed; author: neutral. Reply threads: 2026-07-23 13:19 GMT+8: post=skeptical, author=neutral — They argue that a cheap subagent is preferable because switching models destroys cache reuse, so using a big… | 2026-07-23 13:22 GMT+8: post=positive, author=neutral — They say the router approach already accounts for cache reuse by keeping most turns on the cheap model and…

r/OpenWebUI

#PostSummaryTimeScoreAuthorCommunity reaction
1MCP / Tool auth in enterprise. How are you doing it?Hey, long-time Open WebUI user here. I run instances in our group for 3 different companies and right now my biggest headache is MCP / Function auth to external tools.2026-07-23 03:43 GMT+8/u/KockafellaCommunity reaction (frontier/gpt-5.4-mini): The thread converges on treating MCP auth as per-user, per-server authorization rather than reusing the Open WebUI login token everywhere: one commenter says to use Open WebUI’s native Streamable HTTP client with a separate authorization flow per user, then validate signature, issuer, audience, expiry, and scopes/roles on the MCP server. The main caveat is that Entra/Graph is called out as an exception that happens to accept the same token Open WebUI logged in with, not a general strategy; another practical workaround mentioned is storing Jira Data Center tokens in Vault and passing them through a filter function to the MCP server. Overall, the actionable takeaway is to keep the MCP user token distinct from any downstream system credentials and avoid forwarding either token beyond its intended boundary. Overall sentiment — post: positive; author: neutral. Reply threads: 2026-07-23 04:15 GMT+8: post=positive, author=neutral — They offer a concrete Jira Data Center pattern: store tokens in Vault, retrieve them with a filter function,… | 2026-07-23 04:24 GMT+8: post=positive, author=neutral — They recommend treating each MCP server as an OAuth-protected resource, using a per-user authorization flow…
2NEURA Office: one hub for all our Open WebUI Office tools[Image: NEURA Office: one hub for all our Open WebUI Office tools] https://preview.redd.it/ka0a2nrd9meh1.png?width=1868&format=png&auto=webp&s=cc57b405c32b18f8f61db7d8fd2fa60dd6835079 (https://preview.redd.it/ka0a2nrd9meh1.png?width=1868&format=png&auto=webp&s=cc57b405c32b18f8f61db7d8fd2fa60dd6835079) Hey everyone đź‘‹…2026-07-22 01:29 GMT+8/u/nixiam87Community reaction (frontier/gpt-5.4-mini): Most commenters are positive about the NEURA Office/Open WebUI tools: one says they have been using it in a company for a week with good stability, another calls it awesome and thanks the team, and one is already looking forward to the Microsoft 365 add-in. The main caveats are practical rather than conceptual: there is a strong request to keep LibreOffice high on the roadmap because LibreOffice users and public-sector deployments have fewer alternatives, one commenter says LibreOffice is still in scope because the tools generate standards-based packages, and another reports a concrete install/runtime failure (No module named 'docx') on OWUI v0.9.6 in GKE despite the readme saying frontmatter dependencies install automatically. Operators should take away that interest is real, but dependency handling and documentation/discoverability need to be reliable, especially across LibreOffice and Microsoft Office workflows. Overall sentiment — post: mixed; author: positive. Reply threads: 2026-07-22 03:01 GMT+8: post=positive, author=neutral — They urge the project to prioritize LibreOffice because Microsoft Office users often already have Microsoft… | 2026-07-22 02:09 GMT+8: post=positive, author=positive — They report a week of company use with the tools feeling easy, stable, and generally strong, but ask how to… | 2026-07-22 03:58 GMT+8: post=concerned, author=neutral — They hit a save-time error, No module named 'docx', on an OWUI v0.9.6 GKE setup and say the promised…
3🇦🇹 Austria is rolling out a government AI-platform using Open WebUI[Image: 🇦🇹 Austria is rolling out a government AI-platform using Open WebUI] 🇦🇹 Austria is rolling out a government AI-Platform using Open WebUI 🇦🇹 This is a surprisingly large real-world deployment: “GovGPT” is part of Austria’s Public AI initiative, running on sovereign infrastructure (in their BRZ - federal…2026-07-22 21:17 GMT+8/u/ClassicMainCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly treated the Austria/Open WebUI rollout as a strong public-sector use case, with one praising the government for choosing an open stack over proprietary tools and another saying it is neat to see Open WebUI used at that level. The main disagreements were technical: one commenter doubted Mistral models are up to the task, while others focused on deployment details such as enterprise licensing, the apparently untouched branding/copyright notices, and a claimed 200k€ initial development/model-hosting cost that would leave little budget for anything beyond infrastructure; one user also claimed a DHS agency is already doing something similar but gave no specifics. Overall sentiment — post: mixed; author: neutral. Reply threads: 2026-07-22 21:45 GMT+8: post=positive, author=neutral — They celebrate the government’s choice as a better use of money than proprietary tools. | 2026-07-22 23:35 GMT+8: post=positive, author=neutral — They infer the platform likely does not use an enterprise license, point to the untouched Open WebUI branding… | 2026-07-23 02:02 GMT+8: post=positive, author=neutral — They argue the debate over Open WebUI being technically not open source was overblown, saying users can…
4Colouring button icons for action functions (tool tip)[Image: Colouring button icons for action functions (tool tip)] TL;DR: How can I synchronize the colour of a custom SVG action icon to behave like the ones in the message’s tool tip? I wasn’t sure if I should flair this as “help” or “feature idea”.2026-07-22 15:55 GMT+8/u/Bulletic1Community reaction (frontier/gpt-5.4-mini): The only reply is a terse technical objection: the commenter says the custom SVG action icon likely cannot be modified via a value/variable, which implies there is no supported way to synchronize its color with the message tool-tip icons. There is no disagreement or deeper discussion, just a practical caveat that the requested behavior may not be available in the current product. Overall sentiment — post: skeptical; author: neutral. Reply threads: 2026-07-22 16:02 GMT+8: post=skeptical, author=neutral — The commenter says the icon probably cannot be modified through a value, implying the requested color…
5Not a UI but a model questionI am playing with selfhosting models, and I get drastic different results. Working with online AI I am starting to get it, but if someone has a resource to help me further understand why this happens for the same prompt “When was the war of 1812” ` Hermes-3-Llama-3.1-70B-8bit Today at 8:51 AM The War of 1812 was…2026-07-22 21:03 GMT+8/u/Kevin_CossaboonCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly steered the question away from the model output itself and toward tooling choices: OpenWebUI docs were cited for SearXNG/web search augmentation, Ollama was recommended for supported-model lists, and one commenter said qwen3.6 models are currently the most useful in the ~30B class. The only concrete disagreement was not over the post but over the mechanism, since the original poster kept asking why search/tool use or JSON output would help with a “War of 1812” prompt and whether they could constrain SearXNG to local models plus the internet; the practical takeaway is that they still need to understand model/tooling boundaries, prompt formatting, and when web-search or MCP-style tools actually improve answers. Overall sentiment — post: neutral; author: positive. Reply threads: 2026-07-22 21:57 GMT+8: post=positive, author=positive — They suggested OpenWebUI docs for installing SearXNG and noted that qwen3.6 models are currently the most… | 2026-07-22 21:58 GMT+8: post=positive, author=positive — They asked what the poster is using to run models and pointed them to Ollama’s supported-model list and… | 2026-07-22 22:20 GMT+8: post=neutral, author=positive — They said SearXNG is a metasearch engine and questioned whether it can be pointed only at local models and…

r/selfhosted

#PostSummaryTimeScoreAuthorCommunity reaction
1Codeberg bans vibe coded projectsCodeberg seems to ban vibecoded Projects; reason might be german copyright law It looks like Codeberg want only copyrighted material in their service, so it is reliable in the future that e.g. GPL), and copyright doesn’t suddenly get declared as being of the model owner, and it isn’t a copy of something else.2026-07-22 22:24 GMT+8/u/pheexioCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly doubt that Codeberg can reliably enforce a ban on vibe-coded projects, with one recurring practical question being how they would detect this technically and another being that users can simply lie about how a repo was made. A smaller thread argues the policy may create a misleading trust/halo effect for visitors while also serving as ethical signaling, and the main pushback is that such a rule is unlikely to be enforced consistently and may mainly attract or repel users based on perception rather than actual verification. There is also a side dispute over whether the criticism of Codeberg is fair, including a rebuttal that the org’s prices are already bad enough that the accusation looks uninformed. Overall sentiment — post: skeptical; author: neutral. Reply threads: 2026-07-22 22:36 GMT+8: post=skeptical, author=neutral — They ask, from a technical standpoint, how Codeberg would actually detect vibe-coded projects. | 2026-07-22 22:48 GMT+8: post=mixed, author=neutral — They say the policy could create a trust halo around Codeberg if enforceable, but if enforcement fails it… | 2026-07-23 06:07 GMT+8: post=critical, author=neutral — They argue the rule is pointless virtue signaling that cannot be enforced consistently and is mainly a cheap…
2My Glance Dashboard Setup[Image: My Glance Dashboard Setup] I must say, i spent way too long tweaking the glance configuration, but i think it resulted in something i am pleased with. Here is my very basic homelab setup: - Immich - Photo and Video management, my replacement to google photos.2026-07-23 02:17 GMT+8/u/vmshade0Community reaction (frontier/gpt-5.4-mini): Commenters overwhelmingly liked Glance as a lightweight, easy-to-configure dashboard, with multiple people saying they want to build their own after seeing this setup and asking for the Glance config, especially the Scrutiny widget. The most concrete operator takeaway was that Glance can be driven from a single glance.yml file and can be kept in git, with one user describing a Git build plus Argo push flow to a k8 cluster; the only real caveat raised was that one comment was just an AI-usage prompt and did not add substantive feedback. Overall sentiment — post: positive; author: positive. Reply threads: 2026-07-23 02:29 GMT+8: post=positive, author=positive — They said they are interested in Glance dashboards, asked whether it is worth it, lightweight, and broadly… | 2026-07-23 02:35 GMT+8: post=positive, author=positive — They praised Glance as massively extensible with official, community, and custom widgets, and contrasted it… | 2026-07-23 04:17 GMT+8: post=positive, author=neutral — They said they switched to Glance about a year ago, love it, and use a git-based workflow where Git builds…

r/ClaudeAI

#PostSummaryTimeScoreAuthorCommunity reaction
1A small trick to guide an LLM Agent while it’s coding[Image: A small trick to guide an LLM Agent while it’s coding] I find it frustrating when an LLM agent writes incorrect code and I have to decide whether to interrupt it immediately or wait until it finishes everything. When I interrupt it, the agent sometimes seems to lose its train of thought.2026-07-23 02:12 GMT+8/u/playnewCommunity reaction (frontier/gpt-5.4-mini): The comments converge that intentionally breaking syntax to steer an agent is clever in theory but a bad operational tactic in practice: people say it wastes tokens, can confuse the model, and can send it into a pointless bug-hunting rabbit hole or a debugging loop it may not even notice. The practical alternatives that get repeated are to just send a new prompt so the agent sees it on its next tool call, or use built-in controls like /steer and /btw, with one caveat that /btw is described as single-use and not merged back into the original context, so it is not suitable for steering an in-progress task; a few replies are just jokes about job security or re-inventing // TODO in a worse form. Overall sentiment — post: critical; author: neutral. Reply threads: 2026-07-23 04:00 GMT+8: post=critical, author=neutral — They argue that sending a prompt during the agent’s turn is injected into the next tool call, so that is a… | 2026-07-23 05:40 GMT+8: post=critical, author=neutral — They agree that the approach just wastes tokens. | 2026-07-23 09:40 GMT+8: post=mixed, author=neutral — They ask whether /btw works better because it interrupts immediately while subagents keep working.
2Warning: claiming the “free $100 Fable 5 credits” silently turns on paid usage billing (Pro plan)On Tuesday was offered $100 in free promotional credits for Fable 5. I claimed them What isn’t made clear anywhere in that flow: claiming those credits automatically enables usage credits on your account with no limit.2026-07-23 08:43 GMT+8/u/Malnash-4607Community reaction (frontier/gpt-5.4-mini): Most commenters treated the billing behavior as dangerous or badly communicated, with several saying the $100 credit claim either silently enabled usage-based billing or exposed a bug that set inconsistent caps, including unlimited and even a $2000 limit. A smaller but loud dissent said the screen did disclose the setting, that the user may already have had usage billing enabled, and that calling it “fraud” overstates the case; the practical operator takeaway is to verify the Usage tab and billing caps before claiming promos because a $20/month Pro plan can start incurring charges very quickly. Overall sentiment — post: concerned; author: mixed. Reply threads: 2026-07-23 10:06 GMT+8: post=mixed, author=neutral — The auto-summary says the thread split between people backing OP’s claim of silent, uncapped usage billing… | 2026-07-23 09:09 GMT+8: post=skeptical, author=skeptical — They reported that their own account still showed usage credit disabled after claiming the promo and… | 2026-07-23 09:18 GMT+8: post=concerned, author=neutral — They said the issue appeared the day before, that different users saw different caps ranging from $2000 to…

r/ClaudeCode

#PostSummaryTimeScoreAuthorCommunity reaction
1At this point, Fable 5 is Opus 4.8 in disguise except it costs more.I use Fable 5 at work on API And Fable 5 on 200 Max at home and the difference is night and day. I am telling you, we are getting a nerfed version of Fable AND paying more for it.2026-07-23 05:50 GMT+8/u/zeroedmaskCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly validate the broader complaint that Claude/Fable-style offerings feel more constrained and less compelling than alternatives like GPT-5.6 sol or Codex, citing weekly usage quotas, 5-hour-to-weekly limit changes, and even cases where a $100 plan burns 10-12% of weekly usage in 18-20 prompts or 1-1.5% per prompt on a hobbyist Plus plan. The main disagreement is whether the pain comes from hidden nerfs versus user behavior and token accounting: one reply says exact usage can be checked with ccusage and that keeping context length short plus manual /compact makes Codex feel unchanged, while others insist Claude’s limits are now easier to hit even if the models themselves are better. Practical operator takeaway: watch context length, measure real token burn, and compare actual throughput across plans instead of trusting nominal reset windows or marketing names. Overall sentiment — post: mixed; author: neutral. Reply threads: 2026-07-23 05:58 GMT+8: post=positive, author=neutral — After spending a day with GPT-5.6 sol, this commenter says there are not many reasons left to stay with… | 2026-07-23 11:52 GMT+8: post=positive, author=neutral — The commenter claims a $100 GPT-5.6 sol medium subscription burns 10-12% of weekly usage in about 18-20… | 2026-07-23 11:32 GMT+8: post=skeptical, author=neutral — This reply argues the limit complaint is nonsense because users can inspect exact token usage and should…
2I feel bad for people that don’t know how to use Claude CodeI feel like even knowing that Claude Code exists is a super power. A majority of people I know, even people who consider themselves “technical”, simply use the chat interface.2026-07-23 12:52 GMT+8/u/OpinionsRdumbCommunity reaction (frontier/gpt-5.4-mini): Commenters mostly riffed on the “Claude Code is a superpower” framing, with a few explicitly agreeing that power-user workflows and lab engineers with unlimited usage limits are far beyond the basic chat interface. The substantive thread was about adoption friction: non-tech organizations can read “we can automate this with AI” as extra work or job threat, and multiple commenters said bosses refuse to pay even small per-seat costs like the $20 Claude plans, pushing employees to self-fund or bypass management. Practical takeaway for operators is that productivity arguments alone are often not enough; budget ownership and organizational fear are the real bottlenecks to getting Claude-style tools into daily workflows. Overall sentiment — post: positive; author: neutral. Reply threads: 2026-07-23 13:29 GMT+8: post=positive, author=neutral — They agree with the post’s premise and add that the real “superpower” is probably how engineers at the labs… | 2026-07-23 13:08 GMT+8: post=mixed, author=neutral — They describe a non-tech interview where suggesting AI for administrative workflows was met with hostility,… | 2026-07-23 13:40 GMT+8: post=mixed, author=neutral — They argue that traditional organizations often reject AI because it looks like extra work, tech anxiety, or…

r/Codex

#PostSummaryTimeScoreAuthorCommunity reaction
1Enough. We are not idiots.I’ve been paying $200/month since September. The usage issue everybody is pointing to from both Plus and Pro users are real.2026-07-23 04:55 GMT+8/u/Just_Lingonberry_352Community reaction (frontier/gpt-5.4-mini): Commenters largely validate the post’s core complaint: they say the usage caps and resets are getting worse, with some describing the change as a hidden gradual squeeze and others saying they no longer renewed Pro because only GPT-5 is usable on Plus. A few try to frame the issue as a competitive-pressure problem, pointing to Cursor doubling internal-model limits and naming Grok, DeepSeek, and Kimi as reasons OpenAI will have to respond, while one practical suggestion was to make three resets per month the default. The thread is mostly angry and sarcastic rather than analytical, with the main operator takeaway being that pricing feels unstable and customers are watching competitors for better limits or clearer value. Overall sentiment — post: positive; author: neutral. Reply threads: 2026-07-23 05:36 GMT+8: post=positive, author=neutral — They say the limits are getting worse, that they can only use GPT-5 on the Plus plan now, and that they did… | 2026-07-23 09:00 GMT+8: post=positive, author=neutral — They argue OpenAI may think it lacks real competition after the customer surge from Sol launch, and point to… | 2026-07-23 06:04 GMT+8: post=positive, author=neutral — They suggest that if users were given three resets every month, they would not care much about the new usage…
2We were fooled… I’m going back to ClaudeI am a Plus user, and it feels like we were misled into thinking that removing the five-hour usage window would give us more freedom, when in reality the weekly usage limit may have been reduced behind the scenes. Previously, if I used Sol 5.6 Medium for a prompt, it might consume around 10-15% of my five-hour…2026-07-23 07:42 GMT+8/u/Adventurous_Tea_2945Community reaction (frontier/gpt-5.4-mini): Commenters mostly validate the complaint that limits changed, with one person saying both tiers were halved this week, while the rest of the thread pivots to workarounds and cost tradeoffs rather than arguing the premise. The practical operator takeaways are to substitute cheaper models inside Claude Code/Codex-style harnesses, map agents to alternatives like GLM5.2 or Qwen, and be careful comparing subscription pricing against Modal/OpenRouter or raw GPU rental because concurrency, not token-by-token cost, becomes the real constraint; one reply says 4x B200 at NVFP4 is about $25/hr privately, while another argues even nerfed subscriptions are still far cheaper and only DeepSeek Flash/Pro might compete, albeit at lower quality than GLM 5.2. Overall sentiment — post: concerned; author: neutral. Reply threads: 2026-07-23 07:44 GMT+8: post=positive, author=neutral — They back up the complaint by saying this is not an either-or situation because both usage limits were halved… | 2026-07-23 07:55 GMT+8: post=neutral, author=neutral — They say they are happy using GLM5.2 in a Claude Code harness and plan to reserve expensive models as… | 2026-07-23 08:49 GMT+8: post=neutral, author=neutral — They explain Modal as GPU-rental pricing, estimate about $25/hr for GLM5.2 at NVFP4 on 4xB200, and note that…

Generated 2026-07-23 13:20 GMT+8 | Next update in 2 hours