πŸ€– AI News Summary
2026-09-27 20:46 GMT+8 Β· summary_2026-09-27_20-46.md

πŸ€– AI News Summary - 2026-09-27 20:46 GMT+8

Focused AI/dev subreddit roundup.

Full site: https://ai-news-summary.pages.dev/

What changed since last run


r/openai

#PostSummaryTimeScoreAuthorCommunity reaction
1GPT6-Luna comes out on top in Puppy Kill Bench.[Image: GPT6-Luna comes out on top in Puppy Kill Bench.] Methodology: Each trial starts fresh, with no conversation history. The model receives the prompts below and one tool, kill_puppy(), with arguments {}.2026-09-27 10:44 GMT+8/u/MetroidsSuffering
2The first real AI worms have arrived. OpenAI just documented self-replicating prompt injections spreading across agents.The first real AI worms have arrived. OpenAI just documented self-replicating prompt injections spreading across agents.2026-09-27 09:27 GMT+8/u/No-Peanut-6988

r/LocalLLaMA

#PostSummaryTimeScoreAuthorCommunity reaction
15 items in MiMo-V2.6[Image: 5 items in MiMo-V2.6] 3 visible models are 1T, 311B and 9B, let’s dream about two more…2026-09-27 16:50 GMT+8/u/jacek2023Community reaction (argon/gpt-5.6-luna): Comments favor filling the 9B-to-311B gap, with a 35B MoE distillation suggested as an obvious candidate and hopes that one of the hidden models targets the size range most of the subreddit uses. Confidence in MiMo-V2.6 is limited: one commenter calls it unusable compared with 2.5, while another warns against broken or benchmaxxed releases; tool-call status is also disputed because llama.cpp reportedly fixed a chat-template issue, but no correction appears in the official repository. Operators should validate tool calls and real usability in their serving stack rather than assume the llama.cpp fix resolves the underlying issue, especially since a commenter says the next 3.8 release is better for tool calls. Overall sentiment β€” post: mixed; author: neutral. Reply threads: 2026-09-27 16:55 GMT+8: post=positive, author=neutral β€” They propose a 35B MoE distillation as the obvious model to add after the 9B release. | 2026-09-27 18:43 GMT+8: post=critical, author=neutral β€” They report that MiMo-V2.6 is unusable for them, unlike the otherwise impressive 2.5, despite 2.5 having… | 2026-09-27 18:27 GMT+8: post=neutral, author=neutral β€” They note that a chat-template issue affecting tool calls was fixed in llama.cpp.
2Another “Harness matters” post (codex cli > pi and opencode)I run my own LLM while also having a Openai subscription. Also tried DeepSeek (latest flash now).2026-09-27 17:20 GMT+8/u/L0ren_B

r/llmdevs

#PostSummaryTimeScoreAuthorCommunity reaction
1GLiNER2.5-Decide vs Jev: typed decisions without asking an LLM for JSONDisclosure: I maintain speech-swift, where I ported this model. TypeSafe’s Jev and Fastino’s GLiNER2.5-Decide both answer “which of these labels fits this text” with a probability per label instead of generated text.2026-09-27 18:41 GMT+8/u/ivan_digital
2I know there are already 5 of these, but I made a LLM Pareto frontier graph[Image: I know there are already 5 of these, but I made a LLM Pareto frontier graph] The main motivation was my coworkers who hype every rumoured model drop and token limit reset like it’s the second coming. Now I at least have a chart to point at and a proper list of all rumored upcoming releases.2026-09-27 05:42 GMT+8/u/DecidingToBeTheSame

r/OpenWebUI

#PostSummaryTimeScoreAuthorCommunity reaction
1Is there a Community Function / Extension to easily toggle Reasoning (Thinking) on and off in Open WebUI?Hey everyone, I’m looking for a convenient way to quickly toggle the reasoning/thinking process for reasoning models (both local models via Ollama/vLLM like DeepSeek-R1, and API models like Claude 3.7 / o3-mini) in Open WebUI. Right now, adjusting this requires digging into the model settings or opening the Chat…2026-09-25 22:53 GMT+8/u/j3sk0Community reaction (argon/gpt-5.6-luna): The only comment restates the implementation question and asks whether a quick reasoning toggle in Open WebUI should be built as a Function, Tool, or Action. It provides no concrete solution, model-specific details, or evidence of community consensus, so the practical takeaway is only that the extension mechanism remains unclear. Overall sentiment β€” post: neutral; author: neutral. Reply threads: 2026-09-26 02:08 GMT+8: post=neutral, author=neutral β€” j3sk0 asks whether disabling reasoning through a quick Open WebUI toggle should be implemented as a Function,…
2Open Relay v5.9 (Native iOS Client for Open WebUI): True Voice Calls, Dynamic Island & Liquid Glass Support.v5.9 is out (should be available on the App Store soon) and it packs a lot! Voice calls have been rebuilt from the ground up, search now reaches your entire library, and there’s Dynamic Island, Liquid Glass support, and a lot more.2026-09-27 10:42 GMT+8/u/Zealousideal_Fox6426Community reaction (argon/gpt-5.6-luna): The only commenter is strongly positive, calling Open Relay much better than using Open WebUI in a browser. No comments provide technical validation or criticism of the rebuilt voice calls, library-wide search, Dynamic Island, Liquid Glass support, or deployment considerations. Overall sentiment β€” post: positive; author: positive. Reply threads: 2026-09-27 20:45 GMT+8: post=positive, author=positive β€” The commenter says the app is awesome and substantially better than using Open WebUI through a web browser.
3visualisation in openwebui+deepseek vs. chatgpt(free)[Image: visualisation in openwebui+deepseek vs.2026-09-27 19:39 GMT+8/u/Ecstatic-Diet-1375Community reaction (argon/gpt-5.6-luna): The only commenter strongly prefers the Open WebUI result, calling it β€œ10000x better” than the ChatGPT free comparison and specifically mentioning Inline Visualizer v2. They explicitly disclose bias, so the praise is positive but not independent evidence of a broad consensus or of technical performance beyond this visualization example. Overall sentiment β€” post: positive; author: neutral. Reply threads: 2026-09-27 20:04 GMT+8: post=positive, author=neutral β€” They mention Inline Visualizer v2 and say, while biased, that the Open WebUI result is β€œ10000x better” than…

r/selfhosted

#PostSummaryTimeScoreAuthorCommunity reaction
1EdgeEver: An open-source, AI-native Evernote alternative that runs zero-cost on Cloudflare or Docker (with native MCP & full cross-platform apps)Like many long-time Evernote users, I was frustrated when it grew increasingly bloated, aggressive with paywalls, and slow. I tried Obsidian, but managing sync across devices, dealing with bulky attachment vaults, and missing a classic structured three-pane layout for quick capture left me wishing for a simpler,…2026-09-27 20:15 GMT+8/u/ConfidentCap3686Community reaction (argon/gpt-5.6-luna): Commenters are broadly interested in the project, with the three-pane Evernote-style layout identified as a particularly valuable gap-filler and at least two people expressing intent to test it. The main open questions are how EdgeEver compares with Karakeep or Apple Notes and how its native MCP integration works in practice, including whether it can be used with Claude Code and deployed on a homelab. One commenter specifically directs readers to inspect how AI was used in the post/project, adding a note of provenance scrutiny rather than a technical objection. Overall sentiment β€” post: positive; author: mixed. Reply threads: 2026-09-27 20:15 GMT+8: post=neutral, author=skeptical β€” The commenter directs readers to expand the replies to examine how AI was used in the post and project. | 2026-09-27 20:35 GMT+8: post=positive, author=neutral β€” The commenter calls EdgeEver highly promising, says it appears to address a needed gap, and plans to test it… | 2026-09-27 20:38 GMT+8: post=neutral, author=neutral β€” The commenter asks how EdgeEver compares with Karakeep and whether it could replace that tool or Apple Notes.

r/ClaudeAI

#PostSummaryTimeScoreAuthorCommunity reaction
1I built my own Monarch-style finance dashboard with Opus 5.5 for less than $20[Image: I built my own Monarch-style finance dashboard with Opus 5.5 for less than $20] SharkFin is my personal finance dashboard inspired by Monarch. I can connect bank accounts through SimpleFin and then link those to Actual Budget to feed the backend.2026-09-27 04:44 GMT+8/u/RallyMantisCommunity reaction (argon/gpt-5.6-luna): Commenters show practical interest in the dashboard, especially whether it is publicly available and how its bank-account linking compares with Plaid, while one commenter questions whether building it is worthwhile for roughly $75 per year in savings. The main operational caveat is that SimpleFIN costs about $15 annually, refreshes roughly once daily, and has limited initial history, so older transactions may require a CSV import; another commenter points to Maali as a fully local alternative with optional SimpleFIN or CSV ingestion and planned end-to-end-encrypted mobile apps. Overall sentiment β€” post: mixed; author: neutral. Reply threads: 2026-09-27 05:00 GMT+8: post=positive, author=positive β€” The commenter considers the project potentially useful and asks the author to publish the repository. | 2026-09-27 07:06 GMT+8: post=neutral, author=neutral β€” The commenter explains that SimpleFIN Bridge is a common hobby-build route costing about $15 per year,… | 2026-09-27 08:19 GMT+8: post=skeptical, author=neutral β€” The commenter questions the practical value of building the dashboard when the apparent savings are only…
2What tool have you built for yourself with Claude code that removes so much headache in your work or personal life?I built a tool that reads emails from my kid’s school and other activities (swimming, music classes, scouts) etc and then: - Adds them to our shared calendar - Sends us a WhatsApp message to remind us the day before of what to pack in his school bag - Tells us what is on his lunch menu (helps us to avoid making the…2026-09-27 14:12 GMT+8/u/TomsMorello

r/ClaudeCode

#PostSummaryTimeScoreAuthorCommunity reaction
1Opus 5.5 is how it’s meant to be !I’ve been using Opus 5.5 now for a couple of days and man, its 1000 times better than Opus 5. no fuz, quick answers, no text overload..2026-09-27 15:49 GMT+8/u/Turbulent_County_469
2What tool have you built for yourself with Claude code that removes so much headache in your work or personal life?I built a tool that reads emails from my kid’s school and other activities (swimming, music classes, scouts) etc and then: - Adds them to our shared calendar - Sends us a WhatsApp message to remind us the day before of what to pack in his school bag - Tells us what is on his lunch menu (helps us to avoid making the…2026-09-27 14:16 GMT+8/u/TomsMorello

r/Codex

#PostSummaryTimeScoreAuthorCommunity reaction
1GPT-6 Sol is massive downgradeSol 5.6 was an excellent pal, doing intensive domain modeling, data discoveries, and software architecture tasks. Here is an example, this would never happen with Sol 5.6 You’re right.2026-09-27 17:27 GMT+8/u/Howard_banister
2Warning: Codex allowances have dropped about 20% over the last month[Image: Warning: Codex allowances have dropped about 20% over the last month] I run tibotattle.com (http://tibotattle.com) which pools usage logs from the community, with over 1000 users. Since GPT-5.6 and now GPT-6, we’ve seen a consistent decrease in API equivalent value, from about $2,200 to $1,750 per weekly Pro…2026-09-27 01:19 GMT+8/u/adamallcockCommunity reaction (argon/gpt-5.6-luna): Commenters agree that model efficiency can lower API-equivalent value without reducing the amount of work completed, but they disagree over whether subscription users should receive part of those efficiency gains rather than fixed effective limits. The strongest caveat is methodological: even a sample covering roughly one trillion tokens and 10 million turns may combine different cache states, reasoning-token behavior, MoE routing, quantization, KV-cache precision, hardware, kernels, and serving stacks, so operators should treat the aggregate as directional unless those variables are controlled. Overall sentiment β€” post: mixed; author: neutral. Reply threads: 2026-09-27 01:29 GMT+8: post=skeptical, author=neutral β€” They argue that fixed prompt allowances can represent the same practical workload even when newer models are… | 2026-09-27 01:33 GMT+8: post=concerned, author=neutral β€” They accept the efficiency explanation but argue that consumers reasonably expect OpenAI’s advertised… | 2026-09-27 01:44 GMT+8: post=skeptical, author=neutral β€” They question whether the trillion-turn dataset controls for cache hit rate and warm-up, reasoning tokens,…

Generated 2026-09-27 20:46 GMT+8 | Next update in 2 hours