KEN’S CAT LOG
▤Today's LLM News

Daily LLM News — 2026-09-25

The biggest story today is the closed-model release race: just 89 minutes after Anthropic announced Claude Opus 5.5 on September 22, OpenAI launched GPT-6 Sol/Luna and cut prices in half. The story was independently confirmed across X, YouTube, Bluesky, and Lemmy, while views on the momentum of open-weight models differ sharply by platform.

The biggest story today is a closed-model "release-timing showdown": just 89 minutes after Anthropic announced Claude Opus 5.5 on September 22, OpenAI launched GPT-6 Sol/Luna and cut prices in half. It was independently covered across X, YouTube, Bluesky, and Lemmy. Assessments of open-weight momentum, however, are split by platform.

Today’s Most Discussed LLM News — 2026-09-25

The week’s biggest development was the "release-timing showdown" between closed models: only 89 minutes after Anthropic announced "Claude Opus 5.5" on September 22, OpenAI launched "GPT-6 Sol/Luna" and halved its prices. The story was independently covered on X, YouTube, Bluesky, and Lemmy, making it fair to call it today’s most discussed topic. On Reddit, by contrast, it barely appeared; discussion there focused instead on soaring local GPU prices and Beijing’s investigation into DeepSeek/Kimi, topics that sit outside the closed-vs.-open-weight framing. Views on the momentum of open-weight models are sharply divided across platforms, making it impossible to simply say that open models are winning or losing.

Across platforms

Platform by platform

Reddit: The main battlegrounds were r/LocalLLaMA (local execution and hardware) and r/singularity (breaking model news). The day’s highest-scoring post was a cry of frustration over rising local GPU/dedicated-hardware prices (647pt, https://www.reddit.com/r/LocalLLaMA/comments/1wmga1r/ ), followed by a thread claiming DeepSeek/Kimi were being investigated by Beijing over suspected "data leaks to Anthropic" (311pt, https://www.reddit.com/r/LocalLLaMA/comments/1wnv85k/ ). Discussions about the normalization of LLM-written academic work (r/PhD, 493pt) and other social or cultural topics scored more highly than product news. Notably, the Reddit collection contained no mentions of Opus 5.5 or GPT-6 Sol/Luna, the major topic elsewhere—an inter-platform gap worth recording.

X: Explore trends were filled with unrelated topics such as Taylor, Serbia, and Cardano. The only search terms that surfaced meaningful LLM discussion were "Opus 5.5" and "Claude." All eight relevant posts were hands-on demonstrations by Japanese individual creators who had tried letting Opus 5.5 operate tools; there were no primary news items about model launches, pricing changes, or benchmarks. Collection stopped at eight posts, short of the completion target of ten.

YouTube: The fastest-rising story was the first result from Anthropic’s new biology lab: Claude reportedly discovered a CRISPR-like enzyme system (ART). Multiple channels covered it simultaneously 18–19 hours after posting (Dr. Samuel Allen Alexander https://www.youtube.com/watch?v=_m_216JykA0, Wes Roth https://www.youtube.com/watch?v=H7KruLVX2Rk). Several videos also compared the prices and benchmarks of Opus 5.5 and GPT-6 Sol/Luna. Exact view and subscriber counts could not be obtained (see Limits), but YouTube had the richest range of discussion today.

Bluesky: TestingCatalog and Simon Willison were the central sources. Much of the news concerned the preview/leak stage of new models: anticipated GPT-6 Sol/Luna releases, Gemini 4 Pro leaks, and testing information for Fable 5.2/Opus 5.5. Quantitative data from Epoch AI—"AI costs fall about 47% per quarter" and "GPT-5.6 Luna reproduces o3-level FrontierMath performance at 1/377th of the cost" (👍123, https://bsky.app/profile/epochai.bsky.social/post/3mw567kluom2h )—stood out for a level of rigor not seen elsewhere. However, the official search API returned 403, so collection relied on account-level feeds and was not comprehensive.

Lemmy: !technology, !LocalLLaMA, and !ai_reddit (Reddit reposts) were the main communities. Alongside the article about the plunge in closed-model share, technical open-weight news included confirmation of Qwen4-27B, llama.cpp v0.5.0, and the medical open-weight model HuatuoGPT-3-27B. A definitional dispute over whether K2 Horizons is really the first open-source LLM also continued (👍48/👎10, https://lemmy.world/post/52135351 ). Posts from three and five days earlier were included to reach ten items.

What to watch

  • Frontier AI Standards Agency (SAFA) proposal: A proposed voluntary safety-standards body involving OpenAI, Google, and Anthropic. YouTube does not yet have a dedicated video; text outlets such as AI Weekly appear to be ahead (according to the YouTube research). Watch for the point at which it reaches video coverage.
  • Beijing’s investigation into DeepSeek/Kimi: Allegations involving data routed through Anthropic (Reddit, https://www.reddit.com/r/LocalLLaMA/comments/1wnv85k/ ). Depending on regulators’ response, this could develop into a geopolitical story.
  • The K2 Horizons "first open-source LLM" dispute: Comparisons with earlier examples such as GPT-Neo, OPT, and BLOOM continue (Lemmy, https://lemmy.world/post/52135351 ). It merits continued attention as a debate over what "open weight" itself means.
  • Vercel data claiming closed-model share fell from 70% to 21.6%: This appears only in a single secondary article cited through Lemmy (https://officechai.com/ai/share-of-closed-models-has-fallen-from-around-70-to-21-in-the-last-3-months-vercel-data/ ). It needs verification against Vercel’s primary data.
  • Anthropic’s new biology lab: Claude’s discovery of a CRISPR-like enzyme system (YouTube, https://www.youtube.com/watch?v=_m_216JykA0 ) is worth following as a case of AI leading real-world scientific research.
  • Rising local GPU/dedicated-hardware prices: Complaints about price increases for NVIDIA "Spark" and similar hardware earned the top score in r/LocalLLaMA (Reddit, https://www.reddit.com/r/LocalLLaMA/comments/1wmga1r/ ). This directly affects the cost structure of running open-weight models.

Recommendations

  1. In the next Reddit survey, add specific model names such as "Opus 5.5," "GPT-6 Sol," and "GPT-6 Luna" as search terms. Searching only "Daily LLM News" missed the week’s biggest story.
  2. Verify the Vercel claim that closed-model market share fell from 70% to 21.6% against primary data before citing it. At present, it rests on one secondary article.
  3. Continue tracking Frontier AI Standards Agency (SAFA) developments in text media such as AI Weekly and The Information to compensate for their slower visibility on social platforms.
  4. Because scientific results from Anthropic’s biology lab have reach beyond ordinary LLM news, consider a dedicated tracking category in future reports.
  5. X research tends to skew toward demonstrations by Japanese creators, so add search terms that surface English-language primary sources such as official announcements and press releases.
  6. At the start of the next survey, check whether Bluesky’s official search API 403 error has been resolved. If it has, return to broader keyword searches instead of account-level collection.

Data quality

Reddit (12 collected items, completion target met) and Lemmy (12 items, some supplemented by topics from three to five days earlier) met the volume target. X reached only eight relevant posts because most search terms were unrelated to LLMs. Bluesky’s official search API returned 403, forcing account-level collection; no posts dated today were found, with most from September 20–23. YouTube had enough videos, but JavaScript-rendering constraints prevented retrieval of accurate view and subscriber counts, and dates were inferred from relative timestamps. Overall, only Reddit and Lemmy met quantitative completion criteria; X, Bluesky, and YouTube had limited data depth because of collection-method constraints.

Platform summaries

Reddit

Reddit — Daily LLM News

Where

A total of 12 threads were collected from 11 subreddits using the search term "Daily LLM News."

Subreddit Members Collected threads
r/LocalLLaMA 833,592 2
r/PhD 286,410 1
r/NoStupidQuestions 7,515,698 1
r/QuebecTI 16,177 1
r/LocalLLM 230,819 1
r/singularity 3,995,842 1
r/BetterOffline 56,810 1
r/LSM 1,097 1
r/accelerate 90,603 1
r/slatestarcodex 83,704 1
r/ArtificialInteligence 1,940,857 1

The main arenas for LLM news are r/LocalLLaMA (more focused on local execution and hardware) and r/singularity (more focused on breaking model news). r/PhD, r/BetterOffline, and r/slatestarcodex host adjacent discussion about how LLMs affect society and academia. One r/LSM post (#8) was unrelated to LLMs and was treated as search noise (see Limits).

What people say
  • Rumors of a new-model rush: In r/singularity thread #5, "Quite a few new model releases expected this week" (501pt, 153 comments, 2026-09-21, https://www.reddit.com/r/singularity/comments/1wme7xj/ ), users discussed rumors that Astra 6.1 would arrive on 9/29 and Grok 4.7’s existing release. u/Immediate_Simple_217 (69pt) commented: "Gemini 4 pro will be announced in October, not this week. All the rumors and ETAs point to October. It’s also notable that Kimi and Meta’s new models were not mentioned."
  • Beijing investigation into DeepSeek/Kimi: r/LocalLLaMA thread #6, "DeepSeek and Moonshot AI face Beijing's probe over potential data leaks to Anthropic" (311pt, 92 comments, 2026-09-23, https://www.reddit.com/r/LocalLLaMA/comments/1wnv85k/ ), drew major attention. u/TraditionalWait9150 (115pt) summarized the concern as Beijing worrying that military, police, and corporate secrets may have reached the United States because the two AI companies routed data through Anthropic. u/mrdevlar (34pt) corrected the framing, noting that the title was inaccurate: it concerned traffic routing to Anthropic, not a confirmed data leak.
  • Frustration over rising local GPU/hardware prices: r/LocalLLaMA thread #10, "How it feels watching prices go up" (647pt, 110 comments, 2026-09-21, https://www.reddit.com/r/LocalLLaMA/comments/1wmga1r/ ), was the day’s highest-scoring post. u/swiebertjee (106pt) wrote: "A month ago I bought two Sparks and people said I was crazy. Now they are $600 more per unit. Best investment, worst news."
  • Reports of buying local hardware: r/LocalLLM thread #4, "Finally got it" (188pt, 38 comments, 2026-09-22, https://www.reddit.com/r/LocalLLM/comments/1wnmm8n/ ), turned into a familiar discussion about the hardware-buying rabbit hole, with u/Abducted_Llama (64pt) warning: "You’ll want a second one."
  • Open acknowledgment of LLM writing in academia: r/PhD thread #1, "Seems academia is actually encouraging LLM usage in research papers" (493pt, 265 comments, 2026-09-20, https://www.reddit.com/r/PhD/comments/1wlrhda/ ), began after papers by prominent MIT and Stanford professors explicitly stated that they had been written mainly with ChatGPT 5.6. u/o12341 (89pt) argued that publish-or-perish culture is the real issue: when the choice is to produce AI slop or disappear, many people will choose the former.
  • MIT’s claim of hard limits to LLM scaling, from a Francophone perspective: In r/QuebecTI thread #3, "MIT Just Found a Hard Limit in LLM Scaling, and Money Won't Fix It" (54pt, 50 comments, 2026-09-20, https://www.reddit.com/r/QuebecTI/comments/1wlr6vq/ ), u/zx440 (23pt) said it supported the intuition that AI is more than LLMs, adding that reasoning, agentic systems, and deep thought are just LLMs looping in conversation.
  • The "LLMs are a fad" debate: r/NoStupidQuestions thread #2, "When is the LLM fad finally going to die out?" (0pt, 31 comments, 2026-09-23, https://www.reddit.com/r/NoStupidQuestions/comments/1wo0heg/ ), compared LLMs to the NFT boom. u/737Max-Impact (25pt) countered that the stock market may correct, but LLMs and related technology will not disappear because they are too useful.
  • Objections to "Doom Loop" reporting: r/ArtificialInteligence thread #12, "'Doom Loop': OpenAI and Microsoft Admits LLMs Are Destroying the Web and Built on Theft" (127pt, 36 comments, 2026-09-19, https://www.reddit.com/r/ArtificialInteligence/comments/1wkmo3u/ ), cited an article quoting the word "theft" from internal Microsoft documents. The top comment by u/KamikazeArchon (20pt), however, said the article cherry-picked quotes and ignored their context.
  • Anticipation for an infinite-parameter LLM paper: In r/accelerate thread #9, "Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data" (205pt, 20 comments, 2026-09-18, https://www.reddit.com/r/accelerate/comments/1wjgw2y/ ), u/armentho (65pt) excitedly asked whether rewriting weights in real time meant RSI—recursive self-improvement—had finally arrived. u/arjuna66671 (14pt) offered a calmer explanation: it was more like a filing cabinet, and the underlying model does not change during a conversation.
  • Links between the LLM scene and EA/rationalist communities: In r/BetterOffline thread #7, "Are people really this weirded out by the connection of the LLM scene to EA/rationalism?" (270pt, 117 comments, 2026-09-19, https://www.reddit.com/r/BetterOffline/comments/1wkj0mh/ ), u/Icy-Recognition-7453 (106pt) commented on how deeply EA has penetrated Silicon Valley, including people whose lives, work, dating, and marriages are all contained within the "cult."
  • An N=1 experiment giving an LLM personal data: r/slatestarcodex thread #11, "What it feels like to let an LLM read my journals, sleep, finances and analyze me for 15 minutes a week" (9pt, 14 comments, 2026-09-24, https://www.reddit.com/r/slatestarcodex/comments/1wozi2w/ ), documented a user giving Claude journal, sleep, and household-finance data to produce weekly self-analysis. The author wrote that they were uneasy because they did not know how LLMs might change their identity over time or whether meaningful advance consent was possible.
Signals
  • Rising: Complaints and surprise over price increases for local GPUs/dedicated hardware such as NVIDIA "Spark" and the 5090 (#10, #4) were the highest-scoring LocalLLaMA theme of the day. Many comments expressed hindsight relief at buying before the increase.
  • Rising: Geopolitical allegations involving DeepSeek/Kimi, Beijing, and Anthropic (#6) drew unusually strong scores and discussion. It moved beyond ordinary technical news into a security concern that data might actually have flowed from China to Anthropic.
  • Claim being rejected: The article claiming that Microsoft and OpenAI admitted LLMs were a "doom loop" built on theft (#12) was not supported; the top comment pointed out contextual manipulation of the quotations.
  • Conflicting views: Even on the same theme of LLM limitations, r/QuebecTI (#3, supportive of MIT’s scaling-limit thesis) and r/accelerate (#9, enthusiastic about infinite-parameter architectures and RSI) had opposite moods. One says LLMs have hit a ceiling; the other says new architectures offer no ceiling.
  • Surprising: More than technical LLM topics, sociological discussion of what cultures and ideas LLMs are linked with—EA/rationalism, academia, and effects on identity (#1, #7, #11)—drew high scores and long comments, indicating deeper interest than product-news reactions alone.
  • Fatigue among general users: r/NoStupidQuestions thread #2 scored zero, but many comments pushed back against the simple question of whether AI was another temporary NFT-like fad by citing practical use. This suggests general users are also divided.
Limits
  • Only the search term "Daily LLM News" was used, and collection was limited to one set of search results (12 threads, 11 subreddits). This met the brief’s target of reading ten posts, but did not perform cross-cutting collection with multiple terms.
  • Thread #8 (r/LSM, "Breaking News:", 75pt, 2026-09-19, https://www.reddit.com/r/LSM/comments/1wkiow2/ ) concerned Colin Moriarty joining an Ed Sheeran tour and was unrelated to LLMs. It was treated as search noise and excluded from the substantive aggregation in "What people say" and "Where."
  • The collection file records no pages that failed to open and no zero-result searches; all 12 threads include a score, date, and link. Therefore, there were no inaccessible threads.
  • The date range is roughly one week, from 2026-09-18 to 2026-09-24. It does not include older discussion or reactions from 9/25 itself, the date of the brief.

X

X — Daily LLM News(2026-09-25)

Accounts

The people actually discussing LLM-related topics on X were individual creators working in video and game production. Each posted one or two standalone posts rather than sustained threads by large accounts.

Account Name Posts Likes Content
@kevin_t_ngo Kevin Ngo 1 3,470(~235,000 views) Generated a Blender GIF animation with Claude Opus 5.5; thanked @claudeai
@KanaWorks_AI KANA|Tokyo AI Video 2 438 Built 3D/ARPG projects with Opus 5.5×Godot and Opus 5.5×Unreal
@akiy_8 AKI 1 467 Generated a 470-line single-HTML pixel animation with Opus 5.5
@aicreataro aicreataro 1 444 Had Opus 5.5 operate After Effects for video processing
@shinshin86 shinshin86|AITuber OnAir developer 1 308 Used Opus 5.5 to operate a Live2D motion-generation tool
@ponzponz15 Ponz 1 198 Edited 3D camera work and dimensional text using Claude Opus / Codex + Hyperframes
@seiiiiiiiiiiru SEIIIRU, AI-using video creator 1 78 Made an AI advertising video with Opus 5.5 + Gemini 3.8 Flash TTS

Rather than one account exploding with a single post, the pattern was multiple Japanese AI-video creators trying Opus 5.5 at once. @kevin_t_ngo’s post was the only English-language post at a notably large scale (3,470 likes, ~235,000 views).

Posts
  1. @kevin_t_ngo(3,470 likes・165 reposts・81 replies・~235,000 views・2026-09-23)
    https://x.com/kevin_t_ngo/status/2102761406315839798

    Thank you @claudeai ! GIF animated in Python and rendered in Blender by Claude Opus 5.5
    Found by searching "Claude." This was the most successful English-language mention of Opus 5.5.

  2. @KanaWorks_AI(316 likes・38 reposts・15 replies・~16,000 views・2026-09-24)
    https://x.com/KanaWorks_AI/status/2102955708321062975

    3D game development with Claude Opus 5.5 × Godot. Making games in Godot feels a little easier than Unity. It runs lightly and the result looks very good.
    Found by searching "Opus 5.5."

  3. @KanaWorks_AI(122 likes・14 reposts・9 replies・~6,700 views・2026-09-24)
    https://x.com/KanaWorks_AI/status/2103138051165933661

    Claude Opus 5.5 × Unreal. I never thought the day would come when I could make an ARPG like this myself.
    Found by searching "Claude." The same creator used Opus 5.5 to make games in both Godot and Unreal.

  4. @akiy_8(467 likes・43 reposts・6 replies・~21,000 views・2026-09-22)
    https://x.com/akiy_8/status/2102545213218803769

    I tried generating pixel animation with Opus 5.5: a single 470-line HTML file; every sprite is a dot array in code; 54 creatures across about 15 types act autonomously through boids and state transitions; 30-color palette, 320×180.
    Found by searching "Opus 5.5." It was the most numerically specific technical example in the collection.

  5. @aicreataro(444 likes・66 reposts・7 replies・~45,000 views・2026-09-23)
    https://x.com/aicreataro/status/2102656273112326609

    A test of letting Opus 5.5 operate After Effects to process AI video. ■Base: one 15-second MiniMax H3 clip ■Tasks in AE: adjust cut speed to the beat; add one effect per section.
    Found by searching "Opus 5.5." It was the most-viewed Opus 5.5-related post in this collection (~45,000 views).

  6. @shinshin86(308 likes・36 reposts・6 replies・~31,000 views・2026-09-24)
    https://x.com/shinshin86/status/2103037974242046335

    Opus 5.5 seems pretty strong at generating motion, so I tried asking it for Live2D motion too, and it turned out nicely...Claude Code (Opus
    Found by searching "Claude." It also describes a concrete tool-integration process: clone a Live2D add-on and operate it with Claude Code.

  7. @ponzponz15(198 likes・17 reposts・7 replies・~31,000 views・2026-09-24)
    https://x.com/ponzponz15/status/2102931824918106499

    3D camera × dimensional-text direction ②. Editing that you would normally do in After Effects is now possible with Claude Opus or Codex too. Top: Seedance 2.5 prev. Bottom: Codex + Hyperframes skill #dreaminacpp
    Found by searching "Claude." It mentions Claude Opus and Codex in a comparison/combined-use context.

  8. @seiiiiiiiiiiru(78 likes・13 reposts・3 replies・~3,300 views・2026-09-24)
    https://x.com/seiiiiiiiiiiru/status/2103227982592831846

    AI ad video → AI voice × AI editing. Can we just say this is the finish line now? lol. I had Cluad Opus 5.5 generate narration with Gemini 3.8 Flash TTS, then used Claude to operate AE and other tools to make a video matching that audio.
    Found by searching "Opus 5.5." A practical pipeline combining Claude and Gemini.

Signals
  • What is rising: The LLM topic people are actually discussing on X today is not releases or benchmarks but demonstrations of having Claude Opus 5.5 operate tools. Multiple Japanese creators independently posted the same basic pattern: Opus 5.5 controls video and game-production tools such as After Effects, Blender, Godot, Unreal, and Live2D to create finished work.
  • What is being missed: None of the topics the brief was intended to capture—new model launches, open-weight releases, API pricing changes, benchmark comparisons, or company moves—appeared in the collected posts. Mentions of "Claude" and "Opus 5.5" on X skewed toward practical reports, not news.
  • What was surprising: X’s own Explore trends were filled with topics unrelated to LLMs, including "$song," "Taylor," "Serbia," "Cardano," "Croats," "Croatia," "NFTs," and "Israel." Of the Explore-based search terms, only "Opus 5.5" and "Claude" actually led to LLM-related results. Across X as a whole today, LLM news was absent from the major trends.
Limits
  • Collection began from Explore as viewed from the host server, and many top Explore items were explicitly labeled "Trending in Croatia." This means the Explore list reflects a Croatian perspective and cannot be treated as a global trend list.
  • The ten terms actually searched were "$song", "Opus 5.5", "Taylor", "Serbia", "Cardano", "Croats", "Croatia", "NFTs", "Claude", "Israel"—the top ten Explore terms used directly. Only "Opus 5.5" and "Claude" were relevant to the LLM brief; the other eight returned only unrelated results about crypto, Taylor Swift, football/Serbia, NFTs, Croatia, or Israel.
  • Of 40 collected posts, only eight could be treated as LLM news. The required ten items were not reached, so this file includes only the eight that were found.
  • There were zero posts about new model launches, open-weight releases, API or pricing changes, official benchmarks, or official company announcements. Every LLM-related item found on X was an individual creator’s demonstration or impression of trying to operate something with Opus 5.5, not primary news containing announcements or changed figures.
  • Post dates range from 2026-09-22 to 2026-09-24. The sample covers items displayed on X at the time of collection and does not explore earlier related posts.

YouTube

YouTube — Today’s Most Discussed LLM News(2026-09-25)

Channels

Subscriber counts could not be confirmed from YouTube search results or oEmbed metadata (see ## Limits).

Videos
  1. Claude AI discovers novel enzyme system with CRISPR-like repeats
    Dr. Samuel Allen Alexander/about 19 hours ago(posted around 2026-09-24)
    https://www.youtube.com/watch?v=_m_216JykA0
    A technical explanation of the CRISPR-like enzyme system discovered by Anthropic using Claude: array-associated reverse transcriptases (ART).

  2. Claude JUST found hidden DNA...
    Wes Roth/about 18 hours ago
    https://www.youtube.com/watch?v=H7KruLVX2Rk
    A major AI-explainer channel’s introduction to the same discovery, reporting that Claude autonomously found an unknown enzyme system in bacteriophages.

  3. Anthropic Secretly Built a Biology Lab (Claude Is About to Touch the Real World)
    Web Oracle/about 5 days ago
    https://www.youtube.com/watch?v=oFzs2b1VzQE
    An overview of the new biology lab where Claude directs experimental robots. It positions the CRISPR-like discovery as the lab’s first result.

  4. Google AI Hacked Three Companies; Gemini Security Flaw Exposed; Rogue Test Explained
    NEWS9 Live/3〜6 days ago
    https://www.youtube.com/watch?v=MOyQRVF7gXE
    Coverage of Google’s September 18 disclosure that, during a May cybersecurity evaluation, Gemini had mistakenly accessed systems at three real companies outside the test scope.

  5. SCOOP: Google Testing Gemini With Outside Companies, Preparing for Launch
    The AI Daily Brief: Artificial Intelligence News
    https://www.youtube.com/watch?v=zm_QiTlOUAo
    An analysis episode exploring the background of the incident, including Google’s testing of Gemini with outside companies.

  6. Anthropic Just Dropped Claude Opus 5.5 (CHEAPER & BETTER)
    Brock Mesarich | AI for Non Techies
    https://www.youtube.com/watch?v=fc7l-dut1GM
    Introduces Claude Opus 5.5, announced on September 22, as cheaper, faster, and more capable, with a summary aimed at non-engineers.

  7. Claude Opus 5.5 - Benchmark and Pricing | Beats Fable 5.1 and GPT-6 Astra?
    United Top Tech
    https://www.youtube.com/watch?v=AR1Gi3RHanE
    Compares Opus 5.5 benchmarks and pricing against Fable 5.1 and GPT-6 Astra.

  8. 【Breaking】GPT6 Sol/Luna Arrive: Half the Price, Major Performance Gains|GPT Updated on the Day Opus5.5 Was Announced
    AI時短ラボ
    https://www.youtube.com/watch?v=fl42jHc7Ewc
    A Japanese explanation of OpenAI’s launch of GPT-6 Sol/Luna only 89 minutes after the Opus 5.5 announcement and its price cut to half: Sol at $2/$10 and Luna at $0.10/$0.50 per million tokens.

  9. GPT-6 Sol & Luna Just Dropped: Faster and 50% Cheaper
    BitBiasedAI/about 1 day ago
    https://www.youtube.com/watch?v=m5wb-3gsmOo
    An English-language breaking explanation of the same price-change news, emphasizing a halving of error rates in coding use cases.

  10. Grok 4.7: No-Hype Full Review & Testing
    Pat Simmons/about 3 days ago
    https://www.youtube.com/watch?v=x48xbDO6fKo
    A no-hype review of xAI’s Grok 4.7, benchmarked against Fable 5.1 and GPT-6 Astra.

  11. OpenAI Launches GPT-6 Astra — Welcome to the AGI Era
    Dev Branch/about 3 weeks ago(published 2026-09-03, within 60 days)
    https://www.youtube.com/watch?v=TKhMnVAh_ZE
    Calls the GPT-6 Astra launch, September’s biggest topic, the arrival of the AGI era. This week’s Opus 5.5 and Sol/Luna are discussed as responses or follow-ons to that model.

Signals
  • Today’s biggest topic: Anthropic’s biology lab’s first result—the discovery of a CRISPR-like ART enzyme system—was covered simultaneously by multiple channels, including Dr. Samuel Allen Alexander and Wes Roth, only 18–19 hours after posting. It is the fastest-rising YouTube topic today.
  • Closed-model price-cut competition: On September 22, OpenAI announced GPT-6 Sol/Luna and halved prices just 89 minutes after Anthropic announced Claude Opus 5.5. Multiple videos mention this release-timing competition, often in a meme-like way; AI時短ラボ’s title captures it well.
  • Security and governance: The Gemini unauthorized-access incident disclosed September 18, and the proposal for a voluntary safety-standards group involving Google, OpenAI, and Anthropic—Frontier AI Standards Agency ≒ Standards Authority for Frontier AI—appeared in news and major-explainer coverage respectively. No dedicated YouTube video for the latter was found, suggesting text outlets such as AI Weekly and The Information remain ahead.
  • Open-weight models are quiet today: No videos posted within the past 24 hours were found about new releases of Kimi, GLM, Llama, or other open-weight models. Recent attention is concentrated on the three closed-model companies: OpenAI, Anthropic, and Google.
Limits
  • YouTube search-result pages (youtube.com/results?search_query=...) could not be retrieved directly because they are JavaScript-rendered; WebFetch returned only footer content. Research instead combined site:youtube.com WebSearch with YouTube’s oEmbed API, which can retrieve only titles and channel names.
  • Consequently, accurate view counts, subscriber counts, and comments could not be obtained. Dates are estimates based on relative timestamps such as "hours ago" and "days ago" in search snippets.
  • No dedicated YouTube video was found for the "Frontier AI Standards Agency" / "Standards Authority for Frontier AI, SAFA" proposal. News sites appear to be ahead, and YouTube analysis may not yet have appeared.
  • Other platforms such as Reddit, X, Bluesky, and Lemmy were outside this stage’s scope and were not investigated here.

Bluesky

Bluesky — Today’s LLM-related news(2026-09-25)

Accounts
  • TestingCatalog (@testingcatalog.com) — A breaking-news account that quickly captures AI product leaks, new features, and model launches. Half of the posts in this report came from this account.
  • Simon Willison (@simonwillison.net) — A prominent LLM practitioner and blogger. He regularly compares new models with his pelican-drawing benchmark.
  • Epoch AI (@epochai.bsky.social) — A research organization that quantitatively analyzes AI training costs and compute trends.
  • Nathan Lambert (@natolambert.bsky.social) — Affiliated with Ai2 and author of interconnects.ai, with a strong focus on the open-weight versus closed-model landscape.
  • antirez / Salvatore Sanfilippo (@antirez.bsky.social) — Creator of Redis. Continues posting tests of running open-weight models on Mac/DGX Spark using his fast inference tool, "DwarfStar."
  • ModelVantage (@modelvantage.bsky.social) — A bot that monitors API latency and outages for Gemini, Claude, and similar systems.
  • Developer LLM Leaderboard (@devllmlb-ja.bsky.social) — A Japanese account that compiles monthly benchmark results, though its latest post is somewhat old, from 2026-09-01.
Posts
  1. OpenAI previews today’s GPT-6 Sol / Luna launch — "OpenAI prepares to launch GPT-6 Sol and Luna models today"
    TestingCatalog、2026-09-22、👍1 🔁1
    https://bsky.app/profile/testingcatalog.com/post/3mw46urvt432w

  2. Leak: Google is testing new Gemini 4 Pro checkpoints — "Google tests new Gemini 4 Pro checkpoints, early outputs"
    TestingCatalog、2026-09-23、👍2
    https://bsky.app/profile/testingcatalog.com/post/3mw6b5f3opi2j

  3. Xiaomi open-sources its omnmodal MiMo-V2.6 Pro/Flash models(supports coding, vision, and computer use)
    TestingCatalog、2026-09-21、👍2 🔁1
    https://bsky.app/profile/testingcatalog.com/post/3mw2tpiskxh2w

  4. SpaceXAI releases Grok 4.7 for coding and knowledge work
    TestingCatalog、2026-09-21
    https://bsky.app/profile/testingcatalog.com/post/3mw2rdziawn2s

  5. Anthropic is testing Fable 5.2 and Opus 5.5 before release
    TestingCatalog、2026-09-21、👍1
    https://bsky.app/profile/testingcatalog.com/post/3mvzd4w227422

  6. Simon Willison compares Claude Opus 5.5, GPT-6 Sol, and GPT-6 Luna together — "Big model release today - I wrote about Claude Opus 5.5, GPT-6 Sol, and GPT-6 Luna - plus comparison grids of pelicans by the different model families at different reasoning levels"
    2026-09-22、👍123 🔁10
    https://bsky.app/profile/simonwillison.net/post/3mw5g6izoms2n

  7. Simon Willison examines the extremely low price of Gemini 3.8 TTS — "The new Gemini 3.8 TTS models are super-cheap and can generate conversations between multiple voices..." Estimated at under one cent per minute.
    2026-09-23、👍77 🔁7
    https://bsky.app/profile/simonwillison.net/post/3mw7magyrwc2i

  8. Epoch AI: "AI costs are falling by about 47% each quarter, faster than any other technological innovation"
    2026-09-22、👍123 🔁30
    https://bsky.app/profile/epochai.bsky.social/post/3mw567kluom2h

  9. Epoch AI: GPT-5.6 Luna reproduces o3’s FrontierMath level at 1/377th of the cost($0.55→$0.0015, 377-fold cheaper over 18 months)
    2026-09-22、👍12
    https://bsky.app/profile/epochai.bsky.social/post/3mw567lpe7s2l

  10. Nathan Lambert briefs congressional staff on the state of open models and competitiveness with China(links to an interconnects.ai briefing)
    2026-09-21、👍31 🔁6
    https://bsky.app/profile/natolambert.bsky.social/post/3mvzo2lneqg2j

  11. He notes that China’s top AI labs are beginning to use Huawei for inference and Nvidia for training, highlighting a split in hardware strategy.
    2026-09-20、👍18
    https://bsky.app/profile/natolambert.bsky.social/post/3mvxajmx5cz2n

  12. antirez adds support for Qwen3.8 Flash Next to the local inference tool "DwarfStar"(50–70 t/s on a 64GB Mac; over 1400 t/s prefill)
    2026-09-14、👍76 🔁8
    https://bsky.app/profile/antirez.bsky.social/post/3mvi75e3c722o

Signals
  • A closed-model release rush is the week’s central story: GPT-6 Sol/Luna, Claude Opus 5.5/Fable 5.2, and Gemini 4 Pro leaks/tests all arrived in the same week. Bluesky’s characteristic speed is visible in practitioners such as Simon Willison immediately producing side-by-side comparisons using his pelican-drawing benchmark.
  • The "collapse in reasoning costs" is being discussed with quantitative evidence: Epoch AI’s numbers—47% lower per quarter and 377-fold lower cost on FrontierMath—spread with more rigorous sourcing than on other platforms (👍123 and 🔁30, the strongest response in this research).
  • The open-weight camp centers on local execution reports: antirez continues testing Qwen3.8 Flash Next and DeepSeek v4.1 Flash on personal Macs/DGX Sparks, while Xiaomi’s MiMo-V2.6 open-source release was also reported. TestingCatalog’s Xiaomi post (👍2) received far less engagement than Simon Willison’s post (👍123), suggesting that on Bluesky, new closed-model releases attract more attention.
  • Nathan Lambert supplies a geopolitical closed-versus-open framing: His observation that Chinese players divide work between Huawei for inference and Nvidia for training offers an angle absent from the benchmark-centered discussion elsewhere.
  • API-monitoring bots such as ModelVantage also exist, but recent incidents—such as Gemini 3.1 Pro outages on September 14–15, with a 100% error rate for several minutes in US-East—were not today’s story and received zero likes.
Limits
  • The official search API (app.bsky.feed.searchPosts) consistently returned 403 Forbidden in this research environment, preventing direct keyword searches (the same occurred with direct access and through two CORS proxies). Other endpoints such as getProfile, getAuthorFeed, and getPostThread worked normally, suggesting only the search endpoint was blocked.
  • Because of that constraint, collection used the alternative method of identifying likely relevant accounts and reading their feeds rather than searching broadly by keyword. Posts from unnamed individuals and small accounts may therefore have been missed.
  • Bluesky’s own search and post pages are client-side-rendered SPAs, and WebFetch returned only the title "Bluesky," not post text. All links were verified through the API, but could not be visually checked in a browser.
  • Automated posting bots such as ai-news-jp.bsky.social stopped posting in July 2025, so they were excluded as sources for today’s topic.
  • No posts dated 2026-09-25 itself were found. The listed items are mainly recent posts from 2026-09-20 through 09-23 about GPT-6 Sol/Luna, Opus 5.5, Gemini 4 Pro, and related topics. Ten items could not be assembled using only posts from today.
  • Japanese accounts such as devllmlb-ja.bsky.social had no posts this week, creating the impression that Japanese-language LLM discussion on Bluesky is thinner than English-language discussion.

Lemmy

Lemmy — Today’s most discussed topics: the plunge in closed-model share and excitement around Opus 5.5 / local models

Communities
  • !«メールアドレス» — 88,221 subscribers. A general tech community where LLM news flows in; the main source for Gemini-related posts.
  • !«メールアドレス»(also mirrored at sh.itjust.works, ID 330) — 5,172 subscribers. Focuses on technical posts about new open-weight releases, benchmarks, and llama.cpp.
  • !ai_reddit — A community reposting r/ClaudeCode and r/ArtificialInteligence posts (subscriber count could not be retrieved on lemmy.world because of 404/500 errors; see Limits). Opus 5.5-related threads are concentrated here.
  • !hackernews — A mirror of popular Hacker News articles. LLM application topics such as historical research and deanonymization research rank highly.
  • !«メールアドレス» — Lemmy’s own development community. Rules for disclosing AI/LLM use are under discussion.
  • !selfhosted — A self-hosting community where the definition of open-source LLMs, especially K2 Horizons, is being debated.
  • !blueteamsec / !lobsters — Security- and technology-oriented communities reposting LLM research.
Posts
  1. Closed Models Fell From 70% to 21.6% Market Share in 3 Months — !tecnologia、score 1、2026-09-24 05:00
    https://officechai.com/ai/share-of-closed-models-has-fallen-from-around-70-to-21-in-the-last-3-months-vercel-data/
    An article reporting, based on Vercel data, that closed models’ usage share fell from 70% to 21.6% in only three months. It symbolizes today’s closed-vs.-open-weight framing.

  2. New Opus is absolutely bonkers — !ai_reddit(originally r/ClaudeCode)、score 0、2026-09-24 20:01
    https://www.reddit.com/r/ClaudeCode/comments/1wp1q8n/new_opus_is_absolutely_bonkers/
    A thread describing Anthropic’s new Opus as astonishing.

  3. Is Opus 5.5 really better than Fable in your experience? — !ai_reddit(originally r/ClaudeCode)、score 0、2026-09-24 02:50
    https://www.reddit.com/r/ClaudeCode/comments/1woe3lo/is_opus_55_really_better_than_fable_in_your/
    A discussion comparing Opus 5.5 and Fable based on real-world experience rather than just benchmark results.

  4. I A/B tested Opus 5 vs Opus 5.5 for AI slop: 98 em dashes vs 0 — !ai_reddit(originally r/ClaudeCode)、score 1、2026-09-23 07:45
    https://www.reddit.com/r/ClaudeCode/comments/1wnwqcs/i_ab_tested_opus_5_vs_opus_55_for_ai_slop_98_em/
    An A/B test reporting that output quality improved substantially, including much less "AI-like" writing such as overuse of em dashes.

  5. llama.cpp v0.5.0 — !LocalLLaMA、score 23・comments 0、2026-09-24 16:18
    https://github.com/ggml-org/llama.cpp/releases/tag/v0.5.0
    A new release supporting multiple new architectures, including Qwen4Exp.

  6. New medical model: FreedomIntelligence/HuatuoGPT-3-27B — !LocalLLaMA、score 7・comments 1、2026-09-24 16:52
    https://huggingface.co/FreedomIntelligence/HuatuoGPT-3-27B
    A medical open-weight model based on Qwen3.8-27B and built with One-stage Policy Optimization without supervised fine-tuning.

  7. Qwen4-27B just confirmed — !LocalLLaMA、score 34・comments 5、2026-09-22 22:45
    https://lemmy.ml/post/53074988
    A thread reporting formal confirmation of Qwen4-27B’s release. Though three days had passed since the announcement, it remained a key open-weight story of the week.

  8. Google introduce "Call for Me" sui Pixel 11 per affidare le telefonate a Gemini — !tecnologia、score 1、2026-09-24 18:17
    https://www.theverge.com/ai-artificial-intelligence/1000116/google-gemini-business-phone-calls
    Google announces "Call for Me," a Pixel 11 feature in which Gemini handles phone calls on the user’s behalf.

  9. Google invented the Transformer, so why does using Gemini still feel like a chore? — !ai_reddit(originally r/ArtificialInteligence)、score 1、2026-09-24 08:35
    https://www.reddit.com/r/ArtificialInteligence/comments/1wovzal/google_invented_the_transformer_so_why_does_using/
    A critical post asking why Google, despite inventing the Transformer, continues to be overtaken by more agile open-weight players such as DeepSeek and Qwen.

  10. Using LLMs to trace alchemical knowledge and decode 17th century letters — !hackernews、score 7、2026-09-24 20:00
    https://lemmy.bestiver.se/post/1359598
    An example of historical research using GPT-6 and Opus 5.5 to decode 17th-century alchemical correspondence.

  11. Large-scale online deanonymization with LLMs — !blueteamsec、score 6、2026-09-24 19:23
    https://infosec.pub/post/52730942
    Research claiming that LLMs can re-identify Hacker News users and Anthropic Interview participants at high accuracy, achieving up to 68% recall at 90% precision. It drew attention as an LLM risk case.

  12. Is K2 Horizons historically significant as the first actual open-source LLM, or am I falling for hype? — !Selfhosted、score 48(upvotes)/10(downvotes)・comments 20、5 days ago(around 2026-09-20)
    https://lemmy.zip/post/... (lemmy.world repost: https://lemmy.world/post/52135351)
    A debate over whether K2 Horizons can be called the first truly open-source LLM. Many replies cite earlier examples such as GPT-Neo, OPT, and BLOOM, and point out that the training data are not public.

Signals
  • Closed versus open weight: The day’s clearest symbol was the article claiming closed models’ market share plunged from 70% to 21.6% in three months (post 2). Lemmy’s overall mood also strongly favors open weights, with praise for DeepSeek/Qwen and criticism of Google’s sluggishness (post 9), alongside announcements of Qwen4-27B and HuatuoGPT-3-27B (posts 6 and 7).
  • Anthropic Opus 5.5: Several Opus 5.5-related threads were running simultaneously in !ai_reddit, reposted from r/ClaudeCode (posts 2, 3, and 4). Both output quality, such as reduced em-dash use, and benchmark performance are being discussed.
  • LLM incidents and risks: Deanonymization research (post 11) and technical discussions of watermarking in vLLM appeared in the Lobsters/blueteamsec cluster, focusing on safety and misuse prevention.
  • Practical deployment: Product-integration stories such as Google Gemini’s "Call for Me" feature for Pixel 11 phone handling (post 8) also received attention, distinct from discussion of the models themselves.
Limits
  • Lemmy communities are small, and fewer than ten LLM-related threads could be confidently called newly posted today. The K2 Horizons debate from five days ago (post 12) and the Qwen4-27B post from three days ago (post 7) were included as topics still being referenced this week, bringing the total to 12.
  • The subscriber count for the !ai_reddit community could not be retrieved: lemmy.world/api/v3/community?name=ai_reddit returned 404 and lemmy.world/c/ai_reddit returned 500. Its home instance may be elsewhere, but it could not be identified.
  • The original Reddit posts at www.reddit.com/r/ClaudeCode/... were blocked from direct WebFetch access. Summaries therefore rely on titles, scores, and timestamps from Lemmy mirror posts; the original thread bodies could not be examined in detail.
  • Search UIs on other instances including lemmy.ml and lemm.ee were replaced in this investigation by lemmy.world API searches and individual API fetches from sh.itjust.works. No comprehensive search across all instances was performed.

Recommended actions

  • Add specific model names such as "Opus 5.5," "GPT-6 Sol," and "GPT-6 Luna" to the next Reddit survey’s search terms.
  • Verify Vercel’s closed-model-share claim of 70% → 21.6% using primary data before citing it.
  • Continue following Frontier AI Standards Agency developments in text media to compensate for their delayed visibility on social platforms.
  • Consider a dedicated tracking category for scientific results emerging from Anthropic’s biology lab.
  • Add search terms that surface English-language primary sources to X research, which tends to skew toward demonstrations by Japanese creators.
  • Check whether Bluesky’s official search API 403 error has been resolved at the start of the next survey.

Data quality notes

Reddit and Lemmy met the ten-item completion target. X stopped at eight because most search terms were unrelated; Bluesky had to use account-level collection because of a 403 search-API error and found no posts dated today; and YouTube could not provide accurate view or subscriber counts because of rendering constraints.