KEN’S CAT LOG

Image and Video Generation AI News — 2026-09-17

Closed-source is defined by Sora 2’s de facto exit and a head-to-head battle between GPT Image 2.5 and Nano Banana Pro. In open source, Chinese players (Wan, MiniMax H3, HunyuanVideo, and Z-Image Turbo) are taking the lead, while suspicions that Wan 3.0 may be going closed-source are the biggest wildcard.

Image and Video Generation AI News — 2026-09-17

Wow—the image/video generation AI space is moving insanely fast this time around 🔥 On the closed-source side, Sora 2 has effectively faded out, with Veo 3.1, Kling 2.5 Turbo, and Seedance taking center stage. Meanwhile, OpenAI’s GPT Image 2.5 and Google’s Nano Banana Pro are locked in a real head-to-head battle 👊 On the open-source side, Chinese players such as Wan, MiniMax H3, HunyuanVideo, and Z-Image Turbo have completely taken the initiative, while the ComfyUI community is in festival mode, with new nodes and LoRAs dropping every day 🎉 But Google’s models (Nano Banana Pro/Flow) have also had plenty of mishaps: its safety filter rejects even images generated by its own models, and a Google Earth feature was rolled back after fake disaster images were mass-produced. Reliability is looking a little shaky 😬

Across platforms

  • Sora 2’s exit and the handoff to Google/ByteDance/Kling: YouTube reports that “Sora 2’s API is scheduled for final shutdown on 2026/9/24,” while both Reddit and YouTube independently rate Seedance (2.0→2.5) highly for human realism. It is a strong signal when the same conclusion emerges organically across platforms 💪
  • The GPT Image 2.5 vs Nano Banana Pro debate: The two models are being compared independently across three platforms: YouTube (several review videos immediately after the announcement), Bluesky (@testingcatalog.com’s breaking post on 2026-09-09), and Lemmy (a Berserk coloring comparison in which ChatGPT Images 2.5 came out ahead at preserving linework, while Nano Banana Pro lost points for ignoring instructions).
  • Frequent trouble around Nano Banana Pro/Google Flow: Reddit has a “runaway filter” story in which Flow rejects images made by its own Nano Banana Pro; Lemmy reports that a Nano Banana 2 feature built into Google Earth was withdrawn the day after rollout because fake disaster and military-facility images were mass-produced. The platforms differ, but they paint the same picture: Google’s generative AI is powerful, but its controls have not caught up.
  • Open source is a Chinese-player-plus-MiniMax-H3 battleground: Reddit repeatedly names Wan, MiniMax H3, and LTX Video as the top free options; YouTube’s Wan, Z-Image Turbo, and HunyuanVideo/Image are all Chinese; Lemmy’s !«メールアドレス» has seen a burst of ComfyUI tools for MiniMax H3 in recent days; and Bluesky shows Lightricks LTX-2.5 trending on Hugging Face for three consecutive days. All four platforms are pointing in the same direction.

Platform by platform

Reddit: Collected 11 threads (short of the target of 20). On the closed side, the key themes are the contradiction between Google Flow’s generous free tier and its overzealous safety filters, plus Seedance’s strong reputation for human realism. On the open side, Wan/MiniMax H3/LTX Video keep coming up as local-generation staples. There was also real community texture in arguments over what “AI slop” means and cynicism toward fully automated YouTube-video generators 🗿

X: A total miss this time 😅 The collection method was trend-driven, and because the session was geolocated to Croatia, the 40 collected posts were all about South African politics, Serbia/Kosovo, Israel/Hamas, Zcash governance, and giveaway spam. Not one mentioned image/video generation AI. This is the only platform explicitly named in the brief that can be called a genuine gap—recollection is needed.

YouTube: Identified 13 videos from titles/snippets. On the closed side, GPT Image 2.5 (released 2026-09-08, in Flare and Sunburst variants) and Sora 2’s end stand out, while Meta’s entry with Muse Image/Video (from Meta Superintelligence Labs, using agentic generation with search and code execution) is also significant. On the open side, the main players are Alibaba (Wan, Z-Image Turbo), Tencent (HunyuanVideo 1.5, HunyuanImage 3.0), and Black Forest Labs (Flux.2). However, view counts and channel names were largely blocked behind JavaScript rendering, leaving authority verification weak.

Bluesky: Official brand accounts (Midjourney, Stability AI, ComfyUI, Runway, Kling, Black Forest Labs) have almost all been silent since late 2025. The main source is the independently run news account @testingcatalog.com (860 followers). It surfaced ChatGPT Images 2.5, Meta Muse Video, startup Visko’s real-time video model “Orbis,” and ElevenLabs’ image/video expansion. For open-source developments, the only clear signal was ongoing attention to LTX-2.5 from a Hugging Face trends bot. The search API returned 403, forcing an account-by-account investigation.

Lemmy: Just over 10 relevant posts were found (below the 20-post target; no forced padding). Closed-source topics tend to appear as mishaps or controversies—such as the Google Earth AI rollback or jokes comparing Midjourney’s medical-device venture to Theranos. On the open side, !«メールアドレス» is effectively the hub, where a classic OSS pattern is visible: ComfyUI nodes, LoRAs, and memory-efficient implementations for MiniMax H3 appeared in rapid succession over just a few days.

Pinterest: Visually reviewed 30 of 50 pins. Open-source-related content was almost completely absent—there were zero pins explicitly naming Stable Diffusion, Flux, Wan, or ComfyUI. What appeared instead were closed consumer brands such as ChatGPT, Gemini, Claude, Adobe Firefly, Midjourney, and Canva. Pins with actual news value included the Tilly Norwood “AI actress” controversy, ShengShu Technology’s Vidu S1 (real-time interactive video), and a technical diagram for Microsoft VASA-1. Overall, stock imagery stoking fears about robots and deepfakes dominates: it reflects a visual culture of vague anxiety around AI more than “news.”

What to watch

Recommendations

  • Recollect X (formerly Twitter) using theme-specific terms such as “Sora 2,” “Nano Banana Pro,” “Wan,” and “Flux,” rather than trend-driven discovery.
  • Verify the suspected closure of Wan 3.0 against primary sources, especially Alibaba’s official announcement, and track the implications for the open-source video-generation landscape.
  • Monitor !«メールアドレス» regularly for the freshest developments around MiniMax H3.
  • Because GPT Image 2.5 and Nano Banana Pro comparisons align across several platforms, next time also examine third contenders such as Seedream 5.0 Pro.
  • Reports of Google-model safety-filter overreactions and faulty outputs are recurring, so keep tracking them as reliability news.

Data quality

X’s collection method (geolocated trend-based discovery) was completely misaligned with the topic: zero of 40 posts were relevant, so the platform effectively did not work in this run. Reddit (11 items) and Lemmy (just over 10) did not reach the brief’s completion benchmark of “around 20 items”; no artificial padding was used. On YouTube, JavaScript rendering prevented verification of view counts and channel names, weakening authority checks. Bluesky’s official search API was unavailable due to a 403 response, requiring a substitute method based on known-account feeds and limiting coverage. Pinterest offered only images and titles; linked articles and save counts could not be verified.

Platform summaries

Reddit

Reddit — Image and Video Generation AI News

Where

The 11 collected threads came from seven subreddits. All were found using the search term “Image and Video Generation AI News.”

Subreddit Members Collected threads
r/generativeAI 154,126 5
r/aitubers 28,633 1
r/AIDangers 53,410 1
r/aifilmmaking 7,951 1
r/aivideomaking 7,085 1
r/AIToolTalks 6,307 1
r/NoStupidQuestions 7,494,416 1

Specialist image/video-generation subreddits (r/generativeAI, r/aifilmmaking, r/aivideomaking, r/aitubers) account for 8 of the 11 threads. The other three (r/AIDangers, r/AIToolTalks, r/NoStupidQuestions) discuss AI more broadly and are not limited to image/video generation.

What people say
  • Closed-source (Google Flow/Veo) is offering a very generous free tier, but restrictions are also tightening. Thread 4, “Google Flow is giving creators a surprisingly generous amount of free AI video generation right now” (r/aifilmmaking · 8 points · 2026-09-14) https://www.reddit.com/r/aifilmmaking/comments/1wg7izt/ describes free Veo credits as becoming “a serious production workspace with scene construction, camera controls, and reference images,” rather than merely text-to-video. In the comments, u/MechanicForward2274 (1 point) says, “50 credits a day is actually enough to experiment with a lot if you're not wasting them.”
  • In contrast, thread 7 (r/generativeAI · 2 points · 2026-09-15) https://www.reddit.com/r/generativeAI/comments/1wh1y9d/ complains that Google Flow’s safety filter rejects even images made by Google’s own model, Nano Banana Pro. The poster retried the same image 60 times without success. That friction is pushing RTX 5090 owners toward local generation.
  • Local/open-weight video models repeatedly come up as the top free options: Wan, MiniMax H3, and LTX Video. In thread 1 (r/generativeAI · 8 points · 2026-09-16) https://www.reddit.com/r/generativeAI/comments/1whtesl/, u/tuckken2 (2 points) says, “Wan 2.2 and Minimax H3 are free on cnaps studio, credit resets every 5 hours.” In thread 7, u/f5alcon (3 points) calls “Minimax H3 ... the best local generator,” and u/Liberation2020 (1 point) names “Wan, mini max H3 and LTX Video.”
  • Seedance (closed-source, ByteDance-related) is gaining recognition for human realism. The thread 7 poster writes, “Seedance 2.0 is the top video generator regarding human realism.” In thread 8, “HOW IS THIS MADE?!” (r/generativeAI · 0 points · 2026-09-10) https://www.reddit.com/r/generativeAI/comments/1wced4g/, u/lagbit_original (3 points) explains that Seedance 2.5, a single prompt, suitable reference images, and music/cuts added in editing software can achieve viral-video-level quality. In thread 11, u/pb404 (2 points) also praises the consistency of Seedance 2.5’s reference model.
  • Demand for tools that make YouTube videos automatically from start to finish is a target of ridicule. In thread 5 (r/aitubers · 0 points · 35 comments · 2026-09-14) https://www.reddit.com/r/aitubers/comments/1wg3duk/, a post seeking fully automated script-to-video generation for under €50/month drew a sharp response from u/RangeWilson (7 points): “I'm looking for a unicorn who will slide down a rainbow and give me a pot of gold every day. But I'm not willing to pay more than €50 per month.” u/RobJames007 (12 points, the thread’s highest score) also cautioned, “YouTube is trying to reduce AI slop, not allow more of it onto the platform.”
  • Opinion is divided over what the label “AI slop” means (see Signals for detail). Thread 11, “Is there a way to make AI video that does not look like AI slop?” (r/generativeAI · 0 points · 30 comments · 2026-09-14) https://www.reddit.com/r/generativeAI/comments/1wgjjzu/, is the main discussion. The original poster reports that simply using the same character, voice, and continuing gag format increased watch time ninefold without technical changes—and people stopped calling it slop.
  • There is intense demand for free and low-cost options. In thread 2, “Please help me generate free ai video without subscription” (r/aivideomaking · 0 points · 14 comments · 2026-09-15) https://www.reddit.com/r/aivideomaking/comments/1wh80z3/, u/Fluffybabyyoda (1 point) says, “google flow gives you 50 credits daily sometimes 2 times a day. Google gemini i think gives you 3 videos a day.” u/Bastisheen92 (2 points) adds the basic principle: “Local Generation is free if you already own the Hardware.”
  • Some users are also looking for free tools for NSFW/borderline content. In thread 3, “Any free video platform?” (r/generativeAI · 2 points · 5 comments · 2026-09-16) https://www.reddit.com/r/generativeAI/comments/1whvzu3/, the poster writes, “I was able to find ways to get Grok for free and spicy minimax.” In the same thread, u/Critical_Emu_6565 (2 points) promotes their own tool, “RexCut AI,” so it is worth noting that self-promotional comments are mixed in.
  • Broader anxiety and distrust of AI form part of the background, even when not specific to image/video generation. Thread 9, “Are all of you also anxious about recent news about AI?” (r/AIToolTalks · 12 points · 35 comments · 2026-09-13) https://www.reddit.com/r/AIToolTalks/comments/1wfhbnv/, raises concerns about water use and job losses and mentions actor Andrew Garfield reportedly stopping AI use after portraying Sam Altman in “Artificial.” u/Big-Pops78 (3 points) counters that water use is not uniquely high compared with other technologies.
  • Skepticism about AI’s cost-effectiveness is also prominent. Thread 6, “Biggest AI news of 2026: AI is not a necessary tool...” (r/AIDangers · 26 points · 31 comments · 2026-09-14) https://www.reddit.com/r/AIDangers/comments/1wg5y17/, argues that AI adoption has worsened performance at many companies. u/presentofai (1 point) counters, “most of these "AI made it worse" numbers are measuring clumsy rollouts, not the models,” reflecting genuine disagreement.
Signals
  • Rising: Local/open-weight video models (Wan, MiniMax H3, LTX Video) have clear support as alternatives to closed-source free tiers. The discussion specifically describes high-end GPU owners, such as RTX 5090 users, moving local because of restrictive cloud policies (thread 7).
  • Rising: Seedance (2.0/2.5) is gaining recognition for human realism and consistency, with independent references across multiple threads and comments.
  • Being dismissed/rejected: Expectations for fully automated tools that produce a finished video from a script alone are strongly mocked by the community (thread 5, whose top-scoring comment is sarcastic). The term “AI slop” itself is also being dismissed by some as an insult that has lost its meaning.
  • Disagreement: Thread 11 is split down the middle over what “AI slop” is. u/Tenth_10 (9 points, the thread’s highest score) and u/pb404 (2 points) argue it is criticized because the story lacks stakes, not because of technical shortcomings. u/RioNReedus (3 points), by contrast, argues that it has become a meaningless term used by people who dislike AI for everything. It is notable that readers reached opposite conclusions from the same post.
  • Surprising: Google’s own service, Flow, rejects images created by Google’s own Nano Banana Pro model—an example of safety-filter overreach (thread 7). Generous free credits (thread 4, 50 credits/day) and strict censorship of outputs are happening simultaneously, producing a contradictory user experience.
  • Surprising: Nearly half the threads returned by the search phrase “Image and Video Generation AI News” (three from r/AIDangers, r/AIToolTalks, and r/NoStupidQuestions) were about broader fears and skepticism around AI rather than image/video generation specifically. This suggests that image/video-generation AI is being discussed within the wider question of whether AI can be trusted.
Limits
  • Only 11 threads were collected, short of the “around 20” requested by brief.md. Only one search term was used (“Image and Video Generation AI News”—the title of this research itself, not a natural search query); no follow-up searches used alternative wording such as “Sora 2,” “Midjourney update,” or “open source video model release.”
  • Closed-source discussion is skewed toward Google Flow/Veo and Seedance. No dedicated threads about other major closed-source services—including Sora, Midjourney, Runway, Kling, or Grok Imagine—were collected; they appear only in comments.
  • On the open-source side, Wan, MiniMax H3, and LTX Video are repeatedly mentioned, but no dedicated threads were found covering current developments in Flux, the Stable Diffusion ecosystem, or ComfyUI.
  • Three threads from r/AIDangers, r/AIToolTalks, and r/NoStupidQuestions are not specific to image/video generation AI and are only indirectly related to the theme.
  • Browser access to Reddit was limited to data collected in advance by a worker; this agent did not conduct additional searches or page views, in line with the playbook.

X

X — Image and Video Generation AI News

Accounts

None of the accounts in the collected set are talking about image or video
generation AI, closed-source or open-source. The 38 accounts in
output/x.posts.json are a mix of:

  • Political/commentary accounts arguing about South Africa, Serbia/Kosovo,
    Brexit and football eligibility, and Israel/Hamas (e.g. @TheSirRobotto,
    2 posts totalling 12,537 likes on Serbia/EU football rules; @marklevinshow,
    1 post, 3,250 likes, on Erdogan/Hamas).
  • Entertainment/fan accounts (@eva_kirby21, @PossiblyApollo — Apple TV
    Emmys chatter; @justluvbourzgui, @theholydemi, @acervylaur —
    Lauren Jauregui/DWTS fandom).
  • Crypto accounts on the Zcash (ZEC) governance vote (@CoinMarketCap,
    @zooko — Zcash's own co-founder — and @Hasan_NFTOX running a free
    NFT mint).
  • Low-engagement giveaway/spam accounts repeating the same "Pour PalZ Neon
    Crochet Bears Giveaway" copy (@corrysue1, @PHIFFERJ,
    @lilredshoper — all 0 likes, near-identical text via @gaynycdad).

No account here is an AI, Midjourney/Sora/Runway/Stable Diffusion/etc.
account, researcher, or news outlet.
This is not "who is driving the AI
image/video conversation on X" — it is who is driving Croatia's trending
topics that day.

Posts

No findings — none of the 40 collected posts mention image generation,
video generation, or AI models of either kind (closed- or open-source).
Skipping fabricated relevance rather than stretching these into the brief.

Signals

Nothing can be read off this set as a signal for the research theme. The one
observable pattern in the data itself: it is dominated by regional politics
(South Africa, Serbia/Kosovo, Israel-Hamas), one crypto governance vote
(Zcash cutting block times from 75s to 25s, output/x.posts.md #19), and
fandom/spam noise — none of it AI-related.

Limits

The collection for this stage did not target the research theme. Per
output/x.posts.md and .json, the worker's method was "source": "trending":
it read X's Explore trends list — which is geolocated to the session's
server location and came back "Trending in Croatia" / general "Politics ·
Trending" / "Technology · Trending"
— and then searched those trend names
verbatim: "Africa", "Serbia", "Apple", "Canada", "Zcash", "#sweepstakes",
"#giveaways", "#chance", "Lauren", "Hamas". None of these terms relate to
image or video generation AI, so the resulting 40 posts (38 accounts) contain
zero relevant material.

I did not search X myself — per the playbook, this session has no live X
access and I can only read what was already collected. So the completion
criteria (~20 analyzed posts per platform) cannot be met for X in this
run
: there is no theme-relevant post in the collected file to analyze, and
I'm not able to re-run the collection with the right search terms (e.g.
"Sora 2", "Midjourney", "Stable Diffusion", "Runway", "Nano Banana", "Flux",
"Veo 3", "closed-source image AI", "open-source video model") from here.

This stage should be flagged as failed / needing re-collection with
theme-appropriate search terms rather than trending-topic terms.

YouTube

YouTube — Image and Video Generation AI News (Closed-Source/Open-Source)

Channels

YouTube search-result pages are JavaScript-rendered, and in many cases only the page footer was retrievable; channel names and subscriber counts were not included in the body (see “Limits” for details). Channels/publishers identifiable from search-result titles and snippets include:

  • Official OpenAI channel — Introducing GPT-Image-2.5 in the API (official explainer video)
  • AI news and review channels (specific names could not be identified): channels that regularly publish ComfyUI tutorials and coverage of GPT Image 2.5, Nano Banana Pro, Sora 2’s end, Wan/HunyuanVideo/Z-Image/Flux.2, and related topics. Search results support that specialist AI channels such as MattVidPro AI, TheAIGRID, Theoretically Media, and Matt Wolfe (about 965K subscribers; AI news focused) frequently cover these themes, but a one-to-one mapping between individual videos and channel names could not be verified.
Videos

Closed-source

  1. GPT Image 2.5 Is AMAZING: Full Breakdown of OpenAI's New Image Model — early September 2026 (“1 week ago”) — https://www.youtube.com/watch?v=DY6TkI8XtkQ — A full review of OpenAI’s new image model, “GPT Image 2.5,” released on 2026/9/8 in Flare and Sunburst variants.
  2. Introducing GPT-Image-2.5 in the API — September 2026 — https://www.youtube.com/watch?v=A7MSwdXj86k — An official OpenAI-leaning explainer. It emphasizes sharper, more natural lighting in Sunburst and improved API editing accuracy.
  3. OpenAI's New Image Model Is Brilliant… Until You Look Closer — mid-September 2026 (“5 days ago”) — https://www.youtube.com/watch?v=pEYqorQrc4U — A critical review comparing GPT Image 2.5 with other leading image models and identifying weaknesses such as detail breakdowns.
  4. [NEW] Meta Muse.Ai Image Generation First Reaction — around July 2026 — https://www.youtube.com/watch?v=rAvnO1DWnMs — A first-look review of Meta Superintelligence Labs’ new “Muse Image” model, based on Muse Spark and using agentic generation with search and code execution.
  5. Grok Imagine Video 1.5 vs Seedance 2.0 | Full AI Video Test — 2026-06-08 — https://www.youtube.com/watch?v=5mFkZiRTVu4 — A practical generation comparison between xAI’s “Grok Imagine Video 1.5” and ByteDance’s “Seedance 2.0.”
  6. Grok's New FREE AI Video Generator Is INSANE — 2026-07-06 — https://www.youtube.com/watch?v=lrCjbL51dbk — Focuses on the ability to try Grok Imagine Video 1.5 for free.
  7. BREAKING: Sora 2 is SHUTTING DOWN… So I Tested the Best Alternatives — 2026-03-27 — https://www.youtube.com/watch?v=Zf9979o8RGk — Following OpenAI’s announced shutdown of the Sora 2 app/API on 2026/3/24, this tests alternatives such as Veo 3.1, Kling, and Hailuo, marking a shift in the closed-source landscape.

Open-source

  1. Z Image Turbo (FREE): This Open-Source AI Changed Image Generation Forever! — 2025-11-27 — https://www.youtube.com/watch?v=sPQQPpQ9X4E — Alibaba (Tongyi-MAI)’s 6B-parameter Apache-2.0 model, “Z-Image Turbo,” claimed to rival Nano Banana Pro in quality and speed.
  2. New Open-Source Model DESTROYS Flux 2 — Z-Image + Promptus Tutorial — November–December 2025 — https://www.youtube.com/watch?v=JF4Q5hXh-_Q — A ComfyUI tutorial evaluating Z-Image Turbo as outperforming Flux.2.
  3. Flux.2 Has Landed! Is It the Open Source King? — November 2025 — https://www.youtube.com/watch?v=D0_QGrdtvEg — Reviews Black Forest Labs’ Flux.2 (Pro/Flex/Dev/Apache-2.0 Klein), though later criticism argues that its openness is only superficial.
  4. We have a new #1 open-source AI video generator! (Tencent HunyuanVideo 1.5) — 2025-11-25 — https://www.youtube.com/watch?v=6EQP8-D37bs — A review and installation guide for Tencent’s 8.3B-parameter open-source video-generation model.
  5. HunyuanImage 3.0: Open Source Text To Image Model (Forget Nano Banana & Seedream 4) — September 2025 — https://www.youtube.com/watch?v=32lhiZKhgOY — Introduces Tencent’s MoE open image-generation model with more than 80B parameters as “the largest open-source image model.”
  6. WAN 2.2 Animate in ComfyUI (Step-by-Step Guide) — 2025-09-26 — https://www.youtube.com/watch?v=dyqz3PJ0wvo — A tutorial on Alibaba Wan 2.2’s open-source feature for generating video from a single reference image.
Signals
  • OpenAI’s video generation has effectively exited through Sora — The Sora 2 app/API was halted and ended around March–April 2026 (with final API shutdown reportedly scheduled for 2026/9/24). Several videos frame Google’s Veo 3.1, Kling 2.5 Turbo, and ByteDance’s Seedance 2.0 as the new leaders in closed-source video generation (BREAKING: Sora 2 is SHUTTING DOWN, https://www.youtube.com/watch?v=Zf9979o8RGk).
  • The closed-source image-generation frontier is GPT Image 2.5 versus Nano Banana Pro (Gemini 3 Pro Image). After OpenAI launched GPT Image 2.5 on 2026-09-08, direct comparison videos with Google’s Nano Banana Pro, launched in November 2025, increased rapidly (GPT Image 2.5 Is AMAZING, https://www.youtube.com/watch?v=DY6TkI8XtkQ).
  • Meta is making a serious move into image/video generation — Muse Image / Muse Spark / Muse Video from Meta Superintelligence Labs appeared in July 2026, introduced as a new agentic approach that validates images through web search and code execution.
  • xAI’s Grok Imagine is iterating rapidly — It progressed from its initial version in August 2025 to Video 1.5 between May and July 2026, with many videos emphasizing free access.
  • Chinese models lead the open-source field — Alibaba (Wan, Z-Image Turbo), Tencent (HunyuanVideo, HunyuanImage), and Germany’s Black Forest Labs (Flux.2) are central. Z-Image Turbo (6B, Apache-2.0), in particular, is even being rated as outperforming Flux.2.
  • A reversal signal to watch: Wan was fully open under Apache-2.0 through Wan 2.2, but multiple articles claim that “Wan 3.0,” announced 2026-08-24, switched to an API-only, weight-unavailable closed policy. Though the source is blog reporting, YouTube coverage also expresses concern about the uncertain successor to open-source video generation’s flagship.
  • The ComfyUI ecosystem remains the main battleground for open-source video/image generation — Tutorials for Wan 2.1/2.2, HunyuanVideo, and Z-Image Turbo continue to proliferate.
Limits
  • Direct WebFetch requests to YouTube search-result pages (youtube.com/results?search_query=...) returned only the pre-JavaScript footer—terms of service, copyright notices, and similar content—rather than video titles, channel names, views, or publication dates. Individual video pages (youtube.com/watch?v=...) behaved similarly, preventing access to embedded JSON-LD and OG tags.
  • Therefore, the view counts and subscriber counts above are only included where available from search snippets; precise view counts and channel names could not be verified for most videos (zero videos had confirmed view counts). A separate route such as the YouTube Data API would be needed for those numbers.
  • For Kling 2.5 Turbo, multiple non-YouTube sources confirmed its #1 position on the Artificial Analysis leaderboard, but a corresponding specific YouTube URL could not be identified because results skewed toward blog posts.
  • For Midjourney’s video model, only news from April–June 2025 was found; no major developments within the preceding 60 days could be confirmed. Minor V7.1/V8 updates were only available through information from September 2025.
  • Comment content, including top comments, was not present in the retrieved pages and could not be analyzed.

Bluesky

Bluesky — Image and Video Generation AI News

Accounts
  • 🚨 AI News | TestingCatalog @testingcatalog.com — 860 followers. An independently run AI breaking-news account describing itself as “Latest AI News on AI Agents, Model Releases, Tools, Leaks, and Rumors.” It posts several short headline-style updates daily about releases and leaks and was the most frequently updated, closest-to-current source in this research.
  • Daily HuggingFace Trends @huggingfacetrends.bsky.social — 122 followers. A bot that automatically posts Hugging Face’s trending models daily after 19:00 (GitHub: aegisfleet/hugging-face-trending-to-bluesky). Useful for tracking open-source model activity.
  • Midjourney (official) @midjourneyofficial.bsky.social — 112 followers. Midjourney joined Bluesky in August 2025, but its latest post was on 2025-10-08 and it has been silent since.
  • Stability AI (official) @stabilityai.bsky.social — 241 followers. It has posted announcements such as “Stability AI Image Services,” but its most recent post was in November 2025 and updates have been sporadic.
  • ComfyUI (official) @comfyuiorg.bsky.social — The official account for the open-source image/video workflow tool. Its most recent post dates to June 2025, so 2026 activity could not be tracked.
  • Black Forest Labs (Flux developer) @blackforestlabs.bsky.social / Runway @runwayml.bsky.social / Kling AI @klingai.bsky.social — The accounts exist, but posting is nearly nonexistent: Black Forest Labs has zero posts, Runway has only one from October 2024, and Kling’s feed is empty. Bluesky is not their main communication channel.
Posts
  1. OpenAI announces “ChatGPT Images 2.5” (closed-source) — @testingcatalog.com, 2026-09-09, 👍2 🔁1
    https://bsky.app/profile/testingcatalog.com/post/3mv3hn2whdz27
    Introduces fast image generation and editing, stronger subject consistency, and new API models “Flare” and “Sunburst.”

  2. Early outputs from Meta’s video-generation model “Muse Video” emerge (closed-source) — @testingcatalog.com, 2026-08-19, 👍2
    https://bsky.app/profile/testingcatalog.com/post/3mthtr6iu7p24
    Early information on a video-generation model for “Muse,” an agentic super-app reportedly in development ahead of Meta Connect.

  3. Startup Visko announces its real-time video-generation model, “Orbis” (closed-source) — @testingcatalog.com, 2026-09-01
    https://bsky.app/profile/testingcatalog.com/post/3muhthn6m7i2j
    Noted as a new entrant claiming real-time generation.

  4. ElevenLabs’ MCP expands to generate music, images, and video (a closed/API-oriented service) — @testingcatalog.com, 2026-09-14, 👍3 🔁1
    https://bsky.app/profile/testingcatalog.com/post/3mvj6k7eycz2n
    An example of ElevenLabs, previously voice-focused, expanding into image and video generation.

  5. Tencent releases the open-source preview model Hy4 (open-source) — @testingcatalog.com, 2026-08-29
    https://bsky.app/profile/testingcatalog.com/post/3mu7v3bdkwo2h
    An open release of a preview model seen as a possible successor in the Hunyuan line.

  6. Lightricks/LTX-2.5 trends on Hugging Face (open-source video generation) — @huggingfacetrends.bsky.social, 2026-09-15 (also reposted on 09-13 and 09-11)
    https://bsky.app/profile/huggingfacetrends.bsky.social/post/3mvkun3tjqp2t
    Lightricks’ LTX-2.5 video-generation model was featured as a Hugging Face trending model three days in a row, indicating sustained attention.

  7. ComfyUI integrates FLUX.1 Kontext as an official API node (open-source) — @comfyuiorg.bsky.social, 2025-06-06
    https://bsky.app/profile/comfyuiorg.bsky.social/post/3lqvsxzzjhb2t
    Explicitly states “Open Source coming soon!” and marks a major point of integration between Black Forest Labs’ Flux ecosystem and ComfyUI. The date is older, but it remains an important milestone for open-source image editing.

  8. Wan2.1 launches with ComfyUI support (open-source video generation) — @comfyuiorg.bsky.social, 2025-02-26
    https://bsky.app/profile/comfyuiorg.bsky.social/post/3lj42stgelb2l
    Released four models simultaneously: text-to-video (14B/1.3B) and image-to-video (14B). While older, it is an important milestone for Chinese open video models.

  9. Stability AI launches image-editing API services on Amazon Bedrock (closed/API) — @stabilityai.bsky.social, 2025-09-18
    https://bsky.app/profile/stabilityai.bsky.social/post/3lz5eotkuvk2i
    Enterprise image-editing tools, such as turning sketches into photorealistic product shots, are made available through an API.

  10. Midjourney doubles the number of styles in “Style Explorer” (closed-source) — @midjourneyofficial.bsky.social, 2025-09-18, 👍10 🔁2
    https://bsky.app/profile/midjourneyofficial.bsky.social/post/3lz5cskktzk2c
    One of the few feature-update announcements directly posted by the official account.

Signals
  • Bluesky has a thin ecosystem for breaking image/video-generation AI news. Official brand accounts for Midjourney, Runway, Kling AI, Black Forest Labs, ComfyUI, and Stability AI all exist, but most post infrequently or stopped updating in late 2025. Compared with X (formerly Twitter) or Reddit, it is clearly not the primary venue for this conversation.
  • The only account continually publishing “latest news” was the independent breaking-news account @testingcatalog.com, which posted one to three items per day in September 2026 about OpenAI, Meta, Google, Anthropic, and others. Only some of these were directly related to image/video generation—ChatGPT Images 2.5, Meta Muse Video, Visko Orbis, ElevenLabs’ expansion, and Tencent Hy4—with agent and LLM news dominating overall.
  • Open-source activity was easier to find through automated trend-aggregation bots such as Daily HuggingFace Trends than through news accounts. Lightricks/LTX-2.5 trended repeatedly on September 11, 13, and 15, suggesting sustained interest as an open-source video-generation model.
  • Many individual posts showcase AI artwork made with Stable Diffusion, ComfyUI, or Midjourney, but these are artwork showcases rather than primary news information and were therefore omitted from the Posts section.
Limits
  • Bluesky’s official post-search API (app.bsky.feed.searchPosts) consistently returned HTTP 403 Forbidden in this environment regardless of the keyword used, making it unavailable. getProfile, getAuthorFeed, getPostThread, and actor.searchActors worked normally. The research therefore used an alternative workflow: ① discover accounts through actor.searchActors, ② retrieve each account’s feed through getAuthorFeed, and ③ verify relevant posts individually with getPostThread. This limitation prevented comprehensive discovery of general-user posts that full keyword/hashtag searches could have surfaced.
  • Bluesky’s web search page (https://bsky.app/search?q=...) is a client-side-rendered SPA, so post text is not contained in the HTML and could not be read directly.
  • Discovering Bluesky posts through web search (site:bsky.app) mostly surfaced old posts from 2024 to early 2025; current posts from September 2026 were rarely indexed.
  • As a result, collected posts are skewed toward specific accounts, mainly @testingcatalog.com, @huggingfacetrends.bsky.social, and official brand accounts. The overall volume of discussion or public sentiment on the platform cannot be measured.

Lemmy

Lemmy — Discussion Around Image and Video Generation AI (research conducted 2026-09-17)

Communities
  • !«メールアドレス» — 5.71K subscribers, 1.41K posts, 1.93K comments. Moderators: db0 and Even_Adder. Primarily focused on releases of open-weight tools, LoRAs, and ComfyUI custom nodes; it is the most active community, with posts nearly every day.
  • !«メールアドレス» — 2.36K subscribers (362 local). A gallery-oriented community for work made with Stable Diffusion/open-weight models. One user, Even_Adder, posts almost every hour or every few hours.
  • !«メールアドレス» — 51 subscribers and 25.2K total posts. A community mirroring AI-related Reddit subreddits via RSS. Despite its small subscriber count, it can surface generative-AI comparison topics unavailable on other instances.
  • !«メールアドレス» — Contains Nano Banana-related artwork, but listing retrieval was blocked by Anubis bot protection.
  • !«メールアドレス» — Search results indicate a very small community with 222 subscribers and 11 posts; direct access failed with a 500 error.
  • !«メールアドレス» (also called sneerclub) — A community for satire and criticism of the AI industry in general, which occasionally includes skeptical posts about image-generation AI companies.
Posts
  1. “Same Berserk spread, 4 AI colorizations. Looks like GPT wins again?” — !«メールアドレス», 2026-09-14, score 1
    https://lemmy.durstig.online/post/60037
    A comparison in which four AIs colorize a Berserk panel. ChatGPT Images 2.5 performed best at preserving linework and composition; Wan 2.7 Pro had good lighting but repainted some areas; Nano Banana Pro ignored instructions not to change elements; and Seedream 5.0 Pro maintained composition but was oversaturated. The total test cost was stated as $0.30.

  2. “Taking a traditional filmmaking approach while making the Robo Dad animation” — !«メールアドレス», 2026-08-22, score 0
    https://lemmy.durstig.online/post/55214
    An example hybrid workflow: filming handmade puppets → generating character sheets with Nano Banana → creating video with Seedance 2.0.

  3. “Google quietly discontinues its Earth AI feature a day after its rollout after users made no-no images” — !«メールアドレス», 2026-08-07, score 1
    https://lemmy.durstig.online/post/52162
    A feature integrating Nano Banana 2 into Google Earth was withdrawn the day after rollout after images fabricating disasters and military facilities at real locations were mass-produced.

  4. “It's like an album cover...” — !«メールアドレス», 2026-08-07, score -8
    https://lemmy.ca/post/69025640
    A tropical album-cover-style image made with Nano Banana (Gemini 2.5 Flash). Its negative score reflects a cool reception to AI-generated images.

  5. “Midjourney wants Hollywood studios to reveal AI usage” — !«メールアドレス», 2026-07-05, score 12
    https://lemmy.today/post/56011946
    News that Midjourney wants Hollywood studios to disclose their use of AI.

  6. “Midjourney AI pivots to Theranos: Ultrasonic CT” — !«メールアドレス» (posted by David Gerard), mid-June 2026, score 41
    https://awful.systems/post/8731127
    A satirical post comparing Midjourney’s reported move into an “ultrasonic CT” medical-device business, involving ultrasound-like full-body scanning underwater, to the failed blood-testing company Theranos. Many comments are skeptical: why is a company known for image generation and copyright infringement entering medical devices?

  7. New open-source tools (!«メールアドレス», posted by Even_Adder, 2026-09-14–09-17)

  8. “Dark Wings Silhouette” — !«メールアドレス», 2026-09-17, score 2
    https://lemmy.dbzer0.com/post/75601450
    Artwork made with the Krea2 model plus Qwen Image processing. The comments also discuss how to run Qwen 3.5/3.8-series and GLM 5.2 models locally, as well as hardware constraints—frontier models may not fit even on B300 clusters with 2.1TB of HBM.

Signals
  • Closed-source discussion tends to appear not as “new-model announcements” but as mishaps, controversies, and satire—the Nano Banana Google Earth incident and the Theranos comparison for Midjourney. Many such stories are only found through Reddit-mirror communities (!ai_reddit, !imageai), while native Lemmy communities have relatively little discussion.
  • For open source, !«メールアドレス» is effectively the central hub. A classic OSS development pattern is visible: around a single model, MiniMax H3, the ComfyUI ecosystem of nodes, LoRAs, and memory-efficient variants was assembled in a burst over a few days.
  • Subscriber counts are small across the board, ranging from dozens to thousands, so volume is limited compared with Reddit or X.
  • Search tools are prone to noise. At one point, a post about the Irish custody case hashtag “#BringSoraHome” was almost mistaken for news about OpenAI Sora’s shutdown; checking the body confirmed that Sora was a child’s name, and it was excluded from this report.
Limits
  • Lemmy alone did not meet the completion benchmark of “analyze around 20 posts.” Even combining API searches with community-page reviews yielded only just over 10 directly relevant posts, and no artificial padding was used because Lemmy is a smaller platform.
  • lemmy.world/c/stablediffusion and lemmy.world/c/«メールアドレス» returned 500 errors on direct access.
  • !«メールアドレス» was blocked by Anubis bot protection, preventing list retrieval.
  • Searches for “Grok Imagine,” “Veo 3,” and “Kling AI” returned no relevant posts; discussion of these specific closed-source video-generation services is scarce on Lemmy.
  • API search returns substantial low-relevance noise, including unrelated political and technology stories, so only posts individually opened and verified were included.

Pinterest

Pinterest — Image and Video Generation AI News

Visual themes
  • Chrome/white humanoid android as the visual shorthand for "AI": a glossy robot head or torso with glowing blue eyes/circuitry, often at a news desk or holding a mic, stands in for "AI" across dozens of otherwise unrelated pins [5, 7, 10, 12, 20, 26, 30, 33, 47]. It's the single most repeated motif in the set.
  • "Is journalism dying to AI?" framing: newsroom scenes where robots replace or sit beside human anchors/reporters, usually with an anxious human in frame, paired with headlines like "AI Is Writing the News — Can We Still Trust Journalism?" [2, 7, 12, 17, 19, 26, 33].
  • Deepfake/misinformation anxiety graphics: facial-mesh overlays morphing one face into another, biometric scan HUDs, and explicit "FAKE" stamps over AI-generated content (a cat photo, a hooded figure) [9, 14, 16, 18, 44, 47].
  • Blue/cyan tech-circuitry palette dominates: whether it's a robot, a hand holding a glowing "AI" chip, or a holographic office dashboard, the near-default colour grade is deep blue/cyan with white highlights [6, 16, 20, 21, 30, 44].
  • Faceless-content / UGC commerce ads: Fiverr-style gig ads and "Go Viral Without Showing Your Face" tools selling AI avatars, face-swap, and product-demo video generation to marketers and creators [4, 9, 11, 18].
  • Listicle/roundup infographics for tool discovery: "Top 5 Free AI Tools for Photo Generation," "Top AI News of July 2026," a large logo wall of ChatGPT/Gemini/Claude/Microsoft/Google/Apple/Samsung/Meta/Nvidia — all closed-source, big-brand names [30, 36, 42].
  • Named closed-source products actually surfacing: Adobe Firefly (branding collage) [42], Microsoft's VASA-1 audio-to-face research diagram [14], ShengShu Technology's Vidu S1 real-time video model, shown twice in near-identical cyberpunk-woman key art [29, 38], and the Tilly Norwood AI-actress controversy via a TV chyron screenshot [1].
  • Cute AI-generated animal renders as soft marketing bait: an otter in a lifejacket on a paddleboard, a puppy in flowers — used as demo output for photo/video generators rather than "AI news" per se [13, 36].
Notable pins
  • #1 — "Hollywood Reporter Interviews the First AI Actress...": TV screenshot with a bold "AI-GENERATED IMAGE" chyron over the Tilly Norwood/Instagram photo — the clearest real-world closed-source controversy captured in the set.
  • #7 — untitled robot-in-newsroom: android at an anchor desk surrounded by floating "NEWS" panels while a human reporter looks startled — the "AI takes the news desk" anxiety image in its purest form.
  • #47 — "AI content and social media concerns": robot holding a phone next to a "FAKE" stamp over an AI cat photo and a hooded figure — the most direct visualization of the open/closed debate's shared trust problem.
  • #14 — VASA-1 diagram: single image + audio clip + control signals → live facial animation grid — an actual technical figure, not stock art, tied to Microsoft's audio-driven talking-face research.
  • #29 / #38 — Vidu S1 (ShengShu Technology): "The World's Most Advanced Model for Real-Time AI Interaction," cyberpunk key art repeated across two pins — a genuine named closed-source video-model launch out of China.
  • #42 — Adobe branding collage: red Adobe/Firefly wordmark over generated portrait and nature shots — closed-source incumbent staking a claim in the space.
  • #36 — "Top 5 Free AI Tools for Photo Generation": ranks Midjourney, Bing Image Creator, Canva AI, Leonardo AI, and Playground AI with sample outputs — the closest thing to an open/free-tier roundup in the set.
  • #13 — AI-generated otter on a paddleboard: captioned as a "3D digital render," used as demo footage for a video-gen tool — shows how "AI video news" on Pinterest often means product demos, not journalism.
  • #18 — "Go Viral Without Showing Your Face": hooded creator at a dual-monitor face-clone editing rig — packages AI avatar/face tech as a growth-hacking product for creators.
  • #9 — face-swap ad ("Videos Pro"): two women's faces merging via a tracking mesh, sold as "hyperrealistic face editing" — deepfake tooling marketed directly to consumers.
Signals
  • What's current: the dominant Pinterest narrative right now is not "which model shipped what" but anxiety and opportunism around AI faces — deepfakes, AI news anchors, face-swap apps, and "go faceless" content tools all cluster tightly together. The Tilly Norwood AI-actress story [1] and Vidu S1's real-time interactive video model [29, 38] are the two pins that read as genuine current-events news rather than stock concept art or affiliate content.
  • What's largely absent: open-source-specific news is essentially invisible in this pin set. No pin names Stable Diffusion, Flux, HunyuanVideo, Wan, ComfyUI, or any GitHub-native project — the "AI tool" pins that do exist [30, 36, 42] list closed, hosted products (ChatGPT, Gemini, Claude, Adobe Firefly, Midjourney, Bing Image Creator, Canva, Leonardo). Pinterest's incentive structure (SEO blog thumbnails, affiliate listicles, gig-marketplace ads) appears to select for consumer-facing closed products over open-source community releases.
  • Real technical detail is thin: outside of the VASA-1 diagram [14] and the Vidu S1 branding [29, 38], almost everything else is generic stock-photo/clipart illustrating the concept of AI rather than depicting an actual news event, release, or demo.
Limits
  • This stage used pins the worker had already collected with a headless browser (output/pinterest.pins.md, 50 pins under a single unlabeled search/genre); per the playbook I did not browse Pinterest myself, so this reflects one search pass, not an exhaustive crawl.
  • I opened and visually reviewed 30 of the 50 collected images directly; the remaining 20 were not opened individually, though their titles were reviewed and are consistent with the themes above.
  • None of the pin destination pages were fetched — only the pin image and its saved title were available, so context beyond what's visible in-frame (e.g., full article text, comment discussion, save/repin counts) could not be verified.
  • Open-source image/video-gen news is close to a dead end on this platform for this brief: the pin set simply does not surface named open-source projects, so the "open-source trends" half of the requested report will need to lean on the other platform stages (Reddit, X, YouTube, Bluesky, Lemmy) rather than Pinterest.

Recommended actions

  • Recollect X (formerly Twitter) using theme-specific terms such as “Sora 2,” “Nano Banana Pro,” “Wan,” and “Flux,” rather than trend-driven discovery.
  • Verify the suspected closure of Wan 3.0 against primary sources, especially Alibaba’s official announcement, and track the implications for the open-source video-generation landscape.
  • Monitor !«メールアドレス» regularly for the freshest developments around MiniMax H3.
  • Because GPT Image 2.5 and Nano Banana Pro comparisons align across several platforms, next time also examine third contenders such as Seedream 5.0 Pro.
  • Reports of Google-model safety-filter overreactions and faulty outputs are recurring, so keep tracking them as reliability news.

Collected images

Hollywood Reporter Interviews the First AI Actress...New York's Bold Gambit: Inside the Legislative...imageDiscover how Grok is revolutionizing video...The future of creativity: MrBeast on AI’s impact on YouTubeI will create ai ugc video, ai ugc ads, ugc, ai product ads, ai commercial video adsBest 10 AI YouTube Script Writers in 2026AI Ecosystem interconnected Artificial Intelligence Ecosystem"Why AI Chatbots Fail at Keeping Up with Breaking News"The perfect avatar era: The rise of AI influencersTransform your photos and videos! Akool is the AI-powered tool you’ve been waiting for! 🚀The Snarky AI News RoundUp #8imageNew Microsoft AI Tool Can Now Animate Faces from...imageWhat are the Applications of AI in Video Recognition?AI News Anchors & Creator-Led Coverage Take Over U.S. Media!Free AI-Powered Visual Content Generation Platform – Viva AIAI Is Writing the News – Can We Still Trust Journalism?Go Viral Without Showing Your FaceCreate A News Channel Using AI | News Presenter| AI Avatar | Video Generator Free |ChatGPT & CanvaUncategorized Archives - Expert Graphic International: Photo Editing CompanyimageTurn AI art into physical products. Follow for tips!imageWatermarking AI Content: Proving Origin (2025 Challenge)imageAI Draft vs Human Editing: What Makes Content Feel Real?Revolutionize Your Videos with AI: Top 5 Tools!Top AI News of July 2026: Biggest AI Updates, New Features & Industry TrendimagePhotography — Blog — Glyn DewisThe state of AI in the newsroom | Framing the...Looking for latest world news updates? Here are ten global headlines you should knowAITop 5 Free AI Tools for Photo Generation in 2026 | Best AI Image GeneratorsFuturistic Classroom – Learning About Global TechnologyShengShu Technology Unveils Vidu S1, Bringing Real-Time Interactive Generation to AI VideoCreate Amazing Videos in Minutes with Invideo AI! 🚀🎥XIAOMI 12_ Life is a filmthe invisible wardenMarch 16 Cutoff: Firefly Opens Unlimited AI Images and VideosAI Tool To Turn Text Into ImagesAdvancements in AI Facial Recognition for Headshot...GoDaddy “Make a Different Future 2021”I will create ai art video or animated story using aiAI content and social media concernsHow Artificial Intelligence is changing education trends15 Biggest AI Updates You Need to Know | Latest AI News & TrendsReduce Decision Fatigue with a Better Notion System

Data-quality notes

X’s collection method was misaligned with the topic and returned zero relevant items, making it effectively unusable. Reddit and Lemmy did not reach the completion target of “around 20 items.” YouTube could not verify view counts or channel names, and Bluesky had to rely on alternative methods because its search API was unavailable.

Gallery