KEN’S CAT LOG

Image and Video Generation AI News — 2026-10-01

Open source (Qwen-Image-2.1, LTX 2.5) is gaining momentum around ComfyUI, while major closed-source players (Google Nano Banana, OpenAI Sora, xAI Grok Imagine) are drawing attention for misuse, lawsuits, and service shutdowns.

Image and Video Generation AI News — 2026-10-01

Alright, straight to the point: open source is seriously hot right now 🔥 Alibaba’s Qwen-Image-2.1 (native transparency and 2K output despite being only 7B) and the open-source video model LTX 2.5 are spreading rapidly through the ComfyUI community, making local, free, uncensored environments more powerful by the day 😤 Meanwhile, closed-source players (Google’s Nano Banana lineup, OpenAI Sora, Midjourney, xAI Grok Imagine) have polished features, but the bigger stories have been screw-up news: deepfakes, child sexual abuse material misuse, lawsuits, and service shutdowns 🙄 Also, to be honest, we were barely able to collect any data from X or Bluesky that directly addressed the topic this time, so apologies for that 🙏

Across platforms

Three common themes stood out across multiple platforms!

  • Open source means ComfyUI is the common language: Reddit (Vlo 0.3 received 1,055 points in r/StableDiffusion), YouTube (a steady stream of Qwen-Image-2.1 and LTX 2.5 explainers), and Lemmy (Qwen-Image-2.1 workflow derivatives posted daily in !«メールアドレス») all clearly showed that ComfyUI has become the de facto standard platform 💪
  • Closed source is more likely to make news through misuse and backlash: Reddit (SM Entertainment took legal action over deepfake images of K-pop idols, 730 points), Lemmy (Google Earth’s Nano Banana feature was shut down within a day after fabricated disaster images, score 473), Bluesky (404 Media reporter Emanuel Maiberg continues covering non-consensual generated images), and Pinterest (at least 7 out of 50 pins used anxiety-driven visuals about AI replacing journalists). It is genuinely fascinating that four platforms from different genres show the same pattern 🫣
  • Tighter censorship leads users to move local: Reddit (users seeking alternatives such as Seedream and MiniMax H3 after WAN 3.0 strengthened NSFW filters, and recommendations to switch to Midjourney or VeniceAI after complaints about ChatGPT over-censorship) and Lemmy (support for decentralized, free infrastructure such as AI Horde) both clearly reflected the sentiment that “if regulation gets too strict, people flee to open source” 🏃

Platform by platform

Reddit: Collected 12 threads, short of the target of around 20. A K-pop deepfake lawsuit (730 points) and backlash over a Vocaloid music video (327 points) were overwhelmingly strong; ethics and harm-related posts went more viral than technology-trend threads. Both closed- and open-source topics appeared.

X: This was a real mess 😭 Because the search terms reused Croatian trending words (Spain, Tesla, JubJub, and so on) rather than topic-related keywords, effectively none of the 40 collected posts were relevant. The conclusion of the X stage is that it has nothing to say about image and video generation AI.

YouTube: Reviewed 10 videos. Open-source projects Qwen-Image-2.1 and LTX 2.5 spread immediately through primarily English-language channels, while Japanese-language channels focused more on practical comparisons of closed-source tools such as Nano Banana, Flow, and Sora2. It was also interesting to see open- versus closed-source interest split by language.

Bluesky: The official full-text search API (searchPosts) consistently returned HTTP 403, forcing a switch to individual account tracking. On the closed-source side, dfl.bsky.social posts landscape images generated with Nano Banana 2 almost daily; open-source creators such as doctordiffusion had stopped posting in the first half of 2025, leaving few posts that felt current.

Lemmy: Selected 11 representative items after reviewing more than 60. !«メールアドレス» was by far the densest source of open-source technical information, while closed-source coverage centered on controversy-driven stories such as Google Earth backlash and the xAI Grok Imagine CSAM lawsuit. The overall volume comfortably met the completion target.

Pinterest: Reviewed all 50 pins. Most were SEO-oriented stock images or listicle thumbnails, with Lumalabs’ transformation grid about the only genuine product demo. There were no visuals of open-source tools such as ComfyUI or Stable Diffusion UI. Closed-source brand visuals (Veo2, Gemini Omni) stood out, but the same images were also reused across multiple pins.

What to watch

Recommendations

  1. For the next X stage, recollect data using topic-specific keywords such as Midjourney, Sora, Stable Diffusion, Wan, and Kling. This run’s data is unusable.
  2. If Bluesky’s searchPosts continues returning 403, consider another access method such as an authenticated API.
  3. Continue monitoring adoption of Qwen-Image-2.1 and LTX 2.5 as open-source indicators in the next cycle.
  4. Prioritize misuse and backlash stories involving closed-source companies (Nano Banana/Google Earth and the Grok Imagine lawsuit), as they tend to maximize engagement.
  5. Split Reddit searches across multiple queries to broaden collection and reach the target of around 20 items.
  6. Check next time how the closed-source video market shifts after the Sora API shutdown.

Data quality

X returned only topic-irrelevant data because of poor search-term selection, so it effectively had “no data.” Bluesky’s full-text search API was blocked by 403, forcing an alternative account-tracking method that did not reach the target of about 20 posts. Reddit likewise fell short with only 12 items from a single query. Lemmy (more than 60 items reviewed), Pinterest (all 50 pins reviewed), and YouTube (10 videos examined closely) had broadly sufficient volume.

Platform summaries

Reddit

Reddit — Image and Video Generation AI News

Where

The 12 collected threads came from nine subreddits. Member counts and collected items are below (all matched the search term “Image and Video Generation AI News”).

Subreddit Members Items collected
r/kpop 3,893,359 1
r/youtube 3,462,568 1
r/StableDiffusion 1,027,325 1
r/AI_Agents 452,181 1
r/Vocaloid 218,928 1
r/aiwars 170,021 1
r/generativeAI 159,505 4
r/B2BForHire 139,944 1
r/GenAI4all 49,028 1

r/generativeAI had the most posts, with four, concentrated around question threads such as “Which AI tools can I use for free?” Every other subreddit produced one result. Topics where GenAI was used, or where its use caused problems, in film, music, K-pop, and similar areas were prominent.

What people say
  • Thread #1, “Dreamworks has used GenAI since 2023” (r/aiwars, 17 points, 48 comments, 2026-09-30, https://www.reddit.com/r/aiwars/comments/1wty4m6/): Claims that Dreamworks works such as Trolls Band Together and Kung Fu Panda 4 were made with GenAI. The top comment was skeptical: “Nothing you posted says that these movies are 'produced with gen AI'. ... You have no idea how they used AI in these productions.” (u/Majestic-Coat3855, 25 points)
  • Thread #2, “Looking for adult-oriented image-to-video alternatives after DreamFort strengthens filters” (r/generativeAI, 10 points, 8 comments, 2026-09-30, https://www.reddit.com/r/generativeAI/comments/1wtzefv/): WAN 3.0’s NSFW generation suddenly became much stricter, and users sought alternative closed models such as Seedream 2.0 Unfiltered and MiniMax H3.
  • Thread #3, “AI-generated content is killing YouTube” (r/youtube, 107 points, 21 comments, 2026-09-25, https://www.reddit.com/r/youtube/comments/1wqbl4f/): Points to a rapid increase in narrated AI clip compilations. “We have entered the realm of slop.” (u/three-sense, 2 points), “YouTube CEO says he's pro ai slop.” (u/HotUKTakes, 2 points)
  • Thread #4, “There is no such thing as a completely free AI video-generation tool” (r/generativeAI, 55 points, 36 comments, 2026-09-30, https://www.reddit.com/r/generativeAI/comments/1wty6zj/): A reality check that cloud GPU costs are not charity. “videos are not like cheap LLMs, no one can give it away for free” (u/opengradient, 5 points)
  • Thread #5, “What is the best AI image-generation tool if free use is the top priority?” (r/generativeAI, 4 points, 54 comments, 2026-09-26, https://www.reddit.com/r/generativeAI/comments/1wqfleb/): Recommendations included Yodayo/MoescapeAI (50 credits per 24 hours) and Leonardo (daily free tokens).
  • Thread #6, “Hiring an AI video creator / AI content manager” (r/B2BForHire, 2 points, 5 comments, 2026-09-25, https://www.reddit.com/r/B2BForHire/comments/1wpx42u/): A job listing seeking practical experience with hybrid AI-plus-traditional and fully AI-generated video. A sign that AI video production is already becoming a marketable profession.
  • Thread #7, “AnythingBecomeMoe used GenAI in new MV ‘DisplayHolic’” (r/Vocaloid, 327 points, 105 comments, 2026-09-28, https://www.reddit.com/r/Vocaloid/comments/1wsa4nz/): Opinion was divided. “I don't understand when incredibly talented people embrace ai” (u/AungCowMyat, 297 points) versus “at least they're disclosing it... far better usage than some others who just straight up reject the fact that they used AI” (u/memorie_desu, 200 points)
  • Thread #8, “SM Entertainment takes legal action against AI-generated images targeting Lim Yoona and aespa’s Winter” (r/kpop, 730 points, 52 comments, 2026-09-28, https://www.reddit.com/r/kpop/comments/1wsdr6d/): Strong concern around deepfakes. “the deepfakes are so concerning. glad sm is taking action :(” (u/llamacana, 167 points)
  • Thread #9, “Vlo 0.3 — an open-source, AI-compositing-focused video editor/generator” (r/StableDiffusion, 1,055 points, 71 comments, 2026-09-28, https://www.reddit.com/r/StableDiffusion/comments/1wsmat4/): Well received for workflows integrating Minimax H3, Qwen2.1, Krea2, and LTX2.5. “Great ad, same as before :) I forgot about this app, I will Def try it.” (u/Maskwi2, 22 points)
  • Thread #10, “AI videos that went viral in 2024 look wild now when you see how far quality has progressed” (r/GenAI4all, 199 points, 66 comments, 2026-09-24, https://www.reddit.com/r/GenAI4all/comments/1wou5vb/): A nostalgic thread looking back on quality improvements in just a few years, with many responses echoing the title’s “the leap in quality since then is insane.”
  • Thread #11, “I want an AI image-generation tool without guardrails” (r/AI_Agents, 0 points, 22 comments, 2026-09-29, https://www.reddit.com/r/AI_Agents/comments/1wt7jqr/): Frustration over ChatGPT’s perceived over-censorship led to many recommendations to move to Midjourney, VeniceAI, or Chinese open-source models.
  • Thread #12, “Are there free AI tools that can make animated videos from scripts?” (r/generativeAI, 4 points, 23 comments, 2026-09-27, https://www.reddit.com/r/generativeAI/comments/1wrip4m/): “Minimax H3 with comfyUI” (assuming a local setup, u/Amonsul_, 4 points) was the leading answer. The conclusion was also that quality ultimately requires paying.
Signals
  • Upward trend: Multiple threads (#2, #11, #12) consistently show a desire to shift to open-source/local models (ComfyUI, Minimax H3, SDXL variants) in response to strengthened closed-source filtering (#2). “If you want to avoid censorship, local is the only option” has become a stock conclusion.
  • Signs of professionalization: The hiring thread (#6) and the Dreamworks thread (#1) suggest that GenAI is moving beyond hobby use and being incorporated into the practical infrastructure of film and content production.
  • Claims being dismissed: Expectations for “completely free, unlimited, watermark-free video-generation tools” (#4, #5) are dismissed by the community itself as unrealistic. Replies emphasizing the reality of cloud GPU costs dominate the top comments.
  • A divided debate: Threads #1 and #7 reveal disagreement over whether disclosing GenAI use is enough. In the Vocaloid thread, both the defense that disclosure is at least more honest and the criticism that it is disappointing when talented people rely on AI appear among top comments. A simple pro-versus-con binary does not explain the response.
  • Unexpected finding: The highest-scoring discussions were the legal-action thread about malicious AI images of K-pop idols (#8, 730 points) and the Vocaloid music-video backlash thread (#7, 327 points), not tool-introduction posts (#5, #12, and others). AI-related harm and ethics attracted far stronger responses than tool trends. Interest in misuse and likeness-rights disputes is easier to make visible than interest in technology trends themselves.
Limits
  • The brief’s completion criterion was analysis of around 20 posts per platform, but only 12 threads were collected. There was only one search phrase, “Image and Video Generation AI News”; no additional searches separated “closed source” and “open source.” Because this agent simply read the worker’s collected results, it did not conduct additional searches.
  • There is no record in threads.md of failed or empty searches. All 12 threads were returned by this one query, so the shortfall may stem from the query being too narrow.
  • The opening text was not recorded for threads #7 (Vocaloid), #8 (kpop), #9 (StableDiffusion), and #10 (GenAI4all); assessment is based only on titles and top comments.

X

X — Image and Video Generation AI News

##Important note about the collected data

After reviewing output/x.posts.md, the search terms used for this X collection were:
“Spain,” “Tesla,” “JubJub,” “Holy,” “Italy,” “#bb28,” “Elon,” “Croat,” “Donald,” and “Roman.” These were not the result of searches for the research theme (closed-source/open-source image and video generation AI), but rather X Explore trend terms listed in the “What X says is happening” table at the beginning of output/x.posts.md (trends visible from a session geolocated to Croatia) reused directly as search terms. As a result, the 40 collected posts from 37 accounts mostly covered unrelated topics such as football (Spain versus Croatia), Tesla driver-assistance features, WWE, Big Brother (#bb28), and Donald Trump’s political statements.

The following is an honest report based on the data actually collected, but it should be stated upfront that it does not represent trends in image or video generation AI.

Accounts

Most of the 37 collected accounts were one-off contributors with a single post, and no topic-relevant publishers were found. Only the following posted multiple times, and all were off-topic:

  • @ryu15 (Ryu) — 2 posts, personal technical notes about running VRChat on a “Tesla V100”; this matched Tesla incidentally as a GPU model name, not autonomous driving
  • @grouchoskie (Grouch Oskie) — 2 posts, everyday remarks
  • @reigns_era (Roman Reigns SZN) — 2 posts, WWE fan posts related to Roman Reigns

No account in this dataset posted consistently about image or video generation AI.

Posts

Only a handful of posts included terms somewhat close to the topic, such as “AI” or “generation,” and none concerned image/video generation AI itself. They are listed for reference.

  1. @rammcodes (Ram Maheshwari) — 276 likes, 51 reposts, 4 replies, about 15,000 views, 2026-09-29
    https://x.com/rammcodes/status/2104962267267944848
    “Holy Sh*t... this is like PDF parsing on steroids... A free, open-source PDF parser that converts PDFs into AI-ready Markdown, JSON, and HTML.”
    → An introduction to an open-source PDF parser, a different field from image/video generation AI: document parsing and tools for RAG.

  2. @LuisBizarro (Luis Bizarro) — 930 likes, 48 reposts, 37 replies, about 52,000 views, 2026-09-29
    https://x.com/LuisBizarro/status/2104780688113455342
    “Another AI slop and vibe coded experiment... based on Neon Genesis Evangelion. This took me like 45 minutes and three prompts. This is all Three.js and WebGL.”
    → An interactive 3D experience “vibe coded” with AI using Three.js/WebGL. This concerns code-generation AI, not image or video generation models.

  3. @thomasunise (Thomas Unise) — 732 likes, 48 reposts, 43 replies, about 111,000 views, 2026-09-29
    https://x.com/thomasunise/status/2105003307232059397
    “Elon is diabolical OpenAI drops dots https://Dot.com goes to Grok Bot”
    → A post about a domain-related jab between OpenAI and Grok, not technical news about generative AI models.

  4. @LinusEkenstam (Linus ✦ Ekenstam) — 2,640 likes, 408 reposts, 354 replies, about 344,000 views, 2026-09-30
    https://x.com/LinusEkenstam/status/2105212016583422270
    “This video is going Viral everywhere. A trillionaire and a pile of billionaires are unable to answer a simple question about regular laptop jobs...”
    → A post featuring AI-industry figures such as Dario, but its focus is a discussion of employment, not technical developments in video generation AI.

The other 36 posts — football between Spain and Croatia, complaints about Tesla autonomous driving, fan posts about a pet/character called JubJub, Big Brother #bb28, WWE, and Trump political comments — had no connection whatsoever to image or video generation AI.

Signals
  • No clear topic-related signal was detected. None of the 40 collected posts mentioned image or video generation AI, whether closed source (for example Midjourney, Sora, Runway, Veo, Flux Kontext) or open source (for example Stable Diffusion, Wan, HunyuanVideo, ComfyUI).
  • The sole useful signal concerns the collection process itself: searches reused region-specific X Explore trends for Croatia rather than terms connected to the research brief. This collection-logic failure is the largest finding of this run.
Limits
  • The biggest constraint was that the search terms were unrelated to the topic. The ten terms — “Spain,” “Tesla,” “JubJub,” “Holy,” “Italy,” “#bb28,” “Elon,” “Croat,” “Donald,” and “Roman” — did not originate from the image/video generation AI theme, but were X Explore trend terms (from a Croatia-geolocated session) used as-is. As a result, effectively zero of the 40 posts were relevant.
  • Topic-specific search terms such as “Midjourney,” “Sora,” “Stable Diffusion,” “Runway,” “Flux,” “Veo,” “Kling,” and “HunyuanVideo” were never searched. This X run effectively says nothing about the theme.
  • The Explore trends table itself (Spain / Tesla / JubJub / Holy / Italy / #bb28 / Elon / Croat / Donald / Roman) was also limited to a Croatian-server viewpoint and should not be treated as a global trend.
  • In conclusion, no valid insight about closed-source/open-source image and video generation AI trends on X can be reported from this dataset. A recollection run needs theme-appropriate search terms, including specific model, tool, and company names.

YouTube

YouTube — Latest Image and Video Generation AI News (Closed Source/Open Source)

Channels
  • AI WITH Rithesh (@RitheshSreenivasan) — An international AI-tools channel that tested Qwen-Image-2.1 on the day it launched.
  • Daily AI Roundup (@ai_roundup) — A daily AI-news roundup channel that introduced Qwen-Image-2.1’s release in a news format.
  • Veteran AI (@VeteranAI-y1r) — A channel focused on ComfyUI workflows, covering LTX 2.5 and Qwen Image 2.1 in succession.
  • MaxonShire (@MaxonShire) — A local AI video-generation comparison channel focused on HunyuanVideo and the Wan series.
  • 動画編集の中の人 — Japanese-language channel with strong paid-tool comparisons for image generation AI.
  • AI大学【AI&ChatGPT最新情報】 (@AIAIChatGPT-cj4sh) — Japanese-language channel mainly providing beginner-friendly explainers and guides to AI tools.
  • AIで創作効率化とマネタイズby呪喝蛙 (@AI_aaafrog) — Japanese-language channel covering ComfyUI models and monetization.
  • TBS NEWS DIG — TBS’s official news channel, reporting on Sora 2 in a broadcast-news (news23) format.

Subscriber counts were not visible until after JavaScript rendering on the video pages. They could not be confirmed from the oEmbed/search results collected in this run (see Limits).

Videos
  1. “Qwen-Image-2.1 Is INSANE… Just 7B Parameters?!” — AI WITH Rithesh / late September 2026 (immediately after Qwen-Image-2.1’s 2026-09-20 release) / https://www.youtube.com/watch?v=3qGl0wiv2G0 — Introduces Alibaba’s 7B open-weight model, which supports transparent PNG generation, up to 10 reference images, and native 2K output, as competing with large closed-source models.

  2. “A 7B Model That Beats Nano Banana 2.0? Master Qwen Image 2.1: Native 2K, RGBA & 10-Image Workflows” — Veteran AI / late September 2026 / https://www.youtube.com/watch?v=fwHS0ubmI7I — A ComfyUI workflow explainer for Qwen-Image-2.1. Built around a claim that it beats the closed Nano Banana 2.0 in proprietary benchmarks (also described in Tom’s Hardware reporting as Qwen Image Benchmark 60.2 versus Nano Banana 2.0’s 59.82).

  3. “Alibaba Open-Sources Qwen-Image-2.1, A 7B Model With Native Transparency” — Daily AI Roundup / late September 2026 / https://www.youtube.com/watch?v=cdZ4E7_kiJY — A breaking-news-style introduction to Qwen-Image-2.1’s open-sourcing. Native transparency (alpha-channel) support is the headline feature.

  4. “Master LTX 2.5 in ComfyUI: Multi-Shot, First-Last Frame & 2K Workflows” — Veteran AI / mid-August 2026 (shown in search results as about 40–48 days old) / https://www.youtube.com/watch?v=6B9WSxi6j8E — A practical guide to using the fully open-source video model LTX 2.5 in ComfyUI, covering multi-shot, first/last-frame controls, and 2K-output workflows.

  5. “LTX 2.5 is a New Free and Opensource AI Video Model that is Insanely Fast in Generation” (Shorts) — mid-August 2026 / https://www.youtube.com/shorts/h8qy-MBDzho — A short introduction to LTX 2.5’s fast generation speed.

  6. “使わにゃ損損!お得に生成するならComfyUI一択!2026年おすすめ生成モデルベスト!” — AIで創作効率化とマネタイズby呪喝蛙 / 2026 (shown in search results as about 172 days old, somewhat older) / https://www.youtube.com/watch?v=qmiPya-U2KM — A roundup of recommended open-source ComfyUI models, including image, video, music, voice, and 3D categories.

  7. “【ドS課金して判明】2026年おすすめ画像生成AIはコレ!NanoBananaPro・Seedream4.0・Midjourney・AdobeFireflyを徹底比較します!” — 動画編集の中の人 / 2026 (recent; exact date unconfirmed) / https://www.youtube.com/watch?v=vJLDbXaSKW4 — A paid hands-on review comparing four closed-source services: Nano Banana Pro, Seedream 4.0, Midjourney, and Adobe Firefly.

  8. “【画像生成AIの頂点】Google最新の画像生成AI「Nano Banana 2(ナノバナナ)」を無料で使い倒す!最短スタート完全ガイド!” — AI大学【AI&ChatGPT最新情報】 / 2026 (Nano Banana 2 itself launched 2026-02-26 to 27; the video continued to attract views afterward) / https://www.youtube.com/watch?v=Asuw6jHVcSs — A beginner-focused guide to getting started with Gemini 3 Pro-based Nano Banana 2 (Pro) for free.

  9. “【Google Flow完全ガイド】無料で使えるAI映像制作ツールが超進化!Gemini Omni Flash・ナノバナナを使って動画や画像を作れる!” — AI大学【AI&ChatGPT最新情報】 / after May 2026 (mentions the addition of Gemini Omni Flash) / https://www.youtube.com/watch?v=dBFK1yFJXOA — Explains the latest developments in Google’s closed-source integrated tool Flow, including its ability to handle Nano Banana image generation and Veo-style video generation within one tool.

  10. “最新動画生成AI「Sora2」リアルな映像が約3分で生成可能に…実演でわかったその性能 一方で著作権への懸念も【news23】” — TBS NEWS DIG / 2026 (a terrestrial-TV news feature on Sora 2) / https://www.youtube.com/watch?v=_r9hojDlABM — A news report demonstrating closed-source Sora 2 while raising copyright concerns. Separate research indicates that OpenAI announced Sora’s shutdown on 2026-03-24, ended the app/web version on April 26, and planned to end the Sora 2 API on 2026-09-24. This shows that Sora itself moved toward closure roughly half a year after the news segment.

Signals
  • Open source is gaining momentum: Qwen-Image-2.1 (Alibaba, released 2026-09-20; 7B parameters with native transparency and up to 10 reference images) was covered immediately by multiple channels. Proprietary benchmarks claiming that it matches or exceeds a major closed-source competitor, Nano Banana 2.0, became a focal point for its spread. Comparison videos for locally runnable free video models such as LTX 2.5, HunyuanVideo, and Wan continue to appear, with ComfyUI serving as the de facto shared platform.
  • Closed-source developments: Google products (Nano Banana 2 / Nano Banana Pro, Gemini Omni Flash, and Flow) are repeatedly featured on Japanese-language channels. By contrast, OpenAI’s Sora attracted attention early in 2026 before moving through a phased shutdown from March to September (web/app end → API end), giving discussions of major closed-source companies a boom-and-bust framing: strong immediately after launch, but with doubts about longevity.
  • Language distribution: Japanese channels (AI大学, 動画編集の中の人, 呪喝蛙) more often cover practical comparisons and guides for closed-source products, while English-language channels (Veteran AI, AI WITH Rithesh, Daily AI Roundup, MaxonShire) more often cover open-source ComfyUI workflows and benchmark-breaking news.
Limits
  • YouTube video and search pages (/watch, /results) returned only static HTML through WebFetch, mainly footer navigation links. Dynamically rendered data such as views, likes, subscriber counts, and comments could not be collected directly.
  • Instead, the YouTube oEmbed API (/oembed?url=...&format=json) was used to supplement titles and channel names, while view counts and dates relied on web-search snippets with relative expressions such as “days ago.” Some posting dates are therefore approximate, including “about 40–48 days ago” and “about 172 days ago,” rather than precise dates.
  • Some video IDs, such as yVMs1Z-FWmQ, the Russian-language 03MX2U8ZLDk, and the Spanish-language jH0AjVP49KI, appeared in search results but were not reviewed in detail because of lower language/relevance priority.
  • Top-comment content could not be obtained because WebFetch did not render the comments section.

Bluesky

Bluesky — Image and Video Generation AI News

Note on collection method

The playbook’s https://public.api.bsky.app/xrpc/app.bsky.feed.searchPosts?q=... endpoint was tested with several queries, including Midjourney, "Stable Diffusion", and cat, but returned HTTP 403 every time and could not be used. The web version at bsky.app/search is also an SPA rendered with JavaScript, so its content could not be retrieved. By contrast, app.bsky.actor.getProfile, app.bsky.feed.getAuthorFeed, and app.bsky.unspecced.getTrendingTopics worked normally. The research therefore used WebSearch to identify accounts likely to be relevant, then read each account’s getAuthorFeed directly. Details are in ## Limits.

Accounts
  • dfl.bsky.social (DFL inc.)
    https://bsky.app/profile/dfl.bsky.social
    An individual account posting photo-like Japanese landscape images generated with Google’s “Nano Banana 2” almost daily. It remained actively used in September 2026 and was the most consistently observable real-world user of a closed-source model.

  • doctordiffusion.bsky.social
    https://bsky.app/profile/doctordiffusion.bsky.social
    An open-source creator who makes and distributes LoRAs and models for Stable Diffusion 3.5, Flux Dev, HunyuanVideo, and LTX-Video through Civitai and Hugging Face. However, posts stopped between November 2024 and March 2025, with no 2026 activity confirmed.

  • emanuelmaiberg.bsky.social (404 Media reporter Emanuel Maiberg)
    https://bsky.app/profile/emanuelmaiberg.bsky.social
    Continues reporting on and posting about misuse of image/video generation AI, including deepfakes, non-consensual generated images, and celebrity impersonation posts made with AI music. Posting continued through August and September 2026, making it one of Bluesky’s most consistently relevant accounts on the topic.

  • midjourneyofficial.bsky.social (Midjourney official)
    https://bsky.app/profile/midjourneyofficial.bsky.social
    Announced its arrival on Bluesky in August 2025, but its most recent post was October 8, 2025; no posts from 2026 were confirmed. This is an example of weak Bluesky activity from official closed-source companies.

  • stablediffusion.bsky.social (“Stable Diffusion online AI”)
    https://bsky.app/profile/stablediffusion.bsky.social
    An AI-art community account with around 1,800 followers. Posts stopped after February 2025.

  • wagahai-ai.bsky.social
    https://bsky.app/profile/wagahai-ai.bsky.social
    An individual Japanese-language account posting AI illustrations and AI video. This account likewise had no posts after January 2025.

Posts
  1. @dfl.bsky.social — 2026-09-30, 14 likes, 1 repost
    https://bsky.app/profile/dfl.bsky.social/post/3mwr727gyck26
    “AI image generation #Nano Banana 2 ... a canal touched by golden reflections.” A Japanese canal landscape generated with Google’s closed-source image model Nano Banana 2.

  2. @dfl.bsky.social — 2026-09-25, 51 likes, 2 reposts (the highest engagement among the latest 20 posts)
    https://bsky.app/profile/dfl.bsky.social/post/3mwesrcxg5k2c
    “#Nano Banana 2 ... a sweeping riverside view, stone embankments, and Japanese homes in the distance.”

  3. @dfl.bsky.social — 2026-09-21, 45 likes, 2 reposts
    https://bsky.app/profile/dfl.bsky.social/post/3mw2ny6hsa22p
    “#Nano Banana 2 ... an old wooden bridge, a creek, and a cool twilight.” The daily-posting pace and stable response in the 30-to-50-like range were confirmed.

  4. @doctordiffusion.bsky.social — 2025-01-19, 3 likes, 4 replies
    https://bsky.app/profile/doctordiffusion.bsky.social/post/3lg36y7u2i22x
    “SD3.5 Lives! Absynth 2.0 released.” A free release of a fine-tuned Stable Diffusion 3.5 Large model through Civitai/Hugging Face, a grassroots open-source release.

  5. @doctordiffusion.bsky.social — 2025-01-01, 1 like, 2 replies
    https://bsky.app/profile/doctordiffusion.bsky.social/post/3leofdsx6yc23
    Shares the technical finding that training became stable through text-encoder alignment between Flux Dev and Stable Diffusion 3.5, tagged #opensource.

  6. @doctordiffusion.bsky.social — 2024-11-26, 4 likes
    https://bsky.app/profile/doctordiffusion.bsky.social/post/3lbssdeocz226
    “Extending video in ComfyUI using LTX-Video Img2Vid and custom nodes.” A practical example combining open-source video generation tools LTX-Video and Mochi in ComfyUI.

  7. @emanuelmaiberg.bsky.social — 2026-08-20, 14 likes, 6 reposts
    https://bsky.app/profile/emanuelmaiberg.bsky.social/post/3mtjegiaxpk2h
    A 404 Media report on how non-consensual AI-generated images are becoming increasingly subtle and harder to notice.

  8. @emanuelmaiberg.bsky.social — 2026-08-17, 36 likes, 9 reposts
    https://bsky.app/profile/emanuelmaiberg.bsky.social/post/3mtbvku5wds2p
    A podcast with deepfake researcher Hany Farid discussing the past, present, and future of synthetic media.

  9. @midjourneyofficial.bsky.social — 2025-08-15, 41 likes, 4 reposts, 14 replies
    https://bsky.app/profile/midjourneyofficial.bsky.social/post/3lwi222j4sk2b
    “#Midjourney is now on BlueSky! Who are the best creators we should follow?” The official account’s first post on Bluesky. It posted only sporadically afterward, through October 2025.

Signals
  • AI commentators on Bluesky focus more on LLMs and agents than image/video generation. Recent September 2026 posts from AI commentators with around 10,000 followers, such as Tim Kellogg (@timkellogg.me) and Ethan Mollick (@emollick.bsky.social), centered on GPT-6.1, Gemini 4, and Claude’s agent features, with zero mentions of image/video generation models such as Midjourney, Stable Diffusion, Sora, Wan, or Kling.
  • Closed-source examples appear through routine use by individuals, not official companies. Midjourney’s official account is nearly silent, while accounts like dfl.bsky.social quietly continue posting work made with Google’s Nano Banana 2 every day.
  • The open-source maker community — LoRA distribution and ComfyUI-workflow sharing — appears to have cooled on Bluesky. doctordiffusion, wagahai-ai, and stablediffusion all stopped updating in the first half of 2025, suggesting late 2024 to early 2025 may have been their last active period.
  • The most consistently active image/video generation AI content comes from criticism and journalism. 404 Media reporter Emanuel Maiberg continued posting about deepfakes and non-consensual generated images in August and September 2026, making it the most current-feeling content on the topic.
  • From this, the biggest finding is that Bluesky is not a primary source of active discussion for this topic: news about closed- and open-source image/video generation AI. Like X, it had few current posts directly tied to the theme.
Limits
  • app.bsky.feed.searchPosts (the public search API) always returned HTTP 403. It returned 403 for “Midjourney,” "Stable Diffusion", and cat, suggesting the endpoint was blocked regardless of the q value. app.bsky.unspecced.searchPostsSkeleton returned HTTP 501. By contrast, getProfile, getAuthorFeed, and getTrendingTopics worked normally. In other words, Bluesky’s read-oriented endpoints were available, but its full-text search endpoint was blocked from this environment.
  • The web version of bsky.app/search is also a JavaScript-rendered SPA, so WebFetch could not retrieve post text and returned only title strings.
  • As a result, keyword-based post collection as specified in the playbook could not be conducted. The research switched to discovering accounts through WebSearch and reading individual getAuthorFeed streams. This accurately reads recent posts from discovered accounts, but cannot provide a comprehensive cross-platform count of current posts on the topic.
  • WebSearch indicated that sungkim.bsky.social and ekazakos.bsky.social had discussed MiniMax H3, an open-weight video model. However, the relevant posts could not be found when scrolling their getAuthorFeed streams within the latest 30 items, likely because they dated from late July 2026 and were older than the latest 30 posts by late September. They were omitted in favor of the rule of including only content actually read in this run.
  • doctordiffusion.bsky.social, wagahai-ai.bsky.social, and stablediffusion.bsky.social were the most directly relevant accounts found, but all stopped activity in the first half of 2025 and cannot be treated as current news as of September–October 2026. The only substantially active topic-related accounts identified were dfl.bsky.social (closed-source real-world use) and emanuelmaiberg.bsky.social (critical journalism).
  • The target of analyzing around 20 posts was not reached. With full-text search unavailable, only a little more than a dozen relevant posts were identified, half of which dated from 2024–2025. The honest conclusion is that Bluesky has less to say about this topic than the other platforms.

Lemmy

Lemmy — Image and Video Generation AI (Closed/Open Source) News

Research date: 2026-10-01. The Lemmy API search (/api/v3/search) and community post-list APIs were used to search across communities such as stable_diffusion, imageai, fuck_ai, and technology, as well as keywords including FLUX.2, Qwen-Image, Nano Banana, Midjourney, Grok Imagine, Sora, Veo, and Kling.

Communities
  • !«メールアドレス» — 5,713 subscribers / 1,423 posts. The most active Lemmy community for technical information about image and video generation AI, including ComfyUI workflows, LoRAs, and new models. Announcements for AI Horde, a decentralized free image-generation project operated by dbzer0, also appear here.
  • !«メールアドレス» — 8,532 subscribers / 4,680 posts. A place to post AI-generated images regardless of model. It centers on artwork rather than news, and many posts have negative scores, illustrating Lemmy’s generally cool reception toward AI-generated work.
  • !«メールアドレス» — 8,333 subscribers. A community sharing industry-critical articles from an anti-AI perspective. Closed-source backlash and lawsuit stories tend to gather here.
  • !«メールアドレス» — A general technology community. Major closed-source news, such as the Google Earth incident, earned their highest scores here.
  • !«メールアドレス» — 53 subscribers. A bot community mechanically reposting RSS feeds from Reddit’s r/ArtificialIntelligence, so it should not be considered original Lemmy discussion.
Posts

Open-source leaning (all from !«メールアドレス»)

  1. Qwen-Image-2.1 in ComfyUI: Open-Weight Image… — 2026-09-22, score 13.
    A blog.comfy.org article on ComfyUI support for Alibaba’s open-weight image model Qwen-Image-2.1. This was the highest-scoring post in the survey.
  2. WebUI for qwen-image-2.1 — 2026-09-27, score 2. A Fooocus-style lightweight WebUI project adding support for Qwen-Image-2.1.
  3. Few-step LoRA adapters for Qwen-Image-2.1 — 2026-09-26, score 0. LoRAs that speed generation by reducing the number of steps.
  4. Long, seamless multi-shot video continuations — 2026-09-29, score 4. A video-generation workflow using MiniMax-H3-Latent to connect multiple shots seamlessly.
  5. Free, local, open source video editor with AI — 2026-09-29, score 5. The local-first AI video-editing tool “vlo.”
  6. LanPaint — training-free inpaint for every model — 2026-09-29, score 8. A technique that adds high-quality inpainting to existing models without additional training.
  7. Tiny 100M-parameter text-to-image model (Supra2-IMG) — 2026-09-22, score 7. An ultra-lightweight text-to-image model that runs with low VRAM.
  8. AI Horde — New Interface & Documentation — 2026-09-19, score 8. An announcement of a UI/documentation refresh for dbzer0’s decentralized free image-generation service.

Closed-source leaning

  1. Google Earth's new AI image generation function didn't survive a day after users deepfaked disasters — 2026-08-01, score 473 (!«メールアドレス»). Google Earth’s Nano Banana (Gemini)-powered AI feature was shut down in less than a day after misuse to fabricate disaster images. One of the highest-scoring posts observed in the entire survey.
  2. Tennessee Teens Sue Elon Musk's xAI Over Child Sexual Abuse Images — 2026-03-17, score 13 (Mother Jones community). A lawsuit concerning child sexual abuse images generated by xAI’s Grok Imagine. It was claimed that roughly 23,000 such images were generated over 11 days.
  3. Disney, Universal sue image creator Midjourney for copyright infringement — 2025-06-11, score 88 (!«メールアドレス»). Somewhat older, but Midjourney’s copyright lawsuit continues to be mentioned as background news in an ongoing case.
  4. Same Berserk spread, 4 AI colorizations. Looks like GPT wins again? — 2026-09-14, score 1 (via !«メールアドレス»). A mirrored Reddit thread comparing automated manga colorization across multiple models, including Nano Banana Pro. Note that it is not an original Lemmy post.
Signals
  • The open-source side is all about the ComfyUI ecosystem. Derivative workflows, LoRAs, and node extensions for FLUX.2 (Black Forest Labs) and Qwen-Image, especially 2.1, are posted daily, making this by far Lemmy’s densest AI-generation area in terms of posting frequency and technical detail. Mentions of free decentralized infrastructure such as dbzer0’s AI Horde also align closely with Lemmy’s anti-commercial, self-reliant culture.
  • Video generation remains niche within a niche. Video-related posts about MiniMax-H3 and homegrown local video editors are fewer than image-generation posts and have modest scores, around 4–8 points.
  • Closed-source news enters through controversy and lawsuits, not feature introductions. Google Earth (Nano Banana), Grok Imagine, and Midjourney all entered !technology or !fuck_ai during misuse/litigation moments rather than for feature announcements themselves, earning high scores. The Google Earth post (score 473) stood out even across the entire research set.
  • The overall attitude toward AI-generated content is cold. Nearly half the posts in !«メールアドレス» had negative scores (-18 to -3), meaning the mere fact that something was AI art was itself treated as a demerit. The existence of !fuck_ai, with 8,333 members, also says much about the atmosphere of the Lemmy community.
Limits
  • Lemmy is relatively small, and no closed-source community specifically focused on image/video generation AI was found. There is an open-source-focused !stable_diffusion community, but no symmetrical closed-source equivalent. Closed-source insights are therefore limited to general communities such as !technology and !fuck_ai.
  • Searches for “Veo 3” and “Kling AI” found no substantive Lemmy posts, only irrelevant results. Sora-related results contained only one incidental mention in a meme rather than discussion of the video model itself.
  • Direct retrieval of the lemmy.world/c/stablediffusion community page failed with a 500 error, so API-search information was used instead. Requests to lemmy.today/api/v3/community also returned 404, and it took time to determine that the relevant home instance was lemmy.dbzer0.com.
  • !«メールアドレス» is a 53-member bot community reposting Reddit RSS feeds and is not original Lemmy discussion. Only one item was included as reference.
  • No Japanese-language Lemmy communities or posts were found within the searched scope.
  • This stage focused solely on Lemmy. Although the completion target was around 20 posts per social network, more than 60 posts/comments were reviewed through API searches and community listings before the 11 representative entries above were selected.

Pinterest

Pinterest — Image and Video Generation AI News

The worker collected 50 pins for the query “Image and Video Generation AI News” (no genre split; all were in one batch). All 50 were opened and reviewed.

Visual themes
  • Chrome/white humanoid android as the “face of AI” — glowing blue eyes, exposed metal joints, sometimes in a suit. Used for everything from generic explainers to news-anchor and ethics content. [1, 16, 19, 20, 24, 29, 30, 35, 45]
  • Dark navy/black background with neon cyan-purple circuitry glow — the default palette for AI-explainer and listicle thumbnails; brain-with-circuits or glowing “AI” wordmark as the centerpiece icon. [2, 9, 14, 21, 28, 33, 40, 42, 50]
  • “AI is replacing journalists / fake news” anxiety framing — AI anchors at news desks, “FAKE” stamps, deepfake-detection UI, real broadcaster logos (NBC) watermarked “AI GENERATED.” [1, 6, 23, 31, 35, 36, 44]
  • Human-vs-AI dichotomy graphics — split-screen hands (human fingers vs. robot fingers), magnifying glass over a keyboard, “Human or AI?” captions. [30, 31]
  • Dystopian/villain AI imagery — a horned, red-eyed “AI” demon puppeteering phone-zombies; used for fear-bait captions (“AI is getting scary good”). [17, 23]
  • Tool listicles and comparison charts — named real products (Runway, Pika, Kling AI, Luma/Dream Machine, Canva AI, Hailuo, Midjourney, Adobe Firefly, Recraft, Kaiber, Ideogram, SeaArt, Krea) laid out as “best paid vs. best free” or “6+ best websites” infographics. [5, 15, 40]
  • Actual product-demo screenshots — step-by-step UI mockups for image-to-video workflows (upload → prompt → motion sliders → render) branded to small SaaS tools (JoyFun AI, Blueticks). [8, 10]
  • Photoreal/stylization transformation demos — a single source photo turned into wooden-block, origami, Lego, flower, and toy-brick versions side by side (Lumalabs), plus a composited marble-statue-playing-cello image and a fisherman portrait credited to AMD/“Amuse 3.0.” [4, 15, 38]
  • Big-name model marketing collages — Google Veo 2 (the same official promo collage reused across two unrelated pins), Gemini Omni “turn anything into video,” Meta “Muse Spark” 2026 model. [17, 24, 32, 3]
Notable pins
  1. AI content and social media concerns [1] — robot scrolling a phone next to a “FAKE” stamp and a suspicious cat photo; classic deepfake-anxiety bait.
  2. Meta Unveils Muse Image and Muse Video Generative... [3] — pitches a 2026 “Muse Spark” model as a personal-assistant/“superintelligence” play, not just a content tool.
  3. (untitled, Lumalabs transform grid) [4] — one source video turned into wooden-block, origami, Lego-brick and flower versions of a person, a dog, and a car; the clearest actual before/after capability demo in the set.
  4. Create Videos From Images [8] — full 6-step UI mock (upload → prompt → AI processing → motion sliders → render → output) for a tool called “JoyFun AI.”
  5. Transform your ideas into stunning visuals! [15] — a hyperreal marble statue playing a cello, used as an image-generation showcase image.
  6. (untitled, Veo 2 promo) [17] — Google’s official “Veo 2: our state-of-the-art video generation model” launch collage (dog swimming, flamingos, animated girl); the identical image is reused verbatim under pin 32’s different caption.
  7. Google’s Gemini Omni Turns Anything Into Video [24] — “any input → any output” framed as Google’s big generalization claim for video generation.
  8. ‘Amuse 3.0’, an AI art creation tool that includes... [38] — photoreal portrait of an old fisherman, AMD-branded, suggesting a hardware-accelerated local image-gen tool angle.
  9. Introducing Luma Dream Machine [43] — extreme close-up of an eye over “Dream Machine” wordmark, pure brand imagery rather than a product shot.
  10. AI NEWS: May 20 2026 Google I/O recap infographic [35] — Gemini 3.5 Flash, Gemini Omni, Gemini Spark, Android XR glasses, framed as “biggest AI event” — a dated, concrete news digest rather than evergreen listicle content.
Signals
  • The set skews heavily toward SEO/blog-thumbnail content (stock-photo robots, generic “AI News” wordmarks, tool listicles) rather than genuine generated-art showcases — these dominate by volume over real demo content.
  • Deepfake/misinformation anxiety and “AI replacing journalists” is a clear recurring narrative thread, not a one-off — it shows up in at least 7 of the 50 pins, often paired with real broadcaster branding (NBC) or “FAKE” stamps.
  • Real, named tools do surface repeatedly across listicles and demos: Veo 2/Gemini Omni (Google), Dream Machine/Luma, Runway, Pika, Kling AI, Canva AI, Hailuo, Amuse 3.0, Meta Muse Spark — but mostly as logos/marketing collages, rarely as hands-on workflow screenshots (JoyFun AI and Blueticks are the two exceptions).
  • Absent: no visible open-source model names or local-tooling screenshots (no ComfyUI, Automatic1111, Stable Diffusion UI, or checkpoint/LoRA talk) despite “Stable Diffusion” appearing once in a pin title with no matching screenshot; no anime/illustration-style AI-art community content; no multi-shot narrative/consistency video demos beyond the single Lumalabs transform grid.
  • The same official Veo 2 marketing image was found reused under two different pins with different surrounding text [17, 32] — a sign Pinterest’s AI-news corner is partly recycled/aggregated rather than original commentary.
Limits
  • Per the playbook, Pinterest was not browsed directly — all findings are a visual reading of the 50 pins and images the worker had already collected into output/pinterest.pins.md and images/pinterest/.
  • The pin table has no genre sections (a single unlabeled batch), so the themes above are drawn across the full set rather than genre by genre.
  • Several pins are titled “(untitled)” and contain only generic stock art, providing no attributable source beyond the pin URL and no way to confirm posting date or engagement.
  • Titles/images make brand and headline claims (for example, “Grok is revolutionizing video” and “GPT 5.2 ... full body control”), but the linked images themselves are generic stock graphics rather than screenshots of the actual product or feature. Those specific claims could not be verified from the images alone.

Recommended actions

  • Recollect the X stage using topic-specific keywords such as Midjourney, Sora, Stable Diffusion, Wan, and Kling
  • Consider ways to address Bluesky’s searchPosts API 403 issue, including an alternative search method
  • Continue watching adoption of Qwen-Image-2.1 and LTX 2.5 as indicators for the open-source side
  • Prioritize misuse/backlash news involving closed-source companies, including Nano Banana/Google Earth and the Grok Imagine lawsuit
  • Broaden Reddit collection by splitting searches across multiple queries to reach the target of around 20 items

Collected images

AI content and social media concernsAI News: Anthropic Leak Shows Us The Future of AI...Meta Unveils Muse Image and Muse Video Generative...Google announces ultra-high quality video...Just Nail It 🎥image🎨 AI Image and Video CreationCreate Videos From ImagesGPT 5.2, realtime video editor, AI stereo videos, mobile AI agents, full body control: AI NEWSAI Generated Videos for Content MarketingAI For Image And video GenerationAI Tool To Turn Text Into ImagesimageimageTransform your ideas into stunning visuals! 🎥imageimageProfessional AI Video Creation | Cinematic AI Videos | Custom Video EditingimagePINTERESTimageAi Content Creation Concept With Icons For Text Image Music And Video Generation Artificial Intelligence Images – Browse 66 Stock Photos, Vectors, and Video‎🚨 AI is getting scary good.Google’s Gemini Omni Turns Anything Into VideoimageDiscover how Grok is revolutionizing video...Stability AI’s Stable Diffusions Maker Can Now Build AI Generative VideoAIHow to Make AI Videos: Step-by-Step Guide to Creating Videos With AISpot the AI ✨ Master the Tells!"Why AI Chatbots Fail at Keeping Up with Breaking News"Mastering Video and AI: Key Social Media Marketing Trends 2025Google Unveils Veo 2: Advanced AI Video Generation...imageimageAI news videos blur line between real and fake reportsGoogle's Flow AI Gets a Voice: Adds Speech Generation to Video'Amuse 3.0', an AI art creation tool that includes...AI News: OpenAI Finally Released What We Asked For10 Best Artificial Intelligence (AI) ApplicationsTop AI Video Generators of 2025AI Image Generation and the Entertainment IndustryIntroducing Luma Dream Machine - Next Generation AI VideoAI News Anchors & Creator-Led Coverage Take Over U.S. Media!The Marvels of AI, Astonishing BenefitsAI-Generated Videos: Good or Bad? Here's the Truth! 🤖🎥Stop scrolling… this story will completely change how you see AI!How AI is Changing Video ProductionThe Secret Weapon for Content Creators: AI-Powered Video GenerationTransform Your Ideas into Stunning Visuals with AI Image Technology

Data quality notes

X effectively had no data because of poor search-term selection, which returned only material unrelated to the topic. Bluesky and Reddit also fell short of the target of around 20 items, while Lemmy, Pinterest, and YouTube broadly confirmed sufficient volume.

Gallery