Image and Video Generation AI News — 2026-09-18
Closed-source image generation from Google’s Nano Banana family has become entrenched in personal use, while Sora 2 is fading with its API shutting down (9/24). On the open-source side, MiniMax-H3 and Qwen-Image have become the two dominant forces in the ComfyUI community.
Image and Video Generation AI News Roundup — 2026-09-18
Wow—this one was packed 🔥 Bottom line: on the closed-source side, Google’s Nano Banana family (Gemini image generation) is firmly in the spotlight, while Sora 2 is starting to feel like yesterday’s news following the app shutdown and upcoming API shutdown (9/24) 🤖📉. On the open-source side, MiniMax-H3 (video generation) and Qwen-Image (image generation) are the two clear leaders, and the ComfyUI community in particular was flooded with MiniMax-H3 discussion 😳. One genuinely surprising takeaway: ethics and legal stories—such as legalizing non-consensual generation and the Disney v. Midjourney lawsuit—are getting more attention than tool-performance comparisons ⚖️🔥. X (formerly Twitter) was a complete miss this time, so I’ll be honest and mark it as “no data” 🙏
Across platforms
Here are only the points that could be corroborated across multiple platforms 📝✨
- 🍌 Nano Banana (Gemini image generation) has become an everyday tool for individual creators. On Bluesky, users such as dfl and triflingtree post Nano Banana 2 work daily (https://bsky.app/profile/dfl.bsky.social/post/3mvqfoprrek2i), while YouTube’s “Ultimate Nano Banana Pro Guide 2026” is drawing substantial views through channels with 327K subscribers (channels in the https://www.youtube.com/watch?v=Q7SiipUCjro orbit). Nano Banana was also mentioned as an integration target in a Riverside Reddit thread (https://www.reddit.com/r/RiversideFM/comments/1wiweik/).
- 🌊 Sora 2 is becoming something people look back on rather than look forward to. A Bluesky post advises users to switch to Runway/Kling/Veo because the standalone app is ending (https://bsky.app/profile/beginnersinai.bsky.social/post/3mhvsiik2ij2l), and YouTube increasingly treats it as a baseline in “State of AI Video” comparisons rather than covering new features. The API is scheduled to shut down on 2026-09-24. The tide is clearly turning 😮
- 🧩 Closed-source players are moving toward an “integrated platform” strategy. Reddit (Adobe Firefly Boards and Riverside offering multiple models together) and Pinterest (many tool-comparison infographics listing Runway, Pika, Kling, Luma, Veo, Hailuo, and CapCut) show the same pattern. The era is shifting from promoting one model to letting users choose 🛠️
- 🐉 MiniMax-H3 and Qwen-Image lead open source. Lemmy’s ComfyUI-focused communities were nearly filled with MiniMax-H3 tooling posts from early to mid-September (for example: https://github.com/filliptm/ComfyUI-FL-MiniMaxH3 , score 5). Hands-on user reports say it can run locally even on RTX 3090-class machines (score 12 comment). Qwen-Image appeared on both YouTube (including 700-test benchmark videos) and Bluesky (Qwen3.8-Flash preview: https://bsky.app/profile/primereports.bsky.social/post/3mub3pkd2rf2s).
- ⚖️ Ethics and legal issues are drawing more attention than the technology itself. Reddit (a debate over legalizing non-consensual AI generation, nearly 900 comments total: https://www.reddit.com/r/CriticalState/comments/1wisk2n/), Bluesky (an update on the Disney v. Midjourney lawsuit: https://bsky.app/profile/ai.bots.law/post/3mvqwgixvkm2n), Lemmy (Midjourney’s pivot from image generation to medical devices prompted “Is this Theranos?” reactions: https://awful.systems/post/8731127), and Pinterest (a ByteDance–Hollywood copyright-agreement pin) all showed legal and ethical topics as the hottest discussions.
- 😤 Pushback against “AI slop” remains consistently strong. Reddit (r/aitubers backlash against fully automated faceless video creation) and Lemmy (a r/fuck_ai post titled “companies make AI worse,” score 63) showed the same sentiment. There is no unquestioning enthusiasm anywhere.
Platform by platform
Reddit — Because the search used only one term, image-generation-focused subs such as r/StableDiffusion were missed. Instead, ethics and culture threads—such as DHS misuse of AI imagery and the debate around legalizing non-consensual generation—captured most engagement. Discussion of individual tools (Google Flow, Wan, Minimax) was mixed, with no simple “this is the best” consensus 😐
X — Honestly, sorry: this was a total miss 🙅♂️. All 40 collected posts were about politics, sports betting, or crypto, with not a single post about image or video generation AI. It appears the search captured X trends (Africa, Serbia, Zcash, and similar topics) rather than AI-related terms. Since one of the six platforms specified in the brief failed completely, this needs to be reported as a real gap.
YouTube — Closed-source news arrives on a weekly cadence: OpenAI’s ChatGPT Images 2.5 (released 9/8) and Google’s Nano Banana Pro/2 were rapidly reviewed by mid-sized channels with 30K–330K subscribers and went viral. Open-source coverage centered on technical channels such as SECourses and Japanese Qwen-Image explainers, with view counts roughly one tenth of the closed-source leaders. One unexpected point: ByteDance’s Seedance 2.0/2.5 was the most-covered video-generation model this year 📹
Bluesky — Individual creators post Nano Banana 2 work every day, while news bots also surfaced LTX-2.5 (Lightricks), FLUX LoRAs, and fully open model releases from Abu Dhabi. However, repeated API 403 responses for searches such as “Stable Diffusion,” “Wan 2.5,” and “Qwen Image” made direct discussion of open-source video generation look thinner than it likely is 😑
Lemmy — The best-performing platform this time 🎯. !«メールアドレス» was packed with ComfyUI extension posts related to MiniMax-H3, making it the clearest view of the current open-source video-generation scene. In contrast, the closed-source side was dominated by “off-track” stories such as Midjourney’s medical-device pivot, and Lemmy’s overall tone was notably skeptical.
Pinterest — The worker’s pre-collected, image-based research found a large volume of SEO/affiliate tool-comparison infographics listing Runway, Pika, Kling, Luma, Veo, Hailuo, and CapCut 🖼️. “Can you tell real from AI?” media-literacy content has also become an established category. Open-source content was much thinner, with Stable Video Diffusion appearing mostly as a logo in tool lists.
What to watch
- 🔥 Expansion of the MiniMax-H3 ComfyUI ecosystem — Lemmy, https://github.com/filliptm/ComfyUI-FL-MiniMaxH3
- 📉 Sora 2 API shutdown (scheduled for 2026-09-24) and where users migrate — Bluesky, https://bsky.app/profile/beginnersinai.bsky.social/post/3mhvsiik2ij2l
- 🍌 How firmly Nano Banana Pro/Gemini image generation has taken hold among individual creators — Bluesky, https://bsky.app/profile/dfl.bsky.social/post/3mvqfoprrek2i
- 🎬 ByteDance Seedance 2.0/2.5’s prominence in video generation and its Hollywood copyright agreement — YouTube (Dan Kieft tutorial) and Pinterest’s ByteDance–MPA pin
- ⚖️ Where the debate over legalizing non-consensual AI generation goes next — Reddit, https://www.reddit.com/r/CriticalState/comments/1wisk2n/
- 🐦⬛ Qwen-Image’s growing presence among open-weight models — YouTube (700-test benchmark videos) and Bluesky Qwen3.8-Flash post
Recommendations
- For the next X collection, search AI-specific terms such as Sora, Nano Banana, Kling, Qwen, Flux, and ComfyUI instead of broad trends.
- Continue tracking MiniMax-H3 and Qwen-Image as the hottest indicators in open source right now.
- Track how users move among Runway/Kling/Veo around the Sora 2 API shutdown (9/24).
- Prioritize legal and ethics tracks—including the Disney v. Midjourney lawsuit and debates about legalizing non-consensual generation—over pure tool-performance comparisons.
- Add queries targeting image-generation-focused subreddits such as r/StableDiffusion and r/aiArt to Reddit collection.
- Pinterest research is currently based only on image-level inference, so next time open linked articles live and verify them.
Data quality
X (formerly Twitter) produced 40 collected posts that were all unrelated to the AI theme, effectively leaving no usable data. Reddit missed image-generation-focused communities because it used a single search term; five of the 12 collected threads had a score of 0. About two thirds of YouTube information could not be verified because of 401/CAPTCHA blocks, and view counts for many open-source video items remain unconfirmed. Bluesky’s public API repeatedly returned 403 for phrase searches, preventing direct collection for terms such as “Stable Diffusion” and “Wan 2.5.” Lemmy only covered posts indexed by lemmy.world; other instances such as lemmy.ml were not investigated. Pinterest was based solely on worker-precollected images, with no live verification performed.
Platform summaries
Reddit — Image and Video Generation AI News
Where
A total of 12 threads were collected across 11 subreddits using the search term “Image and Video Generation AI News.”
| Subreddit | Members | Collected threads |
|---|---|---|
| r/generativeAI | 154,407 | 1 |
| r/CriticalState | 35,134 | 2 |
| r/aifilmmaking | 8,038 | 1 |
| r/aitubers | 28,687 | 1 |
| r/agi | 123,187 | 1 |
| r/AIToolTalks | 6,513 | 1 |
| r/ContentCreators | 88,168 | 1 |
| r/AIContentAutomators | 9,387 | 1 |
| r/AIDangers | 53,842 | 1 |
| r/popculturechat | 6,680,224 | 1 |
| r/RiversideFM | 728 | 1 |
Major image-generation-focused subreddits such as r/StableDiffusion and r/aiArt did not appear in the results, so this search was skewed toward AI tools, culture, and ethics-debate communities.
What people say
- Thread 1 (r/generativeAI · 13pt · 41 comments · 2026-09-16) https://www.reddit.com/r/generativeAI/comments/1whtesl/ — In response to “What free AI video-generation tools can I use?”, u/RioMetal said, “If you can run locally, there are tons of Wan and Minimax options,” while u/juzellicious said, “Wan gives 3 seconds a day for free; Google Veo 3 still offered more.” The discussion reflects shrinking cloud free tiers and a return to local/open-source models.
- Thread 3 (r/aifilmmaking · 7pt · 7 comments · 2026-09-14) https://www.reddit.com/r/aifilmmaking/comments/1wg7izt/ — A post says Google Flow’s free credits are “very generous right now.” u/MechanicForward2274 said, “With 50 credits a day you can experiment quite a bit if you don’t waste them.” On the other hand, u/Sea-Temporary-6995 argued, “It’s weak at physics and logic; Minimax is much better,” showing divided evaluations.
- Thread 8 (r/ContentCreators · 0pt · 6 comments · 2026-09-17) https://www.reddit.com/r/ContentCreators/comments/1wj380j/ — A report on making a beauty commercial for a spec portfolio with Adobe Firefly Boards. The poster explicitly chose Firefly Video over Veo and Kling because of commercial-use safety. u/avatar0027 said, “The close-ups are beautiful, but the continuity between cuts makes it look AI-generated.”
- Thread 12 (r/RiversideFM · 0pt · 1 comment · 2026-09-17) https://www.reddit.com/r/RiversideFM/comments/1wiweik/ — Podcast-production tool Riverside released image and video generation features integrating multiple models including Seedance 2.5, Veo 3.1, and Nano Banana. All paid plans receive 2,000 credits.
- Thread 4 (r/aitubers · 0pt · 40 comments · 2026-09-14) https://www.reddit.com/r/aitubers/comments/1wg3duk/ — A request to “make faceless YouTube videos fully automatically from a script” drew strong backlash. u/RobJames007 said, “YouTube is trying to reduce AI slop, yet you’re looking for tools to make more of it,” while u/-NearlyThere- said, “If I find it, I’ll report and downvote it.”
- Thread 9 (r/AIContentAutomators · 0pt · 11 comments · 2026-09-11) https://www.reddit.com/r/AIContentAutomators/comments/1wdm1n4/ — A claim that $300 in free Google Cloud credits for new accounts cut AI-video costs and generated $5,000 per month. u/Truegentlehat replied that the author had written an hour earlier that they were only just starting to earn money at €120/month, making the most-upvoted response a challenge to the claim’s credibility.
- Threads 2 and 5 (r/CriticalState · 0pt/1pt · 421/475 comments · 2026-09-17) https://www.reddit.com/r/CriticalState/comments/1wisk2n/ , https://www.reddit.com/r/CriticalState/comments/1wiwhm4/ — Two poll-style threads asking whether non-consensual AI image generation should be legalized or banned. In both threads, top comments were dominated by votes for banning it. u/AwarenessOk7862 said it “falsifies reality and makes it even harder to distinguish truth from fiction.”
- Thread 11 (r/popculturechat · 448pt · 22 comments · 2026-09-11) https://www.reddit.com/r/popculturechat/comments/1wd49kb/ — The U.S. Department of Homeland Security used an AI-generated Optimus Prime image in deportation-campaign publicity, leaving Hasbro to respond. u/Jerkrollatex said, “Optimus was also a war-refugee character—were the people who made this humans who had just arrived on Earth?”
- Thread 7 (r/AIToolTalks · 12pt · 35 comments · 2026-09-13) https://www.reddit.com/r/AIToolTalks/comments/1wfhbnv/ — A post expressing anxiety about AI broadly, including water consumption and job losses in a context that includes image and video generation. u/EstablishmentRare276 pushed back: “Data centers have existed since the 1950s; a two-hour YouTube video is equivalent to 50 heavy AI requests.”
- Thread 10 (r/AIDangers · 28pt · 31 comments · 2026-09-14) https://www.reddit.com/r/AIDangers/comments/1wg5y17/ — A claim that “AI is not essential, and many companies have seen performance worsen rather than improve.” u/LopsidedSolution criticized the age of the research: “Was that research using 2024 models such as GPT-4o? Re-test it with Astra or Fable first.”
Signals
- Rising: Attention to local/open-source video-generation models (Wan, Minimax). Shrinking cloud free tiers (Thread 1) appear to be driving users toward self-hosting to avoid costs.
- Rising: Multi-model workflow platforms. Adobe Firefly Boards (Thread 8) and Riverside (Thread 12) increasingly let users select closed-source partner models such as Veo, Kling, Seedance, and Nano Banana from a single platform.
- Strong backlash/rejection: Faceless-video automation that promises fully automatic generation from a script is consistently mocked as “slop” and treated as report-worthy in r/aitubers and r/ContentCreators. Income success stories (Thread 9) are also viewed as advertising or exaggeration rather than trusted.
- Visible divide: Practical evaluations of Google Flow and AI in general are mixed. Thread 3 includes claims that Minimax is better, while Thread 10 disputes whether AI improves corporate performance at all. Practitioner comments tend to avoid both simple AI cheerleading and blanket rejection.
- Unexpected: Ethics and political-context threads, such as the legal status of non-consensual generation (Threads 2 and 5, nearly 900 comments combined) and DHS misuse of AI-generated imagery (Thread 11, 448pt), generated orders of magnitude more engagement than performance discussion. Pure tool-comparison threads were mostly low-score (0–13pt).
Limits
- All Reddit searching in this run used one term, “Image and Video Generation AI News.” No separate open-source/closed-source queries or searches for individual models such as Stable Diffusion, ComfyUI, or Flux were performed. As a result, major image-generation subreddits such as r/StableDiffusion and r/aiArt were entirely absent.
- Five of the 12 collected threads had a score of 0pt, so some included threads had little community response. Voices from small communities (r/RiversideFM: 728 members; r/AIToolTalks: 6,513 members; etc.) should be considered limited samples.
- Due to playbook constraints, this agent did not browse Reddit directly and relied entirely on 12 threads precollected by a worker. It therefore does not cover fresh posts from the past few hours or relevant threads outside the original search terms.
- The collection period covers roughly one week, from 2026-09-11 through 2026-09-17, so earlier shifts in discussion are not captured.
X
X — Image and Video Generation AI News
Accounts
The 40 collected posts come from 39 distinct accounts (output/x.posts.md), none of which are accounts associated with image/video generation AI (no Midjourney, OpenAI/Sora, Runway, Stability AI, Black Forest Labs, ByteDance/Kling, Pika, Luma, or similar handles appear anywhere in the set). The accounts present are a mix of political commentary (@GunterFehlinger, @AntiTrumpCanada, @Rusia_HD), sports betting tipsters (@BDreamz37398, @thekasik, @TheRealCEOAmber), crypto/NFT promotion accounts (@Hasan_NFTOX, @0xHeatt, @jubjubcommunity, @patlandcrypto, @AltcoinSherpa), and general commentary/meme accounts (@Keegan59992745, @LeeLovesBey, @Captian_Loden_4). None post about AI image or video generation, and none is a large single-account driver on the theme — because there is no theme-relevant activity in the set to drive.
Posts
No findings — none of the 40 collected posts concern image or video generation AI, closed-source or open-source. Skimming every post and its full text against the research brief (closed-source and open-source image/video generation AI news) turned up zero relevant items. Representative off-topic examples: #1 (https://x.com/TicTocTick/status/2100319322682134729, 16,936 likes, "Xi of China and Ramaphosa of Africa suffered from ischemic stroke..."), #35 (https://x.com/GunterFehlinger/status/2100099276672040971, 13,824 likes, "Russia is collapsing in Donbas"), #38 (https://x.com/Keegan59992745/status/2100583905204326909, 159,030 likes, a burrito/Israel joke) — the three highest-engagement posts in the set, all unrelated to AI.
Signals
Nothing to report about closed-source or open-source image/video generation AI trends on X this run. The only observable pattern is in the search terms themselves: "Africa", "Serbia", "Zcash", "Germany", "$JubJub", "Canada", "Zagreb", "#CyberSecurity", "Russia", "Israel" — these read as X's own Trending/Explore topics for this session's location (per output/x.posts.md, Explore showed Croatia-geolocated and "Politics" trends: Africa, Serbia, Zcash, Germany, $JubJub, Canada, Zagreb, #CyberSecurity, Russia, Israel), not search terms tied to the research brief. Whatever produced this stage's collection searched X's trending topics instead of the AI image/video generation theme.
Limits
- The collected data does not address this run's brief.
output/x.posts.mdandoutput/x.posts.jsoncontain 40 posts found by searching "Africa", "Serbia", "Zcash", "Germany", "$JubJub", "Canada", "Zagreb", "#CyberSecurity", "Russia", "Israel" — general political, sports-betting, and crypto content, matching X's own geolocated (Croatia) Explore trends list at the top of the file, not the theme "Image and Video Generation AI News." - Per the playbook, this stage does not browse X directly (anonymous reads return nothing) and works only from what the worker already collected. Since none of the collected terms relate to closed-source or open-source image/video AI (e.g. Sora, Veo, Midjourney, Runway, Kling, Stable Diffusion, Flux, ComfyUI, LoRA, Nano Banana, etc.), there is no way to answer the brief from this file — a re-collection with search terms tied to the actual theme is needed for this platform.
- No search terms relevant to the brief were used in this collection, so it is not possible to say what X is or isn't saying about image/video generation AI this run.
YouTube
YouTube — Image and Video Generation AI News
Channels
- Paul J Lipsky — general AI-tools reviewer; covered "ChatGPT Images 2.5 Is Here (And It's A LOT Of Fun)" to 277K views in 9 days (watch).
- Kevin Stratvert — large mainstream tech-tutorial channel; "ChatGPT Images 2.5 Tutorial for Beginners" pulled 58K views in 1 day.
- AI Master — 327K subscribers; "Ultimate Nano Banana Pro Guide 2026: How to Use Gemini 3 Image AI", 76,658 views since Jan 17, 2026.
- JohnnyTube — 30.6K subscribers; "Nano Banana 2 Is INSANE! Everything You Need To Know (Full Course)", 26,162 views since Mar 6, 2026.
- AffContent — 262K subscribers; comparison/affiliate-style content, e.g. "BEST AI Image Generator in 2026 (Don't Choose Wrong!)", 14,694 views since Apr 9, 2026.
- Dan Kieft — 295K subscribers, AI-filmmaking focus; "Seedance 2.0 is CRAZY for AI Filmmaking - Full Course", 179,589 views since Apr 24, 2026 — his best-performing recent upload relative to his usual range going by the subscriber count.
- SECourses — 53.7K subscribers, dense local/open-source ComfyUI tutorials; "Wan 2.2, FLUX & Qwen Image Upgraded: Ultimate Tutorial for Open Source SOTA Image & Video Gen Models", 6,649 views (posted Aug 18, 2025 — outside the 60-day window, kept here only as background on an active open-source-tutorial channel).
- Smaller/long-tail channels repeatedly surfacing on Qwen-Image and Wan searches (titles only, metrics not confirmed): Japanese-language AI tool channels covering "Qwen-Image Complete Guide" and "Qwen-Image-2512" walkthroughs, and English channels doing "ComfyUI Text-to-Image Models Compared (Local): Qwen 2512, Z-Image, FLUX.2 klein, Ovis, LongCat".
Videos
- "ChatGPT Images 2.5 Is Here (And It's A LOT Of Fun)" — Paul J Lipsky, ~9 days old, 277K views. https://www.youtube.com/watch?v=Q7SiipUCjro — Reacts to OpenAI's Sept 8, 2026 ChatGPT Images 2.5 launch (sketch-to-image, ~50% lower latency, Flare/Sunburst API models), calling identity-consistency the standout upgrade.
- "ChatGPT Images 2.5 Tutorial for Beginners" — Kevin Stratvert, ~1 day old, 58K views. Walks a mainstream audience through the new sketch feature and multi-step editing that finally holds context between edits.
- "Everyone's Lying to You About GPT Image 2.5" — Joseph Martin, ~2 days old, 10K views. Contrarian take arguing the "leap" is mostly speed/latency, not raw image quality, versus GPT-Image-1.
- "Ultimate Nano Banana Pro Guide 2026: How to Use Gemini 3 Image AI" — AI Master, since Jan 17, 2026, 76,658 views. A 6-part prompt-structure guide (Subject+Action+Environment+Art Style+Lighting+Details) for Google's Gemini 3 Pro Image model, aimed at brand/product-photo use cases.
- "Nano Banana 2 Is INSANE! Everything You Need To Know (Full Course)" — JohnnyTube, since Mar 6, 2026, 26,162 views. Frames Nano Banana Pro 2 as an image+video generator in one tool, benchmarked informally against Midjourney, Flux, Sora 2, and Veo 3.1.
- "BEST AI Image Generator in 2026 (Don't Choose Wrong!)" — AffContent, since Apr 9, 2026, 14,694 views. Broad roundup/affiliate comparison of free vs. paid closed-source generators.
- "Seedance 2.0 is CRAZY for AI Filmmaking - Full Course" — Dan Kieft, since Apr 24, 2026, 179,589 views. Calls out ByteDance's Seedance 2.0 for action-sequence quality, prompt accuracy and visual consistency, with noted limitations; his highest-traction Seedance video versus several same-topic uploads from other creators that week (search turned up at least 9 other Seedance 2.0 tutorials/reviews posted Feb–Jun 2026, e.g. "How Seedance 2.0 is SO GOOD (And Why Hollywood is Shook)").
- "State of AI Video in 2026: Sora 2, Veo 3.1, Kling 3.0 & Cinematic Prompting Workflows" — @StarnationsAI, premiered Mar 20, 2026, 177 views / ~1K subscribers. Low-traction example, included to show the comparison-video format exists even on small channels; description is largely unrelated corporate-training content padded around the AI-video keywords.
- "Wan 2.2, FLUX & Qwen Image Upgraded: Ultimate Tutorial for Open Source SOTA Image & Video Gen Models" — SECourses, 6,649 views (Aug 18, 2025 — outside the 60-day window but the clearest open-source-specific data point found). Week-long test of preset configurations for Wan 2.2 (video), FLUX and Qwen-Image (text-to-image), covering downloads, presets, upscaling, and training/fine-tuning.
- Cluster of unverified-metrics open-source titles found via search but not confirmed by page fetch (YouTube blocked repeated metadata fetches — see Limits): "Qwen-Image Complete Guide: Create Posters and Edit Images in Five Minutes" and "[In-Depth Guide] The Strongest Open Source Model, Qwen-Image-2512" (Japanese-language deep dives on Alibaba's Qwen-Image family), "ComfyUI Text-to-Image Models Compared (Local): Qwen 2512, Z-Image, FLUX.2 klein, Ovis, LongCat" (head-to-head local benchmark of five current open-weight models), and "Qwen Image Dominates Text-to-Image: 700+ Tests Reveal Why It's Better Than FLUX."
Signals
- Closed-source cadence is weekly-scale right now. OpenAI shipped ChatGPT Images 2.5 on Sept 8, 2026 and Google pushed Nano Banana 2 (Gemini 3.1 Flash Image, 4K) on Feb 26, 2026 — each triggered a wave of same-week tutorial/reaction videos from mid-size (30K–330K subscriber) channels, with views in the tens of thousands within days.
- ByteDance's Seedance 2.0 (released Feb 10, 2026) is the single most YouTube-covered video model this year — a plain title search turned up 10+ dedicated tutorials/reviews from Feb through June 2026, more than for Sora 2, Veo 3.1, or Kling 3.0 individually, and the best-performing one found (Dan Kieft, 179.6K views) outperforms the closed-source image tutorials by ~2x.
- Sora 2's YouTube conversation is now retrospective, not forward-looking: search results skew toward "state of AI video" comparison pieces rather than new-feature tutorials, consistent with OpenAI's confirmed shutdown of the consumer app (Apr 26, 2026) and the API sunset date (Sept 24, 2026) — i.e., videos are increasingly using Sora 2 as the baseline other models beat, not as the thing to learn.
- Open-source coverage skews toward dense, tool-focused tutorial channels (SECourses-style) and Japanese-language deep-dives on Qwen-Image, rather than the big comparison-video format that dominates closed-source content — view counts for the confirmed open-source video (6,649) are an order of magnitude below the closed-source leaders, suggesting a smaller but more technical audience.
- Qwen-Image (Alibaba) is the open-weight model getting the most dedicated video attention, spanning both an English 700-generations-vs-FLUX benchmark video and multiple Japanese “impossible to tell it’s AI now” framing videos on the Qwen-Image-2512 checkpoint.
Limits
- YouTube's search-results pages (
/results?search_query=...) and a majority of individual/watchpages returned repeated 401/CAPTCHA blocks when fetched for structured metadata (this run used an HTML-reader proxy since YouTube's search page is client-rendered JS that a plain fetch cannot see). Roughly 1 in 3 fetch attempts succeeded; the rest had to be dropped or reported as title/date-only from search-engine snippets. - Because of that blocking, several open-source-relevant videos (Qwen-Image-2512 deep dives, the Ovis/LongCat/Z-Image ComfyUI comparison, Qwen-Image-Edit-Rapid-AIO) could not be confirmed for exact view counts, subscriber counts, or full descriptions — they're listed above as titles-only with the caveat.
- Could not retrieve top-comment content for any video; the same page-access blocking prevented reading the comments section.
- Could not confirm exact view counts for the general Sora 2 / Veo 3.1 / Kling 3.0 comparison videos beyond the one low-traction example (StarnationsAI) — several other comparison titles were found (Kling 2.6 Pro vs Veo/Sora, Seedance 2.5 vs Sora/Kling/Veo) but their watch pages consistently 401'd across retries.
- No dedicated, well-performing English-language video specifically about a single open-source video generation model (Wan, HunyuanVideo, LTX) with confirmed view counts was found in this run — the one open-source video data point that resolved (SECourses) bundles Wan 2.2 alongside two image models rather than treating video generation alone, and it's over a year old.
Bluesky
Bluesky — Image and Video Generation AI News
Research date: 2026-09-18. Using the Bluesky public search API (https://api.bsky.app/xrpc/app.bsky.feed.searchPosts), more than 70 posts were reviewed using keywords such as “Sora 2,” “Nano Banana,” “Midjourney,” “ComfyUI,” “open weights model release,” and “Runway/Veo,” with up to 15 posts per query.
Accounts (posting accounts)
- dfl.bsky.social — An individual creator posting landscape-photo-style illustrations made with Nano Banana 2 (Gemini image generation) every day.
- triflingtree.bsky.social — A heavy user posting substantial volumes of AI art with the
#NanoBananaProand#GeminiAitags. - ai.bots.law — An automated account tracking AI-related lawsuits and legal developments.
- gen-ai.news / aichina.news / stechtimes.com / channelnewsasia.bsky.social (via bridgy) — News bots/feeds covering generative-AI industry news.
- pitchwall.bsky.social / onestudio1.bsky.social / beginnersinai.bsky.social — Accounts covering AI-tool comparisons and product news.
- primereports.bsky.social / techpresso.bsky.social / theregister.com / socialmedialab.ca — International news accounts tracking open-weight developments around models such as Meta “Muse” and Alibaba Qwen.
- singularityhorizon.bsky.social / heppokopu.bsky.social / misaka-mikoto-17.bsky.social / 404ai.bsky.social — Japanese-language users creating AI art with Midjourney and local environments such as ComfyUI.
Posts (key posts)
Closed-source
- A post introducing the Sora 2 API: “Sora 2 API: A developer-friendly API for Sora 2 and Sora 2 Pro.” 2026-09-15, Likes 0 / Reposts 0.
https://bsky.app/profile/pitchwall.bsky.social/post/3mvjwvr327k2u - A comparison post for “Veo 3.1 vs Sora 2 vs Kling 3.0.” 2026-09-14, Likes 2 / Reposts 0.
https://bsky.app/profile/onestudio1.bsky.social/post/3mvhdlv5vgk2w - A post noting the end of the standalone Sora app and advising users to switch to “Runway, Kling, or Veo” (explaining that OpenAI is winding down the standalone app, not the underlying technology). 2026-03-25, Likes 0.
https://bsky.app/profile/beginnersinai.bsky.social/post/3mhvsiik2ij2l - A post reporting a new filing in the Disney v. Midjourney copyright lawsuit. 2026-09-18, Likes 0.
https://bsky.app/profile/ai.bots.law/post/3mvqwgixvkm2n - A bird’s-eye-view Japanese landscape illustration generated with Nano Banana 2. Likes 25 / Reposts 1, 2026-09-17.
https://bsky.app/profile/dfl.bsky.social/post/3mvqfoprrek2i - A surreal tree-art post tagged
#NanoBananaPro. Likes 15, 2026-09-17.
https://bsky.app/profile/triflingtree.bsky.social/post/3mvpliiqt5c25 - A breaking-news post stating, “Meta now has a rival to Nano Banana 2.” 2026-09-16, Likes 0.
https://bsky.app/profile/laiadesk.bsky.social/post/3mvmiod2iod26
Open source
- A news post reporting that Lightricks deprecated video-generation model LTX-2.3 and changed its recommended model to LTX-2.5. 2026-09-12, Likes 0.
https://bsky.app/profile/gen-ai.news/post/3mvdadb2upx2l - A post introducing an open-weight FLUX LoRA for portraying people of African descent, released via Modelers.cn. 2026-09-09, Likes 0.
https://bsky.app/profile/aichina.news/post/3mv3dcr6hgz2i - A post reporting the release of Alibaba’s multimodal MoE model “Qwen3.8-Flash,” a preview of the Qwen4 architecture. 2026-08-30, Likes 1.
https://bsky.app/profile/primereports.bsky.social/post/3mub3pkd2rf2s - A post reporting that an Abu Dhabi research institution released a fully open family of AI models including weights, training data, and code. 2026-09-03, Likes 0.
https://bsky.app/profile/channelnewsasia.bsky.social/post/3mumk4agbwm2g - An analysis post saying Chinese AI companies are releasing open models in succession amid a price-cutting race with Western rivals. 2026-09-06, Likes 0.
https://bsky.app/profile/stechtimes.com/post/3muuykcvzv62i
Signals
- Nano Banana 2 (Google Gemini image generation) has become routine for individual creators. People such as triflingtree and dfl publish work nearly every day with
#NanoBananaProand#GeminiAi, suggesting actual adoption rather than purely hype-driven attention. - Sora 2 discussion combines anticipation with uncertainty about its future. Alongside developer-oriented API news, posts recommend moving to Runway/Kling/Veo following the standalone app’s closure, making OpenAI’s video-generation strategy reshuffle a topic of discussion.
- Copyright litigation is a major closed-source topic. Bot accounts are tracking new developments in Disney v. Midjourney, as legal conflict between studios and AI image-generation companies continues.
- Open-source news is concentrated in updates distributed by news bots, with few high-energy posts from individual users. LTX-2.5 (Lightricks), FLUX LoRAs, Qwen3.8-Flash, and fully open models from Abu Dhabi are spread across reporting from the U.S., China, and the Middle East.
- Meta’s “Muse” models (Spark/Glimmer 30B) continue to be framed as “coming soon” open-weight releases, with several international news accounts watching them as potential Nano Banana 2 rivals.
- ComfyUI remains a standard tool for local generation among Japanese users (including heppokopu and misaka-mikoto-17). Posts show experimentation aimed at getting local workflows closer to ChatGPT/Nano Banana-level quality.
Limits
- The public search APIs (
public.api.bsky.app/api.bsky.app) intermittently returned HTTP 403 to automated fetching. Basic space-separated queries (such asSora 2 videoandMidjourney) repeatedly succeeded, but quoted phrases (such as"Stable Diffusion") and OR queries almost always failed. No data could be retrieved for “Stable Diffusion,” “Wan 2.5,” or “Qwen Image,” which partly explains the thin direct coverage of open-source video-generation models. - Since
bsky.apppages are client-side rendered, post and profile pages could not provide body text without executing JavaScript. All actual reading was based on text and counts from the JSON API. - Without login, followed feeds, custom feeds, and private lists were unavailable. Findings cover public keyword search only.
- Searches centered on English product and model names; no active Japanese-language search was performed. Japanese posts that appeared were incidental matches to English keywords, so Japanese-language discussion is not guaranteed to be comprehensively covered.
Lemmy
Lemmy — Image and Video Generation AI Topics (as of September 2026)
Communities
- !«メールアドレス» — Focused on open-weight image/video generation AI, including ComfyUI extensions and model releases. This was the most productive source in the run, with concentrated discussion of the MiniMax-H3 video model.
- !«メールアドレス» — About 1.6K subscribers (206 local). A SFW-only community for anime artwork made with Stable Diffusion-family models.
- !stable_diffusion_art / !stable_diffusion_abstract / !stable_diffusion_witches / !stable_diffusion_mycology / !stable_diffusion_furry / !stable_werewolves (all on lemmy.today) — Narrow AI-art communities organized by theme. These are galleries rather than news sources.
- !«メールアドレス» — A mirror-like community reposting r/ArtificialIntelligence content. It contains little original Lemmy-native discussion.
- !«メールアドレス» — 8,235 subscribers. A critical, anti-AI-leaning community with many skeptical posts about image/video generation AI.
- !«メールアドレス» — 2,698 subscribers. A “sneer” community satirizing tech-industry hype. Midjourney’s strategic pivot was most actively discussed here.
- !«メールアドレス» — General technology news, with multiple major Midjourney-related stories.
Posts
- A run of ComfyUI extensions for MiniMax-H3 (the main open-source story) — Roughly 20 posts in !«メールアドレス» between 2026-08-28 and 09-16. Examples:
- “ComfyUI-FL-MiniMaxH3” (score 5, 2026-08-30, github.com/filliptm/ComfyUI-FL-MiniMaxH3), featuring prompt timeline management and shot-composition functionality.
- “H3-World” (score 3, 2026-09-02, huggingface.co/DANNY621/H3-World), self-described as “the first interactive world model built on MiniMax-H3.”
- “ComfyUI-MiniMaxH3-CLSS” (score 0, 2026-09-01, github.com/nazgut/ComfyUI-MiniMaxH3-CLSS), claiming arbitrary-length video generation with audio on consumer hardware with 16GB VRAM.
- “ComfyUI_MinimaxH3_AutoContext” (score 1, 2026-09-16, github.com/supElement/ComfyUI_MinimaxH3_AutoContext), enabling indefinitely continuous video generation even with limited VRAM.
- According to a comment from user brucethemoose (score 12, 2026-09-14), MiniMax-H3 can “combine audio, video, and images as references in arbitrary ways,” and ran locally on a fairly typical RTX 3090 + Ryzen 7800 custom PC.
- Midjourney pivoting from image generation to medical devices (a whole-body ultrasound scanner) — A post on !«メールアドレス» (https://awful.systems/post/8731127 , score 40, 15 comments, 2026-06-19). David Gerard introduced the “Midjourney Medical” announcement, and comments repeatedly compared it sarcastically to Theranos. User YourNetworkIsHaunted noted a technical issue: “This device can only accommodate people up to 60 inches [about 152 cm], so it cannot image anything above the heart.” Terranoid called it “a Hail Mary after the image-generation business hit a wall” (score 10).
- The same news appeared several times in !«メールアドレス»: “Midjourney pivots from AI image generation...” (theregister.com, score 58, 2026-06-18), and “Midjourney wants Hollywood studios to reveal...” (techcrunch.com, score 73, 2026-07-05).
- Runway’s enterprise business reportedly doubled — !«メールアドレス» (score 0, 2026-08-30, cross-posted from reddit.com/r/ArtificialInteligence/comments/1w2hnea). Its enterprise net revenue retention reportedly exceeded 300%.
- Google Earth’s “Nano Banana” AI feature was paused one day after launch — Posts on !«メールアドレス» (fortune.com, 2026-08-07) and !«メールアドレス». Concerns were raised that anyone could make fabricated satellite-photo-like imagery (also referencing a 404media.co article), and Google reportedly stopped the feature immediately.
- Comparison of six AI video models using the same prompt — !«メールアドレス», “I ran the same prompt through 6 AI video models...” (score 1, 2026-09-17, reddit.com/r/ArtificialIntelligence/comments/1wipgwa).
- Comparison of colorizing a black-and-white manga page with ChatGPT Images, Wan, Nano Banana Pro, and Seedream — !«メールアドレス», “Same Berserk spread, 4 AI colorizations. Looks like GPT wins again?” (score 1, 2026-09-14). It compares open-source Wan with Nano Banana Pro, Seedream, and GPT Images; comments generally favored GPT’s result.
- A post reflecting the anti-AI mood — !«メールアドレス», “Why companies like anthropic and open AI make AI even worse” (score 63, 2026-09-14). It is not specific to image/video generation, but serves as an indicator of Lemmy’s broader distrust of major closed-source AI companies.
- A report on running Qwen locally — !«メールアドレス», “Testing Qwen 3.8 27B running locally on a single 5090” (score 1, 2026-09-16, v.redd.it/jdrothgw4vph1). Not image generation itself, but relevant to the broader context of consumer-GPU open-model execution.
- The everyday flow of AI-art posts, including anime and abstract art — !share_anime_art and !stable_diffusion_art communities receive near-daily new posts with images hosted on civitai.com (for example, “Isshiki Clear - Mahjong Fight Girl,” score 15, 2026-09-17, lemmy.dbzer0.com/post/75601449). These lack news value but demonstrate routine use of Stable Diffusion, Flux, and Krea2-family models.
Signals
- The open-source side is overwhelmingly MiniMax-H3. The stable_diffusion community from early to mid-September was nearly full of ComfyUI tools for this video model. Repeated emphasis was placed on combining audio, images, and video locally and on optimization for systems with around 16GB VRAM.
- The closed-source side had little positive news. Midjourney’s hardware pivot and Google Earth’s halted feature both stood out as stories of apparent misdirection. Runway’s growth report was one of the few upbeat items.
- Lemmy’s overall tone is skeptical and critical (!fuck_ai, !techtakes). Unqualified praise is rare, and there is a wide gap in tone between implementation communities (the stable_diffusion ecosystem) and news-criticism communities (techtakes and fuck_ai).
- Artwork posts and cross-posts outweigh original discussion. !ai_reddit is literally a Reddit relay, so it should not be treated as primary, Lemmy-native information.
Limitations
- Lemmy is distributed across instances, and its search API only retrieves posts indexed on lemmy.world. There was not enough time to separately search lemmy.ml or lemm.ee, so this report covers only results available through lemmy.world federation.
- Searches covered multiple keywords including “Sora,” “Veo,” “Gemini image,” and “Flux,” but found no posts about new Sora 2 or Google Veo model announcements themselves. The reality this time was that news of major closed-source product launches was thin on Lemmy.
- The original pivot-to-ai.com article rendered as garbled text in WebFetch and could not be quoted directly, so its content was supplemented using the awful.systems Lemmy post and comments plus external reporting found through search engines, including Bloomberg.
- Community subscriber counts were difficult to retrieve because the API was unstable (400/500 errors). Only !share_anime_art (about 1.6K), !fuck_ai (8,235), and !techtakes (2,698) could be confirmed.
- Against the brief’s completion criterion of analyzing roughly 20 posts per social network, more than 20 post bodies, scores, and dates were reviewed across searches. This file presents only nine selected representative examples.
Pinterest — Image and Video Generation AI News
Visual themes
- Dark neon “AI news roundup” infographic templates dominate. Near-black/navy backgrounds, glowing cyan-magenta-purple lines, bold white headline type, and numbered rows of tool logos inside rounded badges. This single visual language covers most informational pins.
[6, 9, 10, 11, 18, 20, 26, 31, 38, 39, 41, 45, 48] - “Spot the AI” real-vs.-fake challenge formats are everywhere. Split-screen or A/B comparisons challenge viewers to choose the real photo, video, or eye; examples also include a face-swapped red-carpet demo and a hyperreal selfie captioned “This is AI.”
[3, 21, 24, 33, 46] - AI is personified as a chrome or white humanoid robot. Full-body androids—one at a “BREAKING NEWS” anchor desk, one holding a phone beside a “FAKE” stamp, and one touching a glowing globe over a city skyline—stand in for AI itself rather than specific products.
[3, 40, 44] - YouTuber reaction-face thumbnails are a common visual genre for AI news. Shocked, bearded men are framed beside rows of company logos including OpenAI, Google, Meta, Anthropic, and Notion.
[2, 19, 22, 42] - Single-source-video, multiple-style demos showcase image/video consistency. The same walking person and car are rendered as wooden blocks, origami, Lego bricks, flowers, and toy constructions—a “Lumendra/Luminate”-style diffusion showcase.
[5] - The same handful of tool names recur in every listicle: Runway (Gen-3), Pika, Kling AI, Luma/Dream Machine, Google Veo (2 and 3), Canva AI, Hailuo AI (MiniMax), CapCut, ElevenLabs, and Stable Video Diffusion.
[6, 10, 18, 20, 31, 45] - Hyperreal AI portraiture is used as proof of quality. A photoreal old fisherman (AMD “Amuse 3.0”), a marble-and-wood statue playing cello, and “AI baby” face-morph composites are used to sell how convincing current image models appear.
[25, 47, 50] - Fake-broadcast and fake-newspaper framing is used for AI news itself. Mock “BREAKING NEWS” chyrons, an AI news-anchor-avatar generator (Wondershare Virbo), and a fully mocked newspaper front page (“The AI Dispatch,” Aug 15 2026) mimic television and print-journalism conventions.
[14, 36, 40, 41]
Notable pins
[5](untitled, multi-style video transform grid) — A source video of a person and a car is re-rendered in five materials: wood, origami, Lego, flowers, and toy plastic. It is the clearest concrete demonstration in the set of current image/video consistency technology.[9]“4 Major AI Updates You Need to Know in 2026” — Names specific items: OpenAI “GPT-6 Astra,” Gemini becoming agentic across Gmail/Docs/Slides, Adobe Premiere generating video/audio in the timeline, and Cognition (Devin) raising a $2B Series E at a $48B valuation.[28]“ByteDance and Hollywood reach global deal on AI copyright” — A real-news-style pin saying ByteDance reached an agreement with the Motion Picture Association to strengthen IP protection in its AI image/video tools.[45]“Today’s AI News — May 20, 2026” — Recaps Google I/O 2026: Gemini 3.5 Flash, Gemini Omni, an AI-powered Search overhaul, Apple AI accessibility features, and an AI smart-glasses race (Android XR).[41]“The AI Dispatch,” Aug 15 2026 — A mocked newspaper front page listing a broad mix of current AI-adjacent themes: inference speed, sovereign AI, agent harnesses, AI safety, and deep research.[31]“AI Video Creation Tools Changing How Content Is Made” — The most exhaustive tool directory found, naming 20 products across text-to-video, image-to-video, editing, voiceover, stock/B-roll, and publishing/optimization categories.[24](untitled, AI-generated red-carpet photo vs. “original clip”) — A face-swap/deepfake demo pairing a polished AI portrait with a real source clip of another person, illustrating identity-transfer technology.[40]“Discover how Grok is revolutionizing video...” — A humanoid robot news anchor at a “BREAKING NEWS” desk promoting Grok’s video capability.[47]“Amuse 3.0” — AMD-branded local/on-device AI art tooling demonstrated with a photorealistic portrait of an elderly fisherman.[10]“AI Image Video: Best Quality vs Best Free” — A decision-tree infographic dividing tools into paid/best-quality (ChatGPT Plus, Midjourney, Adobe Firefly, Runway Gen-3, Luma) and free (Ideogram, Gemini free tier, SeaArt, Krea, CapCut).
Signals
- Video-generation tool roundups and listicles are the loudest, most repeated content type. Pinterest’s AI-image-generation audience is largely made up of SEO/affiliate creators repackaging the same closed-source tool set—Runway, Pika, Kling, Luma, Veo, Hailuo, and CapCut—rather than breaking news.
- “Can you tell it’s AI?” literacy content, including real-vs.-fake challenges and deepfake demonstrations, is a full genre in its own right. Audience anxiety and curiosity about detection are content drivers independent of any individual model release.
- A few pins reference dated, named events after this collection window opened—Google I/O 2026, “GPT-6 Astra,” an “Aug 15 2026” news digest, and Cognition’s $2B raise. These are the closest things to hard news on the platform.
- Open-source specifics are thin: “Stable Video Diffusion” appears only as one logo inside a 20-tool directory
[31]; no pin discusses an open-weight release, license, or community fork in depth. - Nothing in this set touches training methods, model architecture, or dataset/licensing debates except the ByteDance–MPA copyright-deal pin
[28].
Limits
- This stage describes only images precollected by the worker (per
platforms/pinterest.md); no live Pinterest browsing, searching, or linked-article reading was done, so no claim here was checked against a live page or the article behind each pin. - Several pins are near-duplicate templates from likely repost/scraper accounts: the neon stock “Generate content with AI” graphic appears twice
[11, 26], the Google Veo 2 banner appears twice[7, 8], and the “Best AI Video Generation Websites” listicle appears twice with the same tool list[6, 18]. - Five pins have no title (
(untitled)) in the source table[8, 17, 20, 21, 24, 28... ]—specifically[8, 20, 21, 24, 28]had blank or generic titles—limiting attribution to what is visible in the image alone. - The pin set covers a single genre; no genre breakdown was defined for this run, so all 50 pins came from one undifferentiated search and no per-genre comparison was possible.
Recommended actions
- Re-run X collection using AI-specific terms such as Sora, Nano Banana, Kling, Qwen, Flux, and ComfyUI; this time it mistakenly collected X trend terms and missed the topic entirely.
- Keep MiniMax-H3 and Qwen-Image under observation as the main indicators among open-source players.
- Follow user migration targets—Runway, Kling, and Veo—around Sora 2’s API shutdown (2026-09-24).
- Track legal and ethics issues such as Disney v. Midjourney and the legal-status debate around non-consensual generation at least as closely as tool comparisons.
- Add queries for image-generation-specific communities such as r/StableDiffusion and r/aiArt to Reddit collection.
- Validate Pinterest findings live by opening linked articles.
Collected images


















































Data quality notes
Of the six platforms specified by the brief, X supplied data that was completely off-topic and unusable. Reddit, Bluesky, and YouTube also missed some major topics—including Stable Diffusion and Wan 2.5—because of search-term limitations and API restrictions (403/401/CAPTCHA).



