Image and Video Generation AI News — 2026-09-16
Users frustrated with censorship are moving toward Chinese open-weight video models such as MiniMax H3, Wan, and LTX-2, while closed players including Google, OpenAI, and ByteDance continue to announce new models at a pace of one or two per month via Replicate. At the same time, risks such as Sora’s API wind-down (2026-09-24) and copyright litigation are developing beneath the surface.
Image and Video Generation AI News Roundup — 2026-09-16
On the closed-source side, Google, OpenAI, ByteDance, Kuaishou, and others are still releasing new models at a pace of one or two per month, with Nano Banana Pro, Veo 3.1, Seedance 2.0, and Kling 3.0 among the most prominent recent players. Meanwhile, on the open-source side, locally run Chinese video models such as MiniMax H3, Wan, LTX-2, and HunyuanVideo are becoming a refuge for users tired of censorship and filters from cloud providers, while a tutorial ecosystem around ComfyUI keeps growing. At the same time, the effective downsizing of Sora (with its API also scheduled to end on 2026-09-24) and lawsuits such as Disney v. Midjourney are clearly casting a shadow over the momentum of closed providers. In this run, X (Twitter) collected unrelated trending terms and yielded effectively zero relevant results, while Reddit and Pinterest also fell short of the completion target of analyzing 20 items, leaving gaps in collection coverage.
What emerged across platforms
- Frustration with censorship and restrictions is boosting open-source and local generation. On Reddit (#1, #5, #9), users describe switching to Minimax H3, Wan, LTX Video, and Seedance 2.5 after being blocked by Google Flow and Gemini filters. On Lemmy (!stable_diffusion), posts about MiniMax H3-related tools appeared almost daily from 2026-09-08 through 13. YouTube also continues to feature ComfyUI tutorials for these model families.
- China-based models (ByteDance Seedance/Seedream, Alibaba Wan/Qwen, Tencent Hunyuan, Kuaishou Kling, MiniMax) are becoming more prominent on both the open and closed sides. YouTube, Bluesky (via Replicate), and Pinterest (Vidu/ShengShu) all repeatedly highlighted the U.S.–China comparison.
- Interest is shifting from “generation” to “editing and consistency control.” Reddit (#7, #8) discussed the value of character consistency and fixing the remaining 10% of an image that is already 90% correct. YouTube also covered editing-focused models such as Qwen Image Edit.
- The “shadow news” for closed-source providers is litigation and retrenchment. The effective closure of Sora (Lemmy), Disney v. Midjourney (Bluesky, Lemmy), and Midjourney’s disclosure-demand litigation against Hollywood (Lemmy) made legal battles and withdrawals more visible than feature announcements.
- Confusion around the “open-source” label persists. Both YouTube and Bluesky noted the mismatch between ByteDance Seedream-related videos marketed as “open source” and their reality as primarily closed API offerings.
By platform
Reddit — Collected 11 threads from seven subreddits, including r/generativeAI, r/aivideomaking, and r/aifilmmaking. The review found direct user quotations about migration toward local generation, strong demand for free or low-cost tools, debates over “AI slop,” and image-generation comparisons between Gemini and OpenAI. However, collection was limited to a single search term and did not reach the target of around 20 items.
X — None of the 40 collected posts from 37 accounts mentioned image or video generation AI. This happened because the search terms came not from the topic but directly from X’s trending list (#Uranium, Clarity Act, and others). The four closest posts, related to “Dario,” concerned AI security policy and were outside the topic. This was effectively a collection failure.
YouTube — Reviewed 12 videos from major channels including TheAIGRID, Matt Wolfe, and Theoretically Media, organizing specific announcements from both closed models (Nano Banana Pro, Veo 3.1, Runway Gen-4.5, Seedance 2.0 vs Kling 3.0) and open models (HunyuanImage 3.0, FLUX.2, LTX-2, Wan2.2/2.7, MiniMax H3, Ideogram 4). However, there were few entirely new model announcements within the last 60 days, and many videos predated the target period.
Bluesky — Reviewed 14 posts centered on feeds from Replicate, TestingCatalog, and Simon Willison. Announcements from closed providers (Nano Banana Pro, Gemini 3 Pro, Veo 3.1, ChatGPT Images 2.5, Muse Video, Orbis) dominated. The open side was limited to FLUX.2 and Wan 2.2 Animate, while Stability AI has been silent since November 2025. Frequent 403 errors from the search API skewed collection toward feeds from known accounts.
Lemmy — The !stable_diffusion community ecosystem has effectively become an “open-source-only news feed,” with bot account Even_Adder posting new MiniMax H3- and LTX-2-related models and LoRAs almost daily. On the closed side, generic communities such as !technology carried only one or two monthly stories, typically “controversy or withdrawal” news such as Sora’s closure and Midjourney lawsuits. More than 10 open-side items were found, but only five or six closed-side items, short of the target.
Pinterest — Reviewed all 50 pins collected through a single query. Closed providers were visually dominant, including Veo 2 promotions, tool-ranking infographics, and AI actress Tilly Norwood-related content. Open source was almost absent, with only one mention of “Stable Video Diffusion” in a tool list. The results lacked categorization and contained substantial template reuse, suggesting SEO-oriented padding.
The other five platforms produced posts or works that could be reviewed, but X alone depended on topic-unrelated terms. That gap could not be filled in this run.
Key items to watch
- Nano Banana 2: Only leak-stage information is available; no official announcement has been confirmed. YouTube — https://www.youtube.com/watch?v=SJvnEvOvEYk
- Sora 2 API shutdown (scheduled for 2026-09-24): The standalone app has already ended. Lemmy — https://lemmy.world/post/45568787
- Development pace of MiniMax H3-related tools: Several derivative tools were posted between 2026-09-08 and 13 alone. Lemmy — https://lemmy.dbzer0.com/post/75382764
- Progress of the Disney v. Midjourney lawsuit: New filings are being added almost every day. Bluesky — https://bsky.app/profile/ai.bots.law/post/3mvlvjpygwl27
- Orbis (Visko) real-time video generation: An emerging model claiming 4K/24fps and sub-second updates. Bluesky — https://bsky.app/profile/testingcatalog.com/post/3muf3e5zh3y2f
- Seedream’s “open-source” labeling issue: Requires verification against primary sources. YouTube — https://www.youtube.com/watch?v=3Hmj05sE1V8
Recommended actions
- For a renewed X collection, use topic-derived search terms (for example, "Nano Banana," "Seedance," "Wan 2.2," and "Sora") instead of relying on trending lists.
- Continue monitoring Lemmy’s !stable_diffusion ecosystem as a leading indicator for open-source video models.
- Track Nano Banana 2’s official announcement and the Sora 2 API shutdown (2026-09-24) as near-term turning points for closed providers.
- Monitor ongoing lawsuits such as Disney v. Midjourney as factors that could affect future strategy among closed-source companies.
- In renewed Reddit and Pinterest collection, use multiple topic-specific queries rather than a single query so that the completion threshold of roughly 20 items can be met.
- For models claiming to be “open source,” verify their actual status through primary sources such as Hugging Face and official blogs before citing them.
Data quality
Because X searches relied on a trending list unrelated to the topic, they yielded effectively zero image- and video-generation AI posts, making this the largest data gap in the report. Reddit was limited to 11 threads from a single search term, while Pinterest consisted of 50 uncategorized pins from one query; neither met the “analyze 20 items” completion threshold. Lemmy had only five or six direct closed-source mentions, reflecting the platform’s limited scale. YouTube and Bluesky were relatively well supplied, but YouTube had few entirely new announcements within 60 days and many videos predated the target period, while Bluesky collection depended on feeds from known accounts because of 403 errors in its search API.
Platform summaries
Reddit — Image and Video Generation AI News
Where
| Subreddit | Members | Collected threads |
|---|---|---|
| r/generativeAI | 153,596 | 3 items (#1, #7, #9) |
| r/AIToolTalks | 6,131 | 2 items (#2, #8) |
| r/aivideomaking | 7,038 | 2 items (#4, #5) |
| r/aifilmmaking | 7,833 | 1 item (#3) |
| r/GeminiAI | 373,910 | 1 item (#6) |
| r/AIDangers | 53,008 | 1 item (#11) |
| r/politicalstress | 14 | 1 item (#10) |
Practical discussions of image and video generation AI were concentrated in r/generativeAI, r/aivideomaking, r/aifilmmaking, and r/AIToolTalks. The two items from r/AIDangers and r/politicalstress were adjacent discussions about broad AI anxiety and regulation.
What people say
-
A movement away from cloud censorship toward local generation — Thread #1, “Need a good video generation AI for image to video of people” (r/generativeAI, 2pt, 2026-09-15, https://www.reddit.com/r/generativeAI/comments/1wh1y9d/). The poster reported that Google Flow blocked Nano Banana Pro-generated images 60 times in a row. u/f5alcon said, “Minimax H3 is the best local generator,” while u/Liberation2020 recommended “Wan, mini max H3 and LTX Video.” Local workflows using ComfyUI are supported among RTX 5090 owners.
-
Google Flow’s generous free tier is praised, but its physics is viewed as weak — Thread #3, “Google Flow is giving creators a surprisingly generous amount of free AI video generation right now” (r/aifilmmaking, 8pt, 2026-09-14, https://www.reddit.com/r/aifilmmaking/comments/1wg7izt/). u/Sea-Temporary-6995 said, “It's weak at physics and logic. Even Minimax is far better IME.” In contrast, u/Rollertoaster7 said it was “pretty good” for a free tool.
-
Demand for free and low-cost tools is extremely high — Thread #4 (r/aivideomaking, 0pt, 9 comments, 2026-09-15, https://www.reddit.com/r/aivideomaking/comments/1wh80z3/) and thread #5, “Looking for free/credit-based AI tools for highly realistic image & video generation” (r/aivideomaking, 5pt, 31 comments, 2026-09-09, https://www.reddit.com/r/aivideomaking/comments/1wbxpdf/). u/Fluffybabyyoda said, “google flow gives you 50 credits daily sometimes 2 times a day. Google gemini i think gives you 3 videos a day”.
-
Demand for uncensored models is discussed using specific brand names — In thread #5, u/zeroludesigner said, “You should try cyberbara. ... the Seedance 2.5 on it is uncensored and supports real face.” u/senorfavela even explained how to earn free credits through reRAW.ai’s referral system.
-
Gemini is at a disadvantage in image-generation comparisons with OpenAI — Thread #6, “Image generation Gemini vs OpenAI” (r/GeminiAI, 9pt, 2026-09-14, https://www.reddit.com/r/GeminiAI/comments/1wgjjco/). u/damaged_terror said, “The gemini one looks like it gave up halfway through and just slapped some text on there.”
-
The real issue behind “AI slop” criticism is a lack of story, not technical quality — Thread #7, “Is there a way to make AI video that does not look like AI slop?” (r/generativeAI, 0pt, 24 comments, 2026-09-14, https://www.reddit.com/r/generativeAI/comments/1wgjjzu/). The poster reported that switching to a “4 part thing with the same character + voice” increased watch time by roughly ninefold. Tool comparisons described argil as strong at presenter-face consistency, kling as losing character consistency after several shots, and runway as easiest to use for camera work. u/RioNReedus said, “AI Slop is now a term people who don't like AI use for anything AI - no matter how good it is. It means nothing anymore”.
-
A shift in interest from “generation” to “editing” — Thread #8, “AI image editing is becoming more interesting than image generation” (r/AIToolTalks, 10pt, 2026-09-14, https://www.reddit.com/r/AIToolTalks/comments/1wgbie5/). u/SethWorks20 said, “being able to take an image that's already 90% right and fix the other 10% is probably way more useful day-to-day”.
-
The era of making viral videos with a single prompt plus reference images — Thread #9, “HOW IS THIS MADE?!” (r/generativeAI, 0pt, 31 comments, 2026-09-10, https://www.reddit.com/r/generativeAI/comments/1wced4g/). u/lagbit_original said, “Use you preferable image model for creating the character Seedance 2.5 and a single prompt and the right references”.
-
Broad AI anxiety is spilling into generative-AI subreddits — Thread #2, “Are all of you also anxious about recent news about AI?” (r/AIToolTalks, 13pt, 35 comments, 2026-09-13, https://www.reddit.com/r/AIToolTalks/comments/1wfhbnv/). Responses were divided over concerns about water consumption and jobs; u/Big-Pops78 countered, “A car takes 13-20,000 gallons. A pair of Jeans takes 3,000 gallons.”
-
AGI and existential-risk discussions are reaching general users — Thread #10, “AI news is really hard right now” (r/politicalstress, 4pt, 15 comments, 2026-09-12, https://www.reddit.com/r/politicalstress/comments/1wecrrc/). u/Fil_77 cited the “AI 2027” scenario and argued for urgent regulation.
-
Skepticism that “AI does not improve business performance” is also emerging — Thread #11, “Biggest AI news of 2026: AI is not a necessary tool...” (r/AIDangers, 23pt, 30 comments, 2026-09-14, https://www.reddit.com/r/AIDangers/comments/1wg5y17/). u/presentofai countered that “most of these "AI made it worse" numbers are measuring clumsy rollouts, not the models.”
Signals
- Rising: Against a backdrop of frustration with censorship and restrictions in closed cloud tools such as Google Flow/Veo and Gemini, support for local/open workflows involving Minimax H3, Wan, LTX Video, Seedance 2.5, and ComfyUI appears consistently in threads #1, #5, and #9. The preference for “editing and consistency control” over pure generation (#7, #8) is also a clear trend.
- Being dismissed / facing backlash: One-off, single-prompt text-to-video is easily derided as “slop” (#7), while Gemini’s standalone image-generation capability is explicitly judged inferior to OpenAI in the relevant thread (#6).
- Points of disagreement: Opinions on Google Flow’s free tier are sharply divided between “generous” (u/Rollertoaster7 in #3) and “not usable in real work” (u/Sea-Temporary-6995 in #3). The term “AI slop” itself is also contested in #7 between those treating it as criticism of technical roughness and those seeing it as a meaningless label used by AI opponents (u/RioNReedus, u/Myg0t_0).
- Unexpected finding: Despite searching for image and video generation AI, two of the 11 collected threads (#10, #11) drifted into the entirely different topic of AGI and existential risk. This suggests that users’ search behavior around image and video generation news is closely connected to broader narratives of AI anxiety.
Limits
- The only search phrase was “Image and Video Generation AI News,” yielding 11 threads from seven subreddits. This falls short of the brief’s request for “around 20 items per social network,” and a single search term likely caused substantial omissions (no crawling with additional search terms was performed).
- The body text for threads #4 and #11 was not recorded in the files, so assessment relied only on titles and comments.
- r/politicalstress (14 members) is an overly niche community, and its post concerns general AI anxiety rather than image- and video-generation AI specifically.
X
X — Image and Video Generation AI News
Accounts
Among the 40 collected posts from 37 accounts, not a single account was posting about image or video generation AI. For reference, the accounts with the highest engagement in the collected data are listed below.
| Account | Posts | Total likes | Content |
|---|---|---|---|
| @tonypraysick | 1 | 36,091 | Ed Sheeran’s political stance (music gossip) |
| @Omolomo_o | 2 | 9,157 | Nigerian religious and political gossip |
| @TheSirRobotto | 2 | 10,608 | Football / UK devolution discussion |
| @SenLummis | 1 | 8,909 | U.S. senator; crypto Clarity Act |
| @realwangjyunhao | 1 | 7,671 | Taiwanese relationship influencer; anniversary post |
| @RippleXrpie | 1 | 7,513 | Crypto (XRP) community; Clarity Act |
| @AshCrypto | 1 | 2,508 | Crypto influencer; Clarity Act |
| @NFT_Chen / @wquguru / @bluetouff / @BPIV400 | 1–2 each | 75–1,106 | Dario Amodei (Anthropic CEO), China stance / AI policy discussion |
No account was posting at high volume; all were one-off or occasional viral-post accounts. @Omolomo_o was the only account with two posts earning thousands of likes, but it is a religious-conflict account unrelated to generative AI.
Posts
All 40 collected posts were reviewed, and there were zero posts about image-generation AI or video-generation AI. The closest were four posts found through “Dario,” which mentioned Anthropic CEO Dario Amodei but concerned AI security policy and China policy, not open- or closed-source trends in image/video generation models. They are recorded here for reference.
-
@bluetouff (153 likes · 35 reposts · 19 replies · approximately 7,300 views · 2026-09-14)
https://x.com/bluetouff/status/2099399733127078156The interesting thing about Dario, Sam, and Elon's sudden love affair with the handbrake might not even be the extinction of humanity. It's that AI cybersecurity is turning into a U.S. industrial policy...
→ Discussion of AI security and industrial policy. Image/video generation is not mentioned. -
@BPIV400 (1,106 likes · 46 reposts · 9 replies · approximately 78,000 views · 2026-09-14)
https://x.com/BPIV400/status/2099336990478958724Dario complaining that AI research is progressing dangerously fast
→ Concern about the pace of AI progress. Not about generative AI models themselves. -
@NFT_Chen (75 likes · 21 reposts · 55 replies · approximately 85,000 views · 2026-09-14)
https://x.com/NFT_Chen/status/2099348095612109306Anthropic CEO Dario Amodei openly admits to being a hawk on China, pushing to treat Chinese AI like Soviet nuclear missiles...
→ U.S.–China AI-hegemony debate. Unrelated to image or video generation. -
@wquguru (176 likes · 22 reposts · 22 replies · approximately 84,000 views · 2026-09-14)
https://x.com/wquguru/status/2099340572519460928Dario Amodei's relationship with China in this lifetime is truly a tangled mess...
→ As above, a post about Dario Amodei’s history.
The remaining 36 posts, searched through #Uranium, #sweepstakes, Clarity Act, LMEOW, #chance, Serbia, Ed Sheeran, England, and Church, concerned uranium mining stocks, sweepstakes, U.S. crypto regulation, meme coins, romance/fandom topics, football and UK devolution, controversy over Ed Sheeran leaving a tour, and Nigerian religious conflict. They were wholly unrelated to the topic.
Signals
- Rising / attracting attention: Cannot be determined within the collected data. There is no evidence at all showing how image and video generation AI topics are moving on X.
- Being dismissed: Not applicable, because no posts addressed the topic.
- Unexpected finding: The 10 search terms (#Uranium, #sweepstakes, Clarity Act, LMEOW, #chance, Dario, Serbia, Ed Sheeran, England, Church) closely match the Explore trends shown under “What X says is happening” for a session originating in Croatia. In other words, this collection appears to have searched X’s trending list directly rather than using the brief’s subject terms. The fact that “Dario” happened to pick up an AI-related trend was nearly incidental, and even that concerned AI security policy rather than image and video generation.
Limits
- The 10 terms searched in this session (#Uranium, #sweepstakes, Clarity Act, LMEOW, #chance, Dario, Serbia, Ed Sheeran, England, Church) are all unrelated to the research brief’s topic of “closed-source/open-source image- and video-generation AI.” Consequently, none of the 40 collected posts mentioned image-generation or video-generation AI news or trends.
- Because the search depended on the Explore section’s trends (as seen from a Croatian session), no topic-aligned terms such as “image generation AI,” “Nano Banana,” “Seedance,” “Sora,” “Midjourney,” “Stable Diffusion,” or “open source video model” were searched. The brief’s completion criterion of analyzing “around 20 posts on each social network” was not met for image- and video-generation AI posts on X.
- As a result, this report’s Posts/Signals sections do not present “findings about the topic,” but report that there were effectively zero findings about it. A new collection with topic-specific terms is needed to understand closed- and open-source image/video AI trends on X.
YouTube
YouTube — Latest Image and Video Generation AI News (Closed/Open Source)
Channels
- TheAIGRID(@TheAiGrid)— A breaking-news-style channel that compiles both closed and open image/video models in weekly “AI News” episodes. youtube.com/@TheAiGrid
- Matt Wolfe(@mreflow)— Approximately one million subscribers. Covers AI tools broadly, with frequent explainers on Sora, Veo, and Nano Banana. youtube.com/@mreflow
- Theoretically Media(@TheoreticallyMedia)— An active filmmaker reviewing AI video tools from a production perspective. Continues to cover open-source tools such as LTX-2, Wan, and Hunyuan. youtube.com/@TheoreticallyMedia
- ComfyUI tutorial creators (several channels could not be identified from search results)— Post local-installation guides for MiniMax H3, FLUX.2, Wan2.2, and LTX-2 almost weekly.
Videos
-
“VEO 3.1 Camera Controls, Nano Banana 2 Leaks, Open-Source QWEN Image Edit & Sora 2 Update”
2025-11-10 / youtube.com/watch?v=SJvnEvOvEYk
An episode spanning closed and open models in a single video. It covers added camera controls for Veo 3.1, unconfirmed Nano Banana 2 leaks (a “temporal coherence mapping” note suggesting video-diffusion applications), advances in open-source Qwen Image Edit, and Sora 2 updates. -
“Google won image generation (it's not even close) NANO BANANA PRO BREAKDOWN”
2025-11-26 / youtube.com/watch?v=UV9GqinedQ8
An explainer on Nano Banana Pro, based on Gemini 3 Pro. It emphasizes progress from Nano Banana’s 1024×1024 limit to 2K/4K generation and highly accurate text rendering, arguing that “Google won image generation.” -
“Ultimate Runway Gen 4.5 Tutorial For New Users in 2026”
2026-03-29 / youtube.com/watch?v=IUge3u4ZrqQ
A tutorial for closed-source Runway Gen-4.5 (formerly codenamed Whisper Thunder/David). It presents the model as reaching a level where live-action and AI video are indistinguishable, and discusses its effect on the stock-footage industry. -
“Seedance 2.0 vs Kling AI 3.0” comparison-video series (posted by multiple channels around the same period)
Early 2026 / example: several YouTube explainers linked from a comparison article for Chinese closed models
ByteDance Seedance 2.0 (2K, support for 12 inputs, one-pass audio generation) and Kuaishou Kling 3.0 (4K/60fps, specialized lip sync) were released around the same time, accelerating the closed-provider China-vs.-U.S. framing involving OpenAI and Google. -
“HunyuanImage 3.0 : Open Source Text To Image Model (Forget Nano Banana & Seedream 4)”
2025-09-28 / youtube.com/watch?v=32lhiZKhgOY
Reports that Tencent released an open-source mixture-of-experts image model with more than 80 billion parameters. It is presented as competitive with closed Nano Banana and Seedream models. -
“China's Seedream 4.0 Just Outperformed Google's Nano Banana (And It's Open Source)”
2025-09-24 / youtube.com/watch?v=3Hmj05sE1V8
Covers ByteDance Seedream 4.0. However, while the video title calls it “open source,” Seedream is actually offered primarily through a closed API, so the information is inconsistent and requires caution (noted in Limits). -
“Is FLUX 2.0 The New Open Source King?” / “FLUX 2 Dev Released -GGUF 16GB Vram...”
2025-11-25–27 / youtube.com/watch?v=YQuTkPVkCS8 , youtube.com/watch?v=yR6TYLlPP-s
A wave of reviews immediately following Black Forest Labs’ FLUX.2 (32B, open-weight) announcement. Its ability to run on consumer GPUs with 16GB VRAM via GGUF quantization drew attention. -
“LTX-2 by Lightricks: Open-Source 4K Video Generation With Audio (50 FPS)”
Around 2026-01-10 / youtube.com/watch?v=t5X5JR_CAv4
Introduces Lightricks’ open video foundation model, LTX-2. Native 4K, 50fps, synchronized audio generation, and the release of training code are described as a milestone for open-source video AI. It was quickly integrated into ComfyUI, Fal, and Replicate. -
“Wan2.2 AI Video Tutorial & Demo | The best open-source video generation model to date just dropped!”
2025-07-28 / youtube.com/watch?v=Tqf8OIrImPw
Presents Alibaba’s open-source Wan 2.2 as a fully free alternative to Runway/Pika. It later expanded into pipeline videos combining it with Qwen Image and FLUX (youtube.com/watch?v=qMda40fB_m0). -
“HunyuanVideo 1.5 vs Wan 2.7 Comparison: Which is The Best Free Local Video AI Generation Model?”
2026-06-25 / youtube.com/watch?v=4GNT4-QYekg
A direct comparison between open-source video models, Tencent HunyuanVideo 1.5 and Alibaba Wan 2.7, assessed through the lens of being local and free. -
“How To Use MiniMax H3 in ComfyUI: Best FREE AI Video Model of 2026!”
August 2026 / youtube.com/watch?v=d_wEd-fZcdg
A guide to running Hailuo AI’s open-source video model MiniMax H3 locally in ComfyUI. It emphasizes avoiding subscriptions and completing the workflow on one’s own computer; the model has generated many derivative tutorials, including “Multi-Keyframes” and “Video-to-Video.” -
“Ideogram 4 in ComfyUI: The Best Open-Source AI Image Model for Text, Logos & Design?”
Around 2026-06-05 / youtube.com/watch?v=rRiQDMZTPiQ
Reports that formerly closed Ideogram became open-weight with Ideogram 4. It is described as offering performance comparable to closed frontiers and expanding open-source choices for text and logo generation.
Signals
- The closed-source side reacts extremely quickly through “breaking news → immediate comparison videos.” For Nano Banana Pro, Veo 3.1, Sora 2, Runway Gen-4.5, Seedance 2.0, and Kling 3.0, multiple channels posted side-by-side comparison videos within days of announcements. Weekly news channels such as TheAIGRID act as hubs.
- The open-source side has a powerful tutorial ecosystem around ComfyUI. After releases of Wan 2.2/2.7, Qwen-Image(-Edit), FLUX.2, LTX-2, HunyuanVideo 1.5, and MiniMax H3, “how to use it in ComfyUI” videos proliferate. VRAM requirements, especially the 16GB/24GB range, are major concerns in titles and comments.
- China-based models (Alibaba Wan/Qwen, Tencent Hunyuan, ByteDance Seedance/Seedream, Kuaishou Kling, MiniMax) are rising across both video and image generation, on both the open and closed sides. Comparisons with U.S.-based providers (OpenAI, Google, Runway, Black Forest Labs, Lightricks) recur frequently.
- There is confusion around the “open-source” label. Seedream-related videos are titled “open source” while the actual product is primarily a closed API, and commenters point out the discrepancy. Primary sources such as official blogs and Hugging Face should be used for verification before reporting.
- Sora 2 is in a downsizing phase. The standalone app/web version ended on 2026-04-26, and its API is reportedly scheduled to end on 2026-09-24. “Sora 2 Update” videos are already shifting toward a retrospective tone ahead of shutdown.
Limits
- Because YouTube search-result pages (
youtube.com/results) and watch pages (youtube.com/watch) are rendered with JavaScript, WebFetch retrieved only footer navigation links rather than direct details. Channel names, subscriber counts, and view counts could not be verified directly, so the review relies on WebSearch snippets. Exact channel names, subscriber counts, and view counts were not confirmed for some videos (#4, #6, #7, #8, #11, #12). - Despite instructions to prioritize uploads within the last 60 days, initial videos for major releases such as Sora 2, Nano Banana Pro, HunyuanImage 3.0, Seedream 4.0, FLUX.2, and LTX-2 mostly date from September 2025 through June 2026, outside the target period beginning 2026-07-18. Newer topics found within 60 days were limited to MiniMax H3 (August 2026), Midjourney V8.2 becoming the default on 2026-07-24, and the HunyuanVideo 1.5 vs Wan 2.7 comparison (2026-06-25). No entirely new model announcements from July through September were found.
- Individual video comment sections could not be retrieved because watch pages were not renderable, so direct quotations of viewer reactions were unavailable.
- Nano Banana 2 remains only a leak-stage report; no official announcement has been confirmed and follow-up is needed.
Bluesky
Bluesky — Latest Trends in Image and Video Generation AI
Accounts
- @replicate.com — The official account of model-hosting company Replicate. It posts whenever new image/video generation models, whether closed or open, become available through its API. It was the most information-dense account in this research.
- @testingcatalog.com (🚨 AI News | TestingCatalog) — An AI-news account focused on leaks and new features. Posts almost daily and frequently covers new video/image models.
- @simonwillison.net — A prominent AI experimenter who tests each new model with “pelican image generation tests,” producing engagement far above other accounts.
- @stabilityai.bsky.social — The official account of Stability AI, creator of Stable Diffusion. It has not posted since 2025-11-19, and its posts over the last year were almost entirely about audio models (Stable Audio), with virtually no image-generation content.
- @ai.bots.law (AI Cases Bot) — A bot that automatically posts new filings in AI-related litigation. It can be used to follow image-generation AI copyright cases such as Disney v. Midjourney.
- @mjsrefs.bsky.social, @un1v3rse.bsky.social, and others — Art accounts posting Midjourney works daily. Their posts chiefly show community energy rather than news.
Posts
- Nano Banana Pro (Google) — Replicate posted a report after using it heavily for 24 hours: “pushing Nano Banana Pro to its limit.” 2025-11-21, ♥2 🔁0. link
- Gemini 3 Pro — Replicate announced API availability for “Google's most advanced multi-modal reasoning model.” 2025-11-19, ♥2. link
- Veo 3.1 (Google DeepMind) — Replicate posted its test results: “We've spent some time testing Veo 3.1.” 2025-10-17, ♥3. link
- FLUX.2 [klein] 9b (Black Forest Labs) — A four-step distilled open-weight model was added to Replicate and described as delivering “outstanding” performance. 2026-01-23, ♥1. link
- Wan 2.2 Animate (Alibaba, open source) — A Replace-feature update cut runtime from seven minutes to under two minutes, with an embedded video. 2026-01-29, ♥2. link
- Kling 3.0 (Kuaishou) — Replicate introduced it as “Multi-shot 4K generation with synchronized audio.” 2026-02-17, ♥2. link
- Runway Gen-4.5 — Replicate announced, “Create stunning cinematic videos.” 2026-02-18, ♥2. link
- Seedream 4.5 (ByteDance) — Replicate introduced it with “Rich world knowledge, stronger spatial understanding.” 2025-12-03, ♥2. link
- ChatGPT Images 2.5 (OpenAI) — TestingCatalog reported faster image generation, sharper output, tighter editing control, and additional Flare/Sunburst models for the API. 2026-09-09, ♥2 🔁1. link
- Muse Video (Meta) — TestingCatalog published an exclusive early-access report, describing high-quality 10-second video generation and native audio support, with distribution planned for Instagram/Facebook. 2026-08-19, ♥2. link
- Orbis (Visko, real-time video model) — TestingCatalog introduced a new model from an emerging startup, claiming “4K/24fps with sub-second updates” and the top result in prompt-following tests. 2026-09-01, ♥0 (still relatively unknown). link
- Blender work made with GPT-6 Astra (Simon Willison) — A workflow post describing “generate concept images with ChatGPT Images 2.5 → pass them to Codex to make a Blender file.” 2026-09-05, ♥405 🔁30 — by far the highest-engagement post in this collection. link
- Disney v. Midjourney (lawsuit) — AI Cases Bot automatically posted the latest filing (Doc #195: proof of service). Disney’s ongoing lawsuit accuses Midjourney of copyright infringement and character generation. 2026-09-16, ♥0 (as a court-news bot). link
- Example Midjourney community post — A personal account posted “Personal post with AI-generated imagery,” receiving ♥15 🔁5, relatively high among the individual art posts reviewed. It is not news, but indicates community activity around Midjourney. 2026-09-16. link
Signals
- The closed-source competitive arena is announcement-driven integration through Replicate: Google (Nano Banana Pro, Veo 3.1, Gemini 3 Pro), OpenAI (ChatGPT Images 2.5, GPT Image 2), ByteDance (Seedream/Seedance), Kuaishou (Kling), Runway, and xAI (Grok Imagine Image) continued announcements via the Replicate account at a pace of roughly one or two per month from October 2025 through April 2026. China-based providers, especially ByteDance, Kuaishou, and Alibaba, were particularly frequent in video-model posts.
- The open-source side is led by Black Forest Labs (FLUX.2) and Alibaba (Wan 2.2), but organic discussion on Bluesky is relatively subdued. Established player Stability AI stopped posting after 2025-11-19, and its posts in the past year centered on audio, so its visible role in driving open-source image generation has almost disappeared on Bluesky.
- Copyright litigation is major negative news for closed-source players: Disney v. Midjourney and Alcon Entertainment v. Tesla (alleged unauthorized generation related to Blade Runner 2049) remain in active proceedings, with AI Cases Bot reporting new filings almost daily.
- Like counts are generally modest (from a few to a few hundred), with no viral surge comparable to Twitter or Reddit. The sole clear exception is Simon Willison’s hands-on test post (♥405), suggesting that primary information from actually trying models earns the strongest response on Bluesky.
- General-user posts found through Bluesky’s live search were largely personal accounts sharing Midjourney-made works (#midjourney, #AIart, #AIイラスト), rather than news. Showcasing creations outweighed technical discussion of models.
Limits
- Bluesky’s official post-search API (
app.bsky.feed.searchPosts) intermittently returned 403 errors and could not be used reliably for some keywords. Queries including “Sora,” “Nano Banana,” “Flux.2,” and “Runway AI” often failed, while “Midjourney” worked once. Research therefore used known relevant-account feeds (Replicate, TestingCatalog, Stability AI, Simon Willison, and others) rather than search. - Because
bsky.appis JavaScript-rendered, directly opening pages with WebFetch did not retrieve post text; collection was limited to public API endpoints. - Because individual searches for models such as Wan 2.2 and FLUX.2 could not be used, the organic volume of discussion across Bluesky as a whole could not be measured. The results may be biased toward information from Replicate and TestingCatalog as sources.
- For some notable open-source models, including Qwen-Image, HunyuanImage, and Ideogram, no verifiable Bluesky posts were found during the allotted research time; they were omitted as no applicable findings.
- Japanese-language Bluesky accounts related to image-generation AI, including feeds from gigazine.net and aitry.bsky.social, were checked, but no recent image/video generation posts were found.
Lemmy
Lemmy — Latest Image and Video Generation AI News
Communities
- !«メールアドレス» — 5,711 members. Established in 2023. The main hub for AI image-generation discussion on Lemmy.
- !«メールアドレス» — 2,361 members. Primarily for artwork posts. Bot account “Even_Adder” automatically republishes new Hugging Face / GitHub models and LoRAs almost daily, effectively making it a news feed for open-source image and video generation models.
- !«メールアドレス» (50 members), !«メールアドレス» (104 members), !«メールアドレス» (52 members), !«メールアドレス» (97 members) — Small niche spin-off communities organized around themes.
- !«メールアドレス» — 51 members. A bot community mirroring RSS from Reddit communities such as /r/ArtificialInteligence.
- !«メールアドレス», plus general technology communities across federated instances including lemmy.zip, lemmy.today, awful.systems (!techtakes), and lemmy.bestiver.se (!hackernews) — Closed-source AI company news is concentrated here. Subscriber counts are dispersed across instances and could not be accurately aggregated.
- !«メールアドレス» (mirrored to piefed.social and others) — A dedicated anti-AI community, useful for gauging opposition to generative AI.
Posts
Open-source
- A series of MiniMax H3 derivative-tool posts (2026-09-08–09-13, !«メールアドレス», posted by Even_Adder)
circlestone-labs/MiniMax-H3-Image-Training-Adapter(score 5, 2026-09-12) https://lemmy.dbzer0.com/post/75382765Speach1sdef178/ComfyUI-VDN-H3-24GB(accelerated implementation for 24GB VRAM, score 6, 2026-09-12) https://lemmy.dbzer0.com/post/75382764F16/krea2-turbo-sda(score 2, 2026-09-10) andlvladikov/Krea2-Turbo-Distill-4step-LoRA(score 2, 2026-09-08) https://lemmy.dbzer0.com/post/75284619 / https://lemmy.dbzer0.com/post/75150626- MiniMax H3 is a 33-billion-parameter open-weight video-generation model announced as API-only on 2026-07-31, with weights released on Hugging Face on 08-03. It supports 2K/24fps, up to 15 seconds, and synchronized audio generation, and ranked first in Artificial Analysis’s video-editing category. The stream of Lemmy posts shows continued active development of tools around the model in September.
Lightricks/LTX-2.5-22b-IC-LoRA-Ingredients(score 3, 2026-09-12, !stable_diffusion) https://lemmy.dbzer0.com/post/75382763 — A reference-image-control LoRA for Lightricks’ open video model LTX. A mirrored post also exists on lemmy.world (https://lemmy.world/post/51841873).- Introduction to the “Uncertainty DMD” paper (SIGGRAPH Asia 2026, score 3, 2026-09-12, !stable_diffusion) https://lemmy.dbzer0.com/post/75382766 — Research on restoring diversity in few-step autoregressive video generation.
- Daily artwork posts in !stable_diffusion_art (2026-09-11–09-16, scores 1–21), for example “Watch at the Burning Falls” (score 16, 09-15) https://lemmy.dbzer0.com/post/75490978 and “Shining Bay Steps” (score 21, 09-13) https://lemmy.dbzer0.com/post/75414322 — Descriptions indicate frequent use of Krea2 and Flux-family models.
Closed-source
- Midjourney pivots from image-generation AI to “body-scan medical spas” (score 66 / DL3, 10 comments, around 2026-06-18, !technology, The Register article) https://lemmy.zip/post/66397234 — It began offering full-body scan diagnostics immersed in “golden light,” without FDA approval. Comments compared it to Theranos. Related mirror: https://lemmy.bestiver.se/post/1175129
- Midjourney demands disclosure of AI use from major Hollywood studios (score 76 / DL2, 2 comments, 2026-07-05, !technology, TechCrunch article) https://lemmy.today/post/56011946 — Amid litigation with studios, Midjourney sought disclosure on the argument that “you are using AI too.”
- OpenAI effectively shuts down Sora (multiple posts, 2026-03–05)
- “How OpenAI scrapping Sora video-generation app points to one of the biggest problems facing technology companies” (score 2, !openai) https://lemmy.world/post/45568787
- “Sora losing $1 million daily” (score 1, 2026-03-30, !kotaku) https://ibbit.at/post/214890 — Reports said daily losses of $1 million contributed to shutdown.
- “OpenAI shutting down Sora just killed a $30M AI movie” (score 1, 2026-05-27, !ai_reddit) https://lemmy.durstig.online/post/34669 — A report that an AI film project valued around $30 million and underway on Sora was caught up in the closure.
- Adobe adds AI chatbots and Firefly-generated audio to Photoshop/Premiere (score 22, 2025-10-29, !boycottus) https://alternativeto.net/news/2025/10/adobe-adds-ai-chatbots-to-photoshop-premiere-ai-masking-and-generative-audio-in-firefly/ — Somewhat old, but still discussed in the context of closed providers’ “app integration” strategy.
Signals
- Lemmy communities related to generative AI are effectively close to “open-source only.” The center is the !stable_diffusion ecosystem on lemmy.dbzer0.com, where content consists not of news articles but daily generation posts and automatic reposts by bot account Even_Adder of new Hugging Face/GitHub models and LoRAs.
- In just the most recent week (09-08–09-16), posts about tools around MiniMax H3—an open-weight video model announced on 07-31 with weights released on 08-03—appeared consecutively, making it the largest focus for open-source video-generation users. Acceleration LoRAs for Krea2 Turbo were also active in parallel.
- Closed-source players (OpenAI, Midjourney, Adobe, Google, xAI) barely appear in Lemmy’s core AI communities. They surface only once or twice per month in generic news communities such as !technology, and the stories are more often about failures, controversy, or lawsuits than new features: Sora’s effective closure, Midjourney’s medical-device pivot, and Midjourney’s courtroom battle with Hollywood.
- Opposition to AI-generated material is also strong across Lemmy. !fuck_ai exists as a distinct anti-AI community, and there were multiple reports of users being banned from other communities for AI-related posts, reflecting broad caution toward generative AI.
Limits
- Lemmy is small, with even major AI communities having only several thousand subscribers. It does not have the breaking-news speed or breadth of Reddit or X.
- Its search API intermittently returned 503/500 errors. Some queries, including Midjourney, “Nano Banana,” “Veo 3” OR “Sora 2,” and direct !stable_diffusion_art searches, initially returned no results or errors and needed retries.
- No direct Lemmy posts were found mentioning major closed-source model names current in 2026, including “Nano Banana,” “Seedream,” “Grok Imagine,” “Kling,” and “Veo 3.1” (hits were only external AI-tool introduction sites). These topics appear to have almost no circulation on Lemmy.
- A link post to a Dutch article claiming that Elon Musk/xAI made an AI film of “The Odyssey” with Grok Imagine was found, but the article text on ad.nl could not be retrieved and fact-checked, so it was excluded from the body.
- Against the completion target of analyzing around 20 posts per social network, the total number of directly relevant Lemmy posts did not reach 20. More than 10 open-source items from dbzer0 model/LoRA posts were verified, but direct closed-source mentions numbered only around five or six, reflecting the platform’s small scale.
Pinterest — Image and Video Generation AI News
There was no genre categorization: 50 pins collected through the single query “Image and Video Generation AI News” are displayed flatly (output/pinterest.pins.md). The following is based on a visual review of all 50.
Visual themes
- Dark, neon, sci-fi tech backgrounds dominate — Navy-to-black backgrounds with blue-purple circuit patterns, glowing holograms, and luminous “AI” wordmarks make up nearly half of the material
[1, 3, 4, 7, 13, 14, 17, 23, 24, 25, 26, 27, 32, 34, 37, 38, 40, 42, 44, 46, 50]. Both news and promotional pins lean into this look. - White humanoid android/cyborg faces are repeatedly used as symbols for AI itself — A robot news anchor in a suit
[27], close-ups of female android side profiles[32, 37], a contemplative robot[26], and a robot looking at a phone[3]demonstrate how “human-like AI” has become a visual template. - “Which one is AI?” quiz format — A/B comparisons of mountain scenery
[11], a close-up eye marked REAL vs AI[35], and a cat photo stamped “FAKE”[3]show that authenticity quizzes have become a standard content format. - Large quantities of YouTube-thumbnail “AI news” videos are mixed in directly — Bearded men with surprised or confused expressions alongside OpenAI, Google, Microsoft, Meta, Anthropic, and Notion logos
[4, 18, 41], plus dramatic headlines such as “MIND BLOWING AI NEWS” and “TOP 5 AI NEWS”[19, 40]. Many appear to be direct reposts of YouTube thumbnails rather than original Pinterest content. - “Recommended AI tools” ranking infographics — Nearly identical layouts rank Runway, Pika, Kling AI, Luma (Dream Machine), Canva, and Hailuo from first to sixth
[9, 28](template reuse). A large chart organizes 20 paid tools by function—text-to-video, image-to-video, editing, audio, assets, and publishing[36]—alongside a paid-vs.-free comparison chart[12]. - Showcases of actual generated work — Google Veo 2 promotion visuals (a woman in a yellow raincoat, a dog swimming in a pool, flamingos, an anime-style girl) appear in almost identical layouts twice
[6, 8]; there is a Vidu S1 real-time interactive-generation demo[39]; and a Lumiere-like multi-material transformation demo that converts a person video into “wooden blocks,” “origami,” “Lego,” and “flowers”[10]. - Images warning about fakes and misinformation — A Facebook screenshot with an “AI-GENERATED IMAGES” banner shows a composited giant cat walking through Springfield and a cat-based anti-immigration protest image
[45]; there is also an audio-detection meter stating “97% likely AI generated”[37]and a pros/cons infographic[26]. - References to synthetic newscasters and AI actors — A screenshot of The Hollywood Reporter discussing AI actress “Tilly Norwood” on the TV show The View
[2], a Wondershare Virbo advertisement styled as a live synthetic news-anchor broadcast[22], and an ad for a news-channel generation tool promoting “AI Lip Sync”[48].
Notable pins
[10](untitled) — A grid converting a person video into wooden blocks, origami, Lego, and flowers. It has strong visual impact as a technical demonstration of consistently transforming one video into multiple materials.[45]AI creations give life to harmful Springfield,... — A cautionary introduction to AI-generated content using images of giant cats and AI-composited anti-immigration demonstrations that recall the 2024 Springfield “eating cats” misinformation story.[2]Hollywood Reporter Interviews the First AI Actress... — A screenshot from discussion of AI actress Tilly Norwood on a broadcast talk show, evidence that the AI-talent issue is reaching mainstream news.[39]ShengShu Technology Unveils Vidu S1... — A promotional visual for a China-based model claiming “real-time interactive generation,” showing that the closed-source landscape includes more than U.S. companies.[47]Transform your ideas into stunning visuals! — Hyperrealistic AI art of a marble sculpture playing a cello, showing how far photorealism has progressed.[36]AI Tools for Video Editing — A quick-reference infographic covering 20 tools across generation, editing, audio, and publishing, useful for understanding the current tool stack.[31]Today's AI News (May 20 2026) — A breaking-news-style infographic combining Google I/O 2026’s Gemini 3.5 Flash/Omni, AI smart glasses, and Apple accessibility features.[33]This Week in AI News - Aug 15 2026 — A retro newspaper layout listing buzzword headlines such as Sovereign AI and Agent Harness. The headlines themselves lack specificity, making it representative of templated and padded AI-news content.[6]/[8]Veo 2 promotion (near-duplicate pins) — Google video-generation examples featuring a swimming dog, flamingos, and an anime-style girl repeatedly circulating on Pinterest.[9]/[28]“Best AI Video Generation Websites” (near-identical template pins) — A reused ranking of Runway/Pika/Kling AI/Luma/Canva/Hailuo, suggesting an influx of affiliate-oriented list content on Pinterest.
Signals
- Closed-source providers’ product announcements and comparison articles—Google Veo 2/3, Gemini, OpenAI, Nvidia, Vidu/ShengShu—make up most pins. Pinterest is strong for practical searches focused on tool selection and how-to summaries.
- Open source is almost entirely absent; “Stable Video Diffusion” appears only once in the large tool list
[36]. The atmosphere of OSS/local-generation communities around Wan, HunyuanVideo, and ComfyUI, visible on Reddit and X, is nearly absent on Pinterest. - The “spot the real versus AI” format
[3, 11, 35], misinformation warnings[45], and audio detection[37]show a meaningful amount of skepticism and media-literacy content rather than uncomplicated celebration. - Interest in synthetic AI presenters and AI talent—Tilly Norwood, Virbo news anchors, and AI Lip Sync tools—appears across several pins, suggesting that “video content without human performers” has become a trend phrase.
- Near-duplicate infographics, including Veo 2 promotions
[6][8]and tool rankings[9][28], show how Pinterest-specific repinning and template reuse amplify SEO-oriented content; this requires discounting when interpreting the results. - No Japanese-language or Japan-originating topics were found. All collected results originated from the English-speaking web, primarily the United States.
Limits
- The worker collected 50 items with a single query and did not divide them into genres, such as closed source versus open source. Accordingly, this file cannot follow the playbook’s expected “section by genre” structure and instead describes all results together.
- Around eight pins had little usable information, with blank titles or titles such as “AI” only
[6, 11, 15, 19, 31(legible in practice), 38, 46]; some judgments therefore relied only on image content. - Individual Pinterest pages were not opened to verify saves, comments, or linked articles. Engagement counts and the authenticity of source-article URLs remain unverified.
- No direct browser searching was performed; in accordance with the playbook, only images previously collected by the worker were used.
Recommended actions
- For renewed X collection, use topic-derived terms such as Nano Banana, Seedance, Wan 2.2, and Sora instead of relying on a trending list.
- Continue watching the stableDiffusion-related Lemmy feeds as leading indicators for open-source video models.
- Track Nano Banana 2’s official announcement and the Sora 2 API shutdown (2026-09-24) as turning points for closed providers.
- Monitor ongoing lawsuits such as Disney v. Midjourney as factors that may drive strategic shifts among closed-source companies.
- Use multiple topic-specific queries when recollecting Reddit and Pinterest material to meet the completion threshold of roughly 20 items.
- Verify models that claim to be “open source” against primary sources such as Hugging Face before publication.
Collected images


















































Data quality note
Because X search relied on a trending list, it produced effectively zero image- and video-generation AI posts. Reddit (11 items), Pinterest (50 items from one query), and Lemmy’s closed side (five or six items) also failed to meet the completion target of 20 items. YouTube had few new announcements within 60 days, while Bluesky had to rely on known-account feeds because of search API failures.



