Image and Video Generation AI News — 2026-09-09
Closed-source leaders Google’s Nano Banana Pro and OpenAI’s GPT Image 2 are in a two-horse race, while Sora is winding down. On the open-source side, Qwen-Image, FLUX, Wan, and HunyuanImage are approaching closed-model quality.
Image and Video Generation AI News Roundup — 2026-09-09
Alright, I pulled it all together 🔥 Bottom line: on the closed side, Google’s Nano Banana Pro (Gemini 3 Pro Image) and OpenAI’s GPT Image 2 are battling for the top spot. The standalone Sora app is nearing its end, while video generation has become a main battleground for Chinese players and Google, including Kling, Seedance, and Veo 😤 Meanwhile, on the open-source side, Qwen-Image, Z-Image, FLUX, Wan, and HunyuanImage have reached the point where you wonder, “It can do this for free!?” The ComfyUI and local-generation communities are still going strong 💪 The only real miss this time was X, and I’ll be candid about that.
Across platforms
Closed-source takeaways (10) 🏢
- Nano Banana Pro (Gemini 3 Pro Image) is being praised as having “completely surpassed other image-generation models,” dominating YouTube discussion → https://www.youtube.com/watch?v=UV9GqinedQ8
- OpenAI’s GPT Image 2 has arrived. It reportedly delivers over 99% text accuracy in more than 12 languages, generates up to eight images per prompt in Thinking mode, and is gaining 4K support → https://www.youtube.com/watch?v=blOlUnC75O4
- Google fully retired Imagen on August 17, 2026 and consolidated around the Nano Banana line. Closed image generation is increasingly gravitating toward Google.
- OpenAI ended the Sora app on April 26, with the Sora API also scheduled for retirement on September 24. The Sora brand is effectively nearing its end → https://lemmy.world/post/44721138
- Bluesky user dahara1 has also reported that the free Sora tier no longer works after running out of credits → https://bsky.app/profile/dahara1.bsky.social/post/3miogq762u223
- xAI’s Grok Imagine Image 2.0, released on August 7, ranked a close second to GPT Image 2 on Arena, creating a three-way race.
- ByteDance opened the Seedance 2.5 API to the public on August 7, making it a major contender alongside Veo 3.1 and Kling 3.0.
- Reddit users compared Seedance 2.5 with open Wan 3.0 using the same prompt, with sharply divided opinions → https://www.reddit.com/r/generativeAI/comments/1w60fip/
- Midjourney has unexpectedly pivoted from image generation toward ultrasound medical scanners, drawing backlash and Theranos comparisons → Lemmy (via theregister.com)
- ByteDance and Hollywood’s MPA reached a global agreement on AI copyright protection, signaling a more serious push on copyright compliance → Pinterest research
- On Reddit, techniques for faking “real hand-drawn timelapses” with AI sparked a major backlash with 7,088 points, including calls for mandatory AI labeling → https://www.reddit.com/r/antiai/comments/1w81kdd/
- Bluesky alone, through luok.ai, saw a rapid succession of new Kling, Runway, PixVerse, Vidu, Seedance, and Niji models from December 2025 through February 2026, highlighting the momentum of Chinese players.
Open-source takeaways (10) 🛠️
- Qwen-Image 2512 (Alibaba, Apache-family license) and Z-Image Turbo are the two biggest trends, with comparison videos everywhere → https://www.youtube.com/watch?v=uAGH_yRZ2Gw
- Flux 2.0 is being put through a 15-category stress test against Qwen-Image 2512 and Z-Image → https://www.youtube.com/watch?v=54zvSiAWz2Y
- Lightricks’ LTX-2.3, a 22B-parameter model with audio sync and 4K/50fps support, is being praised as delivering closed-model-grade quality → https://www.youtube.com/watch?v=qQXzlk134Sw
- Some hands-on videos even claim Alibaba’s Wan 2.5 has surpassed closed-source Veo 3 → https://www.youtube.com/watch?v=yvD8TxNsR4Q
- WanGP, a free agent that combines multiple OSS models including LTX-2.3, Qwen, Flux.2 Klein, and Z-Image, has arrived, advancing the shift toward integrated workflows → https://www.youtube.com/watch?v=h-k6xOEITx4
- Tencent’s HunyuanImage 3.0, an 80B+ MoE model, remains a benchmark large open-source image model.
- Black Forest Labs’ FLUX 3 supports images plus 20-second videos with audio in a limited release → https://bfl.ai/blog/flux-3
- Lightweight FLUX.2-klein-9b-kv-fp8 variants and the FLUX Creator Program have launched, with Martin Scorsese joining as an advisor.
- Lemmy’s Stable Diffusion communities, though around 5,000 people in size, continue receiving daily ComfyUI, SDXL, and Krea2 posts, showing deep real-world use.
- Simon Willison reported on Bluesky that he ran Qwen 3.7 27B locally on an M5 MacBook, using a 17GB GGUF to make his best pelican image yet—part of a growing local-LLM image-generation benchmarking culture → https://bsky.app/profile/simonwillison.net/post/3mt2ynhrlk22z
- The open video foundation model SkyReels V1, based on HunyuanVideo, and the free all-in-one OSS workflow Open Creative Studio are also continuing to receive steady updates → https://bsky.app/profile/perilli.com/post/3mdddenkss22v
The cross-platform takeaway that stood out: Reddit, YouTube, and Lemmy repeatedly frame the closed-versus-open split as “easy but censored and expensive closed models” versus “free but more work-intensive open models.” Multiple communities also share skepticism toward tools promising “free and unlimited” use 🙅♂️. Pinterest was the exception: that binary barely showed up there, where tool explainers and news thumbnails dominated.
Platform by platform
Reddit: 12 threads collected. An anti-AI thread (7,088 points) went far more viral than the technical topics combined, illustrating how backlash against AI misuse is easier to make visible than enthusiasm for the technology. Core subreddits such as r/StableDiffusion produced no hits this time, leaving a coverage gap.
X: Frankly, this was a complete miss 😅 The worker’s queries included unrelated terms such as “Harvey” and “$LAPTOP,” so none of the 40 collected posts concerned image or video generation AI. This is the only genuine gap among the six platforms named in the brief.
YouTube: Strong coverage of review and comparison videos. Closed-source coverage spans Google, OpenAI, and xAI, while open-source coverage includes Qwen-Image, FLUX, Wan, LTX-2.3, HunyuanImage, and WanGP-based integrated workflows. View counts and subscriber counts could not be retrieved because they depend on JavaScript, so no quantitative recency comparison was possible.
Bluesky: The search API (searchPosts) returned 403 for every query, so the research switched to reading feeds through accounts. This produced solid coverage of closed video players through luok.ai, but thinner coverage of open video. SkyReels V1 was about all that surfaced.
Lemmy: Closed-source news centered on Midjourney’s medical pivot controversy and jokes about Sora’s shutdown. On the open side, Black Forest Labs and the FLUX line were dominant, with concrete developments continuing every quarter.
Pinterest: All 50 pins were visually inspected. Neon robot-face news thumbnails and tool-list infographics dominated; images representing self-hosted open culture such as Stable Diffusion or ComfyUI appeared in none of the 50. Engagement metrics were generally unavailable.
What to watch
- OpenAI’s scheduled Sora API retirement on September 24, 2026 — Lemmy https://lemmy.world/post/44721138
- Arena ranking changes for Grok Imagine Image 2.0 versus GPT Image 2 — YouTube comparison videos from channels such as Marin Method
- The open-source three-way battle among Qwen-Image, Z-Image, and FLUX—and which becomes the standard — YouTube https://www.youtube.com/watch?v=uAGH_yRZ2Gw
- Whether Midjourney’s ultrasound medical scanner venture can truly succeed — Lemmy (via theregister.com)
- The specific operating rules emerging from the ByteDance × Hollywood copyright agreement — Pinterest research [24]
- How widely OSS integration agents such as WanGP will spread — YouTube https://www.youtube.com/watch?v=h-k6xOEITx4
Recommendations
- Run additional X collection using directly relevant terms such as “Midjourney,” “Stable Diffusion,” “Sora,” “Nano Banana,” and “Qwen-Image.”
- Future Reddit research should always include large core subreddits such as r/StableDiffusion, r/midjourney, and r/comfyui.
- As a replacement for Bluesky search API access, turn the key accounts found this time—such as luok.ai and simonwillison.net—into a watchlist for ongoing monitoring.
- To make the closed/open divide visible, prioritize comment sections on opinion-rich platforms such as Reddit and Lemmy over Pinterest.
- Make OpenAI’s Sora API retirement on September 24 and follow-up reporting on Midjourney’s business pivot top priorities in the next research cycle.
Data quality
Because of mismatched search terms, X collected no image- or video-generation-AI-related posts and effectively contains no usable data. Bluesky and Pinterest could not provide search API access or engagement metrics such as view and like counts, so the spread of discussion cannot be quantitatively substantiated. YouTube video pages and view counts were also inaccessible due to JavaScript dependence, leaving the research reliant on web-search snippets. The remaining platforms, Reddit and Lemmy, produced generally sufficient material in both quantity and quality.
Platform summaries
Reddit — Image and Video Generation AI News
Where
The 12 collected threads are spread across eight subreddits.
| Subreddit | Members | Collected threads |
|---|---|---|
| r/generativeAI | 151,605 | 4 |
| r/antiai | 319,039 | 1 |
| r/RemoteWorkers | 75,979 | 1 |
| r/Seedance_AI | 20,508 | 1 |
| r/freetoolsAI | 7,942 | 1 |
| r/aivideomaking | 6,614 | 2 |
| r/AItips101 | 1,153 | 1 |
| r/Unrouted_AI | 517 | 1 |
The general generative-AI subreddit r/generativeAI had the most threads, with four. However, the single r/antiai thread critical of AI image generation dominated in score and comments, generating substantially more engagement than all other threads combined, as described below.
What people say
- Uncensored, high-flexibility integrated platforms are gaining ground: Thread 1 (r/AItips101, 81 points, 2026-09-07, https://www.reddit.com/r/AItips101/comments/1w9yftp/) introduces pay-as-you-go services without subscriptions, led by “WildOwl.ai,” which offer access to more than 35 models including Qwen Image, Flux, Seedream, Wan, Kling, Seedance, Veo, and GPT Image. In the comments, u/reliablemicrophone84 praised it for solving the problem of choosing between cutting-edge models and creative freedom, while u/Clear-Assistance449 offered Mage.space’s unlimited subscription as an alternative.
- Strong skepticism of “free and unlimited”: Threads 2 (r/aivideomaking, 2 points, 22 comments, 2026-09-06, https://www.reddit.com/r/aivideomaking/comments/1w953eo/) and 5 (r/generativeAI, 1 point, 2026-09-07, https://www.reddit.com/r/generativeAI/comments/1wa4k88/) share the prevailing view that truly free, unlimited services do not exist. u/ai_art_is_art in thread 5 (3 points) put it bluntly: “Unlimited is a lie. You put $550 in, you get $200 out and you have to wait 24 hours between generations.”
- A request for a 45-second video draws laughs: In thread 4 (r/aivideomaking, 13 points, 22 comments, 2026-09-07, https://www.reddit.com/r/aivideomaking/comments/1w9h5a9/), a request to make a free 45-second AI video using only a phone or iPad prompted u/NewLifeWares (2 points) to point out the technical barrier: “45 seconds is reaching data center level requirements of hardware. At best you might find a free 5 second generator...with a watermark.”
- AI fakes of “real hand-drawn” work trigger a huge backlash: Thread 3 (r/antiai, 7,088 points, 1,023 comments, 2026-09-05, https://www.reddit.com/r/antiai/comments/1w81kdd/) discussed GPT-6-class LLMs controlling a mouse and keyboard to fake “real timelapses” inside painting software. It had by far the highest score of all collected threads. u/Nayutantantan (2,512 points) expressed disappointment that fake-timelapse trends were becoming common, while u/Soggy_Supermarket100 (171 points) called for immediate mandatory AI filtering across every online space.
- Deepfakes of real celebrities converge on a “user responsibility” argument: In thread 6 (r/Seedance_AI, 0 points, 22 comments, 2026-09-07, https://www.reddit.com/r/Seedance_AI/comments/1w9zqac/), u/Jean_velvet (7 points) explained that recent court cases place responsibility on the user and that distribution and sharing are the potential crimes. Meanwhile, u/Alchemist42 (1 point) suggested that people use other, less restricted tools instead of Seedance.
- Closed versus open model showdown: Thread 10 (r/generativeAI, 110 points, 31 comments, 2026-09-03, https://www.reddit.com/r/generativeAI/comments/1w60fip/) compares Seedance 2.5, a closed ByteDance model, and Wan 3.0, an Alibaba open-weight model, using the same prompt. u/SpecialistDragonfly9 (5 points) found Seedance “by FAR superior,” while u/Positive-Key6640 (5 points) questioned the methodology itself, noting that same-prompt comparisons mostly measure which model happens to favor the prompt style.
- OpenAI’s new image model becomes a quiet topic of interest: Thread 7 (r/Unrouted_AI, 2 points, 6 comments, 2026-09-08, https://www.reddit.com/r/Unrouted_AI/comments/1waxp14/) links to OpenAI’s official “introducing-chatgpt-images-2-5” blog post. u/Ok_Homework_1859 was especially interested in the Sketch feature, which lets users draw in ChatGPT to guide image generation.
- Frustration with restrictions on copyrighted-character generation: In thread 11 (r/generativeAI, 2 points, 2 comments, 2026-09-08, https://www.reddit.com/r/generativeAI/comments/1waapc1/), users complain that Grok cannot even make personal-use fan art of Disney characters. u/frighten suggests PixelForge on iOS as an alternative with no built-in restrictions.
- Nostalgia and corrections around ten years of AI image generation: Thread 9 (r/generativeAI, 15 points, 9 comments, 2026-09-06, https://www.reddit.com/r/generativeAI/comments/1w8gzt6/) summarizes the evolution from StackGAN in 2016 to ChatGPT-generated images in 2026. u/noctalla notes a historical omission, pointing out that Google DeepDream preceded these around 2015.
- A new “AI side hustle” emerges around image generation: Thread 8 (r/RemoteWorkers, 169 points, 787 comments, 2026-09-04, https://www.reddit.com/r/RemoteWorkers/comments/1w7azwu/) advertises $30-per-hour remote review work validating AI-generated captions. It attracted 787 comments, but the collected top comments were mostly repeated “Images” requests for an application link, leaving little substantive discussion.
- Questions about the sustainability of free generation tools: Thread 12 (r/freetoolsAI, 40 points, 50 comments, 2026-09-04, https://www.reddit.com/r/freetoolsAI/comments/1w6w0rj/) introduces aifreeforever.com as free, unlimited, and registration-free. u/thecragmire questions the business model—“How do you maintain this? It's got to cost you something, right?”—while u/Relevant_Syllabub895 complains about strict NSFW restrictions.
Signals
- Rising: Aggregators emphasizing “uncensored, multi-model integration”—including WildOwl.ai, Venice, Mage.space, and PixelForge—are named in multiple threads as alternatives to subscriptions for individual models. Services promoting broad model choice and looser filters stand out more than individual models.
- Rejected claim: Video and image generation services that promise “free and unlimited” are met with nearly universal skepticism or rejection across r/aivideomaking, r/generativeAI, and r/freetoolsAI. The shared understanding is that either hardware cost or payment must occur somewhere.
- Conflicting views: The Seedance-versus-Wan comparison in thread 10 splits opinion both on which model is better and on whether prompt-matched comparisons are sufficient. Thread 6 also shows a divide between those stressing user responsibility and those expecting the service provider to act.
- Surprise: Of the 12 collected threads, the anti-AI thread (thread 3, 7,088 points / 1,023 comments) drew engagement on an entirely different scale from technical topics such as model comparisons and tool launches. Its score nearly matches the roughly 7,600 combined points of all the other threads. As a raw Reddit signal, this suggests that backlash against AI misuse and deception is more visible than excitement over AI-generation technology.
- The closed/open split: The clearest open-versus-closed framing appears in the Wan open-weight versus Seedance closed comparison in thread 10, and in the contrast between local generation using ComfyUI/Civitai/Hugging Face and paid cloud services in threads 1, 2, and 6. Open source is repeatedly described as freer but more labor-intensive, while closed tools are easier but constrained by censorship and cost.
Limits
- Only 12 threads were collected, below the brief.md target of roughly 20. The only query was “Image and Video Generation AI News”; no additional searches specifically targeted either closed-source or open-source models.
- Large core subreddits for image and video generation AI, including r/StableDiffusion, r/midjourney, r/aivideo, and r/comfyui, do not appear at all. This is a search-coverage gap, so the sample cannot be considered representative of Reddit overall.
- In thread 8 (r/RemoteWorkers), top comments were effectively just repeated “Images” application requests, providing almost no discussion value.
- Because this agent could not directly view Reddit itself or its search-result pages due to WebFetch rejection, the findings are limited entirely to the workers’ pre-collected
output/reddit.threads.mdand.jsonmaterial. The existence of other threads or comments could not be verified.
X
X — Image and Video Generation AI News
Preface: the collected results do not match the topic
Reviewing output/x.posts.md, the worker actually searched for the following ten terms during this stage:
“Harvey,” “#Crypto,” “#sweepstakes,” “#Chance,” “#giveaways,” “Serbia,”
“$LAPTOP,” “Arsenal,” “Germans,” and “Nazis.” None is related to image or video generation AI, and none of the 40 posts from 39 accounts matched the topic. Likewise, X’s own Explore list—trends localized to Croatia, where the server connected from—contained no AI-related items.
| # | Trend | Where |
|---|---|---|
| 1 | Harvey | Trending in Croatia |
| 2 | #Crypto | Business & finance · Trending |
| 3 | #sweepstakes | Trending in Croatia |
| 4 | #Chance | Trending in Croatia |
| 5 | #giveaways | Trending in Croatia |
| 6 | Serbia | Politics · Trending |
| 7 | $LAPTOP | Trending in Croatia |
| 8 | Arsenal | Sports · Trending |
| 9 | Germans | Politics · Trending |
| 10 | Nazis | Politics · Trending |
In accordance with rule 7, which prohibits fabricating sources, unrelated posts will not be presented as topical findings. The following records only facts from what was actually collected.
Accounts
The 39 collected accounts form unrelated clusters around reality-TV fandom, cryptocurrency trading, giveaways and promotions, Balkan political disputes, memecoins, Arsenal fandom, and disputes over historical revisionism and antisemitism. No account was found discussing image-generation AI or video-generation AI, nor posting content explicitly identified as created with those tools. Some posts were labeled “has an image,” but none mentioned generative AI or explicitly identified the image as AI-generated. The highest-engagement accounts were @saintjupitr (64,747 likes, one post) and @Kwin8675309 (42,021 likes, one post), both cases of unrelated political posts going viral.
Posts
None. Zero of the 40 posts matched the topic of image/video generation AI or closed/open-source developments. Individual posts—for example, Serbian political disputes in #21–24, Arsenal football posts in #29–32, and historical disputes about Germans and Nazis in #33–40—had high engagement, but are unrelated to AI image and video generation and are therefore not counted as findings.
Signals
- No topic-related rising or falling signals could be detected because there is no relevant data.
- The only usable observation is procedural: this session’s Explore output was geolocated to Croatia, and the fixed search queries were unrelated to the topic, such as “Harvey” and “$LAPTOP.”
Limits
- The search terms do not match the topic: Every “Found by searching ‘…’” record in
output/x.posts.mduses terms unrelated to AI image and video generation, including Harvey, #Crypto, #sweepstakes, #Chance, #giveaways, Serbia, $LAPTOP, Arsenal, Germans, and Nazis. This appears to be a problem with the worker’s query setup. Following the stage playbook instruction that the worker had already collected the material, no new searches were run. Collection needs to be redone with correct terms such as “AI image generation,” “Midjourney,” “Stable Diffusion,” “Sora,” “Veo,” or “オープンソース 画像生成AI.” - Of the 40 collected posts, zero contain insights related to closed- or open-source image/video-generation AI. The brief’s target of analyzing roughly 20 posts could not be met because no topical data exists.
- X Explore trends 1–10 were also just general Croatia-localized trends and contained no AI-related topics.
- The session itself was valid and posts were successfully collected; there was no “session expired” error. The issue is query selection.
YouTube
YouTube — Latest developments in image and video generation AI (closed source / open source)
Channels
- AICodeKing — Focuses on hands-on AI tool reviews, including early-access coverage of Google Nano Banana Pro (Gemini 3 Pro Image). https://www.youtube.com/@AICodeKing
- Theo - t3.gg — A developer-focused technology explainer channel with a Nano Banana Pro video. https://www.youtube.com/@t3dotgg
- ElevenLabs (official) — The official channel of the voice-AI company, also covering image-generation-AI news such as GPT Image 2. https://www.youtube.com/@Elevenlabs
- AIOnTrend — Specializes in comparative reviews of open-source image models such as Qwen-Image 2512. https://www.youtube.com/@aiontrend
- BMF MEDIA — Reviews open-source video models such as LTX-2.3. https://www.youtube.com/@BeerMoneyForum
- Marin Method — Produces head-to-head comparisons of closed video-generation models including Kling, Grok Imagine, Runway, and Veo. https://www.youtube.com/@MarinMethod
- Jack Vs. AI — Hands-on workflow channel covering models such as Wan 2.5. https://www.youtube.com/@JackVsAI
- Youri van Hofwegen — Compares and tests recent video-generation models such as Seedance 2.0. https://www.youtube.com/@Yourivanhofwegen
- AI Late to Class — Explains local open-source workflows for ComfyUI, WanGP, and related tools. https://www.youtube.com/@ailatetoclass
- AI大学【AI&ChatGPT最新情報】 — Japanese-language channel covering AI developments, including GPT Image 2 versus Nano Banana comparisons. https://www.youtube.com/@AIAIChatGPT-cj4sh
- Reference: Matt Wolfe (@mreflow) is an AI-news channel with roughly 965,000 subscribers that compiles weekly industry news. It was not a direct hit in this collection, but is a notable major channel in the field.
Videos
-
“Google won image generation (it's not even close) NANO BANANA PRO BREAKDOWN”
Channel: Theo - t3.gg|Posted: around November 26, 2025|https://www.youtube.com/watch?v=UV9GqinedQ8
→ A breakdown arguing that Google’s Nano Banana Pro (Gemini 3 Pro Image) overwhelms other image-generation models. A symbolic snapshot of the closed-source frontier. -
“Nano Banana PRO (Gemini-3.0-Pro-Image): I GOT EARLY ACCESS...”
Channel: AICodeKing|https://www.youtube.com/watch?v=13AovEj4oDM
→ An early-access review of Gemini 3 Pro Image, calling it striking and showing numerous examples. -
“GPT Image 2 Is Here — Everything You NEED to Know”
Channel: ElevenLabs|Posted: April 23, 2026|https://www.youtube.com/watch?v=blOlUnC75O4
→ Explains OpenAI’s GPT Image 2 (ChatGPT Images 2.0), including over-99% text accuracy across 12+ languages, up to eight images per prompt in Thinking mode, and 4K API/beta support. -
【Beyond Nano Banana】What Is OpenAI’s New Image Generation AI Model, “GPT Image 2”?
Channel: AI大学【AI&ChatGPT最新情報】 (Japanese)|https://www.youtube.com/watch?v=ZrO7tv-uv1E
→ Japanese-language overview of GPT Image 2 against Nano Banana, covering performance, use, and examples. -
“Kling 3.0 vs Grok Imagine vs Runway Gen-4.5 vs Veo 3.1: The Ultimate AI Video Comparison”
Channel: Marin Method|Posted: March 10, 2026|https://www.youtube.com/watch?v=x5fMQvhKP60
→ A direct comparison of four closed-source video models. It rates Kling 3.0 highest overall and Grok Imagine strongest for image-to-video. -
“Seedance 2.0 Just Destroyed Kling 3.0, Sora 2 & VEO 3.1 - Comparison”
Channel: Youri van Hofwegen|https://www.youtube.com/watch?v=vw_Jk2-phsA
→ A comparison arguing that ByteDance’s Seedance 2.0 surpasses leading existing closed models. -
“New Qwen Image 2512: Better Than Z-Image? (Open Source & Free)”
Channel: AIOnTrend|Posted: early January 2026|https://www.youtube.com/watch?v=uAGH_yRZ2Gw
→ Compares Alibaba’s open-source Qwen-Image 2512 with the trending Z-Image Turbo, highlighting the advantages of free, Apache-family licensing. -
“Is Flux 2.0 DEAD? Qwen image 2512 & Z-Image vs. The King (15 Stress Tests)”
https://www.youtube.com/watch?v=54zvSiAWz2Y
→ Tests Black Forest Labs’ Flux line against Qwen-Image 2512 and Z-Image across 15 stress tests, arguing that the open-source image-generation landscape is shifting. -
“LTX-2.3 Review: The Open-Source AI Video Model Built for Creators”
Channel: BMF MEDIA|Posted: around one month before the search, approximately August 2026|https://www.youtube.com/watch?v=qQXzlk134Sw
→ A creator-focused review of Lightricks’ LTX-2.3, a 22B-parameter open-source model with audio sync and 4K/50fps support. It is a fresh post from within the last 60 days. -
“WAN 2.5 CRUSHES VEO 3 (Higgsfield AI Workflow)”
Channel: Jack Vs. AI|https://www.youtube.com/watch?v=yvD8TxNsR4Q
→ A hands-on workflow video claiming Alibaba Tongyi’s open-source Wan 2.5 outperforms Google’s closed Veo 3. -
“This AI Agent Can Make Movies WanGP Deepy - Open source free Agent for LTX 2.3 Qwen Klein Z-Image”
Channel: AI Late to Class|https://www.youtube.com/watch?v=h-k6xOEITx4
→ Introduces an automated video-production workflow using WanGP, a free agent that combines LTX-2.3, Qwen, Flux.2 Klein, Z-Image, and other open-source models. It is evidence of increasing integration in the open ecosystem. -
“HunyuanImage 3.0 : Open Source Text To Image Model (Forget Nano Banana & Seedream 4)”
https://www.youtube.com/watch?v=32lhiZKhgOY
→ Covers Tencent’s open-source HunyuanImage 3.0, an 80B+-parameter MoE model with 13B active parameters. It claims parity with or superiority to Nano Banana. Though first introduced in September 2025, it remains a continuing reference point for large open-source MoE image models.
Signals
Closed-source developments
- Google fully ended the Imagen line on August 17, 2026 and consolidated around Nano Banana models such as Gemini 3 Pro Image. The core of closed-source image generation is increasingly converging on Google’s Gemini line.
- OpenAI ended the Sora app on April 26, 2026, with the Sora API itself scheduled to end on September 24, 2026. Sora as a standalone video product is effectively winding down, while OpenAI’s development resources appear to be moving toward GPT Image 2 and integration into ChatGPT.
- OpenAI’s GPT Image 2, released in April 2026, ranks near the top in Arena’s text-to-image and editing categories. xAI’s Grok Imagine Image 2.0, released August 7, 2026, is a close second, making closed-source image generation a three-way contest among OpenAI, Google, and xAI.
- In video, ByteDance’s Seedance 2.5 opened its API to the public on August 7, 2026 and is rapidly emerging as a key option alongside Google Veo 3.1 and Kling 3.0.
- Closed-source video comparisons increasingly settle into a division of strengths: Kling 3.0 for overall capability, Veo 3.1 for visual quality, and Runway Gen-4.5 for creative flexibility.
Open-source developments
- Qwen-Image 2512 (Alibaba, Apache-family license) and Z-Image Turbo, which runs in eight steps on 16GB of VRAM, are the two biggest image-generation trends. The number of comparison videos reflects their attention level.
- Black Forest Labs’ Flux.2 Klein, a 4B Apache 2.0 model that runs on RTX 3090-class hardware with 13GB VRAM, is frequently cited as a lightweight open-source option.
- Tencent’s HunyuanImage 3.0, an 80B+ MoE model, remains a reference point for large open-source image models.
- For video generation, Alibaba Tongyi’s Wan 2.2/2.5, including an 8GB GGUF version, and Lightricks’ 22B-parameter LTX-2.3, with audio sync and 4K/50fps support, are rated by multiple channels as having reached closed-source-grade quality.
- Videos about agents and workflows that combine multiple open-source models, such as ComfyUI and WanGP, are increasing. Interest is shifting from comparing individual models to comparing integrated workflows.
- Overall, review and comparison videos broadly share the view that the quality gap between open and closed source is narrowing, and that many creative teams now consider open source first.
Limits
- In this environment, YouTube result pages and video pages render through JavaScript, so WebFetch could not retrieve actual view counts, publication dates, or channel subscriber totals. Channel names were retrieved through YouTube’s oEmbed API, but views and subscriber counts were unavailable there as well.
- Accordingly, views and subscriber figures are treated as unknown. Only dates and contextual information visible in web-search snippets are included, and no numerical freshness comparison—such as views within the past 60 days—was possible.
- Video bodies and top comments could not be directly read because of YouTube’s JavaScript dependence. The discussion is therefore inferred from search-result titles, snippets, and related summaries, rather than direct quotations from the videos.
- The HunyuanImage 3.0 video first appeared in September 2025, outside the recommended 60-day recency window, and is mentioned only as a continuing reference point for large open-source image models.
- Only one Japanese-language channel hit was found, AI大学, so the depth of Japanese YouTube discussion could not be established relative to English-language coverage.
Bluesky
Bluesky — Latest developments in image and video generation AI
Accounts
- Simon Willison @simonwillison.net — A prominent AI developer who frequently compares image-generation ability using a benchmark image of “a pelican riding a bicycle” whenever a new model appears. Many posts receive 100–500+ likes.
- Ethan Mollick @emollick.bsky.social — A Wharton professor at the University of Pennsylvania who frequently posts experiments using models such as GPT-6 Astra and Fable 5.1 to generate 3D work in Blender.
- Alessandro Perilli @perilli.com — Developer of Open Creative Studio, formerly AP Workflow, an open-source all-in-one generation workflow for ComfyUI. He continues to announce a tool that can generate text, images, video, audio, and more.
- luokai @luok.ai — An account that rapidly covers updates to closed-source video-generation AI, including Kling, Runway, Vidu, PixVerse, Seedance, and Niji.
- dahara1 @dahara1.bsky.social — An engineer posting in Japanese about hands-on LLM and generative-AI use, including self-built machine-translation tools. The account covers current topics such as Sora’s end and Nano Banana comparisons. It is small, with 113 followers.
- Midjourney Experience Newsletter @midjourneynews.bsky.social — An unofficial Midjourney fan/newsletter account, not an official account. It posts prompt experiments and fragments of industry news with low engagement.
- Stable Diffusion online AI @stablediffusion.bsky.social — A feed for sharing SD/Midjourney work. Its recent posts stopped in January–February 2025, so it is not useful as a current-news source.
- Official-account search results: official Bluesky accounts for Black Forest Labs, Runway, Stability AI, and Civitai could not be confirmed, whether because they do not exist or their handles are unknown. The apparent official Midjourney account,
midjourney.bsky.social, is empty.
Posts
-
“Sora ends” — dahara1, 2026-04-28, 0 likes/0 reposts
https://bsky.app/profile/dahara1.bsky.social/post/3mkjnla6hnk2m
Comments on the end of OpenAI’s Sora video service, suggesting that OpenAI may not have been particularly invested in it. Web search also corroborated reports that OpenAI ended the Sora app and developer API around March 2026. -
Free Sora becomes unable to generate after credits run out — dahara1, 2026-04-04
https://bsky.app/profile/dahara1.bsky.social/post/3miogq762u223
Reports that video generation no longer works despite one credit remaining, consistent with the shutdown reports above. -
Watermark comparison: Nano Banana free tier versus ChatGPT Images 2.0 — dahara1, 2026-05-04, 1 like
https://bsky.app/profile/dahara1.bsky.social/post/3mkztj6phrk2l
Compares closed image-generation tools, saying ChatGPT Images 2.0 is easy to use because, like Nano Banana’s free tier, it lacks an intrusive watermark. -
Seedance 2 overwhelms other models — luokai, 2026-02-10, 5 likes
https://bsky.app/profile/luok.ai/post/3mehseuzlls2r
Shares an original video made in one day and praises the performance of the Seedance 2 video model. -
Niji 7 release — luokai, 2026-01-10, 3 likes
https://bsky.app/profile/luok.ai/post/3mc2422hgck2m
Announces the launch of anime-focused image model Niji 7, claiming sharper eyes, better consistency, and improved prompt adherence. -
Runway Gen-4.5 — luokai, 2025-12-02, 4 likes/1 repost
https://bsky.app/profile/luok.ai/post/3m6xsakpud226
Describes a higher bar for motion, fidelity, and prompt control, emphasizing its world-modeling basis. -
Kling IMAGE O1 — luokai, 2025-12-03, 3 likes
https://bsky.app/profile/luok.ai/post/3m72caiexw225
Covers Kling’s full-stack image-generation overhaul, promoted as able to accept, understand, and generate anything. -
Kling VIDEO 2.6 — luokai, 2025-12-05, 2 likes
https://bsky.app/profile/luok.ai/post/3m7aio7hntk2f
Promotes “story-first” video generation with audio sync, lip sync, and scene consistency. -
PixVerse V5.5 — luokai, 2025-12-02, 4 likes/2 reposts
https://bsky.app/profile/luok.ai/post/3m6xrxnh47c27
Promotes simplified production through one-tap multishot generation and automatic sound effects. -
Vidu AI Q2 image model — luokai, 2025-12-02, 5 likes/2 reposts
https://bsky.app/profile/luok.ai/post/3m6xrt43ll227
Promotes “4K generation in five seconds” and an unlimited-generation campaign through December 31. -
SkyReels V1, an open-source video foundation model — luokai, 2025-02-18, 3 likes/1 repost
https://bsky.app/profile/luok.ai/post/3lihm6qiuyk27
Markets itself as the first open-source human-centric video foundation model, fine-tuned from HunyuanVideo on more than 10 million videos to reproduce 33 expressions and more than 400 natural motions. Though over a year old, it was the only clear post on an open-source video model found through this account. -
Open Creative Studio v13.0, an open-source integrated ComfyUI workflow — Alessandro Perilli, 2026-01-26, 0 likes
https://bsky.app/profile/perilli.com/post/3mdddenkss22v
Announces a free, unlimited, highly capable all-in-one ComfyUI workflow that can generate text, images, video, audio, music, sound effects, and lyrics. -
Local image generation with Qwen 3.7 27B: pelican — Simon Willison, 2026-08-14, 183 likes/14 reposts
https://bsky.app/profile/simonwillison.net/post/3mt2ynhrlk22z
Reports that Qwen 3.7 27B, running via LM Studio as a 17GB GGUF on an M5 Max MacBook Pro, made the best bicycle pelican he had seen from a model running on his own laptop. It is an example of open-weight models’ local image-generation capability. -
GPT-6 Astra versus GPT-5.6 pelican comparison grid — Simon Willison, 2026-09-04, 202 likes/9 reposts
https://bsky.app/profile/simonwillison.net/post/3mupxubgrls2i
Shares a 12-image pelican grid comparing GPT-6 Astra with GPT-5.6 Sol, Terra, and Luna, continuing to benchmark closed-model image-generation ability. -
GPT-6 Astra generates Blender 3D architecture — Ethan Mollick, 2026-09-05, 163 likes/17 reposts
https://bsky.app/profile/emollick.bsky.social/post/3murua22igs2i
Reports that GPT-6 Astra created a Blender 3D model and narrated walkthrough of Boullée’s unbuilt 18th-century cenotaph from only a sketch and description. Another post from the same account on September 7 also observes that closed-source LLM visual generation is becoming a way to establish presence visually. -
“Midjourney became part of Meta.ai” — Midjourney Experience Newsletter, 2026-02-03, 1 like
https://bsky.app/profile/midjourneynews.bsky.social/post/3mdxymie7pi25
⚠️ Caution: This is an unofficial fan-account post and likely oversimplifies the facts. Web search found that Meta announced a partnership to license Midjourney’s “aesthetic tech” in August 2025; it was not an acquisition, and Midjourney remains independent (reference: TechCrunch, implicator.ai).
Signals
- Closed-source video generation is a crowded fight: In just December 2025 through February 2026, luok.ai’s feed alone shows at least seven model launches or major versions: Runway Gen-4.5, Kling IMAGE O1/VIDEO 2.6/O1, PixVerse V5.5, Vidu AI Q2, Seedance 2, and Niji 7. Chinese models—Kling, Vidu, PixVerse, and Seedance—stand out for their momentum.
- Sora is losing momentum and heading toward closure: dahara1’s posts suggest that free Sora credits had stopped functioning by April 2026, followed by an explicit “Sora ends” post later that month. This aligns with reports that OpenAI ended Sora’s app and developer API around March 2026. Leadership in video generation appears to be shifting toward Chinese model families and Google/OpenAI features embedded in broader general-purpose products.
- Image generation is becoming benchmarked as an LLM capability: Simon Willison and Ethan Mollick’s posts establish an ongoing culture of comparing the image and 3D-generation ability of general LLMs such as GPT-6 Astra, Claude Fable 5.1, and Qwen 3.7/3.8 27B, rather than treating image AI as a standalone tool category. GPT-6 Astra’s visual-generation ability, including Blender-based architectural reconstruction, is cited as a social-media marker of strength.
- Open-source video discussion is thin on Bluesky: Across searches and feed exploration, the only clear open-source video model post was the February 2025 SkyReels V1 post based on HunyuanVideo. No Bluesky references were found to open-weight models that should be prominent in 2026, including Wan2.2, Qwen-Image, and Z-Image. ComfyUI coverage is mostly limited to Alessandro Perilli’s recurring Open Creative Studio announcements, closer to one project’s updates than a broader wave of discussion.
- The Midjourney–Meta relationship is easy to misstate: An unofficial account simplified the relationship into “Midjourney became part of Meta.ai,” which differs from the actual technical-licensing partnership rather than acquisition. It is a useful example of why social-media claims require careful reading.
Limits
- Bluesky’s official search API (
app.bsky.feed.searchPosts) returned 403 Forbidden for every query in this research, despite the playbook procedure. Other public API endpoints such asapp.bsky.actor.getProfile,app.bsky.feed.getAuthorFeed, andapp.bsky.feed.getPostThreadworked normally, suggesting a search-endpoint-specific access restriction. - The web search page at
https://bsky.app/search?q=...is also client-side rendered, so fetching it did not retrieve post content. - Because of these limitations, discovery switched from keyword search to finding relevant accounts through web search and reading their feeds with
getAuthorFeed. This does not offer complete keyword coverage and depends on accounts surfaced by search engines, likely contributing to biased coverage—especially for open-source video generation. - Web searches for Bluesky posts about Wan2.2, Qwen-Image, Z-Image, and HunyuanVideo mostly returned arXiv papers and unrelated accounts, rather than relevant posts.
- Official Bluesky accounts for major image/video-generation AI companies including Black Forest Labs, Runway, Stability AI, and Civitai could not be located; handle searches either returned 400 errors or no relevant results.
- The
stablediffusion.bsky.socialfeed stopped in January–February 2025 and could not serve as a current source. - Against the completion criterion of analyzing about 20 posts per social network, more than 100 posts from six accounts were reviewed, from which 16 topical posts were selected.
Lemmy
Lemmy — Latest developments in image and video generation AI
Because Lemmy is a federated platform with distributed instances, research centered on lemmy.world API search (/api/v3/search) and lemmy.today, which mirrors communities from lemmy.dbzer0.com, supplemented by web searches such as site:lemmy.world. Approximately 25 posts were reviewed in total.
Communities
- !«メールアドレス» — Roughly 5,050 subscribers. The central community for open-source image generation in the Stable Diffusion ecosystem.
- !«メールアドレス» (
lemmy.todaydisplay name “Stable Diffusion Art”) — 2,362 subscribers. Receives multiple daily pieces made with SDXL, ComfyUI, and Krea2, often linking to civitai.com. - Stable Diffusion Anime, Stable Diffusion Abstraction, Stable Diffusion Witches, Stable Diffusion Furry, and Stable Werewolves—genre-focused communities in the
lemmy.dbzer0.comecosystem. Subscriber counts are not public, but these communities remained active with daily posts in September 2026. - c/«メールアドレス» — General technology community where major closed-source news, such as Midjourney’s medical-scanner controversy and Sora news, appears.
- c/genart — A small community with 47 subscribers that explicitly supports open-source and open-license generative-AI models.
- c/«メールアドレス» — A community centered on anti-AI sentiment. It is not dedicated to image or video generation AI, but a post about pushing back against dependence on AI tools reached a relatively high score of 83.
- c/euroai — A niche community posting lists of European open-source AI and LLM providers.
Posts
Closed-source leaning
- “Midjourney pivots from AI image generation to body scanning medical spa” (2026-06-18, score 58, technology) https://lemmy.world/post/... via theregister.com — Discussion of Midjourney’s apparent move away from image generation into ultrasound CT scanner technology for medical use. The Register notes that the underlying technology appears real, but the company has not disclosed partners.
- “Midjourney AI pivots to Theranos: Ultrasonic CT” (2026-06-19, score 40, techtakes) via pivot-to-ai.com — A sarcastic post comparing the move to Theranos.
- “I was wrong about the Midjourney ultrasound scanner” (2026-06-21, hackernews) https://twitter.com/MattZirwas/status/2068365802491834541 — A previously skeptical poster retracts their earlier criticism after confirming the facts.
- “Midjourney wants Hollywood studios to reveal details of their AI usage” (2026-07-05, score 73, technology) via techcrunch.com — In litigation with three studios, Midjourney seeks disclosure about how Hollywood itself uses AI.
- A reaction to OpenAI ending Sora: “OpenAI: decides to shut down their AI slop video generating model Sora” https://lemmy.world/post/44721138 — Mocks the end of the standalone Sora app, announced March 24, 2026 and stopped April 26, with the Sora 2 API scheduled for retirement September 24, 2026.
- “OpenAI collapses media reality with Sora AI video generator” https://lemmy.world/post/12056737 — Shares an article concerned that anonymous video will become less trustworthy.
- “OpenAI's new Sora video generator to require copyright holders to opt out, WSJ reports” https://lemmy.world/post/36665644 — Criticizes a policy requiring copyright holders to opt out.
- “Elon Musk maakt complete AI-film met zijn versie van The Odyssey” (2026-07-22, films) via ad.nl — Musk announces plans to make a feature-length AI film of The Odyssey using Grok Imagine.
- A shared ChatGPT revenue-analysis post (2026-07-15, ai_reddit) — Argues that most of Grok/xAI’s revenue comes from paid NSFW image and video generation content.
- “Video AI” — A question about Adobe Firefly video generation’s free allowance. Firefly offers two free uses, with Hailuo, Kling, Runway, Pixverse, and Hunyuan compared as alternatives.
Open-source leaning
- “Black Forest Labs launches FLUX 3 capable of generating images and 20-second video with audio” (2026-07-24, score 1, tech) https://bfl.ai/blog/flux-3 — A limited release of the multimodal FLUX 3 model, capable of producing images and 20-second video with audio.
- “FLUX.2-klein-9b-kv-fp8 optimized variant” (2026-03-14, score 8, stable_diffusion) via huggingface.co — A lightweight FLUX.2 derivative with KV-cache support.
- “Black Forest Labs Opening up the first FLUX Creator Program” (2026-05-06, score 3, stable_diffusion) — Announcement of a creator program.
- “Regisseur Martin Scorsese wordt adviseur bij AI-bedrijf” (2026-06-03, films) via nu.nl — Film director Martin Scorsese becomes an advisor to Black Forest Labs.
- “FLUX Chroma (The Perchance T2i model)” (about one year ago, Perchance community) via huggingface.co(lodestones/Chroma) — A derivative using NAG, or Normalized Attentive Guidance, with comments praising lip-sync depiction.
- Daily Stable Diffusion community examples from 2026-09-06 through 09, all made locally with SDXL/ComfyUI/Krea2: “Erika - Mahouka Koukou No Rettousei” (score 16), “Colors in Quiet Water” (score 14), “Giant Wire Walker” (score 11), and “Karin - Umi Monogatari” (score 13), all with civitai.com links.
- “Opendream = NEW! Stable Diffusion Automatic1111/ComfyUI alternative that uses nondestructive layering like Gimp” https://lemmy.world/post/3243205 — Introduction to an OSS UI supporting GIMP-like non-destructive layered editing.
- “Europäische KI- & LLM-Anbieter” (2026-02-15, euroai) — A list of European open-source AI providers, including image-generation models, emphasizing technological sovereignty and data protection.
- “Your Lemmy Crash Course to Free Open-Source AI” https://lemmy.world/post/76020 — An introductory guide to local-generation tools including text-generation-webui, Automatic1111, and ComfyUI + Pinokio.
Signals
- The biggest closed-source topic is Midjourney’s medical-pivot controversy (scores 40–73): Its business move from image generation to ultrasound diagnostic equipment prompted disagreement in technology, hackernews, and techtakes communities. Sarcastic reactions comparing it to fraud are prominent, but a previously critical poster later corrected their position, leaving the discussion unsettled.
- OpenAI Sora is repeatedly targeted as an emblem of “AI slop.” Discussion continues after the app’s March–April 2026 end, including the scheduled September 24, 2026 retirement of the Sora 2 API.
- Black Forest Labs and FLUX dominate the open-source side. FLUX 3, the lightweight FLUX.2-klein variant, the Creator Program, and Scorsese’s advisory role provide concrete developments each quarter.
- Everyday creative communities, including Stable Diffusion Art, Anime, and Abstraction, remain small at roughly 2,000–5,000 subscribers, but posts using SDXL, ComfyUI, and Krea2 locally continue daily in September 2026. This demonstrates the depth of open-source practical use.
- Anti-AI sentiment is also a meaningful signal: A c/fuck_ai post about wanting to break away from AI dependence reached score 83, above many ordinary AI-related posts. This indicates a persistent aversion to AI-generated material across Lemmy.
- Closed-source news on Lemmy is generally not discovered there first; it mostly consists of reposts and reactions to existing media coverage from TechCrunch, The Register, the WSJ, and similar outlets. Primary information lies outside Lemmy.
Limits
- Lemmy is small in both users and posts, and dedicated discussion of image and video generation AI is thinner than on Reddit. No deep-dive community focused specifically on video generation tools such as Sora, Veo, Kling, or Runway was found.
- Full-text search through
lemmy.world’s API is imprecise. Broad queries such as “AI image generation” surfaced unrelated posts, including UN maps and owl breeding. More focused searches for individual models such as Sora, Midjourney, and FLUX produced relevant material. - Some pages, including
lemmy.world/post/44721138and a DLSS 5-related post atlemmy.world/post/44347792, returned HTTP 500 errors or failed to provide content through WebFetch. Their titles and web-search snippets were used instead of their full bodies or comment details. - Direct Lemmy mentions of major current models including Grok Imagine, Qwen Image, HunyuanVideo, Wan2.2, Nano Banana/Gemini, and Veo were scarce. Grok Imagine appeared only once in a film-production context. These models seem to have limited visibility on Lemmy.
- Individual searches of lemmy.ml and lemm.ee largely duplicated lemmy.world results or surfaced the same federated posts mirrored across instances, without yielding distinct additional information.
Pinterest — Image and Video Generation AI News
All 50 collected pins were opened and visually checked. They were gathered as one undivided result set from a single “Image and Video Generation AI News” query.
Visual themes(visual themes)
- A flood of “robot face + breaking news” templates: Repeated thumbnails combine close-up halves of white humanoid robot faces with red and blue neon text such as “AI NEWS” or “BREAKING NEWS.” Examples include a robot at a news anchor desk [22], a surprised male YouTuber composited with a robot face [29], and a female robot captioned “THE FUTURE IS AI” [39]. Dark navy with neon cyan, magenta, and orange is the dominant palette.
- YouTuber reaction-face thumbnails: Several pins repeat the composition of surprised or skeptical white male faces alongside circular, neon-rimmed icons for OpenAI, Google, Meta, Anthropic, Grok, Notion, and other brands [4, 18]. They appear to be reposts of YouTube video thumbnails.
- Numbered listicle infographics: The most frequent format is vertical dark purple-to-black infographics with numbered cards listing tools. Runway, Pika, Kling AI, Luma AI/Dream Machine, Canva AI, and Hailuo AI recur in nearly the same lineup [9, 10, 27, 36, 50].
- Before/after and comparison-quiz formats: Side-by-side mountain images asking “Which is AI?” [13], plus grids converting source videos into characters made of wooden blocks, origami, LEGO, or flowers [11], emphasize both how difficult AI can be to identify and how flexible it is at transforming materials.
- Robot–human collaboration and medical scenes: A quieter visual style also appears, including photorealistic CGI of white humanoid robots operating holograms alongside healthcare workers [3, 33], as well as illustrations of doctors with stethoscopes next to robots [32].
- Dark, neon “tech” styling: Indigo-to-black backgrounds overlaid with cyan, magenta, and purple neon circuitry and glowing particles appear across roughly 70% of the results, suggesting a standard Pinterest aesthetic for technology-news imagery.
- AI-video-generation advertising using avatar or icon personalities: Several pins promote freelance AI video explainers in a Fiverr-like style, including a composite image using Jeff Bezos’s face [38], and AMD’s Amuse 3.0 photorealistic-generation demo featuring an elderly fisherman [46]. Public figures and photorealistic portraits are repeatedly used to signal output quality.
- Images about whether AI generation can be detected: Several pins visualize authenticity concerns, including a robot viewing a phone with a “FAKE” stamp and a cat photo, or comparing it with a masked figure silhouette [3], as well as AI-detection tool explainers [37].
Notable pins(notable pins)
- [11] Luma AI-style material-conversion demo: A grid compares five transformations of the same source video—a person, panda plush, and car—into wooden blocks, origami, LEGO bricks, and flowers. It is the most concrete pin demonstrating video-generation AI’s ability to transform texture and material.
- [24] ByteDance and Hollywood AI copyright agreement: A screenshot with the headline “ByteDance and Hollywood reach global deal on AI copyright.” It reports that ByteDance and the Motion Picture Association reached a worldwide agreement to strengthen intellectual-property protection around AI image and video tools. It is close to primary-source material on a major closed-source company’s copyright response.
- [35] AI news roundup dated August 12, 2026: Three stories: “Grok 4.6 drops,” describing autonomous work across dozens of steps with self-verification; “Gemini hits 1 billion users,” including claims that Google reached 100 million users faster than any product in its history, two of three users use voice input, and 150 million images are generated per day; and “Honor’s robot phone,” with a titanium arm extending from the rear to automatically track subjects. This is a rare pin containing explicit dates and figures.
- [36] “Best AI Video Generators of 2026” ranking: Google Veo 3.1, Runway Gen 4.5, and Kling 3.0 occupy the top three spots with crown icons, followed by Adobe Firefly, Seedance 1.5 Pro, Pika 2.5, and HeyGen. It lists current-generation model version numbers.
- [42] Vidu S1 by ShengShu Technology: Product visual for a Chinese real-time interactive AI video model described as “The World's Most Advanced Model for Real-Time AI Interaction.” One of the few examples showing a player outside the usual Western-model focus.
- [45] OpenAI new image-generation-tool preview article: An article image titled “OpenAI's Secret Image Generation Tool to Debut Soon,” designed to build anticipation for an unannounced tool.
- [25] Google I/O 2026 roundup: A May 20, 2026 overview of Gemini 3.5 Flash, Gemini Omni, Gemini Spark, the shift toward AI search, and Apple-versus-Google XR smart glasses. It shows surrounding ecosystem movement rather than image generation alone.
- [9] “Just Nail It” video-generation-tool comparison infographic: Organizes Runway, Pika, Kling AI, Luma AI, Canva AI, and Hailuo AI by function and intended user, and includes the statistic that 80% of online content will contain video by 2026.
- [8] Stock-photo AI content-creation icon set: An Adobe Stock-style “Ai Content Creation Concept” image with icons for generating text, images, music, and video. It says more about demand for stock materials than actual news.
- [43] Generative-AI history infographic: A purple-toned timeline from ELIZA in the 1950s to GPT in the 2020s, placing the current generative-AI boom in historical context.
Signals(signals)
- The open/closed binary is difficult to read from images alone: Pinterest pins largely consist of tool explainers, news thumbnails, and how-to summaries; no explicit open-versus-closed comparison like Reddit’s Seedance-versus-Wan framing was found. At most, Vidu S1 [42], from China’s ShengShu, and AMD-adjacent, more local-generation-oriented Amuse 3.0 [46] stand out as alternatives to major closed Western tools such as Google Veo, OpenAI Sora, and Runway.
- Roundups, comparisons, and how-tos dominate; breaking news is a minority: Of the 50 pins, only [24], [25], [35], and [45] are close to primary-news content with dates or concrete announcements. Most are evergreen, Pinterest-friendly listicles or affiliate-style content such as “Best AI Video Generators” and “10 Smart Ways.”
- Concern about authenticity and detection is visible: Multiple images, including [3], [22], and [37], feature “FAKE” stamps, robot journalists, or AI-detection tools, indicating ongoing interest and skepticism around the authenticity of generative-AI content among Pinterest users.
- Model generations are updating rapidly: Pin [36] lists current 2026 version numbers—Veo 3.1, Runway Gen 4.5, Kling 3.0, and Seedance 1.5 Pro. Together with [35]’s Grok 4.6 and Gemini 1B users, the images support the observation that version updates have become routine on a quarterly cadence.
- Absence: None of the 50 results contained imagery representing the local, self-hosted open-weight culture that is common on Reddit—Stable Diffusion, ComfyUI, or Hugging Face. Pinterest image results are heavily skewed toward commercial-tool and SaaS promotional material, with almost no visibility for genuinely community-driven open-source publishing.
Limits
- Collection was performed beforehand by a worker using a headless browser; this agent did not directly browse Pinterest, in accordance with the playbook.
- No genre split using
## <genre> (n)sections was made. All 50 pins are bundled results from a single query, so comparison by genre was not possible. - Five pins had the title “(untitled)”—[5, 13, 15, 20, and 24]. For [24], the article’s content could still be inferred from the headline shown in the image.
- Pin URLs in the form
pinterest.com/pin/...were recorded, but their destinations were not opened and verified. Judgments rely only on images and titles. Pinterest did not provide posting dates or engagement data such as likes, except dates embedded directly in images like [25] and [35]. - Against the completion target of at least 10 specific takeaways each for closed and open source, Pinterest alone contains little clearly classifiable open-source material—mainly [42], the Chinese Vidu model, and [46], AMD-related material. This gap needs to be filled through other platforms such as Reddit.
Recommended actions
- Recollect X using directly relevant terms such as Midjourney, Stable Diffusion, Sora, Nano Banana, and Qwen-Image.
- Include core large subreddits such as r/StableDiffusion, r/midjourney, and r/comfyui in future Reddit research.
- Use the key accounts discovered here as a watchlist for ongoing monitoring, as an alternative to Bluesky’s unavailable search API.
- Treat OpenAI’s scheduled Sora API retirement on September 24, 2026 and follow-up reporting on Midjourney’s business pivot as top priorities next time.
Collected images


















































Data-quality note
Because of mismatched queries, X collected no topic-related posts and contains no usable data. Bluesky, Pinterest, and YouTube had partial limitations around engagement metrics and search APIs, leaving some areas with weak quantitative support.



