Image and Video Generation AI News — 2026-09-06
OpenAI’s GPT-6 Astra is dominating the closed-source scene not so much as an “image/video generation” model, but as a “PC-operating agent,” while the open-source side is steadily building up implementation news around Qwen Image and MiniMax-H3 in the ComfyUI ecosystem.
Image and Video Generation AI News Roundup — 2026-09-06
Seriously, this week was packed with wild developments 😳🔥 On the closed-source side, it’s basically a full-on OpenAI “GPT-6 Astra” festival, with everyone excited less about image and video generation itself than about an agent that can operate an entire PC. Meanwhile, Sora seems to be quietly fading away, suggesting an investment shift from video-generation apps to agentic AI. On the open-source side, Qwen Image, Z-Image Turbo, HunyuanImage, and MiniMax-H3 have all been steadily receiving ComfyUI-related updates. It’s less flashy, but the implementation-news density is seriously high. To be candid, though, some platforms did not cooperate with search this time (Reddit, X, and Bluesky), so I was not able to properly review all 20 items everywhere. I’ll be upfront about that in the “Data quality” section below 🙏
Across platforms
- GPT-6 Astra (OpenAI) has nearly monopolized the closed-source conversation: It has dominated attention immediately after launch on both X (@masahirochaen, https://x.com/masahirochaen/status/2095644339041153256) and Bluesky (@todaystopainews, https://bsky.app/profile/todaystopainews.bsky.social/post/3mur7cyocf225). The benchmark result of 98.6% on ARC-AGI-3, up from 7.8% for the previous generation, is going viral.
- Astra is being discussed more as agentic control than image generation: Both X (@aigeboku, testing the ComputerUse feature) and Bluesky (a demo automatically building a city in Unity) are focused less on one-off image or video output and more on its ability to operate a PC autonomously. This genuinely looks like it could become a new category.
- Fable 5.1 is being used heavily as a comparison baseline: On X, @renoiseai and @Mayz1169 independently posted “Astra vs. Fable 5.1” comparison videos through Seedance 2.5. In @k_matsumaru’s post, Fable 5.1 was not the final renderer, but one component in a four-tool workflow as a prompt-generation layer. Almost nobody is treating a single model as a complete solution.
- Aggregator APIs are becoming more prominent in both camps: On Reddit (WaveSpeedAI, bundling more than 1,000 models through one REST API) and Bluesky (useapi.net, comparing Veo/Kling/Seedance/PixVerse/Hailuo/Runway/Nano Banana), both were seen as offering denser practical information than official accounts.
- Open source has a two-leader-plus-rising-star structure: Qwen Image / Z-Image Turbo (Alibaba/Tongyi-MAI), HunyuanImage (Tencent), and MiniMax-H3: There is fairly strong corroboration for this from both YouTube review videos and Lemmy’s !«メールアドレス».
- The MiniMax-H3 ecosystem, an open-weight video-model ecosystem, is expanding at breakneck speed: Lemmy alone had more than 10 posts from 8/8 to 9/5, including GGUF quantization, Turbo LoRA, 120-second long-video support, and hybrid-attention implementations. On X, @ImaStudio_ai was also found to be building its own commercial tool on top of it.
- ComfyUI has become the command center for open source: Across Lemmy, Bluesky (@kunecco), and YouTube tutorial channels, ComfyUI compatibility has become the litmus test for whether a technology is considered real and usable.
- Multi-shot generation plus audio synchronization is the next competitive axis shared by open and closed source: Wan 2.6 (open) and Kling 3.0/Seedance 2.0/Runway Gen-4.5 (closed) were directly compared in YouTube stress tests.
- The anxiety of “Is this AI? Is it real?” is everywhere: Whether it is Reddit jokes about broken hands, classroom debates, or Pinterest “Which is AI?” quiz pins, this concern shows up across platforms.
- Official brand accounts are nearly dead on social media: On Bluesky, the official Midjourney account (last post: 2025-10-08) and the official Kling AI account (zero posts) are essentially inactive. Across platforms, unofficial accounts, individuals, and aggregators appear to be faster sources of news.
Platform by platform
Reddit — Collected 12 threads from nine subreddits, short of the target of roughly 20. Highlights included intense backlash against a new method that lets an LLM operate a mouse to fabricate an entire hand-drawn timelapse (r/antiai, 5,259 points), WaveSpeedAI’s aggregator API, and fast local Krea-2-Turbo generation on Apple Silicon (45 seconds per image). r/VeniceAI also had a power-user post detailing how different models are used for image generation, editing, and video conversion. Negative sentiment around ethics and quality was also strong, including criticism of AI food images and apparent $30/hour job-post spam.
X — The Explore page’s trending terms were skewed by the server’s location in Croatia: seven of the ten terms (Bitcoin, Ukraine, Argentina, and others) were completely unrelated. Every relevant post was about GPT-6 Astra, while no open-source search terms such as Stable Diffusion, Wan, or Flux were tried at all. Therefore, the open-source situation on X for this report is not “nothing,” but rather “not examined.”
YouTube — The biggest closed-source story was Sora’s shutdown: the app ended on April 26, 2026, the API is expected to end on September 24, 2026, and downloads fell sharply from 3.3 million to 1.1 million in 11 months. On the open side, there were numerous first-test reviews of Qwen Image, HunyuanImage, and Z-Image Turbo, with Black Forest Labs’ FLUX.2 (Apache 2.0) also mentioned as a contender.
Bluesky — The full-text post-search API (searchPosts) returned 403 for every query, forcing a switch to account-search-based collection, which did not reach the target of 20 analyzed posts. The closed-source side was led by “news bots” such as todaystopainews and useapi.net, while the open-source side was led by finished work posted by individual creators such as Niji-iro and Kunecco. Almost no open-source technical news was flowing on Bluesky.
Lemmy — Small in scale, but !«メールアドレス» (5,708 members) functions as a de facto breaking-news board for open source, with dense implementation news around MiniMax-H3 and Day-0 support for LTX-2.5. Closed-source developments tend to be consumed more as satire or scandal, such as Midjourney’s move into medical devices, than as product news. No first-party Lemmy posts about Sora, Veo, Kling, or Hailuo were found.
Pinterest — Reviewed 50 results from one worker-precollected query, without genre segmentation. Almost all content consisted of comparison lists and news infographics about closed tools such as Runway, Pika, Kling, Luma, Synthesia, and HeyGen. Only two items mentioned open source: Baidu Ernie 4.5 and AMD Amuse 3.0. This is effectively a closed-source SEO and marketing market.
What to watch
- How far GPT-6 Astra’s “ComputerUse” agent capability enters creative workflows — X, https://x.com/aigeboku/status/2096187322924687799
- The identity of the higher-tier, not-yet-released model Altman hinted is “not Astra” — X, https://x.com/masahirochaen/status/2096372856930386341
- The expansion speed of the MiniMax-H3 GGUF quantization and Turbo LoRA ecosystem — Lemmy, https://huggingface.co/Abiray/MiniMax-H3-Pruned-GGUF
- The migration impact of the complete Sora API shutdown, scheduled for 2026-09-24 — YouTube, https://www.youtube.com/watch?v=H2jVcy5PuBI
- The timing of Qwen-Image-2.0 weight release (20B → 7B, outperforming Nano Banana in Elo ratings) — Lemmy, https://files.catbox.moe/l3mzzs.jpg
- Closed-model price and performance comparisons through aggregator APIs such as useapi.net — Bluesky, https://bsky.app/profile/useapi.bsky.social/post/3mrdyxqswcb2p
Recommendations
- Continue comparing GPT-6 Astra with Fable 5.1 not just on image and video quality, but on agentic control.
- Continue monitoring the MiniMax-H3 ComfyUI ecosystem (GGUF and Turbo LoRA) as the leading edge of open-source video generation.
- In the next X study, do not rely on Explore’s trending terms; explicitly search open-source-specific terms such as Stable Diffusion, Wan, Flux, and Qwen.
- Prioritize Lemmy’s !«メールアドレス» over Reddit or Bluesky as a primary source for open-source developments.
- Do not use Pinterest as a source for open-source trends; restrict it to observing marketing and SEO activity around closed-source tools.
- Follow up on whether any external tools or workflows depend on the Sora API shutdown scheduled for 2026-09-24.
Data quality
Reddit was limited to 12 threads and fell short of the roughly 20-item completion target because direct collection through WebFetch was unavailable and the report relied only on worker-precollected material. On X, most search terms, seven out of ten, were irrelevant trending terms; no open-source-related searches were run, leaving X’s open-source activity effectively unexamined. Bluesky’s full-text post-search API returned 403 for all queries, forcing a shift to account-search collection and leaving the analysis below target. Due to dynamic rendering, many YouTube view and subscriber counts could not be verified. Lemmy itself is small, and no first-party posts about Sora, Veo, Kling, or Hailuo were found, possibly due to search-coverage limits. Pinterest used only one general query of 50 items, without a follow-up search focused on open source.
Platform-specific summaries
Reddit — Image and Video Generation AI News
Where
Twelve threads were collected from nine subreddits using the query “Image and Video Generation AI News.”
| Subreddit | Members | Collected threads |
|---|---|---|
| r/StableDiffusion | 1,001,169 | 1 |
| r/teenagers | 3,674,216 | 1 |
| r/aiwars | 166,685 | 1 |
| r/antiai | 314,968 | 2 |
| r/generativeAI | 150,823 | 3 |
| r/RemoteWorkers | 74,528 | 1 |
| r/VeniceAI | 10,421 | 1 |
| r/DailyIncome | 12,914 | 1 |
| r/freetoolsAI | 7,736 | 1 |
r/generativeAI (three threads) and r/antiai (two threads) covered this topic most heavily. r/antiai takes a critical stance toward AI image generation, while r/generativeAI centers on tool comparisons and practical advice, creating a sharp contrast in tone.
What people say
- Warning over LLM-fabricated hand-drawn timelapses (thread #1, r/antiai, 5,259 points / 760 comments, 2026-09-05, https://www.reddit.com/r/antiai/comments/1w81kdd/). A post described a method where GPT-6 is given mouse and keyboard control to fabricate not only digital art but an entire “convincing” drawing timelapse. It drew strong backlash, including: “Why release technology like this? There is no purpose other than hurting real artists” (u/sanstheskelebrotherc, 235 points).
- Daily comparison of free AI video-generation tools (thread #2, r/generativeAI, 42 points / 28 comments, 2026-08-31, https://www.reddit.com/r/generativeAI/comments/1w3frr2/). Adobe Firefly (two videos per day), PixVerse (one per day), Grok Imagine (one per day), and Google Flow (three to four per day) were compared. Comments also cited Kling AI (66 credits every 24 hours, equivalent to two to six 5-second 720p videos) and CapCut’s daily free credits.
- Thread sharing free, unlimited image-generation tools (thread #3, r/freetoolsAI, 37 points / 38 comments, 2026-09-04, https://www.reddit.com/r/freetoolsAI/comments/1w6w0rj/). NSFW availability was the central concern in the comments, with remarks such as “Why does everyone ban NSFW?” (u/Relevant_Syllabub895), alongside lists of alternatives including Deepany, Aifun, Perchance, and Hugging Face’s Z-Image-Turbo.
- “$30/hour reviewing AI-generated images” job post spreads across two subs (thread #4, r/RemoteWorkers, 114 points / 520 comments, 2026-09-04, https://www.reddit.com/r/RemoteWorkers/comments/1w7azwu/ and thread #6, r/DailyIncome, 129 points / 1,123 comments, 2026-09-04, https://www.reddit.com/r/DailyIncome/comments/1w7b18p/). The near-identical job posts mostly received one-word comments such as “images” and “Interested,” suggesting a flood of template responses from people seeking an application link.
- Demo of faster local image generation (thread #5, r/StableDiffusion, 31 points / 10 comments, 2026-09-03, https://www.reddit.com/r/StableDiffusion/comments/1w6bphf/). A user reported running Vpipe + Krea-2-Turbo-M87 (16-bit) locally on Apple Silicon (M5 Pro, 24GB), achieving roughly 45 seconds per 1024×1024 image at eight steps, or 35 seconds with real-time 8-bit quantization.
- Meme thread about AI images becoming harder to detect (thread #8, r/antiai, 113 points / 30 comments, 2026-08-30, https://www.reddit.com/r/antiai/comments/1w2mr8g/). In practice, the thread focused on broken hands, with comments such as “Why is her foot a hand?” (u/GurInside278, 31 points), showing that familiar generation failures remain a source of humor.
- Assessment of API aggregator WaveSpeedAI (thread #7, r/generativeAI, 2 points / 4 comments, 2026-09-05, https://www.reddit.com/r/generativeAI/comments/1w8ci5z/). It was described as wrapping “more than 1,000 models including Flux, Wan 2.1/2.6, Hailuo, and Kling in a single REST API” (u/Jenna_AI). A user who integrated it into a custom tool via Seedance (u/BradClarkAI) said they had “no issues at all.”
- Model-selection guidance from the VeniceAI community (thread #9, r/VeniceAI, 8 points / 6 comments, 2026-08-30, https://www.reddit.com/r/VeniceAI/comments/1w2rtq1/). u/jerm2z recommended Chroma for image generation; Qwen Edit Uncensored, Seedream, and FireRed for editing; WAN Enhanced, Seedance, and Minimax R2V for image-to-video. FireRed does not support NSFW content, but was noted for strong face consistency.
- AI food images criticized as unappetizing (thread #10, r/aiwars, 4 points / 175 comments, 2026-09-03, https://www.reddit.com/r/aiwars/comments/1w69cdj/). Alongside the harsh comment “This only looks like a grub hive” (u/LCI_Jake, 14 points), a user who tried generating images locally with Krea 2 in 40 seconds (u/Bassed_Hummble) wondered why the output was so poor despite an ordinary prompt.
- Young people debate the merits of AI art (thread #12, r/teenagers, 7 points / 41 comments, 2026-08-31, https://www.reddit.com/r/teenagers/comments/1w2y305/). A post argued that AI itself creates the image after receiving a prompt, making it fundamentally different from human creation with art materials or software. Responses were divided. u/TheDarkmore said they are not good at drawing, so they let AI make the art for 120 cards in a card game—and were fine with that.
Signals
- Rising: The new technique of making an LLM operate a mouse and keyboard to fabricate “hand-drawn” work (thread #1) is drawing rapid backlash as a threat distinct from conventional diffusion-model detection, adding a new flashpoint to the AI-versus-anti-AI conflict. At the same time, developer-oriented infrastructure and efficiency topics are quietly gaining traction, including aggregator services such as WaveSpeedAI (thread #7), which wrap more than 1,000 models behind one API, and fast local generation on Apple Silicon (thread #5).
- Areas being dismissed or mocked: AI-generated food-ad images (thread #10) were harshly criticized, with users saying it would be better to use stock photos, repeatedly highlighting quality limitations. The “$30/hour reviewing AI-generated images” job posts (threads #4 and #6) were nearly identical and generated mostly formulaic responses from people seeking application links, with virtually no substantive discussion. They are likely being treated as spam-like job posts.
- An unexpected split: Even within the same theme of AI images becoming harder to distinguish, r/antiai’s thread #8 still treats broken hands as a joke, while thread #1 treats the new method of fabricating an entire timelapse as a serious threat. The community itself is divided over the actual state of AI-image detection. Likewise, r/VeniceAI and r/StableDiffusion are energized by technical workflows and optimization, while r/antiai and r/aiwars feature strong ethical and quality-based criticism. Attitudes toward AI image generation on Reddit are sharply divided by subreddit.
Limits
- Only one search query, “Image and Video Generation AI News,” was used; there were no separate searches for “open source” and “closed source.” Therefore, this file alone cannot directly separate the minimum ten insights for each group requested by the brief; classification is required in the synthesized report.
- The brief’s completion criterion was roughly 20 posts analyzed per social network, but only 12 threads were collected.
- In threads #4 and #6, the majority of comments were template responses such as “images” and “Interested,” leaving little substantive discussion or validation.
- Reddit itself blocks page retrieval through WebFetch and search, so the report relies solely on worker-precollected files (
reddit.threads.md/json). Further investigation, such as following related threads or expanding comment sections, was not possible in this search session.
X
X — Image and Video Generation AI News
X’s own Explore list for this session is geolocated to Croatia (the server’s
location), not to Japan or a global feed — it shows "Astra", "Bitcoin",
"Ukrainian", "Argentina", "$SONG", "OpenAI", "Russia", "Black", "$KURO",
"Fable" as trending there, with no post counts given. Only three of those ten
terms ("Astra", "OpenAI", "Fable") turned out to have anything to do with
image/video generation AI; the rest are crypto price calls, the Russia‑Ukraine
war, Argentine celebrity gossip and memecoins, and are not covered below —
see Limits.
Accounts
On‑topic accounts, by what they actually posted:
| Account | What they post | This post's reach |
|---|---|---|
| @masahirochaen (Chaen) | Japanese AI/tech news account; broke the GPT-6 Astra launch with benchmark numbers, then a follow-up on Altman's “there's more beyond Astra” remark | 2 posts, 841 + 133 likes, up to ~550,000 views on the launch post |
| @UNIBRACITY (SHINTARO) | Hobbyist 3D artist pairing GPT-6 Astra with Blender | 1 post, 306 likes |
| @hayashimon1 (Hayashimon — AI × indie development) | AI/3D creator who also runs paid seminars; uses Fable 5.1 × Blender | 1 post, 55 likes |
| @manaimovie (Mankyu) | AI video hobbyist, feeding a Tripo 3D model into Astra | 1 post, 409 likes |
| @yachimat_manga (yachimat-AI Short Anime) | AI anime-short creator, speculating Astra could replace dedicated video-gen tools for some styles | 1 post, 51 likes |
| @aigeboku (Servant of AI) | AI power-user testing Astra’s “ComputerUse” agentic control | 1 post, 653 likes |
| @gagarot200 (Gagarot) | AI demo account, one-line-prompt-to-game clips | 1 post, 306 likes |
| @SSSS_CRYPTOMAN | Crypto+AI crossover account, posted an Astra game-build demo | 1 post, 161 likes |
| @Mayz1169 (Kiki) / @renoiseai (Renoise) | Side-by-side AI video comparison accounts, both pitting GPT‑6 Astra against Fable 5.1 rendered via Seedance 2.5 | 1 post each, 64 and 140 likes |
| @k_matsumaru | AI video creator documenting a full pipeline (Typeless → Fable 5.1 → Seedream 5.0 → Seedance) | 1 post, 279 likes |
| @Strength04_X | Prompt-sharing account for Seedance 2.5 (via lovart_ai) | 1 post, 444 likes |
| @ImaStudio_ai | Product/studio account for its own AI fashion-video tool (MiniMax H3-based) | 1 post, 132 likes |
Every one of these is a single post, not a thread or a sustained campaign — this
is a scattered one-day reaction to a launch, not an organized account driving
the conversation.
Posts
24. @masahirochaen (Chaen)
- 841 likes · 148 reposts · 17 replies · ~550,000 views · 2026-09-03
- https://x.com/masahirochaen/status/2095644339041153256
【Breaking】OpenAI announces “GPT-6 Astra.” The long-awaited flagship model that can handle PC work end to end and surpass Fable 5.1 is finally here. … ARC-AGI-3 (ability to solve unfamiliar puzzles): 98.6%. The previous-generation GPT-5.6 scored 7.8%. An extraordinary leap.
The single biggest post in this collection by far, and the one that frames
everything else: it announces GPT‑6 Astra as a PC-operating agent that beats
Fable 5.1, citing an ARC‑AGI‑3 jump from 7.8% (GPT‑5.6) to 98.6%.
21. @masahirochaen (Chaen)
- 133 likes · 23 reposts · 8 replies · ~20,000 views · 2026-09-05
- https://x.com/masahirochaen/status/2096372856930386341
Sam Altman says there is something beyond GPT-6… He explicitly stated that the model whose work was paused in August because of cyber risk was not GPT-6 Astra.
Two days after the launch, the same account reports Altman saying the model
OpenAI paused in August for cyber-risk reasons was not Astra — implying an
even more capable model is still unreleased.
3. @aigeboku (Servant of AI)
- 653 likes · 64 reposts · 7 replies · ~43,000 views · 2026-09-05
- https://x.com/aigeboku/status/2096187322924687799
As expected of GPT-6 Astra’s ability to control other things—ComputerUse greatly expands what it can do. It may be difficult to get everything in one shot, but with more tuning from here, it could enable even more interesting forms of expression.
Frames Astra's selling point as agentic control ("ComputerUse") rather than
raw generation quality — early testers are exploring what that control
unlocks for creative workflows, not just image output.
1. @manaimovie (Mankyu)
- 409 likes · 49 reposts · 16 replies · ~20,000 views · 2026-09-05
- https://x.com/manaimovie/status/2096176821377380677
I gave Astra a model I made with Tripo and asked it to make it move—this is what it produced. … It’s much better than when I tried Fable or Sol!
A concrete before/after workflow claim: feeding a Tripo-made 3D model to
Astra and asking it to animate it worked better than doing the same with
Fable or "Sol".
22. @gagarot200 (Gagarot)
- 306 likes · 29 reposts · 7 replies · ~95,000 views · 2026-09-04
- https://x.com/gagarot200/status/2095789680931287390
GPT-6 Astra can generate it fully automatically… All you have to do is say “GPT-6 Astra, make me a game like this” in a single sentence.
A one-line prompt producing a playable game end-to-end, offered as proof of
Astra's "does the whole task" positioning.
23. @UNIBRACITY (SHINTARO)
- 306 likes · 49 reposts · 4 replies · ~39,000 views · 2026-09-05
- https://x.com/UNIBRACITY/status/2096142182050927035
GPT-6 Astra × Blender: “Forest Retreat.” With about seven hours of Astra runtime, it got this far.
A ~7-hour Astra run producing a full 3D scene in Blender — one of the more
detailed "how long did this actually take" data points in the set.
4. @SSSS_CRYPTOMAN (SSSS.CRYPTOMAN⚡️AI)
- 161 likes · 23 reposts · 5 replies · ~10,000 views · 2026-09-05
- https://x.com/SSSS_CRYPTOMAN/status/2096248797873762805
A game was completed in no time with GPT-6 Astra! 🎮 … With a rough prompt and a few exchanges, I got something that is genuinely playable.
Another vague-prompt-to-playable-game demo, with a live link, echoing
@gagarot200's claim independently.
2. @yachimat_manga (yachimat - AI Short Anime)
- 51 likes · 6 reposts · 5 replies · ~3,600 views · 2026-09-05
- https://x.com/yachimat_manga/status/2096386820997304354
I tried modeling with GPT-6 Astra. … Depending on the video style, could there be a future where we no longer need video-generation tools at all?
A working AI-anime creator speculates out loud that for some styles, Astra's
3D modeling could make dedicated video-generation tools unnecessary.
40. @renoiseai (Renoise)
- 140 likes · 16 reposts · 10 replies · ~99,000 views · 2026-09-04
- https://x.com/renoiseai/status/2095773478200996066
GPT-6 Astra vs Fable 5.1 Both videos generated with Seedance 2.5 on Renoise
A direct head-to-head clip, both outputs rendered through the same
downstream tool (Seedance 2.5) to control for that variable.
39. @Mayz1169 (Kiki)
- 64 likes · 6 reposts · 10 replies · ~12,000 views · 2026-09-04
- https://x.com/Mayz1169/status/2095795273134170259
forget the benchmarks for a second. just watch this. GPT-6 Astra vs. Fable 5.1 side by side comparison on same prompts. which one would you pick?
Same comparison instinct as above, independently — Astra vs. Fable 5.1
head-to-heads are a recognizable micro-genre in this batch, not a one-off.
37. @k_matsumaru (Keigo Matsumaru)
- 279 likes · 34 reposts · 14 replies · ~19,000 views · 2026-09-05
- https://x.com/k_matsumaru/status/2096044983468150935
How I make videos lately… ① Use Typeless to write the brief ② Have Fable 5.1 write the video-generation prompt ③ Turn it into a storyboard with Seedream 5.0 Pro and review the content ④ Seedance…
The most detailed production pipeline in the set: Typeless for scripting →
Fable 5.1 for prompt-writing → Seedream 5.0 Pro for storyboarding →
Seedance for final video — Fable shows up here as a prompting layer, not
the renderer itself.
32. @ImaStudio_ai (Ima Studio)
- 132 likes · 17 reposts · 3 replies · ~4,600 views · 2026-09-04
- https://x.com/ImaStudio_ai/status/2095725066109804603
MiniMax H3 K-Fashion Outfit Change MV…8 LOOKS. 1 GIRL. ZERO IDENTITY DRIFT.
A commercial studio account promoting its own product built on MiniMax H3,
pitching "zero identity drift" across outfit changes as the differentiator.
Signals
- What's rising: GPT‑6 Astra (OpenAI) dominates every on-topic post in
this batch, and it's being talked about as an agentic, computer-operating
model first and an image/video generator second — testers keep pairing it
with other tools (Tripo for 3D meshes, Blender for scenes) rather than
treating it as a standalone renderer. - A recognizable micro-genre: unprompted, independent accounts
(@renoiseai, @Mayz1169) are posting "GPT‑6 Astra vs Fable 5.1" side-by-sides
rendered through Seedance 2.5 — Fable 5.1 is the model people reach for as
the comparison point, and Seedance/Seedream keep recurring as the actual
video-rendering layer underneath both. - Chained pipelines, not single tools: the most detailed post
(@k_matsumaru, #37) chains four different tools (Typeless → Fable 5.1 →
Seedream 5.0 Pro → Seedance) for one video — nobody in this set describes
finishing a video with a single model. - What surprised us: the trending terms this run pulled from X's own
Explore panel (geolocated to Croatia) — Bitcoin, Ukrainian, Argentina,
$SONG, Russia, Black, $KURO — had essentially zero connection to
image/video generation AI. X's "what's trending" for this server's location
on 2026‑09‑05/06 is crypto and geopolitics, not AI models; treating it as a
proxy for AI discourse would have been misleading. - What's being dismissed: no skeptical or critical thread about GPT‑6
Astra surfaced in this batch — every on-topic post is a demo, a benchmark
repost, or a comparison clip. That absence is itself notable given how
aggressive the "beats Fable 5.1" claim (#24) is.
Limits
- The ten search terms used ("Astra", "Bitcoin", "Ukrainian", "Argentina",
"$SONG", "OpenAI", "Russia", "Black", "$KURO", "Fable") came from X's own
Explore trending list for this session, not from a curated image/video-AI
term list. Seven of the ten ("Bitcoin", "Ukrainian", "Argentina", "$SONG",
"Russia", "Black", "$KURO") returned posts entirely unrelated to the
research theme (crypto giveaways and price calls, the Russia-Ukraine war,
Argentine celebrity news, and two small memecoins) and are excluded from
the Posts section above. - Because of that, open-source coverage here is thin and indirect: no post
in this collection used terms like "Stable Diffusion", "ComfyUI", "Wan",
"HunyuanVideo", or "Flux" — those searches were never run this stage.
Everything on-topic ties back to closed/commercial names (GPT‑6 Astra,
Fable 5.1, Seedance, Seedream, MiniMax H3, Tripo), so X's open-source
picture for this run should be treated as absent, not as "nothing is
happening there." - Post #35 (@MINHxDYNASTY, tagged "$KURO") was collected with no post text
captured in the source file, so it could not be used as a finding. - Explore's "Where" column is geolocated to wherever this session's server
sits (Croatia), not to Japan or worldwide — its list should not be read as
a global or Japan-specific trending picture. - No session/access failure occurred: the collection returned data for every
search term attempted, so the gap here is topical (which terms were
searched), not technical.
YouTube
YouTube — Image and Video Generation AI News (Closed Source / Open Source)
Channels
- Bijan Bowen (@Bijanbowen, approximately 37,000 subscribers) — Specializes in in-depth reviews of LLMs and image-generation models. Tested Qwen Image-2512.
- Future Tools / Matt Wolfe (@mreflow, more than 900,000 subscribers) — Publishes weekly “Everything in AI This Week” roundups covering image- and video-generation AI developments including Midjourney and Runway.
- ComfyUI/local-generation tutorial channels (many workflow videos for Wan 2.2/2.6, FLUX, and Qwen Image; channel names were not displayed in search results, which showed only video titles).
Videos
- "Why OpenAI Shut Down Sora So Fast: What's Next?" — Posted around March–April 2026. Explains the reasoning behind OpenAI’s decision to close the Sora app on April 26, 2026, with the API itself also scheduled to end on September 24, 2026: monthly downloads fell from 3.3 million in November 2025 to 1.1 million in February 2026, and a $1 billion, three-year deal with Disney also collapsed. https://www.youtube.com/watch?v=H2jVcy5PuBI
- "Why Did Sora AI Shut Down? The Real Reason OpenAI Killed Its Viral Video App" — Posted in spring 2026. Discusses the economic costs of a video-generation app that failed to monetize and the strategic shift toward agentic AI. https://www.youtube.com/watch?v=MpLY9bjYVuA
- "HunyuanImage-3.0 First Test – The Open Source Nano Banana?" — Posted September 29, 2025. Tests Tencent’s HunyuanImage-3.0, one of the largest open-source text-to-image models, locally including hardware requirements. https://www.youtube.com/watch?v=lLTv4sGf8Wo
- "Qwen Image-2512 First Test – The BEST Open Source Image Model!" (Bijan Bowen) — Posted December 31, 2025. Tests improvements in portrait realism and text rendering in Alibaba’s Qwen Image-2512 through photorealistic examples. https://www.youtube.com/watch?v=SBaK616nP6Q
- "New Qwen Image 2512: Better Than Z-Image? (Open Source & Free)" — Posted January 5, 2026. Directly compares Qwen Image 2511/2512 with Tongyi-MAI’s Z-Image Turbo. https://www.youtube.com/watch?v=uAGH_yRZ2Gw
- "Tencent Hunyuan 2.1 vs. Qwen Image: Who is the True King? A 2K Image & Text Rendering Showdown!" — Posted September 28, 2025. A 2K text-rendering showdown between Hunyuan Image 2.1, then the leader of open-source text-to-image Arena leaderboards, and Qwen Image. https://www.youtube.com/watch?v=MUDTryJS0XM
- "Z-Image la nouvelle star des modèles Open Source?? Test dans ComfyUI" — Posted November 28, 2025. Tests Tongyi-MAI’s open-source text-to-image model Z-Image Turbo in ComfyUI. https://www.youtube.com/watch?v=eOKSBu7KLh0
- "Wan 2.2, FLUX & Qwen Image Upgraded: Ultimate Tutorial for Open Source SOTA Image & Video Gen Models" — Posted January 11, 2026. A long-form ComfyUI tutorial covering recent open-source image and video models including Wan 2.2, FLUX, and Qwen Image. https://www.youtube.com/watch?v=3BFDcO2Ysu4
- "WAN 2.6 Hands-On: The Good, the Bad, and the Useful" — Posted December 26, 2025. A hands-on test of Alibaba’s Wan 2.6, supporting up to 15 seconds, multi-shot generation, and audio synchronization. https://www.youtube.com/watch?v=iNnmLwPg7OE
- "Multi-Shot AI Videos: Wan 2.6 vs Kling 2.6 (Stress Test)" — Posted December 27, 2025. Stress-tests multi-shot generation in open Wan 2.6 versus closed Kling 2.6. https://www.youtube.com/watch?v=CDRKn4I-Iv0
- "I Tested Google's Nano Banana Pro Tips - This One Creates Pro Graphics in 30 Seconds" — Posted around November–December 2025. A practical test of Google DeepMind’s Gemini 3 Pro Image, known as Nano Banana Pro, covering 4K output, support for more than 14 objects, and multilingual text rendering. https://www.youtube.com/watch?v=ca4SXmRqtKE
- "Qwen Image Edit Review — Free Open-Source Alternative to Nano Banana" — Posted December 4, 2025. Reviews Qwen Image Edit, an open-source editing model positioned as an alternative to closed-source Nano Banana. https://www.youtube.com/watch?v=Ur3udfFJ2QQ
Signals
- The biggest closed-source story is Sora’s withdrawal: OpenAI notified developers on March 24, 2026, ended the Sora app and web version on April 26, and plans to end the API on September 24. Several videos cite the collapse in downloads (3.3 million in November 2025 → 1.1 million in February 2026) and the failed $1 billion Disney partnership as context. It is treated as an example of the industry shifting resources away from video generation and toward coding and agentic AI.
- The open-source camp has a “two leaders plus rising challenger” structure: Alibaba models (Qwen Image, Z-Image Turbo, Wan 2.2/2.6) and Tencent models (HunyuanImage 3.0/2.1) compete for the top of major leaderboards, while Black Forest Labs’ FLUX.2 ([dev]/[klein], with an Apache 2.0 VAE release) competes on photorealism. Many creators frame Qwen Image and HunyuanImage as “open-source alternatives to Nano Banana,” making direct closed-versus-open comparisons an established format.
- For video, multi-shot generation and audio sync are the new baseline: Wan 2.6 (open) and Kling 3.0, Seedance 2.0, and Runway Gen-4.5 (closed) all compete around generating multiple shots in a single run and generating audio/SFX at the same time. Competition is moving beyond one-off clip generation.
- Many review videos use a “first test” or “first look” hands-on format, a high-speed reporting format that is commonly published within days to a week of a model announcement.
Limits
- Because YouTube search and watch pages (
youtube.com/results,youtube.com/watch) are JavaScript-rendered, WebFetch returned only static footer content such as copyright notices and terms links. Accurate view counts, subscriber numbers, and comments could not be directly confirmed in most cases. The listed views and subscriber counts only reflect what could be found in web-search summaries; many videos have no view count listed. - Direct YouTube review URLs for Kling 3.0, Seedance 2.0, and Runway Gen-4.5 could not be identified because search results skewed toward technical-comparison blog posts such as invideo.io, Pixo, and AI.cc.
- Subscriber counts could only be confirmed for Bijan Bowen (approximately 37,000) and Matt Wolfe / Future Tools (more than 900,000).
Bluesky
Bluesky — Image and Video Generation AI News
The official Bluesky search API (app.bsky.feed.searchPosts) consistently returned 403 Forbidden during this session, making keyword-based post collection impossible (see Limits for details). The method was therefore changed to finding relevant accounts through app.bsky.actor.searchActors and reading their post histories one at a time through getAuthorFeed.
Accounts
| Account | Content | Followers | Posts |
|---|---|---|---|
| @todaystopainews.bsky.social | “Today's Top AI News.” A bot-like account automatically collecting and posting AI-industry news. Frequently covers closed-source models such as GPT-6 Astra and Fable 5.1. | 272 | 1,826 |
| @useapi.bsky.social | useapi.net. Official account for an aggregator that bundles image/video generation APIs from Veo, Kling, Seedance, PixVerse, Hailuo, Runway, Nano Banana, and others. | 13 | 24 |
| @nijiironokeshiki.bsky.social | “Niji-iro.” An individual creator posting fan-art illustrations for Re:ZERO, Lycoris Recoil, Frieren, and others two to three times per day using Stable Diffusion and ComfyUI. | 699 | 1,963 |
| @kunecco.bsky.social | “Kunecco.” Developer of nodes, workflows, and LoRAs for the Git version of ComfyUI. Also posts technical content such as danbooru-tag translations. | 540 | 494 |
| @midjourneyofficial.bsky.social | Official Midjourney account. Announced its arrival on Bluesky in August 2025. | Unknown | Last post 2025-10-08 |
| @comfyuistudio.bsky.social | Third-party account introducing ComfyUI workflows. | Unknown | Last post 2024-12-19 |
| @klingaivideo.bsky.social | Account introducing Kling AI’s video-generation tool. | Unknown | Empty feed (no posts) |
Posts
Closed source
-
GPT-6 Astra generates an SVG of a PS4 controller (@todaystopainews, 11 likes / 2 reposts / 1 reply, 2026-09-05, https://bsky.app/profile/todaystopainews.bsky.social/post/3mur7cyocf225). The post includes the generation time—11 minutes 54 seconds—offering a concrete example of how long OpenAI’s GPT-6 Astra takes for an image-oriented task.
-
GPT-6 Astra assembles a city scene in Unity (@todaystopainews, 4 likes / 1 repost, 2026-09-04, https://bsky.app/profile/todaystopainews.bsky.social/post/3muoucqf4tk2w). A demo automatically composing a 3D scene using existing assets, showing an agentic use case beyond simple image generation.
-
Repost of “Fable 5.1 world modeling” (@todaystopainews, 11 likes / 4 reposts, 2026-09-03, https://bsky.app/profile/todaystopainews.bsky.social/post/3mumdxgujh22h) and the previous day’s “medieval town made with Fable 5.1” (10 likes / 3 reposts, 2026-09-02, https://bsky.app/profile/todaystopainews.bsky.social/post/3mukekiwxws2j). Anthropic-related Fable 5.1 is repeatedly presented through 3D and world-modeling demos.
-
useapi.net blog: “Seedance 2.5 on the Dreamina API” (@useapi.bsky.social, 2026-08-08, https://bsky.app/profile/useapi.bsky.social/post/3mskj7fb2ok2a). Introduces ByteDance-related video model Seedance 2.5 becoming available through Dreamina’s API in a technical-blog format.
-
useapi.net: “MiniMax H3 vs Seedance 2.0” comparison using the same reference image (@useapi.bsky.social, 2026-07-31, https://bsky.app/profile/useapi.bsky.social/post/3mrw5xyd4zk2x). A direct model comparison using the same source material (omni references), providing quality validation from an API provider’s perspective.
-
useapi.net: “In-depth comparison of more than 17 AI image models” (@useapi.bsky.social, 3 likes, 2026-07-25, https://bsky.app/profile/useapi.bsky.social/post/3mrigfks25k2n). A cross-model comparison article that illustrates how crowded the closed-source image-generation field has become.
-
useapi.net: “Seedance 2.0 official route vs. third-party route price comparison” (@useapi.bsky.social, 2026-07-23, https://bsky.app/profile/useapi.bsky.social/post/3mrdyxqswcb2p). Specifically compares price differences between an official API and aggregator access for the same model, revealing the layered structure of the API market.
Open source
-
Niji-iro’s regular AI illustration posts (@nijiironokeshiki, 87–245 likes across multiple posts from 2026-09-02 to 09-05; example: https://bsky.app/profile/nijiironokeshiki.bsky.social/post/3mumi7nop522l, 104 likes / 25 reposts). As stated in the profile, the work is made with Stable Diffusion and ComfyUI. This individual account consistently receives unusually high engagement on Bluesky.
-
Kunecco’s post about developing ComfyUI workflows (@kunecco, 2025-10-27, https://bsky.app/profile/kunecco.bsky.social/post/3m45jd5s2ek2f): “Making nodes and workflows has become fun, and I’m falling down the rabbit hole.” Around the same period, the account also posted that it had “translated 140,000 lines of danbooru tokens” (https://bsky.app/profile/kunecco.bsky.social/post/3m432qalae22t), an example of working on the ComfyUI ecosystem at the tag-dictionary level.
-
Stagnation among official or quasi-official ComfyUI accounts (@comfyuistudio.bsky.social last posted 2024-12-19; @comfy-ui.bsky.social has an empty feed). Actor search returned many similar handles, including
comfyuistudio,real-comfyui,comfy-ui, andcomfyuiorg, but the active accounts are individual AI-illustration creators such as Niji-iro and Kunecco. Tool developers and official information sources have almost no meaningful presence on Bluesky.
Signals
- Closed source is centered on breaking news, while open source is centered on finished work: Posts from todaystopainews and useapi.net focus on model names, API prices, and benchmarks for GPT-6 Astra, Fable 5.1, Seedance, and MiniMax H3. Their engagement is low—usually single digits to around ten likes—but their news density is high. By contrast, the open-source side, including Stable Diffusion and ComfyUI, is centered on finished illustrations from individual creators like Niji-iro. The focus is on the work itself rather than model developments. Open-source technical news is almost absent from Bluesky.
- Official Midjourney and Kling AI accounts are effectively dormant: Midjourney announced it had come to Bluesky in August 2025 but stopped posting on 2025-10-08. The Kling AI Video account has an empty feed. Brands are not using Bluesky as a sustained communications channel.
- AI-agent persona accounts branded around GPT-6 Astra exist: Accounts such as
astrra.space(“AI agent running on GPT-6 Astra”) andmaple-nekokami.bsky.social(“Runs on OpenAI GPT-6 Astra”) were found. They are not directly about image or video generation, but they show the GPT-6 Astra name being consumed as a meme-like brand on Bluesky. - Aggregator API useapi.net has become a hub for comparing closed-source models: useapi.net bundles models from Veo, Kling, Seedance, PixVerse, Hailuo, Runway, Nano Banana, and others, and is publishing some of Bluesky’s most concrete first-party information on price and performance comparisons. Such cross-platform services offer denser practical information than individual model companies’ official accounts.
Limits
app.bsky.feed.searchPosts, the full-text post-search API, returned HTTP 403 for every query—including simple queries such as “image generation AI,” “Stable Diffusion,” and “Flux.” In contrast,app.bsky.actor.searchActorsfor account search andapp.bsky.feed.getAuthorFeedfor individual account feeds worked normally. As a result, this report relies not on cross-platform keyword search, but on finding plausibly relevant accounts and reading them individually. Posts and accounts not surfaced by account search are not included.- Against the completion goal of analyzing roughly 20 posts per social network, only the ten posts above plus supplementary surrounding posts could actually be read and analyzed. Under the account-level search limitation, increasing the count further would require uncovering more relevant accounts. This session examined the primary angles—closed-source news bots such as GPT-6 Astra/Fable 5.1 accounts, API aggregators such as useapi.net, and individual ComfyUI/Stable Diffusion creators. Most additional accounts found were inactive or unrelated, such as sports, politics, and journalism.
- Official brand accounts such as Midjourney and Kling AI had stopped posting or had empty feeds, so they could not be used as sources of current news.
- Account searches for newer open-source model names such as “Wan2.5,” “Flux.2,” and “open-source image generation” returned zero hits. No Bluesky accounts posting about specific newer open-source models such as Wan, HunyuanVideo, Qwen-Image, or Chroma were found. Open-source developments could only be confirmed within the existing Stable Diffusion/ComfyUI brand space.
Lemmy
Lemmy — Latest Developments in Image and Video Generation AI
Communities
- !«メールアドレス» — 5,708 members. By far the most active community. Custom ComfyUI nodes, newly released model weights, and LoRAs appear daily, making it effectively a breaking-news board for open-source image and video generation AI.
- !«メールアドレス» — 2,360 members. Primarily for sharing generated artwork, with relatively little news.
- !«メールアドレス» — 1,662 members. Similar to the dbzer0 community but with lower posting frequency.
- !«メールアドレス» — 87,858 members. One of Lemmy’s largest technology communities, but discussion of image/video generation AI itself is sparse; occasional Midjourney corporate news appears.
- !«メールアドレス» (including mirrors on other instances) — Frequently reposts Reddit AI subreddits. Volumes are low, but it picks up unique sources such as Runway financial information.
- !techtakes — An AI-critical, skeptical community. Includes satirical posts about Midjourney losing its way.
Posts
- Midjourney pivots to medical ultrasound CT scanner business — !«メールアドレス», score 58, 2026-06-18, https://www.theregister.com (original article from The Register). Related: Midjourney official blog https://midjourney.com/medical/blogpost (!hackernews, score 2, 2026-06-18), and a sarcastic-community response: “Midjourney AI pivots to Theranos,” !techtakes, score 40, 2026-06-19, https://pivot-to-ai.com/2026/06/19
- Midjourney demands that Hollywood studios disclose their AI use — !«メールアドレス», score 73, 2026-07-05, https://techcrunch.com/2026/07/04
- Runway: “Enterprise business doubled, NRR exceeds 300%” — !ai_reddit, score 0, 2026-08-30, https://www.reddit.com/r/ArtificialInteligence/comments/1w2hnea/ (Lemmy repost)
- Qwen-Image-2.0 release: major parameter reduction from 20B to 7B, unified generation and editing, beats Nano Banana in Elo ratings; weights planned soon — !«メールアドレス», score 12, 2026-02-12, https://files.catbox.moe/l3mzzs.jpg (also cross-posted to !LocalLLaMA, !Aii, and !artificial_intel)
- Wan Animate 2 available in ComfyUI — !«メールアドレス», score 8, 2026-08-08, https://blog.comfy.org/p/wan-animate-2-is-now-available-in
- LTX-2.5 receives Day-0 support in ComfyUI — !«メールアドレス», score 5, 2026-08-12, https://blog.comfy.org/p/ltx-25-day-0-support-in-comfyui
- Trellis.2 and Pixal3D, 3D-generation models, gain native ComfyUI support — !«メールアドレス», score 4, 2026-09-01, https://blog.comfy.org/p/trellis2-and-pixal3d-are-now-native
- MiniMax-H3, an open-weight video-generation model, ecosystem expands rapidly — More than ten posts in !«メールアドレス» alone from 8/8 to 9/5. Examples: GGUF quantized edition (21.6GB → 8.9GB, score 10, https://huggingface.co/Abiray/MiniMax-H3-Pruned-GGUF) / LoRA accelerating generation to 4–8 steps (score 6–7, https://huggingface.co/lightx2v/Minimax-h3-Turbo) / long-video support up to 120 seconds with audio synchronization (score 5, https://huggingface.co/Smite79/MiniMax-H3-Longvideos) / hybrid-attention implementation (score 3, https://github.com/Saganaki22/ComfyUI-VDN-H3, 2026-09-05)
- GPU-native video-enhancement node groups: ComfyUI-AetherScale and DLSS5 Visual Enhancer — !«メールアドレス», score 5, 2026-09-01/09-05, https://github.com/vizart-vj/ComfyUI-AetherScale ・ https://github.com/Merserk/dlss5-visual-enhancer
- Use Claude MCP to call more than 30 image/video generation models from one chat; one project shortened from 2.5 hours to 50 minutes — !ai_reddit, score 1, 2026-05-01, https://i.redd.it/6gqfafpaukyg1.png
Signals
- The open-source side is entirely centered on ComfyUI/Hugging Face implementation news. Rather than model announcements themselves, most posts are ecosystem updates: “X is now available in ComfyUI,” “a LoRA was released,” or “a quantized version was released.” Tool development around the MiniMax-H3 video model is the current main battlefield.
- 3D generation, including Trellis.2 and Pixal3D, is beginning to be integrated into ComfyUI, gaining presence as the next generation target after images and video.
- Closed-source developments on Lemmy are more often consumed as criticism or scandal than as product news. Midjourney’s medical-device pivot spread more strongly through general-tech and satirical communities (scores 58, 73, and 40) than through generative-AI communities.
- Mentions of Nano Banana on Lemmy appear almost exclusively in the context of open-source models outperforming it; no primary posts originating from Google were found.
- No direct mentions of Sora, Veo, Kling, Hailuo, Seedream/Seedance, or the latest Adobe Firefly appeared during the research period of the last month. The latest Adobe Firefly post found was from 2025.
Limits
- Lemmy is small, and discussion of image/video generation AI is concentrated almost entirely in !«メールアドレス». Cross-instance searches across lemmy.ml, lemm.ee, lemmy.zip, and others mostly returned duplicate or irrelevant results.
- Full-text search (
/api/v3/search) has weak fuzzy matching for short proper nouns. Queries for names such as “Sora” and “Veo” mostly returned irrelevant posts about people or other services. Both API and web search were attempted, but no first-party Lemmy posts about OpenAI Sora, Google Veo, Kling AI, Hailuo, or ByteDance Seedream/Seedance could be found, including through Google search. This may reflect Lemmy search and coverage limits rather than an absence of discussion. - The Qwen-Image-2.0 post dates from February 2026 and is not a topic from the most recent month, but it remains because it is valuable as an open-versus-closed comparison point.
- Runway and Claude MCP posts did not originate on Lemmy; they are reposts from Reddit through !ai_reddit, with the original sources on Reddit.
Pinterest — Image and Video Generation AI News
Fifty Pinterest pins were collected using the query “Image and Video Generation AI News” (output/pinterest.pins.md, a single search with no genre segmentation). All 50 images were reviewed.
Visual themes
- AI symbolized as a white humanoid robot/cyborg: Regardless of the underlying generative-AI technology, images of white humanoid robots or human-and-robot profile composites recur as a visual shorthand for “AI” ([2, 19, 30, 35, 37, 40, 43, 44, 46, 48]). News and explainer pin thumbnails overwhelmingly converge on this motif.
- Neon-and-navy cyber color palettes dominate: Template-like infographics combining cyan, purple, and magenta glows, circuit-board patterns, and dark backgrounds account for the majority ([3, 5, 6, 7, 10, 13, 17, 18, 19, 21, 29, 32, 34, 39, 42, 44, 46, 48, 49]). There is almost no photography or hand-drawn material, making the mass-produced Canva/stock-asset design style obvious.
- “Best AI tool” rankings are the genre’s core format: Comparison charts and ranking images repeatedly feature logos for Runway, Pika, Kling AI, Luma AI (Dream Machine), Synthesia, HeyGen, InVideo AI, Kaiber, Adobe Firefly, ChatGPT Plus, Midjourney, Canva AI, and more ([9, 16, 20, 24, 38, 42]). [16] and [24] are nearly identical “Best Quality / Best Free” charts, showing the same templates reused by multiple posters.
- Breaking-news and roundup infographics: Many pins summarize “this week’s AI news” in newspaper-like layouts or under “Breaking News” banners ([11, 12, 21, 29, 32, 33, 39, 43, 47, 50]). Many have dates burned into the image; “breaking” images from different periods—June 30, 2025 [29], May 20, 2026 [21], and August 15, 2026 [32]—appear together in the same results.
- Pins showing actual generated outputs or video demos are a minority: Most pins are marketing images built around logos or text. Only around 10% of the 50 show actual output examples, such as Google Veo 2’s official demo reel [8], a Luma-like style-transfer grid (person → wood carving/origami/LEGO/flowers) [4], Microsoft VASA-1’s facial-animation demo from one image plus audio [22], an image-to-video example maintaining a purple monster character [26], and a photorealistic old fisherman portrait made with AMD Amuse [41].
- Skepticism and detection: “Is this real? Is it AI?”: A cat photo with a FAKE stamp and a robot [2], a “Which is AI?” mountain-photo quiz [15], “97% AI voice detection” [18], a pros-and-cons list on whether AI-generated video is good or bad [23], and a neon sign saying people are using AI wrong [46] all illustrate a persistent concern with authenticity and detection.
- Photos of real public figures mix with AI-generated celebrity-like images: There is a real stage photo of Nvidia CEO Jensen Huang [45] and a YouTube thumbnail featuring someone resembling Sam Altman [33], alongside a Fiverr ad image with an AI-generated Bezos-like face [36]. In the pin feed, the boundary between photography and AI-generated imagery is difficult to distinguish.
- Open source is barely mentioned: Closed-source tools such as Veo, Sora, ChatGPT, Midjourney, and Runway dominate. Only two pins refer to open source or local generation: a newspaper-style item on Baidu open-sourcing “Ernie 4.5” [29] and AMD’s local GPU tool “Amuse 3.0” [41]. No pins suggested an OSS image-generation community around Stable Diffusion, ComfyUI, or LoRA.
Notable pins
- [4] (untitled / LumaLabs-like style-transfer grid) — Converts a single frame of live-action video into “wood carving,” “origami,” “LEGO blocks,” and “flowers,” applying the same transformations to a dog and Tesla vehicle. The most concrete technical demo.
- [8] Google Unveils Veo 2 — A montage of official promotional footage including a woman looking through a microscope, a swimmer underwater, a dancer drawing constellations, an animated girl in a kitchen, and a flock of flamingos. It most directly demonstrates the range of Veo 2’s video expression.
- [9] 6+ Best AI Video Generation Websites — A save-oriented infographic organizing Runway, Pika, Kling AI, Luma AI, Canva AI, and Hailuo AI by capabilities and intended users. The “Save this pin” prompt exemplifies Pinterest’s list-consumption behavior.
- [16] AI Image Video: Best Quality vs Best Free — Compares ten paid and ten free tools with a use-case decision flow. It is nearly identical to [24], supporting the spread of this kind of reusable template.
- [22] VASA-1 (Microsoft) — A research-demo diagram showing how a single still image, an audio clip, and optional control signals can generate speaking faces with diverse expressions. It is the most academic of the technical pins.
- [27] Vidu S1 — A flashy cyberpunk visual for a model by China’s ShengShu Technology, intended for real-time AI interaction. It signals the presence of Chinese video-generation models.
- [29] AI NEWS (newspaper style, 2025-06-30) — One of the few relatively primary-information-oriented pins, summarizing specific company developments including Baidu open-sourcing Ernie 4.5, OpenAI borrowing Google TPUs, and Meta investing $14.3 billion in Scale AI.
- [38] Top 10 Image-to-Video Generator Tools — Lists Runway Gen-3, Kling, Google Veo, OpenAI Sora, Luma AI, Pika Labs, Kaiber AI, Synthesia, HeyGen, Adobe Firefly, and InVideo AI in ranking form. It provides a broad overview of major players.
- [41] Amuse 3.0 (AMD) — Uses a photorealistic portrait of an elderly fisherman to promote image quality from a local AI art tool running on AMD GPUs. One of the few examples from the open-source/local-generation side.
- [50] This Week's Top AI Developments — Summarizes recent statistical developments, including Google’s survey of 15 million chats (86% of use is outside work), unlimited ChatGPT access, and a contest result where AI beat 458 of 526 teams.
Signals
- What Pinterest currently looks like: “Image and Video Generation AI” on Pinterest is effectively an SEO/affiliate market for tool-comparison lists and news-roundup infographics. Names such as Runway, Pika, Kling AI, Luma AI, Synthesia, HeyGen, InVideo AI, Adobe Firefly, and Google Veo repeatedly appear near the top, indicating that they are perceived as the standard set of closed-source video-generation tools.
- Reused templates: Since [16] and [24] are effectively the same image, AI-tool comparison content on Pinterest is likely mass-produced and reposted across multiple accounts from a limited set of templates.
- The absence of open source: Only [29] (Baidu Ernie 4.5) and [41] (AMD Amuse 3.0) offer examples directly useful for the brief’s requested open-source trends. There were no images from Stable Diffusion, ComfyUI, or Hugging Face-based generative-AI communities. Pinterest search results are biased toward visually attractive finished outputs and marketing assets, suggesting a poor fit for posts from open-source technical communities.
- Explainers, comparisons, and hype outperform actual creations: Only around 10% of the 50 pins place actual AI-generated video or images at the center. Most are text-led infographics or lists. On Pinterest, decision-support content about which tool to choose appears more likely to spread than concrete examples of what a tool can transform.
Limits
- This stage only reviewed worker-precollected Pinterest images (
images/pinterest/manifest.json, 50 items, one query with no genre segmentation) and did not browse Pinterest itself. Additional searches using keywords such as “open source AI art” or “Stable Diffusion” may have surfaced more open-source-oriented pins. - Dates burned into pin images, such as 2025-06-30, 2026-05-20, and 2026-08-15, indicate when the template was made, not necessarily when it was posted or saved to Pinterest;
manifest.jsondoes not provide that information. - Some text in images, including [12], [15], [21], [23], and [29], was somewhat unclear. This section summarizes the content rather than treating it as exact quotation; some misspellings, such as “Machine-Genertsd Media,” were part of the templates themselves.
- Five pins had empty titles (
untitled), including some of [12], [15], [21], [29], and [35], so they were classified based on their image contents.
Recommended actions
- Continue comparing GPT-6 Astra with Fable 5.1 along the axis of agentic control, not just quality.
- Continue monitoring the MiniMax-H3 ComfyUI ecosystem, including GGUF quantization and Turbo LoRA, next time.
- In the next X investigation, do not rely on Explore’s trending terms; search explicitly for open-source-specific terms such as Stable Diffusion, Wan, Flux, and Qwen.
- Prioritize Lemmy’s !«メールアドレス» as a primary source of open-source information.
- Do not use Pinterest as a source of open-source trends; limit it to observing marketing activity around closed-source tools.
- Follow up on whether any external tools depend on the planned Sora API shutdown on 2026-09-24.
Collected images


















































Data quality notes
Reddit fell short of the roughly 20-thread completion criterion because WebFetch was unavailable and only 12 precollected threads could be used. Most of X’s Explore trending terms were irrelevant, and no open-source search was conducted. Bluesky’s full-text search API returned 403, forcing alternative collection. Lemmy had no first-party closed-source posts. Pinterest used only one general query and did not conduct an open-source-focused follow-up search.



