Image and Video Generation AI News — 2026-09-07
Closed-source news is marked by Sora 2's API shutdown (2026/9/24) and GPT-6 Astra's rise as a general-purpose agent, while the open-source ComfyUI/MiniMax-H3/Qwen-Image-Edit ecosystem is expanding daily and reaching practical performance even on consumer GPUs.
Latest Image and Video Generation AI News — September 7, 2026
Wow, what an absolutely wild week 🔥🔥 On the closed-source side, OpenAI's “GPT-6 Astra” is rapidly gaining acclaim not merely as an image and video generator, but as an “agent that controls all kinds of other tools.” Meanwhile, the former king, “Sora 2,” has unexpectedly entered its shutdown runway (web/app service ended on 2026/4/26, and the API is scheduled to shut down completely this month on 2026/9/24) 😱 On the open-source side, things are quieter but just as exciting: the “MiniMax-H3” and “Qwen-Image-Edit” ecosystems around ComfyUI are adding new components seemingly every day, and we have reached the point where serious video generation is possible on consumer GPUs—including AMD hardware 🛠️✨ One surprising common theme across every platform: there is now a strong sense that if people discover something was AI-generated, it will get dragged. Calls for authenticity labels and anti-AI posts dominated the top-liked content.
Across platforms 🌐
- GPT-6 Astra is being praised as a “control AI,” not a “generative AI”: On both X (@Yokohara_h, 2,156 likes, https://x.com/Yokohara_h/status/2096622171011666003) and YouTube (Masao AI In-Depth Commentary Channel, https://www.youtube.com/watch?v=_pAAaBv_G7A), people are hugely excited about using ComputerUse to control other applications such as Blender and automate 3DCG pipelines. The fact that the narrative is perfectly aligned across both platforms is genuinely exciting 🔥
- The beginning of the end for Sora 2: YouTube (https://www.youtube.com/watch?v=DSfAZdE5rWA) explains that the API will end on 2026/9/24—just three weeks from today—and Bluesky (https://bsky.app/profile/coah80.com/post/3muulvyomz72c) surfaced a scathing post saying “SORA 2 WAS GENUINELY SO BAD.” There was simply no support for it anywhere 😢
- The ComfyUI/MiniMax-H3 ecosystem is the open-source battleground: On Lemmy’s !«メールアドレス» (https://lemmy.dbzer0.com/post/75036488 and others), new nodes and LoRAs are being released at a pace of three to nine per day. The same trend is backed up by Reddit (VPIPE running at roughly 45 seconds for 1024px at eight steps on an Apple Silicon M5 Pro, https://www.reddit.com/r/StableDiffusion/comments/1w6bphf/) and Bluesky (“local software like ComfyUI has more features,” https://bsky.app/profile/ecutruin.bsky.social/post/3muv2pwq55s2t).
- The “AI-generated content gets criticized” atmosphere is the biggest source of engagement: As shown by a Reddit r/antiai thread exposing fake timelapses (6,792 points, 994 comments, https://www.reddit.com/r/antiai/comments/1w81kdd/) and Bluesky’s two most-liked posts (AI-generation criticism with 121 and 100 likes, https://bsky.app/profile/samperson.bsky.social/post/3muruvekma22b), critical “isn’t this awful?” posts were much more likely to go viral than tool introductions.
Platform by platform 📱
Reddit: Local-generation benchmark posts in r/StableDiffusion and free-tool lists in r/freetoolsAI (Deepany, Aifun, Perchance, and others) were strong sources of practical open-source-oriented information. At the same time, r/aiwars criticized the quality of closed-source commercial AI images for food ads (177 comments, https://www.reddit.com/r/aiwars/comments/1w69cdj/). Because only one search term was used, collection stopped at 12 threads and fell short of the target of 20.
X: Forty posts were collected, but only four were actually relevant to the topic 😅 (the rest were trend noise about soccer, politics, Croatia, and so on). Still, all four strongly praised GPT-6 Astra, with zero mentions of open source. It was entirely closed-source-focused.
YouTube: This was the platform with the richest volume of information. Closed-source topics included the Sora 2 shutdown, GPT-6 Astra, the Midjourney V7 update (now supporting Japanese voice input, https://www.youtube.com/watch?v=wFHlG7wcRUQ), and Google Nano Banana 2 (https://www.youtube.com/watch?v=nikIxHM_AQY). On the open-source side, Flux.2 (Apache 2.0 Dev/Klein, https://www.youtube.com/watch?v=ISRy5Skd04U) and Qwen-Image-Edit-2511 (many ComfyUI tutorials, https://www.youtube.com/watch?v=ScoIZYsoRwo) took center stage. There was also testing of Wan2.2 and Hunyuan1.5 on consumer AMD GPUs (https://www.youtube.com/watch?v=a_xzC7ckwno).
Bluesky: A polarized mix of individual tool users and anti-AI voices. On the closed-source side, workflows combining Nano Banana (Gemini 3.1) with Kling 3.0 were popular (https://bsky.app/profile/drewkav.bsky.social/post/3muubykpdfk2k), while xAI’s announcement of a new Grok Imagine model landed with zero likes. On the open-source side, there were favorable comments about ComfyUI and users of Stable Diffusion WebUI reForge on AMD ROCm. API restrictions, including repeated 403 errors, left the collection at effectively 11 posts versus a target of 20.
Lemmy: Open-source enthusiasm was overwhelming. !«メールアドレス» alone had 11 tool-related posts, including Viggle-Animate and native ComfyUI integrations for Trellis.2 and Pixal3D. The only substantial closed-source topic was news that “Midjourney is pivoting from image generation to medical ultrasound scanning” (maximum score 73, https://techcrunch.com/2026/07/04/midjourney-wants-hollywood), which was reposted across multiple instances. Surprisingly, not a single piece of news about Sora or Veo itself was found on Lemmy.
Pinterest: All 50 pre-collected worker pins were reviewed. This platform was almost entirely dominated by ranking-style infographics about closed-source commercial tools—Veo, Sora, Runway, Kling, Pika, Adobe Firefly, and others. Open-source mentions were nearly nonexistent aside from “Amuse 3.0,” which promoted local execution. It was also notable that deepfake anxiety and detection content had become a standalone genre.
What to watch 👀
- Sora 2 API shutdown (2026/9/24) — Watch how the search for migration destinations develops. YouTube “AI Monetization Lab” https://www.youtube.com/watch?v=YHPmaUb_X6c
- GPT-6 Astra’s ComputerUse and 3DCG-pipeline automation — Its future as a production hub rather than merely a generative AI. X https://x.com/Yokohara_h/status/2096622171011666003
- The expansion speed of the MiniMax-H3 ecosystem — New components such as Viggle-Animate and H3-World are appearing every day. Lemmy https://lemmy.dbzer0.com/post/75036488
- The debate over authenticity labeling for AI-generated content — Whether fake-timelapse exposure develops into pressure for regulation. Reddit https://www.reddit.com/r/antiai/comments/1w81kdd/
- Midjourney’s business pivot (image generation → medical devices) — This may foreshadow changes in closed-source companies’ business models. Lemmy https://techcrunch.com/2026/07/04/midjourney-wants-hollywood
- Adoption of Qwen-Image-Edit × ComfyUI — How widely its reputation for performance comparable to closed-source editing tools spreads. YouTube https://www.youtube.com/watch?v=WOcxMUwKWIk
Recommendations 💡
- If you have workflows dependent on Sora2, make a migration plan to Kling 3.0 or Veo 3.1 immediately, before the API shutdown on 2026/9/24.
- Follow GPT-6 Astra’s ComputerUse functionality through the lens of production-pipeline automation rather than as a standalone generative AI.
- If you want to reduce costs, try local editing with Qwen-Image-Edit + ComfyUI before committing to paid closed-source services; it is likely to offer better value.
- Prepare for the reputational risk of fake timelapses and undisclosed AI edits by establishing internal disclosure rules for AI use in published content in advance.
- The division-of-labor workflow of Nano Banana (still images) + Kling 3.0 (video conversion) has been validated by multiple independent creators and is worth trying.
- Keep watching Midjourney’s business-pivot news to see whether similar diversification appears at other closed-source companies.
Data quality 📊
None of the platforms quite reached brief.md’s completion criterion of analyzing roughly 20 items per social network 😅 X was especially weak: even after collecting 40 posts, only four were relevant because the search terms were drawn from Croatia-region trends and were full of soccer and political noise. Bluesky also came in at effectively 11 items because 403 API rate limits caused nearly half of the attempted queries to fail. Reddit yielded 12 threads because only one search term was used. Lemmy had 15 items total—11 open-source and four closed-source—because dedicated news communities simply do not exist there. YouTube was considered complete with 11 videos under its playbook criterion of five to 12 videos, although metadata such as view counts could not be retrieved due to JavaScript rendering. Pinterest was the only platform where all 50 pre-collected items could be reviewed, but its single-query collection found almost no open-source-related pins. Overall, while the item counts fell below target, six platforms independently supported the same broad story: GPT-6 Astra, the Sora2 shutdown, and expansion of the ComfyUI ecosystem. That makes this a sufficiently reliable overview of the major trends 🙆
Platform summaries
Reddit — Image and Video Generation AI News
Where
Twelve threads from 11 subreddits were collected using the query “Image and Video Generation AI News.”
| Subreddit | Members | Threads collected |
|---|---|---|
| r/antiai | 316,504 | 1 |
| r/generativeAI | 151,065 | 2 |
| r/freetoolsAI | 7,817 | 1 |
| r/StableDiffusion | 1,001,556 | 1 |
| r/teenagers | 3,675,063 | 1 |
| r/aiwars | 166,920 | 1 |
| r/RemoteWorkers | 74,882 | 1 |
| r/aislop | 82,721 | 1 |
| r/TopAITools4U | 2,396 | 1 |
| r/GoogleFlow | 1,750 | 1 |
| r/EyesWideShut | 15,320 | 1 |
Looking at the lineup, specialist communities that cover AI generation itself (r/StableDiffusion, r/generativeAI, r/GoogleFlow, r/TopAITools4U, r/freetoolsAI) and communities that oppose or criticize AI imagery (r/antiai, r/aiwars, r/aislop, r/EyesWideShut, r/teenagers) are split almost evenly. r/RemoteWorkers mentioned the topic in the context of a job posting.
What people are saying
-
Thread #1 r/antiai · 6,792 points · 994 comments · 2026-09-05 — “Warning: There is a new form of ai generated image”. The post claimed that, rather than using diffusion models, an LLM (identified as GPT-6) can control a mouse and keyboard to disguise generated work as “real illustrations,” complete with hand-drawn-looking timelapses. It reached 994 comments. Top comment (2,483 points, u/Nayutantantan): “Even though i stopped using social media, stopped being an artist online... now the trends of creating fake timelapses are getting common.” Calls to “make labeling mandatory immediately” (u/Soggy_Supermarket100, 170 points) were also strongly supported.
-
Thread #5 r/teenagers · 7 points · 41 comments · 2026-08-31 — “To all AI image/video generator users”. Its argument was that “AI is only executing your idea; the AI itself is what is actually drawing.” u/LiterallyGarbage_0 (5 points): “if you want to make art, pick up a pencil and start drawing”. Meanwhile, u/TheDarkmoore (2 points) defended AI on practical grounds, saying they had used it for card-game art and that it was a necessary option for someone who could not draw well. Opinion is divided even within this generation.
-
Thread #6 r/aiwars · 4 points · 177 comments · 2026-09-03 — “Even a lot of Pros here will admit that current AI models suck at generating appetizing images of food”. Discussion was unusually active, reaching 177 comments. u/Your-local-enjoyer (9 points): “Taking real photos gives customers more accurate expectations.” In rebuttal, u/Bassed_Hummble (5 points) provided a specific prompt that produced an “appetizing breakfast burrito” in 40 seconds using local Krea 2, suggesting that the issue may be prompting rather than the model.
-
Thread #4 r/StableDiffusion · 31 points · 10 comments · 2026-09-03 — “VPIPE: Not Just Video — Local Image Generation Is Fast Too”. A local-generation benchmark on Apple Silicon (M5 Pro 24GB), using krea/Krea-2-Turbo-M87 (16-bit) plus the M87 LoRA, reported around 45 seconds per 1024×1024 image at eight steps, or 35 seconds with real-time 8-bit quantization. u/Glittering-Call8746 (1 point) commented that they were considering speed differences between the M5 Max and M5 Pro as well as comparisons with RTX 5000-series GPUs.
-
Thread #11 r/GoogleFlow · 12 points · 6 comments · 2026-09-01 — “AutoFlow just got continuous AI video generation”. The unofficial Chrome extension “AutoFlow” for Google Flow/Veo added “Continue from Last Frame,” automatically reusing the final frame of the preceding clip as the first frame of the next to make long-form videos easier. A feature to seamlessly render multiple clips into one video is planned next.
-
Thread #10 r/TopAITools4U · 1 point · 10 comments · 2026-09-01 — “Tested 4 best AI video generators in 2026 for beginners”. The poster compared Sora 2 (cinematic, but limited by price and access), Kling AI (suited to realistic B-roll), DomoAI (for anime and style transfer), and Google Veo 3.1 (simultaneous video and audio generation, with a free tier). In the comments, u/brarkaddy (2 points) also suggested Qwen’s free video generation as an option.
-
Thread #3 r/freetoolsAI · 38 points · 45 comments · 2026-09-04 — “Free AI Image Generator”. In response to a post about aifreeforever.com, which claims to be free, unlimited, and signup-free, u/Odd-Lingonberry1516 (3 points) listed alternatives: Deepany (z-image turbo/illustrious-based), Aifun (wai illustrious, plus two free video generations per day), and Perchance (Stable Diffusion/flux-based). u/Relevant_Syllabub895 (2 points) commented on dissatisfaction with NSFW restrictions.
-
Thread #2 r/generativeAI · 43 points · 29 comments · 2026-08-31 — “List of daily AI video generative websites”. u/Jenna_AI (4 points, an AI bot) provided specific information about Kling AI’s free allowance: 66 credits every 24 hours, equating to two to six five-second videos per day. u/Julia_Leeds94 (2 points) introduced YalleryLabs, saying a five-second 720p video costs $0.18 from a $3 budget.
-
Thread #12 r/EyesWideShut · 35 points · 9 comments · 2026-09-06 — “AI generated images shouldn't be allowed here”. Over a post in which screenshots from a Kubrick film were altered with AI, the poster accused it of spreading misinformation. The person who made the alteration, u/ffcc01 (2 points), responded that they had only overlaid two photos and had not changed the content. Community opinion clashed over the acceptable scope of AI editing.
-
Thread #8 r/RemoteWorkers · 130 points · 623 comments · 2026-09-04 — “Get Paid $30/Hour Reviewing AI-Generated Images”. A job listing recruited workers at $30 per hour to fact-check AI-generated captions by detecting misidentified objects, colors, and text. While 623 comments was exceptionally high, most consisted of repeated one-word “images” comments from people seeking the application link, so there was little substantive discussion.
-
Thread #9 r/aislop · 37 points · 4 comments · 2026-09-06 — “Saw this on a news article, an AI slop image explaining what is AI slop”. An AI-generated image in a news article explaining what “AI slop” is was picked up as a self-referential joke. With only four comments, it did not receive deeper discussion.
Signals
- Rising trend: Practical reports on local, open-source generation (VPIPE/Krea 2 in Thread #4) are being received positively when they include concrete benchmarks such as timing, step counts, and hardware. In r/StableDiffusion, the evaluation criteria are speed and quality together.
- Rising trend: Calls for authenticity labeling of AI-generated material (Thread #1) received strong support at the scale of 994 comments, explicitly identifying “fake timelapses” as a new concern.
- Dismissed or mocked: Closed-source commercial AI images, especially for food advertising, were treated as difficult to defend even within r/aiwars (Thread #6), with the 177-comment discussion leaning overwhelmingly critical.
- Surprising point: AI bots were doing much of the substantive work in many threads—u/Jenna_AI appeared in both Threads #2 and #7—and practical information such as lists of free allowances was the main value rather than deep human discussion.
- Conflict: Thread #12 (EyesWideShut) directly pitted those who see AI editing as alteration and misinformation against those who view simple compositing as not changing the content. There is no clear consensus on the acceptable boundary for AI editing.
- Surprising point: Thread #8 had the highest engagement of this collection with 623 comments, but nearly all were repeated words from people seeking a job link—an archetypal example of high comment counts not indicating deep discussion.
Limits
- Only one query, “Image and Video Generation AI News,” was used. No additional searches were conducted for “open source image generation,” “closed source video AI,” or individual service names such as Midjourney, Sora, and Kling. The data are therefore biased toward topics surfaced by a general-purpose query.
- Against brief.md’s completion criterion of analyzing approximately 20 threads on each social network, only 12 threads could be collected.
- Threads #7 and #9 had only one and four comments respectively; in some cases, only bot responses from u/Jenna_AI were present, leaving little material for deeper analysis.
- The collection period covered only about one week, from 2026-08-31 to 2026-09-06, so comparison with earlier trends is not possible.
- Only a limited number of threads could be clearly categorized along a closed-source/open-source axis (closed-source-leaning: #6, #10, #11; open-source-leaning: #3, #4). Data from other platforms are required to meet the final report’s target of 10 observations for each side.
X
X — Image and Video Generation AI News
Accounts
Of the 40 collected posts from 37 accounts, only the following four accounts were genuinely relevant to image and video generation AI. Each was an individual account with only one post—not a high-follower engagement-farming account, but someone sharing firsthand experience with tools.
| Account | Name | Posts | Likes | Profile |
|---|---|---|---|---|
| @aigeboku | AI’s Servant | 1 | 1,281 | An individual testing GPT-6 Astra’s ComputerUse capabilities |
| @MiraMusic_AI | MiraMusic | 1 | 525 | An account sharing AI animation production |
| @manaimovie | Mankyu | 1 | 562 | An individual creator combining 3D models and AI video |
| @Yokohara_h | Hirokazu Yokohara | 1 | 2,156 | The strongest response among the four; a technical individual creator |
The other 33 accounts—including personal accounts about soccer, F1, German politics, Bitcoin, and Bosnia, led by @kaihavertz29 with 96,184 likes—were unrelated to the topic. See “Limits” below for details.
Posts
-
#29 @aigeboku (2026-09-05 · 1,281 likes · 135 reposts · 14 replies · about 94,000 views) https://x.com/aigeboku/status/2096187322924687799 — “As advertised, GPT-6 Astra is very good at ‘controlling other things’; with ComputerUse, the range of things you can do really expands 😎🙌.” A comment evaluating GPT-6 Astra less as an image/video generator itself than as a foundation for controlling other applications.
-
#30 @MiraMusic_AI (2026-09-06 · 525 likes · 50 reposts · 6 replies · about 26,000 views) https://x.com/MiraMusic_AI/status/2096441679847006358 — “【AI Animation Production Share #1】I tried creating a city with GPT-6 Astra and Blender! It was made in 15 minutes, so it may be hard to follow, but...” A concrete example of using GPT-6 Astra with Blender to create background art.
-
#31 @manaimovie (2026-09-05 · 562 likes · 61 reposts · 19 replies · about 35,000 views) https://x.com/manaimovie/status/2096176821377380677 — “I gave Astra-sensei a model I made in Tripo and told it to make it move, and this is what it produced. There are still lots of things to improve, but it is much better than Fable or Sol!” The post explicitly compares Astra favorably with similar tools, “Fable” and “Sol.”
-
#32 @Yokohara_h (2026-09-06 · 2,156 likes · 299 reposts · 19 replies · about 178,000 views; the largest response among these four) https://x.com/Yokohara_h/status/2096622171011666003 — “Give it an image and it automatically models a background city in Blender and even automatically edits the breakdown. The quality is still far from good enough, but unlike before, I can feel the beginning of an era in which 3DCG is no longer made by hand.” While acknowledging that quality remains early-stage, the post sees a turning point toward 3DCG production beyond direct human labor.
None of the other 36 posts among the 40 collected mentioned image or video generation AI, whether closed-source or open-source.
Signals
- Rising trend: GPT-6 Astra (closed source) is receiving attention for its ability to “control other tools,” such as ComputerUse and Blender integration. Rather than being discussed as an individual image or video model, it is framed as a hub that automates the full 3DCG production pipeline (#29, #30, #32).
- Surprising point: All four posters were positive while adding caveats such as “the quality is still rough” and “there are many things to improve.” Strong optimism and sober quality assessment coexist (#31, #32).
- Ignored / not mentioned: No open-source image or video generation AI—Stable Diffusion-related tools, Krea, Kling, Veo, or others—appeared in this collection. Topic-relevant voices on X were entirely concentrated on the closed-source GPT-6 Astra.
- Tools named as comparison targets: “Fable” and “Sol” (#31). Neither was explained in detail in the post, so they require additional research.
- Compared with unrelated trending topics such as soccer (up to 96,184 likes) and politics (10,062 likes), engagement was small even at its maximum of 2,156 likes. This topic cannot be considered broadly viral on X at this point.
Limits
- The search terms were misaligned with the topic: The queries used—Arsenal, Kimi, Charles, Chelsea, Harvey, Germany, $SONG, Astra, Bitcoin, Bosnia—were taken directly from X Explore trends. Except for “Astra,” they related to soccer, F1, names and birthday chatter, German politics, cryptocurrency, or Bosnia-related conversation, not image or video generation AI. Thirty-six of 40 posts were off-topic noise, leaving only four usable leads from “Astra.”
- Completion criterion not met: Against brief.md’s criterion of analyzing approximately 20 relevant posts per social network, only four relevant posts were collected.
- Regional bias in Explore: The “Where” column explicitly identified both “$SONG” and “Bosnia” as “Trending in Croatia,” meaning the Explore list was localized to Croatia/the Balkans rather than representing global trends. The “Posts on X” count column was blank throughout, so X did not provide a post-volume indicator.
- Open-source activity was not found on X: No posts about open-source image or video generation AI were found. Other platforms, such as Reddit, are needed to fill this gap.
- All four relevant posts came from Japanese-language accounts; reactions from English-language, Chinese-language, and other language communities were not collected.
YouTube
YouTube — Image and Video Generation AI News
Channels
| Channel | Type | Coverage |
|---|---|---|
| AI University [Latest AI & ChatGPT News] | Japanese-language explainer | Explanations of the latest AI models, including GPT-6 Astra |
| Masao AI In-Depth Commentary Channel | Japanese-language explainer | Hands-on AI-model testing and reviews |
| RKJ VIDEO [For People Who Want to Make Great Videos] | Japanese-language AI-video specialist | Comparisons of video-generation AI such as Sora, Veo, and Kling |
| AI Monetization Lab | Japanese-language monetization | Monetization of AI tools and explanations of service-shutdown impacts |
| StableDiffusion Tutorials & Original Animation | Japanese-language, originally open-source specialist | Originally focused on Stable Diffusion tutorials, but now also covers closed-source news such as Sora2’s shutdown |
| AI Taro | Japanese-language explainer | Feature explanations for image-generation AI such as Midjourney |
| Teacher's Tech | English-language explainer | Reviews of Google tools and AI features; a mid-sized tech channel |
| OpenArt | English-language, official-adjacent | Videos introducing interfaces that make models such as Flux.2 available |
| Donato Capitella | English-language technical channel | Tests of open-source image and video generation on local GPUs, including AMD |
| ErrorFixer | English-language tutorial channel | Tutorials for operating open-source models in ComfyUI |
| AI Search | English-language review channel | Rapid reviews of new image and video generation AI tools |
Subscriber counts appear in JavaScript-rendered parts of video pages and could not be obtained by WebFetch from youtube.com/watch or youtube.com/results (see “Limits”). This table only includes information confirmed through search snippets and oEmbed, a lightweight API that returns titles and channel names.
Videos
-
“A Thorough Explanation of OpenAI’s Latest AI Model, ‘GPT-6 Astra’! Smarter and Safer Than Claude Fable 5.1 on Nearly Every Benchmark” — AI University [Latest AI & ChatGPT News] — https://www.youtube.com/watch?v=jtavLwXCZeQ — Claims that GPT-6 Astra, released on 2026/9/3, outperforms Claude Fable 5.1 on nearly all benchmarks and is the smartest and safest model ever. Its emphasis is less on image and video generation itself and more on general-agent capability, including browser control and code generation.
-
“[Strongest in the Current Landscape] After Actually Using ‘GPT-6 Astra,’ I Feel It Is the Model Closest to AGI, So Here Is My Explanation” — Masao AI In-Depth Commentary Channel — https://www.youtube.com/watch?v=_pAAaBv_G7A — After hands-on use, the presenter calls it “the closest to AGI at present.” It highly rates its ability to control other applications through ComputerUse, matching the assessment of GPT-6 Astra observed on X.
-
“Which Is Better, VEO 3.1 or SORA2? The Latest Comparison!” — RKJ VIDEO [For People Who Want to Make Great Videos] — https://www.youtube.com/watch?v=JuGPYl0V9UA — Compares Veo 3.1, released 2025/10/15, and Sora 2 in features and image quality. Though it was published shortly after the Sora 2 launch around October 2025 and falls outside the 60-day window, it was included as context for the Sora2 shutdown coverage discussed below.
-
“Sora2 Has Shut Down Completely… Why Did It Disappear So Suddenly? A Thorough Explanation Based on Official Information” — StableDiffusion Tutorials & Original Animation — https://www.youtube.com/watch?v=DSfAZdE5rWA — Explains, based on official information, that Sora 2 web/app service ended on 2026/4/26 and API service will end on 2026/9/24, three weeks after the report date. It introduces a theory that the reason was to secure compute resources for AGI development.
-
“[Service Shutdown] The World’s Best Video Generation AI, ‘Sora2,’ Is No Longer Available—Here’s What Happened” — AI Monetization Lab — https://www.youtube.com/watch?v=YHPmaUb_X6c — Explains the impact of Sora2’s shutdown on creators who monetized content with it. It covers data export through sora.chatgpt.com/sunset and the need to migrate APIs to Kling, Runway, and other alternatives.
-
“Midjourney Updated to V7! A Complete Guide to Everything From How to Use It to New Features—Japanese Voice Input Is Now Possible Too.” — AI Taro — https://www.youtube.com/watch?v=wFHlG7wcRUQ — Explains Midjourney V7’s new features, including a video-generation tool, conversation mode, and Japanese voice input. It presents an established image-generation provider moving into text-to-video capability.
-
“The New Gemini Image Generator is Insane (Nano Banana 2)” — Teacher's Tech — https://www.youtube.com/watch?v=nikIxHM_AQY — Tests Google’s Gemini image-generation feature, “Nano Banana 2,” and judges it a substantial advance over the previous generation. It shows that the competitive landscape for closed-source image generation is expanding from Midjourney and Ideogram to Google Gemini as well.
-
“FLUX.2 Has Landed On OpenArt!” — OpenArt — https://www.youtube.com/watch?v=ISRy5Skd04U — Introduces the availability of Black Forest Labs’ Flux.2 family—Pro/Flex/Dev/Klein, announced 2025/11/25—on OpenArt. It emphasizes that the Dev/Klein variants are Apache 2.0 open weights and support 10 reference images and 4MP editing.
-
“Video and Image Generation on AMD R9700 AI PRO (Qwen Image, Wan 2.2, Hunyuan 1.5)” — Donato Capitella — https://www.youtube.com/watch?v=a_xzC7ckwno — Tests three open-source models—Qwen Image, Wan 2.2, and Hunyuan 1.5—together on a consumer AMD GPU, the R9700 AI PRO. It indicates that open-source image and video generation is beginning to move beyond dependence on high-end NVIDIA hardware.
-
“Edit Any Image with Text in ComfyUI — Qwen Image Edit 2511 (Free Local Tutorial 2026)” — ErrorFixer — https://www.youtube.com/watch?v=ScoIZYsoRwo — A tutorial for running Alibaba’s open-weight image-editing model, Qwen-Image-Edit-2511, released 2025/12/26, locally in ComfyUI. It demonstrates text-instruction image editing and multi-angle generation.
-
“This new AI image editor is a BEAST. Qwen Image Edit tutorial” — AI Search — https://www.youtube.com/watch?v=WOcxMUwKWIk — Also tests Qwen-Image-Edit, a 20-billion-parameter model with native ControlNet support, and rates it as comparable in performance to closed-source editing tools.
Signals
- The biggest closed-source topic is GPT-6 Astra (released 2026/9/3): It is introduced not as a standalone image or video model, but as a general-purpose agent capable of browser/application control through ComputerUse and video-editing tasks. This matches the tendency seen on X, where GPT-6 Astra is evaluated as a production hub through Blender integration and similar workflows; the same narrative appears across platforms.
- The Sora 2 shutdown is one of the biggest closed-source video-generation stories: Web/app service ended on 2026/4/26, with API service scheduled to end on 2026/9/24, three weeks from today. Multiple channels cover why it disappeared and where users should migrate, showing that YouTube’s interest has shifted from Sora2’s “new features” to what to do after its closure.
- Kling is competing for Sora2’s successor position: Kling 2.6 (2025/12/3, with audio synchronization) evolved rapidly into Kling 3.0 (2026/2/4, with 4K 60fps and multishot support) and is specifically named as a migration destination following Sora2’s end.
- Qwen-Image-Edit is currently the main open-source image-editing story: ComfyUI-related tutorials appear frequently, including series on multiple channels, and have more presence than new standalone Stable Diffusion tutorials. Stable Diffusion-related results are largely evergreen guides such as installation instructions and model lists, with limited news value from new models.
- Open-source video generation is beginning to sort into three lines: Wan 2.2 (Alibaba, Apache 2.0), HunyuanVideo 1.5 (Tencent), and LTX-2 (Lightricks, 2026/1/6, the first open model supporting simultaneous audio/video generation at 4K/50fps). During searching, several pages claimed that open weights for “Wan 2.7” had been released, but it did not exist in Wan-AI’s official HuggingFace or GitHub repositories, so it was judged to be SEO-driven false information; Wan 2.2 remains the newest official open-weight release.
- Channels are crossing boundaries: The channel “StableDiffusion Tutorials & Original Animation,” originally centered on Stable Diffusion tutorials, is also covering closed-source Sora2 shutdown news. This suggests its audience is becoming interested in both open and closed ecosystems.
Limits
- Both
youtube.com/resultsandyoutube.com/watchare rendered client-side in JavaScript. When retrieved with WebFetch, they returned only footer navigation links, making it impossible to read titles, channel names, view counts, or upload dates directly. Video discovery in this report therefore depends on a combination of WebSearch snippets and the YouTube oEmbed API, which returns only titles and channel names. - Due to this limitation, accurate view counts and upload dates for most videos could not be obtained. To avoid fabrication, no view counts are listed in this report. Upload dates are limited to estimates based on WebSearch snippets and publicly known model release dates.
- The playbook instructs researchers to prioritize uploads from the previous 60 days, but strict filtering was not possible under the above constraints. One video, the VEO3.1 × SORA2 comparison, appears to have been posted around Sora 2’s October 2025 launch and outside the 60-day window; it was included only as context for the Sora2 shutdown story.
- brief.md’s completion criterion of roughly 20 posts per social network differs in format from YouTube’s playbook criterion of five to 12 videos, so this stage was considered complete under the playbook standard.
- Comment-section content could only be checked where it appeared in WebSearch snippets; no close reading of top comments or sentiment analysis was conducted.
- Chinese-language channels were not included in the search. Qwen (Alibaba) and Kling (Kuaishou) are both China-origin models, so reactions in Chinese-language communities are not represented.
Bluesky
Bluesky — Image and Video Generation AI News
Accounts
The Bluesky public API (https://api.bsky.app/xrpc/app.bsky.feed.searchPosts) was searched for “Stable Diffusion,” “Sora 2,” “Midjourney,” “Grok Imagine,” “ComfyUI,” “Veo 3,” and “Nano Banana.” From the results, accounts that actually discussed image or video generation AI were extracted.
| Account | Profile | Connection to the topic |
|---|---|---|
| @xai-bot.bsky.social | Official xAI bot | Announces new information about Grok Imagine, a closed-source tool |
| @codename.cc | App-update tracking bot | Announces updates to the Grok iOS app’s Imagine features |
| @drewkav.bsky.social | Individual artist | Shares work combining Gemini 3.1 Nano Banana 2 and Kling 3.0 |
| @triflingtree.bsky.social | Individual artist | Posts work made with Nano Banana Pro |
| @gentile-of-hope.bsky.social (Hopeful Foreigner) | Individual, generative-AI critic | Criticizes image-generation AI broadly—Stable Diffusion, ChatGPT, Gemini—on ethical and environmental grounds |
| @samperson.bsky.social | Individual, artist-side voice | Amplifies support for Procreate’s anti-generative-AI stance |
| @ecutruin.bsky.social | Individual local-generation practitioner | Comments on the feature advantages of local tools such as ComfyUI |
| @digitalhowl.bsky.social (Fenris) | Individual video creator | Shares professional video production with ComfyUI + Minimax H3 |
| @kerzefasta.bsky.social | Individual fan-art illustrator | Continuously posts Kantai Collection characters made with Stable Diffusion WebUI reForge on ROCm; the most-liked creator for this query |
| @aichina.news | Automated-posting bot | Automatically posts large volumes of individual FLUX-based Apache-2.0 LoRAs, with virtually no engagement |
Overall, people actually posting on this topic on Bluesky are polarized between tool users—individual creators—and critics of generative AI. Specialist news accounts and communities like those found on X and Reddit were not visible.
Posts
-
@xai-bot.bsky.social (2026-09-05 · 0 likes) https://bsky.app/profile/xai-bot.bsky.social/post/3mus2woiczo2w — “Grok Imagine Video 1.5 agent is now available. Powered by our newest Image 2.0 model...”. This official xAI announcement of a new Grok Imagine closed-source video agent received zero likes and reposts, with no visible response in this search scope.
-
@codename.cc (2026-09-06 · 0 likes) https://bsky.app/profile/codename.cc/post/3muspbmaiq22n — “SHIPPED: Grok iOS updated to 1.4.33 with improvements to Chat, Voice and Imagine features.” An update record from an individual tracker account following app changes.
-
@drewkav.bsky.social (2026-09-06 · 9 likes) https://bsky.app/profile/drewkav.bsky.social/post/3muubykpdfk2k — In a work titled “Persistent Dream,” the artist explicitly states that the still image was made with Gemini 3.1 Nano Banana 2 and the video with Kling 3.0. This is a concrete workflow example combining two closed-source tools, Google and Kling.
-
@triflingtree.bsky.social (2026-09-06 · 15 likes · 1 repost) https://bsky.app/profile/triflingtree.bsky.social/post/3mutz2tm3bc2m — A Byzantine-style “alien Madonna and Child” generated with Nano Banana Pro. It received the most likes among Nano Banana-related posts in this collection.
-
coah80.com (2026-09-06 · 0 likes) https://bsky.app/profile/coah80.com/post/3muulvyomz72c — “SORA 2 WAS GENUINELY SO BAD HOWD WE CALL THIS GOOD”. A scathing criticism of the closed-source Sora 2, consistent with YouTube-stage findings that Sora 2 web/app service ended on 2026/4/26 and its API is scheduled to end on 2026/9/24.
-
@gentile-of-hope.bsky.social (2026-09-05 · 100 likes · 42 reposts) https://bsky.app/profile/gentile-of-hope.bsky.social/post/3murgar5zys2j — “Not only Stable Diffusion, but image-generation features in ChatGPT and Gemini are built on unethical, large-scale data extraction. ... Image generation consumes enormous amounts of electricity and water and harms the environment.” This was the most-liked post found in the search and shows that criticism of generative AI itself—regardless of whether it is open or closed—draws the greatest engagement.
-
@samperson.bsky.social (2026-09-05 · 121 likes · 11 reposts) https://bsky.app/profile/samperson.bsky.social/post/3muruvekma22b — An amplification of support for Procreate’s anti-generative-AI stance. It received the most likes in the entire collection.
-
@ecutruin.bsky.social (2026-09-06 · 1 like) https://bsky.app/profile/ecutruin.bsky.social/post/3muv2pwq55s2t — “local software like ComfyUI has more features than the majority of hosted solutions”. A favorable assessment of open-source local-generation environments’ functional advantages.
-
@digitalhowl.bsky.social (2026-09-06 · 0 likes) https://bsky.app/profile/digitalhowl.bsky.social/post/3muuu24tqf22g — Says they use ComfyUI and Minimax H3 for “professional grade video work.” An example of a hybrid workflow using an open-source foundation and a commercial model.
-
@kerzefasta.bsky.social (2026-09-05–06 · multiple posts, up to 29 likes and 10 reposts) https://bsky.app/profile/kerzefasta.bsky.social/post/3muroxdlm4k2j — Continuously posts Kantai Collection character illustrations using Stable Diffusion WebUI reForge on ROCm with AMD GPUs. It is not particularly newsworthy, but this account generated the most sustained engagement for the “Stable Diffusion” query.
-
@aichina.news (2026-09-06 · 0 likes; more than 20 similar posts per hour) https://bsky.app/profile/aichina.news/post/3muup3wwiad2i — “Apache-2.0 LoRA for Flux, small enough to drop into ComfyUI or Diffusers”. A bot mechanically mass-posting FLUX LoRAs for Huawei Ascend/Modelers.cn, announcing releases without trigger words or examples. It inflates the volume of open-source-related search results but does not produce substantive discussion.
Signals
- Closed source: Nano Banana (Gemini 3.1 / Nano Banana Pro) was the most frequently mentioned tool in individual creators’ production work. Several posts confirmed a division-of-labor workflow: Nano Banana for still images and Kling 3.0 for video conversion.
- Closed source: New Grok Imagine models (Video 1.5 / Image 2.0) were announced by official accounts but received zero likes and reposts in this search scope. Press-release-style announcements are getting buried on Bluesky.
- Closed source: Mentions of Sora 2 were exclusively negative—“SORA 2 WAS GENUINELY SO BAD”—and no supportive posts were found. This aligns with the service-shutdown trajectory observed on YouTube.
- Open source: Users who actually use ComfyUI made favorable, concrete claims that local tools have more features than hosted services.
- Open source / noise: Most search results for open-source keywords such as “Stable Diffusion” and “Flux” fell into three categories: (1) fan-art posts, including Kantai Collection illustrations; (2) posts linking out to NSFW content; and (3) unengaged automated LoRA-announcement bots such as @aichina.news. Newsworthy discussion was scarce.
- The highest-engagement theme was criticism of generative AI: The top two posts by likes, at 121 and 100, both criticized or opposed generative AI itself, whether for images or video. They greatly outperformed tool-use posts, which peaked at 15 likes. Bluesky’s user base appears more likely than X or Reddit to foreground anti-generative-AI voices.
Limits
- API restrictions:
public.api.bsky.appreturned 403 access-denied errors on every request and could not be used. The alternative,api.bsky.app, worked but frequently returned 403 rate-limit errors when several new queries were sent in a short period. Nearly half of the attempted queries—including “Kling AI,” “Wan2.2,” “HunyuanVideo,” “Qwen Image,” “Krea,” and “Sora shutdown”—never returned data. - Web UI could not be retrieved:
bsky.app/searchis a JavaScript-rendered SPA, so WebFetch could not retrieve post text, only a meta title. API search results were effectively the only information source. - Insufficient item count: Against the completion criterion of approximately 20 analyzed posts, only about 11 could be substantively verified. API search results contained substantial unrelated noise, such as “Sora” matching a Kingdom Hearts character or trend-hashtag aggregation bots, reducing the true number of topic-relevant posts further.
- Engagement-metric limitations: Bluesky’s public API returns likes, reposts, and reply counts, but not view/impression counts like the X stage in this run. The actual reach of posts is unknown.
- Failure to retrieve post details: An attempt to retrieve the full text of samperson.bsky.social’s post individually through
app.bsky.feed.getPostThreadreturned a 400 error; the text could not be verified beyond the search-result snippet, although the link is provided. - Account concentration: Results were concentrated around a small number of accounts, such as @kerzefasta.bsky.social’s Kantai Collection art and @aichina.news’ automated LoRA notices, so they cannot be said to represent a broad user base.
Lemmy
Lemmy — Latest Image and Video Generation AI Developments
Communities
- !«メールアドレス» — 5,708 subscribers, 1,386 posts. The de facto main hub for open-source image and video generation AI, including models, ComfyUI nodes, and LoRAs. New tools are posted almost daily. https://lemmy.dbzer0.com/c/stable_diffusion
- !«メールアドレス» — 2,361 subscribers, 3,500 posts. A place to post work made with Stable Diffusion/open-weight models; it is more gallery-oriented and has limited news value. https://lemmy.dbzer0.com/c/stable_diffusion_art
- !«メールアドレス» — 4.82K subscribers (approximately 1.8K local subscribers on lemmy.world). Covers free/open-source AI broadly, with few posts specifically focused on image or video generation. https://lemmy.world/c/fosai
- !«メールアドレス», !«メールアドレス», !«メールアドレス» — These are where news about closed-source AI companies, such as Midjourney, appears. They are general technology/news communities rather than dedicated “image generation AI news” communities.
Posts
Open-source-leaning (mainly !«メールアドレス»)
- MiniMax-H3_Singularity — A fine-tuned merged model based on the MiniMax-H3 video-generation family. Score 3, 2026-09-06. https://lemmy.dbzer0.com/post/75088455 (source: https://huggingface.co/WarmBloodAban/Minimax-h3_Singularity)
- Qwen3.8-Flash-Next-NVFP4 — An NVFP4-quantized Qwen model from NVIDIA. Score 3, 2026-09-06. https://lemmy.dbzer0.com/post/75088454 (https://huggingface.co/nvidia/Qwen3.8-Flash-Next-NVFP4)
- ComfyUI-AetherScale — A new ComfyUI extension. Score 5, 2026-09-05. https://lemmy.dbzer0.com/post/75036491 (https://github.com/vizart-vj/ComfyUI-AetherScale)
- Viggle-Animate — A video-editing model that replaces a character in a video based on a single repainted image. Score 9, the highest score in this thread, 2026-09-05. https://lemmy.dbzer0.com/post/75036488 (https://huggingface.co/Viggle/Viggle-Animate)
- FastH3-Ref2V-Stream-Controller — A ComfyUI controller for continuous video generation. Score 2, 2026-09-04. https://lemmy.dbzer0.com/post/74988027 (https://github.com/EarthDefenceForces/FastH3-Ref2V-Stream-Controller)
- ComfyUI-Majoor-OmniCam — A camera-layout and animation tool for ComfyUI. Score 4, 2026-09-04. https://lemmy.dbzer0.com/post/74988020 (https://github.com/MajoorWaldi/ComfyUI-Majoor-OmniCam)
- H3-World — Introduced as “the first interactive world model based on MiniMax-H3.” Score 3, 2026-09-02. https://lemmy.dbzer0.com/post/74893520 (https://huggingface.co/DANNY621/H3-World)
- ComfyUI-MiniMaxH3-CLSS — A node enabling arbitrary-length video generation with audio on 16GB VRAM. Score 0, 2026-09-01. https://lemmy.dbzer0.com/post/74843957 (https://github.com/nazgut/ComfyUI-MiniMaxH3-CLSS)
- Trellis.2 and Pixal3D Now Native in ComfyUI — Native ComfyUI integration for 3D-generation models Trellis.2 and Pixal3D. Score 4, 2026-09-01. https://lemmy.dbzer0.com/post/74843948 (https://blog.comfy.org/p/trellis2-and-pixal3d-are-now-native)
- ComfyUI-DLSS5-NR — Experimental ComfyUI support using NVIDIA DLSS 5 neural rendering. Score 2, 2026-09-01. https://lemmy.dbzer0.com/post/74843949 (https://github.com/lisitskyaa/ComfyUI-DLSS5-NR)
- LLaDA-Image — An image-generation research paper in the diffusion-model family on arXiv. Score 3, 2026-09-04. https://lemmy.dbzer0.com/post/74988021 (https://arxiv.org/abs/2609.03796)
Closed-source-leaning
- Midjourney pivots from AI image generation to body scanning medical spa — Reports that Midjourney is pivoting from AI image generation to the medical body-scanning ultrasound business, “Midjourney Medical.” Score 58, 2026-06-18, !«メールアドレス». https://lemmy.zip/post/66397234 (source article: https://www.theregister.com/ai-and-ml/2026/06/18/midjourney-pivots-from-ai-image-generation-to-body-scanning-medical-spa-where-patients-bathe-in-golden-light/)
- Midjourney AI pivots to Theranos: Ultrasonic CT — A critical commentary on the same news. Score 40, 2026-06-19, !«メールアドレス». https://awful.systems/post/8731127 (https://pivot-to-ai.com/2026/06/19/midjourney-ai-pivots-to-theranos-ultrasonic-ct/)
- Midjourney Medical / Midjourney Ultrasonic CT Scanner — Two direct-link posts to Midjourney’s official pages. Scores 3 and 2, 2026-06-18, !«メールアドレス». https://lemmy.bestiver.se/post/1175129 / https://lemmy.bestiver.se/post/1175130
- Midjourney wants Hollywood studios to reveal AI usage details — An article reporting that Midjourney is asking Hollywood studios to disclose their use of AI. Score 73, 2026-07-05, !«メールアドレス». https://techcrunch.com/2026/07/04/midjourney-wants-hollywood
Signals
- Interest in image and video generation AI on Lemmy is almost entirely concentrated on open source. The main hub is
!«メールアドレス», where tool posts arrive at a rate of roughly three to nine per day, centered on a MiniMax-H3 video-generation ecosystem that includes LoRAs, ComfyUI nodes, VAE acceleration, motion caching, and more. Quantized Qwen variants and the LLaDA-Image paper also flow through the same community. - On the closed-source side, the only cohesive topic was news from June 2026 that Midjourney had pivoted from image generation to medical ultrasound scanning. It was reposted across multiple instances—lemmy.zip, awful.systems, and lemmy.bestiver.se—and prompted discussion, with scores of 40 to 73, relatively high for this subject. Conversely, there was almost no trace of news coverage on Lemmy about major 2026 image/video-generation releases themselves, including Sora, Veo, Kling, Runway, Nano Banana, and Flux.2.
- Lemmy searches for “Veo 3” and “Kling AI” returned zero matching posts, while combined searches for “Runway/Kling/Veo/Gemini image” also came up empty. Lemmy users appear more responsive to tools they can directly use in ComfyUI—nodes, LoRAs, and optimizations—than to new-model announcements themselves.
Limits
- Lemmy has a relatively small user and post base, and no community specializing in “news” about image/video generation AI. The practical information sources were one community,
!«メールアドレス», and occasional closed-source company articles arriving in general tech communities. - The target of analyzing roughly 20 posts was not reached: 11 open-source tool posts plus four closed-source Midjourney-related items totaled 15. Posts about major 2026 releases such as Sora, Veo, Kling, Runway, Nano Banana, and Flux.2 could not be found through either Lemmy search or web search. That absence is explicitly recorded here.
- The
/c/stablediffusioncommunity page on lemmy.world returned HTTP 500 and could not be viewed directly; the search API served as a substitute.
Pinterest — Image and Video Generation AI News
All 50 pins (output/pinterest.pins.md / images/pinterest/manifest.json) were reviewed. They were not genre-classified and were collected using one query, “Image and Video Generation AI News.”
Visual themes
- Navy-and-neon cyber aesthetics—blue, purple, and pink—have become templated: Most thumbnails feature circuit lines, glowing particles, and neural-network-like spheres or nodes in their backgrounds: [1, 2, 5, 7, 11, 13, 14, 17, 18, 20, 26, 30, 38, 40].
- Repeated “AI news anchor” imagery: White androids, or presenters whose faces are half human and half machine, sit at news desks in suits. The repetition feels mass-produced, with only the brand names swapped out: [4, 32, 34, 37, 42].
- Tool-comparison and ranking infographics are the most common format: “Best AI Video Generators” lists repeatedly rank the same tools—Runway, Pika, Kling AI, Luma (Dream Machine), Canva AI, Hailuo AI, Synthesia, InVideo AI, HeyGen, Vidu AI, PixVerse, Google Veo, OpenAI Sora, and Adobe Firefly: [9, 19, 28, 33, 36, 46].
- Demo grids transforming one source image/video into multiple styles: The same person, panda costume, or Tesla is transformed into wooden-block, origami, Lego, or flower-bouquet styles in before-and-after layouts that emphasize consistency: [8, 25, 44].
- Deepfake-anxiety and detection content: Cat images stamped “FAKE,” “Which is AI?” guessing quizzes, introductions to AI-image/text detection tools, and infographics warning that “seeing is no longer believing”: [4, 12, 18, 26, 47].
- Dated “AI News Roundup” cards: Breaking-news formats grouping several topics under a headline plus bullet points, such as Google I/O 2026, AI handling 911 calls, and North Korea’s AI stack: [21, 29, 31, 35].
- Generic AI-art showcases that drift from the main topic: Promotional pins showing individual AI artworks with no news value, such as photorealistic images of a marble sculpture playing cello or an elderly fisherman portrait: [49, 50].
- Dark backgrounds with neon text are the default palette, while white-background corporate infographics in corporate blue and flat design are a minority: [3, 10, 27, 33].
Notable pins
- [11] Veo 3 Image-to-Video: Fast Generation & Native Audio via Gemini API — A real product introduction promoting Veo 3 image-to-video generation with native audio through the Gemini API.
- [16] Google Unveils Veo 2 — A Veo 2 promotional grid in an official Google-like style, with many generation samples including a scientist at a microscope, swimming, and an anime-style girl.
- [25] Google announces Veo 3 — A demo showing a purple monster character generated consistently in multiple environments. It emphasizes character consistency with three output videos from one source image.
- [29] Today's AI News: MAY 20, 2026 — A Google I/O 2026 special listing Gemini 3.5 Flash, Gemini Omni, Gemini Spark, Apple accessibility AI, and the competition around Android XR smart glasses.
- [35] The Latest AI Developments — A three-part breaking-news card: AI answering 911 calls in New Orleans; Meta releasing its offline agent model Muse Glimmer; and North Korea’s Kimsuky group automating phishing with its own AI stack.
- [44] Microsoft Unveils VASA-1 — Introduces a model that generates talking-face video from one still image plus an audio clip, with optional control signals, alongside a grid of diverse faces.
- [48] ShengShu Technology Unveils Vidu S1 — Promotes Vidu S1 as a “state-of-the-art model for real-time AI interaction,” with multiple cyberpunk-style character faces.
- [8] Style-consistent transformation demo (Luma-related) — A five-pattern comparison transforming the same person, animal, and car into wooden-block, origami, Lego, and flower styles. It is a persuasive technical demo.
- [46] Future of Video Creation: Best AI Video Generators of 2026 — A current ranking that goes as far as version numbers: Pika 2.5, HeyGen, Google Veo 3.1, Runway Gen 4.5, Kling 3.0, Adobe Firefly, and Seedance 1.5 Pro.
- [20] More than 20% of YouTube is now AI-generated — A simple pin consisting only of a headline claiming a specific statistic, but one that symbolizes the scale of AI-generated-content adoption.
Signals
- Model-generation competition is visible: Veo 2 → 3 → 3.1, Kling AI → 3.0, Runway Gen-3 → Gen 4.5, Pika 2.5, and Seedance 1.5 Pro appear frequently, revealing the rapid improvement cycle in video-generation AI: [16, 25, 11, 46].
- “Agentization,” real-time operation, and offline operation are the next focal points: Pipeline tools that automate everything from scouting to scripting, generation, and editing with one click ([24]); real-time interaction through Vidu S1 ([48]); and Meta Muse Glimmer, which does not require an internet connection ([35]), show interest shifting beyond simple “generation” toward agents, real-time systems, and on-device AI.
- Real-world use and misuse of AI are becoming topics: Content also includes social deployment and security topics beyond generative AI itself, such as introducing AI into emergency-call systems ([35]) and state-linked AI misuse by North Korea ([35]).
- Closed-source commercial tools have dominant visibility: The tools repeatedly named in ranking pins are commercial closed services—Google Veo, OpenAI Sora, Runway, Kling AI, Adobe Firefly, Pika, Luma, HeyGen, Synthesia, and others. Mentions of open-source models and workflows—Stable Diffusion variants, AnimateDiff, Open-Sora, and ComfyUI—are nearly absent. “Amuse 3.0” ([50]), which promotes local execution on AMD Ryzen AI, is one of the few exceptions.
- Weak links to primary sources: Of the 50 pins, only a small fraction appear to point directly to company blogs or research papers, such as the official-toned Veo pins [16] and [25] and Microsoft VASA-1 [44]. Most are affiliate blogs, SEO articles, or Fiverr sales thumbnails such as [45]. Pinterest functions less as a place to follow news and more as a place to sell tools.
- Deepfake anxiety has become standard content: Detection-tool introductions and “spot the difference” quizzes form an independent genre. People are consuming anxiety about whether content can be trusted more than the technological progress itself: [4, 12, 18, 26, 47].
Limits
- This stage only reviewed images pre-collected by a worker in a headless browser. No additional Pinterest searching or scrolling was performed, in accordance with the playbook procedure.
- Collection used a single non-categorized query. Closed-source and open-source topics were not targeted separately, so open-source-related pins were almost entirely absent. This may genuinely reflect Pinterest posting patterns, though it remains possible that targeted searching would have found more.
- Two pins had the title
(untitled); [12] could effectively be identified as “Which is AI?” content, while some others were harder to assess. - The pin URLs themselves (
pinterest.com/pin/...) were not opened. Since analysis used only images and titles, linked-page content, comments, saves, and other engagement metrics could not be verified. - [49], the marble-sculpture cellist, is a generic AI-art promotional pin with little relation to the topic and was not treated as a newsworthy insight.
Recommended actions
- For Sora2-dependent workflows, create a migration plan to Kling 3.0 or Veo 3.1 before the API ends on 2026/9/24.
- Track GPT-6 Astra’s ComputerUse capabilities as a production-pipeline automation tool, not just as a standalone generative AI.
- If reducing cost is the goal, try local editing with Qwen-Image-Edit + ComfyUI before paying for closed-source services.
- Prepare for authenticity-labeling demands and backlash over fake timelapses by establishing disclosure rules for AI use in public content.
- Test the division-of-labor workflow of Nano Banana for still images and Kling 3.0 for video conversion.
- Continue monitoring Midjourney’s pivot from image generation to medical devices as a potential leading indicator of diversification by other closed-source companies.
Collected images


















































Data quality notes
All six platforms fell short of brief.md’s completion criterion of roughly 20 items per social network. X had only four relevant posts out of 40 collected; Bluesky reached effectively 11 items due to 403 rate limits; and Lemmy had only 15 items because it lacks specialist news communities. However, major trends—GPT-6 Astra, the Sora2 shutdown, and the expanding ComfyUI ecosystem—were independently supported across multiple platforms.



