Image and Video Generation AI News — 2026-09-24
Closed-source image/video AI is winning on price and speed (GPT-6 Sol/Luna, ChatGPT Images 2.5, Kling 3.0), while open-weight models are starting to claim outright wins in video quality (Wan 3.0) even as several 'open source' releases (Wan 3.0's later weights, FLUX 3, MiniMax H3's licensing) turn out to be more closed or restricted than advertised.
画像・動画生成AIニュースまとめ🔥🎨🎬 — 2026-09-24
ヤバい、今週の画像・動画生成AI界隈もめちゃカオスだったわ😂 一言でいうと「🔒クローズド勢は値下げ&スピード競争でゴリ押し、🔓オープン勢は動画生成でガチ王者になりつつある」って感じ!GPT-6 Sol/Lunaみたいな新モデルがバンバン出て安くなってる一方で、Wan 3.0みたいなオープンウェイトモデルが動画部門トップに立ったって話もあってマジで熱い🔥。でも同時に「AI画像バレて炎上」「訴訟まみれ」「オープンソース詐称」みたいなダークサイドも盛り沢山で、単純に「AIすげー」じゃ済まされない空気になってきてる😅。
Across platforms 🌐
- 🔒 値下げ・高速化競争がガチで加速中。 X発のGPT-6 Sol/Luna発表は性能アップより「安く速く」がウリで、YouTubeのChatGPT Images 2.5(生成速度-50%)やKling 3.0のAI Director(1回の生成で複数カット=コスト圧縮)とも同じ流れ。プラットフォーム跨いでみんな同じこと言ってるから、これはガチのトレンドっぽい💸
- 🔓🔒 画像は閉じてる方が強い、動画はオープンが逆転してるかも。 Blueskyの追跡bot(olud.ai)のブラインド評価だと、画像生成・編集どっちもGPT Image 2.5 Sunburst🔒が首位で、一番強いダウンロード可能モデルQwen-Image-3.0-Pro🔓は100+ELOも離されてる。でも動画部門だけはWan 3.0🔓がクローズド勢を全部抜いてトップって報告があった、ここ超注目ポイント👀
- ⚖️ 訴訟&規制の圧力がマジで強まってる。 Bluesky経由でDisney v. Midjourney、Andersen v. Stability AIの裁判が9/23に同時に動いてるのが確認できたし、Reddit r/CriticalStateの watermark 義務化・AI開示義務化の法案スレは合わせて1.8万コメント級の盛り上がり。みんなAI生成物に「これ本物?」ってキレ気味😤
- 🎭 「バレて笑われるAI」と「バレずに拡散するAI」が同時に存在してる。 Reddit(不動産写真、Nikon顕微鏡動画、大学の壁画像)やPinterest(NBCの「AI GENERATED」表示、フェイク疑惑ミーム)は総ツッコミ案件多数。なのにr/InstagramrealityではAI画像がバレずに何千アップボート集めてたケースも報告されてて、同じ空間で真逆のことが起きてる😳
- 🔓 「オープンソース」を名乗るけど実は違う問題が多発。 YouTubeで指摘されてたFLUX 3🔒(タイトルは"Open Source"だけど実態は早期アクセスのクローズド版のみ)、Wan 3.0(過去バージョンはオープンウェイトだったのにv3.0で重み非公開のAPI限定に転換)、MiniMax H3🔓(Hugging Faceで重み公開してるのに米・EU・英・韓では使用禁止のライセンス)——「オープン」の定義がガバガバになってきてるの、これマジ要注意案件⚠️
Platform by platform 📱
- Reddit — 12スレッドしか集まらず目標の20に届かず😢。内容もr/StableDiffusionとかガチ勢コミュニティはゼロで、r/mildlyinfuriating(AI不動産写真、2.6万pt)とかr/labrats(AI「盛った」顕微鏡動画がコンペ優勝→炎上)みたいな一般人の「AIバレて草」系リアクションが中心。製品ニュース自体はほぼナシ。
- X — セッションがクロアチア地域判定だったせいでExploreトレンドの大半(暗号資産、テイラー・スウィフト等)が無関係🙄。それでもGPT-6 Sol/Luna発表と、Opus 5.5がHiggsfield/After Effects/Blenderを"操作するオーケストレーター"として使われてる投稿群(例)はガッツリ拾えた。ただしオープンソース系の話題は今回マジでゼロ件🔓❌。
- YouTube — 個別動画のチャンネル名・再生数がページ取得できずゼロ😩(後述)。でも内容面は最強で、Nano Banana Pro🔒、Kling 3.0🔒、Wan 3.0🔓(重み非公開化)、LTX-2.3🔓(音声+映像同時生成)、Qwen-Image-2512🔓、FLUX 3🔒/🔓のブランドと実態のズレなど、開閉両陣営の具体的な動きを一番網羅してた。
- Bluesky — 検索API(searchPosts)が全クエリ403で弾かれたけど、アカウント発見+getAuthorFeedでゴリ押し突破💪。olud.aiのELOリーダーボード(GPT Image 2.5 Sunburst🔒首位、Wan 3.0🔓が動画首位)とai.bots.lawの訴訟速報bot、ComfyUI🔓が13.4万star達成した話も拾えて、数字つきの情報が一番濃かった。
- Lemmy — 規模は激小(ai_redditコミュニティは購読者52人)だけど、
[email protected]はガチのComfyUI/オープンウェイト勢の巣窟🔓。Sora APIサービス終了 vs Kling 30億ドル調達🔒の対比記事や、Midjourneyの超音波スキャナー事業への「Theranos化」批判みたいな経営ゴシップ系も面白かった。 - Pinterest — エンゲージメント数値(いいね・保存数)は一切取得できず、画像内容ベースの分析のみ。50件中クローズドブランド(ChatGPT、Gemini、Veo、Kling、Midjourney、Adobe Firefly等)ゴリゴリで、オープン/ローカル系を匂わせるピンは1件(AMD向けAmuse 3.0)だけ🔓😶。「AI生成って信じていいの?」系の不安訴求ピンも4件独立して出てきてた。
What to watch 👀
- Wan 3.0の重み公開の行方 — Blueskyのリーダーボードでは動画部門トップ🔓なのに、YouTubeでは「重み非公開のAPI限定ベータに転換した」と報告されてて話が矛盾してる。どっちがホントか要追跡(Bluesky投稿 / YouTube)
- Disney v. Midjourney / Andersen v. Stability AIの訴訟の進捗 — 9/23に同時に動きあり、判決次第で業界全体の学習データの扱いが変わる可能性(Bluesky/ai.bots.law)
- Sora API終了 vs Kling 30億ドル調達の対比 — コンシューマー向けAI動画がマネタイズ厳しい一方で企業向けが本命になってきてるかも(Lemmy)
- MiniMax H3のCommunity Licenseの地域制限 — 「オープンウェイト」でも米・EU・英・韓で使えないパターンが今後の標準になるか注視(YouTube)
- AI開示・watermark義務化の法案動向 — Redditで1.8万コメント級の盛り上がり、規制がどこまで進むか(r/CriticalState)
- GPT-6 Sol/Lunaの実運用での性能評価 — 発表直後は「値下げがメインで性能はそこまで」という懐疑的な反応が出てる(X)
Recommendations ✅
- Wan 3.0の重み公開状況は情報が食い違ってるから、実際にHugging Face/公式サイトで最新ステータスを確認してから記事化・意思決定すること。
- 「オープンソース」を謳うモデル(FLUX 3など)はライセンス文言と実際のダウンロード可否を必ず確認し、ブランディングを鵜呑みにしないこと。
- MiniMax H3みたいな地域制限つきオープンウェイトモデルを使う前に、対象国・地域がCommunity Licenseで除外されてないかチェックすること。
- 動画生成AIをビジネス活用するなら、コンシューマー向け無料枠より企業向け(研修・商品デモ)の方が今アツいという流れを踏まえて検討すること。
- Disney/Andersen訴訟の判決が出るタイミングを継続監視し、学習データ関連のリスク方針をアップデートしておくこと。
- Reddit・Pinterestで見えた「AIバレて炎上」パターンを踏まえ、AI生成コンテンツを公開する際は事前の開示表示を徹底すること。
Data quality 📊
- Reddit: 目標20件に対し12件のみ。しかも検索語が単一("Image and Video Generation AI News")で、r/StableDiffusion等のガチ勢サブがゼロ、一般人の「AIバレた」反応に偏った。
- X: セッションが地域(クロアチア)依存のExploreトレンドしか見れず、テーマ関連は10件中3トレンドのみ。オープンソース系の言及は実質ゼロ件🔓❌。
- YouTube: WebFetchがCookie同意ページ(consent.youtube.com、地域=HR)にリダイレクトされ続け、再生数・チャンネル名・コメントが一切取得不可。内容はWebSearchのスニペット頼み。
- Bluesky: 公式検索API(searchPosts)が全クエリで403拒否。アカウント発見経由の代替収集のため、全プラットフォーム網羅の検索結果ではない。
- Lemmy: プラットフォーム自体が超小規模(ai_redditコミュニティの購読者52人)で、動画生成AI単体の検索はヒットゼロ。20件目標に対し実質14件程度。
- Pinterest: いいね・保存数などのエンゲージメント指標が一切取得できず、画像とタイトルの内容分析のみ。50件中30件しか画像本体を開けていない。
プラットフォーム別まとめ
Reddit — Image and Video Generation AI News
Where
12 threads collected across 10 subreddits, all found via the single search term "Image and Video Generation AI News":
| Subreddit | Members | Threads collected |
|---|---|---|
| r/mildlyinfuriating | 12,590,273 | 1 |
| r/youtube | 3,460,141 | 1 |
| r/Instagramreality | 1,155,685 | 1 |
| r/LinkedInLunatics | 1,081,995 | 1 |
| r/labrats | 726,552 | 1 |
| r/generativeAI | 156,580 | 2 |
| r/Purdue | 100,756 | 1 |
| r/CriticalState | 39,323 | 2 |
| r/perchance | 30,990 | 1 |
| r/aivideomaking | 7,568 | 1 |
None of these are dedicated generative-AI-tool communities (no r/StableDiffusion, r/midjourney, r/aivideo, etc. turned up) — the sample is general-audience subreddits reacting to AI-generated content, plus two small tool-support subs (r/generativeAI, r/aivideomaking, r/perchance).
What people say
- #1 (r/labrats, 933 points, 123 comments, 2026-09-23) — Nikon's Small World in Motion competition-winning "microscopy" video turned out to be AI-generated/"enhanced," with biologically impossible cell movements. Nikon confirmed AI was used after microscopists pushed back. Top comment (u/Barkinsons, 551 pts): "AI tools have no place in scientific microscopy images because we have the obligation to reproduce our images as truthfully as possible, and not the way we would like them to look."
- #4 (r/mildlyinfuriating, 26,469 points, 714 comments, 2026-09-21) — AI-staged real-estate listing photos that alter rooms unrealistically (e.g., removing radiators, adding a pool table that couldn't exist). Top comment (u/slothboy, 5,706 pts): "Turning that dinette into a bathroom would be a five figure proposition." This is the highest-scoring thread in the whole collection.
- #5 (r/Instagramreality, 2,179 points, 348 comments, 2026-09-18) — AI-generated images racking up thousands of upvotes on Reddit itself before being caught via shifting baseboard trim and tile lines between photos. u/Habibti-Mimi81 (150 pts): "I'm obviously either too old or too dumb, because I wouldn't have known on my own that this is A I... Sad and dangerous times we live in."
- #10 and #12 (r/CriticalState) — two separate law-proposal threads pushing mandatory AI watermarking (13,835 comments, 1,256 pts, 2026-09-21) and mandatory AI-disclosure by companies (4,537 comments, 214 pts, 2026-09-18). Top comments are dominated by "I voted Yea ✅" (this looks like a poll-bot-driven sub), but the volume itself signals strong appetite for disclosure regulation.
- #2, #6, #7 (r/aivideomaking and r/generativeAI x2) — three near-duplicate "looking for a free AI video generator" threads (12 pts/36 comments, 2 pts/24 comments, 4 pts/29 comments) all posted within days of each other (Sep 19-20). Consistent answer across all three: free tiers cap out around 3-6 second clips; a repeat commenter (u/Jenna_AI, an AI-support bot) explains free tiers can't sustain temporal consistency past ~6 seconds. Tools named repeatedly: ComfyUI + FramePack (local, needs 16GB+ VRAM), Muse (Meta, free, no watermark), Grok ($10/mo), Kling/Hailuo/Pixverse, Nanobanana/Veo via Google Flow.
- #9 (r/perchance, 31 points, 21 comments, 2026-09-19) — thread asking why Perchance AI has no image-to-image mode. Top comment (u/Dack_Blick, 32 pts): "Image to image poses huge legal risks and liability." Others are blunter: u/DoctaRoboto (8 pts) says it's "because of illegal porn and deepfakes," u/Separate-Prior-5985 says "Pedos." This is a rare thread where the builders of an open/free tool explain a deliberate capability limit.
- #8 (r/Purdue, 150 points, 20 comments, 2026-09-21) — AI-generated promotional image on a university engineering building wall, mocked for a nonsensical PCB layout. u/549013 (7 pts): "half of the lecture videos at one of the best engineering colleges in the nation are ai generated like they seriously couldn't reuse older ones."
- #11 (r/LinkedInLunatics, 978 points, 116 comments, 2026-09-21) — AI-generated image of "famous people agreeing with a nobody" LinkedIn post. u/Humble_Daikon (318 pts): "I like how the AI made Musk smoking - must have been trained on all the reddit comments calling those guys a nightmare blunt rotation."
- #3 (r/youtube, 5 points, 22 comments, 2026-09-17) — complaint about YouTube's AI-generated video chapters appearing on every video. Split reaction: u/buildingduck and u/meatmobile682 agree they're intrusive and irrelevant, while u/ickN, u/lieutenatdan and u/davesaunders say they're helpful/unobtrusive.
Signals
- Rising: backlash against undisclosed AI imagery in non-hobbyist contexts. Every general-audience thread this run surfaced (Nikon science competition, real estate listings, a university building, a LinkedIn post) is critical or mocking, not celebratory — the pattern is "AI slop caught in the wild," not "look at this cool generation."
- Rising: appetite for regulation. Two r/CriticalState law-proposal threads pulled a combined ~18,000 comments pushing mandatory watermarking/disclosure, dwarfing every other thread's comment count in this set.
- Dismissed: the idea that good free AI video generation exists. Across three separate threads, the consensus answer to "is there a free option" is a hard no — free tiers are capped at a few seconds, and running locally needs hardware most posters don't have.
- Surprising / disagreement: Reddit is simultaneously a place where AI images get exposed and mocked (r/mildlyinfuriating, r/LinkedInLunatics, r/Purdue) and a place where undetected AI images rack up thousands of upvotes unnoticed (r/Instagramreality thread, #5, describing an AI post inside Reddit's own r/OUTFITS). The same platform is both the debunker and, at times, the unwitting distributor.
- Notable absence: no thread in this set is actual open-source vs. closed-source product news (no model releases, benchmarks, or version comparisons for e.g. Stable Diffusion/Flux/Sora/Veo/Kling as products) — the closest is #9's discussion of why a free tool (Perchance) withholds image-to-image, which is a capability/liability story rather than a release story.
Limits
- Only 12 threads were collected against the brief's ~20-per-platform target; the worker used a single search term, "Image and Video Generation AI News," with no variants tried (e.g., no separate searches for "Stable Diffusion," "Sora," "Midjourney," "open source video generation").
- No dedicated generative-AI-tool subreddits (r/StableDiffusion, r/midjourney, r/aivideo, r/LocalLLaMA-style communities for image/video) appear in the collected set, so closed-source vs. open-source product-specific news is thin — this collection skews toward general-audience reaction/backlash rather than practitioner discussion.
- Per the playbook, I did not browse Reddit myself (WebFetch/search are refused by Reddit); I worked only from
output/reddit.threads.mdand.jsonas collected by the worker, so I cannot say whether a broader search would have surfaced more targeted news.
X
X — 画像・動画生成AIニュース
X の Explore(クロアチアからのセッションで観測、地域依存)は「Cardano」「Opus 5.5」「$SONG」「Taylor」「Astra」「England」「#uranium」「OpenAI」「Croats」「NFTs」の10件をトレンドとして表示していた。このうちテーマ(画像・動画生成AI)に関係するのは Opus 5.5 / Astra / OpenAI の3件のみで、残り7件(暗号資産、テイラー・スウィフト関連ゴシップ、サッカー賭博の勧誘、ウラン鉱株、地域ネタ、NFTミント告知)はテーマと無関係だった。以下はこの3トレンド配下で集まった投稿の分析。
Accounts
- @OpenAI(公式, 52,200 likes / 1投稿)— GPT-6 Sol・Lunaのローンチを発表した唯一の公式アカウント。この場では圧倒的最大のエンゲージメント源。
- @npaka123(布留川英一, 283 likes / 1投稿)— 著名なAI解説アカウント。Opus 5.5への単発の検証投稿。
- @seiiiiiiiiiiru(476 likes / 1投稿)— 映像クリエイター。Opus 5.5でHiggsfield+After Effectsを操作したローンチ映像1本。
- @kevin_t_ngo(2,451 likes / 1投稿)— 単発だが今回のAI関連投稿で最大級の反応(~152,000 views)。Opus 5.5によるBlenderレンダリング。
- @aicreataro(325 likes / 1投稿)— MiniMax H3素材をOpus 5.5でAfter Effects加工する検証系1投稿。
- @sonia_code(377 likes / 1投稿)— Astraでブラウザ動作の3Dシーンを生成した単発投稿。
- @yachimat_manga(73 likes / 1投稿)— AIショートアニメ制作者。Astra時代のワークフロー考察1本。
- @mask_3dcg(1,813 likes 合計 / 2投稿)— 3DCGクリエイター。2投稿とも「AIはアニメーターの代わりにならない」という同一論調で、単発の大ヒットではなく主張を繰り返す発信者。
- @umiyuki_ai(184 likes / 1投稿)— OpenAIとClaudeの性能比較を行う単発の分析投稿。
- @so_ainsight(26 likes / 1投稿)— Claude Code運用で知られるアカウントによるGPT-6 Astra活用事例の紹介1本。
大半は「1投稿だけ拾われた」単発の発信者で、このテーマで複数投稿していたのは@mask_3dcgのみ。
Posts
-
@OpenAI — 52,200 likes・8,156 reposts・2,091 replies・約870万views・2026-09-23
https://x.com/OpenAI/status/2102460975790137662Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. [...] bringing much of its [GPT-6 Astra's] strengths into faster and more affordable models
クローズドソース最大のニュース。GPT-6 Astraの能力を高速・低価格化した派生モデル2種の公式発表。 -
@umiyuki_ai — 184 likes・29 reposts・約47,000 views・2026-09-22
https://x.com/umiyuki_ai/status/2102485745328165103OpenAIがOpus5.5に対してGPT-6SolとLunaをぶつけてきた![...] SolもLunaも5.6からあんま性能上がってないっぽい。ただし、価格は安くなってる[...] Opus5.5がFable5.1超えてきたからGPT-6SolもAstra
公式発表への反応。性能向上より値下げが本質、という懐疑的な読み。 -
@so_ainsight — 26 likes・4 reposts・約3,400 views・2026-09-23
https://x.com/so_ainsight/status/2102641613612990946GPT-6 Astra x 物件サイト→3D空間、えぐい [...] 入力はその物件の写真20枚だけ ・屋内はGPT-6 Astra(OpenAIのAI)とUnreal Engine
写真20枚だけから歩き回れる3D空間を生成する事例。GPT-6 Astraの実用ユースケース紹介。 -
@npaka123 — 283 likes・40 reposts・約31,000 views・2026-09-23
https://x.com/npaka123/status/2102594385758347548Opus 5.5 をおためし中 プロンプト: 3Dのシューティングゲームを作って。[...] うまくいけば1分ほどでクリアできる。BGMとSEもつけて
Claude Opus 5.5に1プロンプトで3Dシューティングゲーム一式(BGM・SE込み)を作らせた検証。 -
@seiiiiiiiiiiru — 476 likes・49 reposts・約27,000 views・2026-09-23
https://x.com/seiiiiiiiiiiru/status/2102636308707287201もう手作業でキーフレーム打つなんてアホらしい。これはClaude Opus 5.5で、HiggsfieldとAfterEffectsを操作してAIが作ったローンチビデオです。[...] 150円くらいでできました
Opus 5.5がHiggsfieldとAfter Effectsを直接操作してローンチ動画を制作、コストは約150円と明言。 -
@kevin_t_ngo — 2,451 likes・115 reposts・約152,000 views・2026-09-23
https://x.com/kevin_t_ngo/status/2102761406315839798Thank you @claudeai ! GIF animated in Python and rendered in Blender by Claude Opus 5.5
今回集まった投稿の中でAI関連トピックとしては最大のエンゲージメント。PythonアニメーションをBlenderでレンダリングまでOpus 5.5が担当。 -
@aicreataro — 325 likes・46 reposts・約29,000 views・2026-09-23
https://x.com/aicreataro/status/2102656273112326609Opus 5.5にAfter Effectsを操作させて、AI動画を加工するテスト。■ベース MiniMax H3 で15秒を1本 ■AEでやらせたこと ・カットの切れ目を拍に合わせて速さを調整 ・区間ごとにエフェクトを1つずつ
MiniMax H3で生成した動画をOpus 5.5がAfter Effectsで拍合わせ・エフェクト加工する、生成後工程の自動化事例。 -
@sonia_code — 377 likes・33 reposts・約19,000 views・2026-09-22
https://x.com/sonia_code/status/2102367557303021766これ、めっちゃエモくない?Astraに夕暮れの湿原を作らせてみたんだけど [...] ブラウザで動いてるから、マウスで好きな方向を見渡せる
Astraでブラウザ上で自由視点操作できる3D風景シーンを生成。 -
@yachimat_manga — 73 likes・13 reposts・約7,700 views・2026-09-21
https://x.com/yachimat_manga/status/2102181794057420889頑張って背景参照カットごとに指定するっていうのもAstra時代にはもはや最適解かも?[...] 既存のアニメの工程をパワーで再現するっていう方向がワンチャンつよいかも
Astra登場後もアニメ制作は「工程を力技で再現する」方向が強いのでは、という現場感覚の考察。 -
@mask_3dcg(2投稿, 224 likes / 約22,000 views・2026-09-22 と 1,589 likes / 約246,000 views・2026-09-23)
https://x.com/mask_3dcg/status/2102234787184537669
https://x.com/mask_3dcg/status/2102715480494752222こういうユニークなキャラクターのアニメーションはaiには本当に無理なので人間がいつまでも必要だと思います
即採用です。AI 時代に一番必要なのは アニメーターです。CGのアニメーションはAIじゃあ、決められたテンプレート以外はうまく作れません
生成AI礼賛への逆張り。2本目は22万超viewsと今回のAI関連投稿で2番目に大きく伸びており、「AIは非定型アニメーションを代替できない」論への共感がそれなりに大きいことを示す。
Signals
- 上昇: Claude Opus 5.5が「画像・動画生成そのもの」ではなく、Higgsfield/After Effects/Blender/MiniMax H3といった既存ツールを操作するオーケストレーターとして使われる投稿が集中(#4〜#7)。生成後の加工・仕上げ工程をAIに任せる使い方が目立つ。
- 上昇: 写真数枚から3Dウォークスルー空間を作る事例(#3, GPT-6 Astra + Unreal Engine)。生成AIと既存ゲームエンジンを組み合わせる方向性。
- 値下げ競争: GPT-6 Sol/Lunaの発表は「性能向上」より「高速・低価格化」というフレーミングで受け止められている(#2)。
- 却下・逆張り: プロの3DCGクリエイターがAIによるキャラクターアニメーション代替を明確に否定する投稿が高い反応を得た(#10、22万view超)。生成AI万能論への反発が一定の支持を集めている。
- 意外だった点: このセッションでXが表示したトレンド10件のうち、画像・動画生成AIに触れていたのは3件だけで、残り7件(暗号資産・テイラー・スウィフト関連ゴシップ・サッカー賭博勧誘・ウラン鉱株・地域ネタ・NFTミント告知)は完全に無関係だった。つまりこの瞬間・この地域では、画像・動画生成AIはXの「今何が起きているか」の主役ではなく、汎用の技術系トレンド語の下にわずかに滲み出ていたに過ぎない。
- オープンソースのモデルやツール(例: Flux、Stable Diffusion系、HunyuanVideo、Wan、ComfyUIなど)への言及は今回収集された投稿には一切なかった。
Limits
- 検索語はテーマから選んだものではなく、X Explore(クロアチアから到達したセッションに地域最適化された一覧)に出ていたトレンド語をそのまま使っている。10語中7語(Cardano, $SONG, Taylor, England, #uranium, Croats, NFTs)はテーマと無関係で、それらの検索結果は今回のレポートから除外した。
- Exploreのトレンドはあくまで「クロアチアから見えるこのセッション」のものであり、世界的なトレンドではない。
@videoai_otakuの投稿(#30, https://x.com/videoai_otaku/status/2102264981807153234)はテキストも画像もなく、内容を引用できなかった(動画のみの投稿とみられる)。- 「Sora」「Veo」「Midjourney」「Runway」といった具体的なクローズドソース生成ツール名や、「Flux」「Wan」「HunyuanVideo」「ComfyUI」といったオープンソース生成ツール名では検索していないため、これらに関するX上の言及は本レポートに反映されていない。
- 上記の理由により、オープンソース動向についてはXから拾える気づきがほぼゼロ(今回の収集データにオープンソースモデル・ツールへの直接言及なし)。オープンソース側の10件の気づきは他のプラットフォームのステージで補う必要がある。
YouTube
YouTube — 画像・動画生成AIニュース
Channels
このステージでは個別動画のチャンネル名・登録者数がページから直接取得できなかった(下記Limits参照)ため、検索結果のタイトル・説明文から確認できた発信者と、このジャンルを継続的にカバーしていると一般的に知られるチャンネルを分けて記載する。
- OpenAI(公式チャンネル)— ChatGPT Images 2.5のローンチ動画「Introducing ChatGPT Images 2.5」を公開。企業公式チャンネルとして今回最大の一次情報源。
- Muapi(チュートリアル系チャンネル)— 「GPT Image 2.5 API (ChatGPT Images 2.5) — What's New + How to Access It」でAPI利用者向けに新モデル(Flare / Sunburst)の使い方を解説。
- 検索結果には他に「Kling 3.0 and Omni FULL guide」「Seedance 2.5 vs Minimax H3 vs Google Omni vs Kling 3.0」「China Did It AGAIN? – 100+ Wan 3.0 AI Videos」「LTX-2.3 Review」「New Qwen Image 2512」などレビュー・比較系動画が多数ヒットしたが、チャンネル名はページ取得不可のため個別に確認できなかった。
- 一般的な補足情報として、この分野(AI画像・動画生成のニュースとレビュー)を継続的に扱うチャンネルとして Curious Refuge(AI映像制作・ワークフロー教育に特化)、Matt Wolfe(毎週のAIニュースまとめ)、The AIGRID(生成AIの速報系)が web 検索で確認できたが、これらのチャンネルが今回ヒットした個別動画の投稿元であるとは確認できていない。
Videos
-
Introducing ChatGPT Images 2.5 — OpenAI(公式) — 2026-09-08ごろ — 再生数不明
https://www.youtube.com/watch?v=6l7ble9P74o
ChatGPT Images 2.5(新モデル名 GPT-Image-2.5 Flare / Sunburst)のローンチ公式動画。生成速度を最大50%短縮、編集の一貫性向上、コメントベースの編集機能などを紹介。 -
GPT Image 2.5 API (ChatGPT Images 2.5) — What's New + How to Access It — Muapi — 2026年9月上旬 — 再生数不明
https://www.youtube.com/watch?v=nk2u6ENW85Q
API経由でのFlare/Sunburstモデルの違いと使い方を解説するチュートリアル。Sunburstは事前計算(プランニング)機能を追加した上位版と説明。 -
Nano Banana Pro: Hands-on with the World's Most Powerful Image Model — 2025-11-25 — 再生数不明
https://www.youtube.com/watch?v=hk6gwiZmSWA
Google DeepMindのGemini 3 Proを基盤とするNano Banana Proのハンズオン。テキストレンダリングやインフォグラフィック生成の精度向上を実演。公開から10ヶ月近く経つが、後続の関連動画が継続的に出ており息の長いトピック。 -
Kling 3.0 and Omni FULL guide 2026 — 2026-04-04 — 再生数不明
https://www.youtube.com/watch?v=tEdHJohIUlM
快手(Kuaishou)のKling 3.0/Kling 3.0 Omniの総合ガイド。「AI Director」機能で1回の生成につき最大6カットの一連シーケンスを作れる点を解説。 -
Seedance 2.5 vs Minimax H3 vs Google Omni vs Kling 3.0 | Full Comparison — 2026-08-14 — 再生数不明
https://www.youtube.com/watch?v=G5D053drKB8
クローズド系トップモデル4つ(ByteDance Seedance 2.5、MiniMax H3、Google Omni、Kling 3.0)の横並び比較。2026年8月時点でのクローズドソース動画生成AIの勢力図を示す代表的な比較動画。 -
FREE Wan 3.0: what actually shipped (there are no weights) and How to use it for FREE
https://www.youtube.com/watch?v=gbI8XV4Szrk
AlibabaのWan 3.0が「重み(weights)非公開のAPI限定ベータ」として登場したことを指摘する動画。Wan 2.2まではオープンウェイトだったが、Wan 3.0でクローズド路線に転換したことを解説(一般web検索でも同様の指摘を複数確認)。 -
China Did It AGAIN? – 100+ Wan 3.0 AI Videos — 2026-08-10 — 再生数不明
https://www.youtube.com/watch?v=fNvv2u-1cGw
Wan 3.0のパブリックベータ(2026年8月6日開始)で作った100本超の生成動画をまとめたショーケース。最大30秒のワンショット生成(Wan 2.7の15秒から倍増)を強調。 -
LTX-2.3 Review: The Open-Source AI Video Model Built for Creators — 2026-08-11 — 再生数不明
https://www.youtube.com/watch?v=qQXzlk134Sw
Lightricks製LTX-2.3のレビュー。オープンソース動画生成モデルの中で唯一、音声と映像を単一の推論パスで同時生成できる点(4K/50fps対応)を評価。 -
LTX 2 Full Tutorial | Run on Low VRAM PC + Video & Audio AI Generation (2026 Strongest Open-Source?) — 2026-01-11 — 再生数不明
https://www.youtube.com/watch?v=ytxbq9ZSgxo
低VRAM環境でのLTX-2ローカル実行チュートリアル。ComfyUI/SwarmUI経由でのセットアップを解説し、オープンソース動画生成モデルの中でも扱いやすさをアピール。 -
New Qwen Image 2512: Better Than Z-Image? (Open Source & Free)
https://www.youtube.com/watch?v=uAGH_yRZ2Gw
Alibaba Tongyi ChatチームのQwen-Image-2512(Apache 2.0ライセンス、200億パラメータ)とZ-Image Turboを比較。オープンソース画像生成のトップ争いとして複数のレビュー動画が同時多発的にヒットした。 -
Can MiniMax H3 Beat Seedance 2.5? My ComfyUI Workflow
https://www.youtube.com/watch?v=zwam3nJTNrI
2026年8月3日にHugging Faceで公開されたMiniMax H3(330億パラメータ、Community License)をComfyUIワークフローで検証。ただしこのライセンスは米国・EU・英国・韓国での利用を明示的に除外しており、その点への言及があるかは本動画のみでは確認できず。 -
FLUX 3 Might Be the Open Source Sora 2 We Deserve — 2026-07-24 — 再生数不明
https://www.youtube.com/watch?v=1s3zslFOSh0
Black Forest LabsのFLUX 3(動画・音声・画像・ロボティクスに対応するマルチモーダルモデル)を紹介。タイトルは「オープンソース」を謳うが、2026年7月時点でFLUX 3自体はクローズドの早期アクセス版のみで、オープンウェイト版(FLUX 3 Dev)はロードマップ表明のみで未公開 — 「オープンソースになる予定」と「すでにオープンソース」を混同した紹介動画が多い点は要注意。
Signals
- クローズドソース陣営の値下げ・高速化競争: ChatGPT Images 2.5(生成速度-50%)、Kling 3.0のAI Director(1回の生成で複数カットを生成しコストを圧縮)など、性能向上よりも「速く・安く」を前面に出す発表が目立つ。X(
output/x.md)で観測されたGPT-6 Sol/Lunaの値下げフレーミングと同じ傾向がYouTubeのレビュー動画群でも一貫している。 - 「オープンソースを名乗る/名乗らせたがる」動きと実態のズレ: Wan 3.0はAlibabaの過去モデル(Wan 2.2まで)がオープンウェイトだったにもかかわらずAPI限定・重み非公開に転換し、複数の解説動画が「これは実質クローズドだ」と指摘。逆にFLUX 3は動画タイトルで「Open Source」を冠しつつ実態はクローズドの早期アクセス版のみ、というブランドと実態の乖離が両モデルで対照的に見られる。
- オープンウェイトでも「使える地域」が限定される新パターン: MiniMax H3はHugging Faceで重みを公開した稀な大型動画生成モデルだが、コミュニティライセンスが米国・EU・英国・韓国でのローカル実行を明示的に除外。「オープンウェイト=誰でも自由に使える」という前提が崩れつつある事例。
- 比較・横並びレビューがジャンルの主流フォーマット: 「Seedance 2.5 vs MiniMax H3 vs Wan 3.0」のような多モデル比較動画が非常に多く、単独モデルの深掘りよりも「どれが一番強いか」を検証する動画がこのジャンルのYouTube上での主要フォーマットになっている。
- 中国発モデルの存在感: Kling(快手)、Seedance(ByteDance)、Wan(Alibaba)、MiniMax H3、Qwen-Image(Alibaba Tongyi)と、動画・画像生成AIの話題の中心に中国発モデルが多数を占め、米国発(GPT Image、Nano Banana Pro、FLUX)と並ぶ二極構造になっている。
Limits
- このステージで使用したWebFetchツールは、YouTubeの検索結果ページ(
/results?search_query=…)および動画ページ(/watch?v=…)の両方で、実データではなくフッターのナビゲーションリンク(コピーライト表記など)しか取得できなかった。動画ページを/@channel/aboutで試した際はconsent.youtube.comへの302リダイレクト(gl=HR=クロアチア地域のCookie同意ページ)が確認され、このセッションの地域・Cookie同意状態がJS未実行のHTML取得を妨げていると見られる。X(output/x.md)のExploreトレンドが同じくクロアチア地域だったことと符合する。 - 上記の理由により、再生回数・チャンネル登録者数・具体的なコメントは今回一件も直接確認できなかった。動画のタイトル・公開時期・内容はWebSearchのスニペットおよび外部メディア記事から裏付けたもので、プレイブックが求める「チャンネル登録者数つきのChannelsセクション」「動画ごとの再生数」は満たせていない。
- 個別動画の投稿チャンネル名も、検索結果のタイトルからは判別できないケースが大半だった(レビュー動画の多くがチャンネル名をタイトルに含めていない)。
- コメント欄の内容は一切確認できていない(プレイブックの「上位コメントも読む」を満たせず)。
- 検索は英語キーワード(Kling 3.0 / Nano Banana Pro / Wan 3.0 / LTX-2 / Qwen-Image / Flux 3 / Seedance 2.5 / MiniMax H3 / GPT Image 2.5)を中心に行った。日本語チャンネル(日本語話者によるAIツール解説動画)は検索していないため、日本語圏のYouTube上の反応はこのレポートに含まれていない。
- 完了基準(「約20件の投稿を解析」)については、YouTubeはX/Redditのような「投稿」単位ではなく動画単位のプラットフォームであるため、12件の動画(うち再生数等の定量データが取得できたものはゼロ)をもって打ち切った。これ以上検索を広げても、WebFetchの技術的制約により定量データが取得できる見込みは低いと判断した。
Bluesky
Bluesky — Image and Video Generation AI News
Accounts
- oludai.bsky.social ("🚀 olud.ai") — a bot (1,543 posts, 948 followers) that posts hourly AI tracking updates: new model launches, pricing changes, GitHub star movements, and — most relevant here — a recurring "media leaderboard" comparing image/video generation models on blind human preference.
- ai.bots.law — a legal-filings bot that posts a new entry every time there's docket activity in AI copyright lawsuits, including Disney v. Midjourney and Andersen v. Stability AI.
- midjourneyofficial.bsky.social — Midjourney's official account. Low-frequency and casual ("How was your weekend?"); not a source of product news.
- Japanese AI-illustration creators posting daily Stable Diffusion/SDXL output with hashtags like #AIイラスト #StableDiffusion: kazanomiya.bsky.social, michi84.bsky.social, garakuta76.bsky.social, blackmagnet3400.bsky.social.
- Midjourney hobbyist art community: johndoesmidjourney.bsky.social, dalyrceri.bsky.social, ttaships.airminded.org, un1v3rse.bsky.social, nevyn79.bsky.social, stumblegirl.bsky.social — daily output, not news, but shows the platform's everyday AI-art usage skews toward mature tools.
- index.photoshoproadmap.com.ap.brid.gy — a fediverse-bridged (Bridgy Fed) account discussing practical Photoshop/Nano Banana workflow issues.
Posts
- oludai.bsky.social, 2026-09-23 (1 like) — "🎨 Where open-source stands in image generation: #14 Qwen-Image-3.0-Pro — the best model you can actually download. 108 ELO points behind GPT Image 2.5 Sunburst (max) (OpenAI). Blind human preference, not marketing." → post
- oludai.bsky.social, 2026-09-22 (2 likes) — "🏆 An open-weight model leads video generation. Wan 3.0 tops the human-preference ranking — ahead of every closed model. You can download it and run it yourself." → post
- oludai.bsky.social, 2026-09-22 (0 likes) — "🎨 Where open-source stands in image editing: #16 Qwen-Image-3.0-Pro — the best model you can actually download. 101 ELO points behind GPT Image 2.5 Sunburst (max) (OpenAI)." → post
- oludai.bsky.social, 2026-09-21 (1 like) — "🚀 ComfyUI v0.37.0 is out … ⭐ 134,258 stars. All open-source AI releases →olud.ai/releases.html" → post
- ai.bots.law, 2026-09-23 (0 likes) — "New filing: Disney v. Midjourney (Plaintiffs sue Midjourney over training and generating copyrighted characters). Doc #206: (IN CHAMBERS) ORDER RE HEARING ON MOTION FOR JUDGMENT ON THE PLEADINGS." → post
- ai.bots.law, 2026-09-23 (0 likes) — "New filing: Andersen v. Stability AI (Artists sue over AI image training). Doc #759: Miscellaneous Relief." → post
- index.photoshoproadmap.com.ap.brid.gy, 2026-09-24 (0 likes) — "Generative Fill with Nano Banana on a large canvas often produces stretched or squashed results with visible selection edges. The problem is a resolution mismatch — the model generates into a small area but gets scaled up to fit a much larger document." → post
- kazanomiya.bsky.social, 2026-09-23 (40 likes, 5 reposts) — 「おはよう」#AIイラスト #AIart #StableDiffusion #SDXL — highest-engagement post found in the general AI-art feed, illustrating routine daily SD/SDXL use in the Japanese community. → post
- midjourneyofficial.bsky.social, 2025-08-15 (41 likes, 4 reposts) — "#Midjourney is now on BlueSky! Who are the best creators to follow here?" — the account's launch post, still its most-engaged. → post
Signals
- Closed source still wins on raw quality, in both directions of image work. olud.ai's blind-preference leaderboard has "GPT Image 2.5 Sunburst" (OpenAI) on top for both image generation and image editing, with the best downloadable open model (Qwen-Image-3.0-Pro) trailing by 100+ ELO points in each category (#14 in gen, #16 in editing).
- Video generation inverts that pattern. The same tracker reports an open-weight model, Wan 3.0, leading the human-preference video ranking ahead of every closed competitor — the one category where open source is currently reported as #1, not catching up.
- Legal pressure on closed/semi-open incumbents kept moving this week. Fresh docket activity landed the same day (Sept 23) in both Disney v. Midjourney (training + generating copyrighted characters) and Andersen v. Stability AI (artists suing over image-training data) — copyright-training liability remains squarely a story about the vendors that shipped hosted/commercial products.
- Open-source tooling layer keeps shipping fast even where model quality trails. ComfyUI, the dominant node-based UI for running open diffusion/video models locally, pushed v0.37.0 this week and sits at 134k+ GitHub stars — infrastructure momentum is strong independent of leaderboard position.
- The organic, everyday Bluesky AI-art community skews toward mature consumer tools, not frontier releases. The most active posters found (mostly Japanese "AIイラスト" creators and Midjourney hobbyists) are using Stable Diffusion/SDXL and Midjourney for daily output; no organic chatter surfaced naming Sora, Veo, Kling, Runway, or Grok Imagine specifically.
- Nano Banana (Google's image model) has reached the "troubleshooting it in production" stage, not just novelty demos — the Photoshop-workflow post is about diagnosing a specific Generative Fill failure mode (resolution mismatch causing stretched output), which reads as mainstream integration rather than early hype.
Limits
- The playbook's documented public search endpoint,
app.bsky.feed.searchPosts, returned HTTP 403 Forbidden on every query attempted (single keywords, multi-word phrases, with and withoutsort=latest) — this fetch client could not perform literal keyword search of Bluesky's post index at all. Other public API actions on the same host (getProfile,getPostThread,getAuthorFeed,getFeed) worked normally, so the block appears specific to the search action. - Worked around this by discovering candidate accounts/feeds via web search, then pulling their real content and engagement numbers through
getAuthorFeed/getPostThread/getFeed(a public "AI art" custom feed generator). This is a narrower, discovery-biased sample, not an exhaustive search — so "20 posts on the topic" here means 20 posts from the reachable sample, not from a full-platform search. - No Bluesky-native discussion specifically naming Sora, Sora 2, Veo, Kling, Runway Gen-4, Grok Imagine, or Seedream turned up in the reachable sample. Given the search-API block, this is more likely a sampling gap than genuine platform silence — treat their absence here as unconfirmed rather than as a finding.
bsky.appitself is a client-rendered single-page app; fetching its web pages directly returns no post content, so all "reading" in this stage was done against the JSON API rather than the rendered site.
Lemmy
Lemmy — 画像生成・動画生成AIの動き
Lemmyは規模が小さいフェディバースSNSだけど、[email protected] を中心にローカルAI勢がガチで話してて、lemmy.durstig.online の ai_reddit コミュニティ(Reddit r/ArtificialIntelligence のミラー/ブリッジ)経由でクローズドソース系のニュースも流れてくる。API検索 (https://lemmy.world/api/v3/search) で "Stable Diffusion" "Midjourney" "Sora" "Flux" "ComfyUI" "Kling AI" "Nano Banana" "Black Forest Labs" "Wan 2.2" "Grok Imagine" などのキーワードを横断的に調べた。
Communities
[email protected]— 5,710 subscribers。ComfyUI・オープンウェイトモデルのニュース/ツール投稿が集まるメインハブ。[email protected]— AI生成アート作品の投稿場所(ニュース性は薄い)。[email protected]— アニメ系AI生成アート専用。[email protected]— 抽象画AI生成アート専用。[email protected]— 52 subscribers(ローカル1人)、投稿25.8K件・月間155ユーザー。r/ArtificialIntelligence の投稿をブリッジしてくるコミュニティで、Nano Banana、Seedream、Sora/Kling系のニュース投稿が多い。!techtakes(dbzer0系) — AI業界批判寄りのコミュニティ。Midjourneyネタあり。!technology— 一般テック系コミュニティ、Midjourney関連の大型ニュースがヒット。!tech— Black Forest Labs関連ニュースがヒット。
Posts
-
Qwen-Image-2.1 in ComfyUI: Open-Weight Image Generation and Editing, Now with Transparency(オープンソース)
[email protected]/ 13 upvotes / 2026-09-22
https://lemmy.dbzer0.com/post/75874773 -
chanon/comfyui-obvpm-timeline: ComfyUI timeline and clip extension nodes for Minimax H3(オープンソース、動画編集ワークフロー)
[email protected]/ 6 upvotes / 2026-09-22
https://lemmy.dbzer0.com/post/75874770 -
SparknightLLC/ComfyUI-NodeSnapshots(ComfyUIフロントエンド高速化、2-3倍FPS向上)
[email protected]/ 3 upvotes / 2026-09-22
https://lemmy.dbzer0.com/post/75874771 -
SupraLabs/Supra2-IMG — 100Mパラメータの超軽量text-to-imageモデル(オープンソース)
[email protected]/ 7 upvotes / 約2026-09-23 -
The AI Horde has a new Interface, a new Image generation frontend, new backend, and all new documentation!(オープンソースの分散画像生成プロジェクトAI Horde刷新)
[email protected]/ 12 upvotes(3 downvotes)/ 約2026-09-20 -
OpenAI is shutting down Sora's API next week while a chinese competitor just raised $3 billion. that's the whole story of ai video right now(クローズド vs 中国勢の対比)
[email protected]/ 1 upvote / 2026-09-20
https://lemmy.durstig.online/post/61333
→ OpenAIがSora APIを9/24に停止する一方、Klingが30億ドル調達。コンシューマー向けAI動画は儲からず、企業向け(研修・商品デモ動画)が本命という分析。 -
Same Berserk spread, 4 AI colorizations. Looks like GPT wins again?(Nano Banana Proの性能比較)
[email protected]/ 1 upvote / 2026-09-14
https://lemmy.durstig.online/post/60037 -
Black Forest Labs launches FLUX 3 capable of generating images and 20-second video with audio — but in limited release to start(クローズド寄りの限定リリース)
!tech/ 3 upvotes(2 down) / 2026-07-24
https://bfl.ai/blog/flux-3 -
Midjourney wants Hollywood studios to reveal the details of their AI usage(クローズドソース、著作権紛争絡み)
!technology/ 75 upvotes / 2026-07-05
https://techcrunch.com/2026/07/04/ -
Elon Musk maakt complete AI-film met zijn versie van The Odyssey(Grok Imagineでオデュッセイア全編AI映画化宣言、クローズドソース)
!films/ 1 upvote / 2026-07-22
https://lemy.nl/post/4342454 -
Seedream 5.0 Pro is here: an honest comparison and technical breakdown(Nano Banana Proとの比較ベンチマーク)
[email protected]/ 1 upvote / 2026-07-10
https://lemmy.durstig.online/post/46063 -
black-forest-labs/FLUX.2-klein-9b-kv-fp8(オープンウェイト公開)
!stable_diffusion/ 8 upvotes / 2026-03-14
https://huggingface.co/black-forest-labs/FLUX.2-klein-9b-kv-fp8 -
Midjourney AI pivots to Theranos: Ultrasonic CT(Midjourneyが超音波スキャナー事業に手を出したことへの批判記事)
!techtakes/ 41 upvotes / 2026-06-19
https://pivot-to-ai.com/2026/06/19/ -
Google quietly discontinues its Earth AI feature a day after its rollout(Nano Banana統合機能が悪用され即撤回)
[email protected]/ 1 upvote / 2026-08-07
https://lemmy.durstig.online/post/52162
Signals
- オープンソース側の主戦場はComfyUI周辺のツール/ノード開発:新モデル発表そのものより、既存モデル(Qwen-Image、Minimax H3、FLUXなど)を動かすワークフロー・高速化ツールの投稿が多い。地味だけど継続的に動いている。
- クローズドソース勢は「金の話」と「炎上」がセット:Sora API終了 vs Kling 30億ドル調達、Midjourneyの超音波スキャナー事業への「Theranos化」批判など、技術そのものより経営・資本・倫理面のニュースが目立つ。
- Nano Banana (Google) がクローズド画像生成の比較対象として定番化:Seedream 5.0やGPTの色付けと比較されるベンチマーク的存在になっている一方、Google Earthでの悪用騒動など負の面も報じられている。
- Black Forest Labs (FLUX) はオープン/クローズドの中間戦略:FLUX.2-klein-9bのようなオープンウェイト小型モデルを出す一方、FLUX 3(画像+20秒音声付き動画)は「limited release」というクローズド寄りの出し方。
- Lemmy独自の切り口はほぼゼロ:見つかったニュース系投稿の多くはRedditやTechCrunch、HuggingFace、bfl.aiなど外部ソースへのリンク投稿で、Lemmyユーザー独自の一次情報・分析はほとんどない。純粋な創作物(アート作品投稿)は多いが、これは「ニュース」ではない。
Limits
- Lemmyは規模が非常に小さい(
ai_redditコミュニティでさえ購読者52人、ローカルユーザー1人)。動画生成AI関連の検索("video generation AI"単体、"Runway AI"、"Kling AI"、"Wan 2.2")はヒットゼロだった。 - 20件の投稿解析という完了基準に対し、実際に見つかった「ニュース性のある」投稿は14件程度。他はComfyUI連携ツールのような周辺的な投稿や、AI生成アート作品(
stable_diffusion_art、stable_diffusion_abstract、share_anime_art)で、テーマ本体(業界ニュース)とは言えないため除外した。無理に20件に water down するより、実在する投稿だけを報告する。 lemmy.mlとlemm.eeの検索は明示的に試みたが、lemmy.worldのAPI検索で得られる連合結果と重複が多く、追加の一次情報は見つからなかった。- WebFetchでの取得内容はAI要約経由のため、スコアや日付が相対表記(「1 day ago」等)から絶対日付への変換に若干のズレが生じている可能性がある。
Pinterest — Image and Video Generation AI News
50 pins collected for the query "Image and Video Generation AI News" (single ungenred set). 30 of the 50 images were opened and read directly; titles for the rest were read from pinterest.pins.md.
Visual themes
- YouTube-thumbnail "AI NEWS" format: a bearded white man with a shocked/wide-eyed expression next to a row of glowing 3D tech-brand logos (OpenAI, Google, Microsoft, Meta, Anthropic-style burst icon, Notion). This exact template repeats with different logo sets [1, 27, 48] — it's clearly a recurring news-recap channel's thumbnail style, not organic user content.
- Cyborg / half-human-half-robot face: a face split down the middle, human on one side and chrome/circuit robot on the other, used as generic "AI is here" imagery [32, 41, 47]. Also appears as a full chrome android portrait with glowing blue eyes [12].
- Neon cyberpunk UI mockups: dark navy background, pink/purple/blue glowing rounded boxes, a central "AI" speech-bubble icon branching to "Generate Image / Music / Video / Text / Link" icons. The identical Adobe Stock watermarked graphic appears twice under different pin IDs [5, 16] — stock imagery being reused/re-pinned rather than fresh content.
- Deepfake/misinformation anxiety: a chrome robot holding a phone next to a cat photo stamped "FAKE" with like/view counters [3]; an NBC News-branded clip of an "AI GENERATED" reporter standing in snow [33]; a "Human Journalism vs Machine-Generated Media" split-screen desk graphic [22]; a plain "AI GENERATED NEWS CLIP" caption over a talking-head video still [50]. Four separate pins independently push the same "can you trust what you're watching" anxiety.
- Tool comparison/decision charts: dense infographic grids listing closed tools side by side — ChatGPT Plus, Gemini Advanced, Midjourney, Adobe Firefly, Recraft, Runway Gen-3, Luma AI, Pika Pro, Kaiber AI, Synthesia on the "paid/best quality" side vs. Ideogram, Gemini Free, SeaArt AI, Playground AI, Krea AI, Canva AI, CapCut AI on the "free" side [9]. A second infographic recaps a "Google I/O 2026" news day (Gemini 3.5 Flash, Gemini Omni, Android XR smart glasses) [30].
- Style-transfer/consistency demos: a Luma-branded grid turning one source video of a person into wooden-block, origami, Lego-brick and flower-covered versions, plus a car turned into different material styles — showcasing image-to-image style consistency across frames [2].
- Meme format: a "144p vs 4K" reaction-face meme used to joke about AI upscaling/generation quality jumps [39] — the one clearly humor-driven, non-promotional pin in the sample.
- Yellow/orange "cheerful SaaS ad" palette: friendly cartoon robot mascots pitching "AI Video Generation — turn your ideas into stunning videos in minutes" with feature icon rows (script-to-video, AI voice, subtitles, editor) [8] — a different register from the moody neon/cyberpunk pins, aimed at small-business/creator marketing.
Notable pins
- [1] "AI News: Anthropic Leak Shows Us The Future of AI..." — the recurring shocked-reactor thumbnail template that shows up three times in this set; a strong signal that AI-news recap channels are actively pinning to Pinterest for traffic.
- [2] Luma-branded style-transfer grid — the clearest concrete demo of current image/video-to-image consistency capability (turning one video into wooden-block/origami/Lego/floral versions of the same subject and pose).
- [9] "Best AI Video Generators in 2026: Veo, Kling..." decision chart — the single densest list of named closed-source tools (16+ brands) in the whole set, useful as a checklist of what's considered current.
- [13] "Veo 3 Image-to-Video: Fast Generation & Native Audio via Gemini API" — names a specific recent capability (native audio in image-to-video) rather than generic hype.
- [25] "Google's Gemini Omni Turns Anything Into Video" — Osiz Technologies-branded post specifically about the Gemini Omni multimodal push, echoed independently in the Google I/O recap [30].
- [33] NBC News "AI GENERATED" reporter clip — a real broadcaster's on-screen disclosure label used as the pin's whole hook, showing the trust/labeling debate has crossed into mainstream news branding.
- [37] "Introducing Luma Dream Machine - Next Generation AI Video" — an extreme macro eye shot as the launch visual for a still-referenced video model.
- [39] "144p vs 4K" meme — the only openly humorous pin, riffing on generation/upscale quality jumps rather than promoting a tool.
- [40] "'Amuse 3.0', an AI art creation tool..." (AMD-branded) — effectively the only pin in the set with any open/local-tooling angle (an AMD GPU-oriented art app), everything else skews closed/cloud SaaS.
- [49] "7 Best AI Baby Generators: Predict Your Child's Face (2026)" — a novelty face-generation use case (age/relationship morphing) distinct from the news/video-tool cluster, suggesting a separate consumer-fun sub-trend.
Signals
- Closed-source, brand-name tools dominate the visual language. Nearly every titled pin names a specific commercial product (ChatGPT, Gemini/Gemini Omni, Veo 3, Kling, Midjourney, Adobe Firefly, Runway, Luma, Pika, Synthesia, Recraft, Ideogram, Krea, Canva). Only one pin in the full 50-title list ([40], Amuse 3.0/AMD) gestures at anything open or locally-run — no Stable Diffusion, ComfyUI, Flux, Wan, or HunyuanVideo branding appeared anywhere in the titles or the 30 images opened. On Pinterest, "AI image/video generation" currently reads as a closed-SaaS consumer category, not a developer/open-source one.
- Trust and authenticity anxiety is a real, recurring sub-theme, not a one-off: four independent pins [3, 22, 33, 50] build their entire visual around "is this real or AI," including one real broadcaster (NBC) using it as an on-air disclosure graphic. This tracks with the AI-news-anchor pin [47] showing newsroom automation as a headline topic in its own right.
- Two distinct visual registers compete for the same query: cold cyberpunk/neon "AI is here" imagery (androids, glowing UI mockups) versus warm cartoon-mascot SaaS-ad imagery (friendly robots, yellow backgrounds). Both are selling tools, but to visibly different audiences (enthusiast/tech vs. small-business/creator).
- Recap-channel thumbnails are a Pinterest-native genre of their own — the same shocked-face-plus-logo-grid template recurs identically in style [1, 27, 48], meaning a chunk of "AI news" content on Pinterest is downstream reposts of YouTube thumbnails rather than native Pinterest content.
- Absent: no pins in the sample surfaced pricing backlash, lawsuits/legal fights over training data, or artist-community protest imagery — despite that being a common theme on other platforms for this brief. The closest is the ByteDance/Hollywood copyright-deal news pin [7], which frames IP protection as a resolved deal rather than a controversy.
Limits
- The pin table has no genre split (one flat list of 50), so this report is not organized by genre as the playbook's genre-branch would call for.
- 30 of the 50 collected images were opened and described directly; the remaining 20 are represented only by their pin title from
pinterest.pins.md, not by their image content, so some visual patterns among those 20 may be under-counted. - Several pins are duplicated/near-duplicate stock assets under different pin IDs (e.g. [5] and [16] are the identical Adobe Stock graphic), which inflates the apparent frequency of the "neon AI icon branch" motif slightly.
- No pin metadata included engagement numbers (likes/saves/comments), so popularity/virality could not be assessed — only content and recency of the pinned material.
- This is Pinterest only, per the assigned stage; Reddit, X, YouTube, Bluesky and Lemmy are covered in their own stage outputs.
推奨アクション
- Verify Wan 3.0's actual weight-release status directly with Alibaba/Hugging Face before citing it as open, since Bluesky and YouTube sources disagree.
- Check license terms, not just marketing claims, before calling a model 'open source' (e.g. FLUX 3, MiniMax H3's geo-restricted Community License).
- Confirm MiniMax H3 deployment region is not excluded (US/EU/UK/KR) under its Community License before using it commercially.
- Favor enterprise/commercial use cases over consumer-facing free tiers for AI video, given Sora's API shutdown and Kling's $3B raise pointing that direction.
- Track the Disney v. Midjourney and Andersen v. Stability AI dockets, as rulings could reshape industry training-data practices.
- Add explicit AI-disclosure labeling to generated content given the repeated backlash pattern on Reddit and Pinterest when undisclosed AI content is discovered.
収集画像


















































データ品質メモ
Reddit (12/20 threads) and Lemmy (~14/20 posts) fell short of the per-platform target and skewed toward general reaction rather than product news; X's region-locked trends yielded almost no open-source signal; YouTube and Pinterest could not retrieve any engagement metrics (views, subscribers, likes) due to fetch/consent-page restrictions; Bluesky's search API was blocked (403) and relied on account discovery instead of full-platform search.



