AI Articles / Weekly AI News

This Week's Personal AI News:
An AGI Declaration, and TRELLIS.2 Landing in ComfyUI Core
(Aug 31 - Sep 6)
今週の私的なAIニュース:
AGI宣言と、TRELLIS.2がComfyUI本体に入った話
(08-31〜09-06)

curated 2026-09 · sources 2026-08-31 → 2026-09-06 · 週次

A big story and a small one arrived the same week.
The big one is OpenAI's AGI declaration, where the claim and its verification diverge cleanly.
The small one — larger, if you measure by what reaches your own machine — is TRELLIS.2 being merged into ComfyUI core.
Also covered: real-time video generation, and research on models controlling their own attention.
Two replies came back from hatori this week and are carried here as notes.

大きい話と小さい話が同じ週に来ました。
大きいほうはOpenAIのAGI宣言で、これは主張とその検証がきれいに食い違っています。
小さいほう、しかし手元にとっては大きいほうが、TRELLIS.2のComfyUI本体への統合です。
ほかにリアルタイム動画生成と、モデルが自分の注意を制御するという研究を取り上げます。
今回はhatoriさんからの返信が2件届いたので、記事の注記として反映しました。

"Welcome to the AGI Era" — and the Benchmark That Disagreed 「AGIの時代へようこそ」と、それを否定したベンチマーク

OpenAI released GPT-6 Astra on September 3, 2026, and company president Greg Brockman told a press call, "Welcome to the AGI era."
The official introduction video is titled "the most intelligent and aligned model in the world."
The declaration was contradicted, however, by the very benchmark cited to support it.
ARC Prize, which built the ARC-AGI benchmark behind the headline 99.9% score, ran the same model on its own provider-neutral harness and got 62.7% — and stated that it is not claiming AGI.
A figure moving by more than thirty points says something prior to which number is right: without published evaluation conditions, a comparison means nothing.
Sam Altman himself has repeatedly called the term AGI poorly defined, so even internally the usage isn't settled.
Shortly after launch, paying subscribers were locked out and the company spent the following day apologizing on X. What a working artist should take from this is not whether AGI has arrived.
It is the fact that the number a vendor publishes and the number a third party gets from the same model can diverge this far.
Benchmark leaderboards will come up more and more when choosing tools, and without the habit of asking whose harness produced the figure, a leaderboard becomes indistinguishable from an advertisement.

OpenAIが2026年9月3日にGPT-6 Astraを公開し、同社プレジデントのGreg Brockmanが記者向けの電話会見で「AGIの時代へようこそ」と述べました。
公式の紹介動画も「世界で最も知的でアラインされたモデル」と題されています。
ところがこの宣言は、その根拠とされたベンチマークの側から否定されています。
見出しになった99.9%というスコアの元になったARC-AGIを作ったARC Prizeが、同じモデルを自前の提供元中立な評価環境で走らせたところ62.7%であり、AGIだとは主張していない、と表明しました。
数字が3割以上動くというのは、どちらが正しいかという以前に、評価の条件が公開されていなければ比較の意味がないということです。
さらにSam Altman自身は以前からAGIという語を「定義が雑だ」と繰り返しており、社内でも用語の扱いが揃っていません。
公開直後には有料契約者がログインできなくなり、翌日にXで謝罪する事態にもなりました。
制作者の立場でこの一件から持ち帰るべきなのは、AGIが来たかどうかではありません。
提供元が出す数字と、第三者が同じモデルを回した数字が、これだけずれうるという事実のほうです。
ツールを選ぶときにベンチマークの順位表を見る機会は今後も増えますが、誰の環境で測ったのかを確かめる癖がないと、順位表は広告と区別がつかなくなります。

TRELLIS.2 Is Now in ComfyUI Core TRELLIS.2がComfyUI本体に入りました

The item that lands closest to the workbench this week.
Image-to-3D generation with TRELLIS.2 and Pixal3D was merged into ComfyUI core on August 22, 2026 — no custom node pack required.
Search the template library for "Pixal3D & TRELLIS.2: Image to Model" and the official workflow comes up.
The 3D pipeline itself was rebuilt, with new Load / Preview / Save 3D nodes, mesh post-processing nodes, and an extended PBR texturing stage that bakes normal and ambient occlusion maps.
GeekatPlay has an episode testing this new batch of templates locally — TRELLIS 2 image-to-3D, motion capture, MiniMax H3 and Music 3 together.
For anyone who has been running this through a custom node pack, the significance is less about quality than about maintenance cost: the cycle of updating ComfyUI, watching the isolated environment break, and repairing dependencies should structurally shrink once the implementation lives in core.
That said, there is no guarantee that node behavior and preprocessing match between an existing wrapper and the core implementation, so if you switch, it is safer to run the same source material through both and compare outputs first.
The video is unwatched, so how the templates actually handle will be added later.

今週いちばん手元に効く話です。
TRELLIS.2とPixal3Dの画像→3D生成が、2026年8月22日にComfyUI本体へマージされました。
カスタムノードパックが不要になります。
テンプレートライブラリで「Pixal3D & TRELLIS.2: Image to Model」を検索すれば公式ワークフローが出てきますし、3Dパイプライン自体が作り直されていて、Load / Preview / Save の3Dノード、メッシュの後処理ノード、そしてノーマルとアンビエントオクルージョンを焼くところまで拡張されたPBRテクスチャリング段が入っています。
GeekatPlayが、この新しいテンプレート群をTRELLIS 2の3D化・モーションキャプチャ・MiniMax H3・Music 3までまとめてローカルで試す回を出しています。
カスタムノードで運用してきた立場からすると、意味は品質より保守コストの側にあります。
ComfyUIを更新するたびに隔離環境が壊れて依存を直す、という作業が、本体側の統合で構造的に減るはずです。
もっとも、既存のラッパーと本体実装でノードの挙動や前処理が同じである保証はないので、切り替えるなら同じ素材で出力を突き合わせてからにするのが安全です。
本編は未視聴のため、テンプレートの実際の使い勝手は視聴後に追補します。

Video Generation Broke Through Real Time 動画生成がリアルタイムを割った

Theoretically Media has an episode on AI video breaking through real time, and AI Search's weekly roundup lists MiniMax going real-time as well — the same direction appearing from several lineages at once.
Neither has been watched, so no figures here, but it is worth noting what this line changes.
Once generation runs faster than the footage it produces, video generation shifts from something you order and wait for into something you decide while it moves.
In environment terms, producing ten look candidates and picking one becomes picking while dragging a slider.
The use case described here two roundups ago — running dozens of few-second clips to check layout — sits directly in the path of this change, because it is a step where iteration count matters more than quality.
At the same time, being real-time also pushes in the direction of making it harder to reproduce the same result later.
The habit of recording seeds and parameters should become more important, not less.

Theoretically Mediaが「AI動画がリアルタイムを割った」という回を出しています。
AI Searchの週次まとめでもMiniMaxのリアルタイム化が項目に挙がっており、複数の系統から同じ方向の話が出ている状況です。
いずれも本編は未視聴なので具体的な数値には踏み込みませんが、この線が何を変えるかは書いておく価値があります。
生成が生成時間より速くなると、動画生成は「発注して待つもの」から「動かしながら決めるもの」に変わります。
背景制作でいえば、ルックの候補を10本出して選ぶ作業が、スライダーを動かしながら選ぶ作業になるということです。
前々回に書いた「レイアウト確認用の数秒のクリップを何十本も回す」という用途は、まさにこの変化の直撃を受けます。
品質より試行回数が効く工程だからです。
一方で、リアルタイムであることは同時に、後から同じものを再現するのが難しくなる方向でもあります。
シードとパラメータを控える習慣は、むしろ今より重要になるはずです。

Breaking Down Research — Models Steering Their Own Attention 研究を噛み砕く — モデルが自分の注意を操作する

AI Era Compass covers "Language Models Can Control Their Own Attention" (2609.02737).
Attention is normally treated as a passive mechanism determined by the input; this work heads in the direction of a model actively controlling where it looks.
What to discard and what to keep when handling long context is the flip side of Proteus, covered last time, which unlocked memory capacity as context grew: one is about when to widen the container, this is about where to direct the gaze inside a limited one.
The video and the paper's details are unverified, so the method itself is left alone here.
What matters from a production standpoint is that behaviors like drifting off-topic partway through a long task, or losing sight of the original instruction, are starting to be treated as problems solvable in the model's design — moving from something you suppress by crafting prompts to something addressed structurally.

AI時代の羅針盤が「Language Models Can Control Their Own Attention」(2609.02737)を取り上げています。
アテンションは通常、入力の内容から決まる受動的な仕組みとして扱われますが、この研究はモデルが自分でどこを見るかを能動的に制御できるという方向を扱っています。
長い文脈を扱うときに何を捨てて何を残すかは、前回取り上げたProteusの「記憶容量をあとから開く」という話と裏表の関係にあります。
片方は入れ物の大きさをいつ広げるかの話で、こちらは限られた入れ物のどこへ目を向けるかの話です。
本編と論文の詳細は未確認なので、手法の中身には踏み込みません。
制作の現場から見て意味があるのは、長い作業を任せたときの「途中で話が逸れる」「最初の指示を見失う」といった挙動が、モデル側の設計として改善されうる問題として扱われはじめていることです。
プロンプトを工夫して押さえ込む対象から、モデルの構造で解く対象へ移りつつある、と読めます。

Building a House in Blender From a Photo of an Empty Lot 空き地の写真からBlenderで家を建てる

A walkthrough of Claude Fable 5.1 includes a demonstration of handing it a photograph and having it assemble a house inside Blender.
Agents operating a DCC have come up here before — driving Unity, handling Blender and 3D — but when the input is a photo of an empty lot and the output is a house that holds up to viewing, the framing changes.
The earlier examples automated operation; this is closer to raising geometry from a site photograph.
Handing over location-scouting photos and having what should stand there raised as a model is the daily substance of environment work.
The video is unwatched and how much manual intervention is involved is unclear.
But if this direction becomes standard within a few years, a background artist's job shifts another step from building toward specifying how it should be built and then correcting it.
A reply came back from hatori on this one; it appears in the note below.

Claude Fable 5.1の解説動画に、写真を渡すとBlender上で家を組み上げる、という実演が入っています。
エージェントがDCCを操作する話はこの欄で何度か扱ってきましたが(Unityを操作する例、Blenderと3Dを扱う例)、入力が「空き地の写真」で出力が「鑑賞に耐える家」というところまで来ると、扱いが変わります。
これまでの例は操作の自動化でしたが、こちらは現場写真からの起こしに近い。
ロケハン写真を渡して、そこに建つはずのものをモデルとして起こさせる、という工程は背景制作の日常そのものです。
本編は未視聴で、どこまで手を入れているかは分かりません。
ただ、この方向が数年で標準になるとすれば、背景アーティストの仕事は「建てる」から「建て方を指定して直す」へ、もう一段寄ることになります。
hatoriさんからこの件への返信が届いているので、下の注記に載せています。

What This Means for Creators 制作者にとっての意味

This was a week that made the distance between marketing language and measurement easy to see.
The AGI declaration shifted by more than thirty points under third-party measurement, while the unmarketable change — TRELLIS.2 entering core and lightening maintenance — is the one that affects actual working hours.
This asymmetry shows up nearly every week.
Chasing leaderboards and "surpasses X" headlines saves you no time at the desk; dependencies no longer breaking definitely does.
That said, the big stories can't simply be ignored either.
Real-time video, and raising a house in Blender from a photograph, look like flashy demos at the stage we're seeing them — but once they hold up, they change the order of the process itself.
The criteria stay what they have been: can it be brought into your own setup, and can you explain it.
One more has joined them: check whose harness produced the number on the leaderboard.

今週は、宣伝の言葉と実測の距離がよく見えた週でした。
AGI宣言は第三者の測定で3割以上ずれ、その一方で、宣伝にならない地味な変化——TRELLIS.2が本体に入って保守が軽くなる——のほうが、実際の作業時間には効きます。
この非対称は毎週のように現れます。
順位表や「◯◯を超えた」という見出しは追いかけても手元の作業は1分も減りませんが、依存関係が壊れなくなることは確実に減らします。
とはいえ、大きい話を無視してよいわけでもありません。
リアルタイム動画も、写真からBlenderで家を建てる話も、いま見ている段階では派手なデモですが、成立してしまえば工程の順番そのものを変えます。
判断の基準は先週までと同じです。
手元に落とし込めるか、そして自分で説明できるか。
順位表の数字がどの環境で出たものかを確かめる、という一手間も、そこに加わりました。

(Production note) Parts of this write-up are based on titles and primary sources without having watched the videos. GPT-6 Astra's release date, Brockman's remark, ARC Prize's 62.7% re-measurement and its statement that it is not claiming AGI, and the post-launch lockout and apology are all verified against reporting. The merge of TRELLIS.2 and Pixal3D into ComfyUI core (merged 2026-08-22, template name, 3D node lineup) follows ComfyUI's own announcement and documentation. No figures are claimed for real-time video generation. The Claude Fable 5.1 Blender demonstration is described from the video's framing and has not been reproduced hands-on. For "Language Models Can Control Their Own Attention," the paper itself could not be verified, so the method is not described. As before, most of the _memo/_news material this week consists of general AI / business-use videos (e.g. Julian Goldie SEO), outside this site's scope and therefore unused.
(制作メモ)本文の一部は、動画本編を未視聴のまま件名と一次情報でまとめています。GPT-6 Astraの公開日、Brockman氏の発言、ARC Prizeによる62.7%という再測定とAGIを主張しないという表明、公開直後のログイン障害と謝罪は、いずれも報道で裏を取りました。TRELLIS.2とPixal3DのComfyUI本体への統合(2026-08-22マージ、テンプレート名、3Dノードの構成)はComfyUI公式の告知とドキュメントによります。リアルタイム動画生成は具体的な数値に踏み込んでいません。Claude Fable 5.1のBlender実演は動画の紹介にもとづくもので、実機での再現は行っていません。「Language Models Can Control Their Own Attention」は論文本文を確認できていないため、手法の詳細は書いていません。なお _memo/_news の材料は今週もAI一般・ビジネス活用寄りの動画(Julian Goldie SEO等)が大半で、本サイトの対象から外れるため不採用としています。
hatori's note · 2026-09

Handing over a photo of an empty lot and having a house built in Blender is not something I can let pass.
Watching the video, the construction isn't crude — it holds up to viewing.
Two years from now, I expect combining this with 3D model generation and finishing in UE6 will be ordinary work.

空き地の写真を渡してBlenderで家を作成させるというのは聞き捨てならない。
動画を見ると家の作りは雑ではなくて、ちゃんと鑑賞できるクオリティ。
今から2年後、3Dモデル生成と組み合わせてUE6で整えるという作業が普通になると思う。

Further viewing その他の参照動画