AI で画像生成

昔、四国の高松に一度行ったことがありました。
学会に出席したあとすぐに帰ったのですが、栗林公園に寄ればよかったと今でも悔やんでいます。

栗林公園と言えばこの写真ですが、じつは先ほど chatGPT で再生したもの。

画像生成用プロンプトは次の通り。

A high-angle landscape photograph of a traditional Japanese strolling garden, Ritsurin Garden style. A prominent wooden crescent arched bridge spans across a calm, dark green pond. In the middle of the bridge, a tiny figure of a person with dark hair in a white dress with a small black crossbody bag walks away. In the middle distance on the water, a small traditional wooden boat carrying people is steered with a pole. The pond is surrounded by manicured pine trees, stone arrangements, and blooming azalea bushes. In the background, traditional wooden teahouse pavilions sit along the water’s edge beneath a lush, forested mountain hillside. Bright natural daylight, realistic photography, sharp focus, serene atmosphere, high detail.

このプロンプトを Gemini に読ませると、

少しアングルが違いますが、この公園は両者ともにほぼ同じ構造をしているように見えます。

栗林公園のこの写真からプロンプトを抽出したのは Gemini。
それを chatGPT に読ませたのが上の図。
Gemini で新たにチャットを開いて、このプロンプトから画像生成したのが、下の図。

プロンプトを見ると、そこまで詳しく記述しているようには見えないのですが、どうやってここまで立体構造を再現できるのでしょうか。

特に Gemini で作ったこの数行のプロンプトで chatGPT がここまで再現できるのがすごいです。

「百億の昼と千億の夜」の阿修羅王なら「あやつらはつるんでおるのじゃ」と言うでしょう。

ほんとうのところは AI たちともっと仲良くなって訊いてみるしかありませんね。

###

 

コメントを残す

メールアドレスが公開されることはありません。 ※ が付いている欄は必須項目です