第1986話 Sunoは「意味」を歌っているのか?──音韻・母音・人工言語による歌唱生成実験

Sunoは「意味」を歌っているのか?──音韻・母音・人工言語比較実験

前回の実験では、

「Sunoは文字をどのように処理しているのか?」

をテーマに、

  • 改行

  • スペース

  • ハイフン

  • アンダーバー

  • 句読点

  • 大文字・小文字

  • 音節分割

など、文字構造の違いが歌唱へ与える影響を比較しました。

その結果、

  • 改行

  • ハイフン

  • 音節分割

では歌唱の区切りに変化が見られました。

一方で、

  • 大文字・小文字

  • CamelCase

  • 句読点

などは大きな変化が確認できませんでした。

今回はさらに一歩進め、

「Sunoは歌詞の意味を理解して歌っているのか?」

それとも、

「文字から音韻情報を抽出して歌唱を生成しているのか?」

という点を検証しました。


実験条件

Styleは全パターン共通。

Style(全パターン共通)

Experimental electronic cinematic fusion, futuristic AI music laboratory, intelligent music generation, emotional cyber pop, modern electronic pop, cinematic atmosphere, evolving synthesizers, deep warm analog bass, precise electronic drums, expressive emotional vocals, vocal as rhythmic instrument, phonetic vocal textures, syllable-driven melodic phrasing, dynamic rhythmic articulation, layered electronic percussion, evolving vocal harmonies, immersive stereo production, professional studio quality, progressive musical journey, subtle glitch textures, intelligent arrangement, balanced commercial sound, memorable melodic hooks, hybrid human and artificial creativity, modern J-pop influenced melodic sensibility, polished modern mix, evolving vocal phrasing, expressive legato and subtle staccato contrast, cinematic electronic anthem

歌詞のみを変更しました。

比較した内容は以下の6種類です。


Pattern1:意味のある英文

[Verse]

Beyond the silent ocean
I remember your voice

Hidden inside the darkness
A melody appears

Every heartbeat
Every moment

Moves together
Through the night


[Chorus]

Music remembers

Every dream

Every voice

Becomes one light

Beyond the sound

Beyond the time

We are connected

Through the melody

通常の歌詞。

例:

Beyond the silent ocean
I remember your voice

一般的な英語歌詞として生成。

結果:

  • 王道的なエレクトロニックサウンド

  • 自然な曲展開

  • 普通の歌モノとして成立

基準パターンとなりました。


Pattern2:擬似英語化

(意味を消して英語風の音だけ残す)

[Verse]

Belonda the silenta oceana
Ai rememora yura voisa

Hidina insidea the darkena
A melodia appeara

Every heartbea
Every momenta

Movesa togethera
Througa the nighta


[Chorus]

Musica remembara

Every dreama

Every voisa

Becomesa one lighta

Belonda the sounda

Belonda the timea

We area connecteda

Througa the melodia

英語らしい音を残しながら、意味を崩した歌詞。

例:

Belonda the silenta oceana
Ai rememora yura voisa

結果:

Pattern1との差はほとんど感じられませんでした。

意味を失った文章でも、

  • 音節

  • 母音配置

  • 英語風の発音感

が維持されることで、自然な歌唱になりました。

この結果から、

Sunoは歌詞の意味だけではなく、音としての情報を利用している可能性が考えられます。


Pattern3:母音中心

(意味を完全排除)

[Verse]

Aeou ia eea oaeoa
Ai eeoa uoa oia

Iaia iaia
Aeoea iaea

Every heartbeat

Aa ee ai

Oo ua ei

Ia oea


[Chorus]

Aa ee ai

Oo ua ei

Every dream

Every voice

Ae ae ae

Ooo ia

Music

Remember

意味を完全に排除し、母音主体の歌詞。

例:

Aa ee ai
Oo ua ei

結果:

予想していたようなスタッカート感は強くありませんでした。

むしろ、

  • 伸ばし気味

  • 声楽的

  • 滑らかな歌唱

になりました。

母音はメロディラインと相性が良く、音の伸びとして処理されている可能性があります。


Pattern4:音節分割

[Verse]

Be yond the si lent o cean

I re mem ber your voice

Hid den in side the dark ness

A mel o dy ap pears


Ev ery heart beat

Ev ery mo ment

Moves to geth er

Through the night


[Chorus]

Mu sic re mem bers

Ev ery dream

Ev ery voice

Be comes one light

Be yond the sound

Be yond the time

単語を細かく分割。

例:

Be
yond
the
sound

結果:

これは明確に変化が感じられました。

  • 発音の区切り

  • リズム感

  • 歌唱タイミング

に影響しているように感じられました。

前回の改行・ハイフン実験とも一致する結果です。

Sunoでは「区切り情報」が歌唱表現に影響する可能性があります。


Pattern5:子音強調

(母音を減らしリズム成分を見る)

[Verse]

Bnd th slnt cn

Rmmbr yr vcs

Hddn nstd th drk

Mldy prs


Evry hrtbt

Evry mment

Mvs tgethr

Thrgh th nght


[Chorus]

Msc rmmbers

Evry drm

Evry vcs

Bcmes wn lght

母音を減らし、子音中心にしたパターン。

例:

Bnd th slnt cn

結果:

意外にも流れるような歌唱になりました。

発音の制約が減り、ボーカルラインを優先した処理になった可能性があります。


Pattern6:人工言語

(完全な未知言語)

[Verse]

Nalera solamai venora

Lunari temora vesai

Kalena moritai

Selora navani

Every rhythm

Every pulse

Nalera flows

Beyond the sky


[Chorus]

Solamai

Nalera

Temora

Vesai

One emotion

Many voices

One melody

Forever

意味を持たない完全な造語。

例:

Nalera solamai
Temora vesai

結果:

過去の人工言語実験と同じく、Sunoは違和感なく楽曲化しました。

意味が存在しない言葉でも、

  • 音節

  • 母音

  • リズム

が成立していれば、歌として処理できる可能性があります。


実験結果まとめ

今回の結果を整理すると、

条件 結果 
通常英文 自然な歌唱 
擬似英語 ほぼ同じように成立 
母音中心 伸びる歌唱 
音節分割 区切りが強くなる 
子音強調 流れる歌唱 
人工言語 違和感なく成立


今回見えてきた仮説

今回の実験だけでSuno内部の処理を断定することはできません。

しかし、今回の条件では、

「文章の意味」

よりも、

  • 音節構造

  • 母音配置

  • 発音リズム

  • 区切り情報

が歌唱生成へ影響している可能性が見えてきました。

仮説モデルとしては、

文字情報
 ↓
音韻情報へ変換
 ↓
歌唱リズム生成
 ↓
メロディ・伴奏と統合
 ↓
楽曲化

という流れが考えられます。


今回の統合曲

最後に今回の実験結果を反映した統合曲を制作しました。

今回の統合曲は、実験結果を反映して、

  • 通常英語

  • 擬似英語

  • 母音

  • 音節分割

  • 人工言語

  • 日本語的な音韻感

を混ぜながら、Sunoが自然な楽曲として統合できるかを見る形にします。

テーマは、

「Different forms, one voice(違う形でも、声はひとつになる)」

です。


Style

Experimental electronic cinematic fusion, futuristic AI music laboratory, intelligent music generation, emotional cyber pop, modern electronic pop, cinematic atmosphere, evolving synthesizers, deep warm analog bass, precise electronic drums, expressive emotional vocals, vocal as rhythmic instrument, phonetic vocal textures, syllable-driven melodic phrasing, dynamic rhythmic articulation, layered electronic percussion, evolving vocal harmonies, immersive stereo production, professional studio quality, progressive musical journey, subtle glitch textures, intelligent arrangement, balanced commercial sound, memorable melodic hooks, hybrid human and artificial creativity, modern J-pop influenced melodic sensibility, polished modern mix, evolving vocal phrasing, expressive legato and subtle staccato contrast, cinematic electronic anthem, futuristic vocal experiment, multilingual phonetic fusion, emotional artificial language performance

Lyrics

[Intro]

Before the sound

There is a voice

Before the word

There is a rhythm


Letters can change

Symbols can move

But the melody

Finds a way


[Verse]

Beyond the silent ocean

I remember your voice

Hidden inside the darkness

A melody appears


Belonda the silenta oceana

Ai rememora yura voisa


Every heartbeat

Every moment

Moves together

Through the night


[Pre-Chorus]

Be
yond
the
sound


Be yond the line

Be yond the time


Different letters

Different forms

Still the emotion remains


[Chorus]

Music remembers

Every dream

Every voice

Becomes one light


Mu sic re mem bers

Ev ery voice

Ev ery dream


Aa ee ai

Oo ua ei


Nalera solamai

Temora vesai


One feeling

Many sounds

One melody

Forever


[Instrumental]

Na na na

Ta da ra

Aa ee ai

Oh oh oh


Nalera

Solamai

Beyond

Beyond the sound


[Verse 2]

Space or no space

Line after line

Symbol after symbol

Time after time


Bnd th slnt cn

Rmmbr yr vcs


Even without the words

The song continues


[Bridge]

Oto no kanata

Koe no hikari

Yume no naka de

Oto wa tsunagaru


Aeou ia eea

Nalera moritai

Solamai vesai


Many languages

Many voices

One heartbeat


[Final Chorus]

Beyond the sound

Beyond the letters

Beyond the signs


Different structures

Same emotion


Beyond-the-sound

Beyond_the_sound


Be

yond

the

sound


Music remembers

More than words

More than language

More than form


Every voice

Every dream

Every rhythm

Becomes music


[Outro]

The experiment continues

The melody remains


Different sounds

Different worlds


One voice

One music

今回の統合版では、あえて完全な人工言語だけにはせず、

人間が意味を感じられる領域

意味が崩れる領域

純粋な音になる領域

を連続的につなげています。

歌詞には、

  • 通常英文

  • 擬似英語

  • 母音列

  • 音節分割

  • 人工言語

  • 日本語的な音韻

を組み込みました。

目的は、

「意味が変化しても、Sunoはどこまで音楽として統合できるのか」

を見ることです。

結果として、複数の異なる文字・音声要素を混ぜても、Sunoは全体を自然な楽曲として成立させました。


まとめ

今回の実験では、

  • 意味を失っても歌唱は成立する

  • 擬似言語でも自然な楽曲になる

  • 母音は伸びる歌唱へ影響する

  • 音節分割はリズム制御に影響する

  • 人工言語でも音楽として成立する

という傾向が確認できました。

AI音楽生成では、単純な「文章理解」だけではなく、

文字



リズム

感情表現

という複数の情報処理が行われているのかもしれません。

今後も条件を一つずつ変えながら、AIがどのように音楽を構築しているのかを実験していきます。


タグ

#Suno #AI音楽 #AI作曲 #生成AI #音楽実験 #AI研究 #プロンプト研究 #音響解析 #人工知能 #音楽生成 #DTM #テクノロジー #実験記録

いいなと思ったら応援しよう!