第1986話 Sunoは「意味」を歌っているのか?──音韻・母音・人工言語による歌唱生成実験
Sunoは「意味」を歌っているのか?──音韻・母音・人工言語比較実験
前回の実験では、
「Sunoは文字をどのように処理しているのか?」
をテーマに、
改行
スペース
ハイフン
アンダーバー
句読点
大文字・小文字
音節分割
など、文字構造の違いが歌唱へ与える影響を比較しました。
その結果、
改行
ハイフン
音節分割
では歌唱の区切りに変化が見られました。
一方で、
大文字・小文字
CamelCase
句読点
などは大きな変化が確認できませんでした。
今回はさらに一歩進め、
「Sunoは歌詞の意味を理解して歌っているのか?」
それとも、
「文字から音韻情報を抽出して歌唱を生成しているのか?」
という点を検証しました。
実験条件
Styleは全パターン共通。
Style(全パターン共通)
Experimental electronic cinematic fusion, futuristic AI music laboratory, intelligent music generation, emotional cyber pop, modern electronic pop, cinematic atmosphere, evolving synthesizers, deep warm analog bass, precise electronic drums, expressive emotional vocals, vocal as rhythmic instrument, phonetic vocal textures, syllable-driven melodic phrasing, dynamic rhythmic articulation, layered electronic percussion, evolving vocal harmonies, immersive stereo production, professional studio quality, progressive musical journey, subtle glitch textures, intelligent arrangement, balanced commercial sound, memorable melodic hooks, hybrid human and artificial creativity, modern J-pop influenced melodic sensibility, polished modern mix, evolving vocal phrasing, expressive legato and subtle staccato contrast, cinematic electronic anthem歌詞のみを変更しました。
比較した内容は以下の6種類です。
Pattern1:意味のある英文
[Verse]
Beyond the silent ocean
I remember your voice
Hidden inside the darkness
A melody appears
Every heartbeat
Every moment
Moves together
Through the night
[Chorus]
Music remembers
Every dream
Every voice
Becomes one light
Beyond the sound
Beyond the time
We are connected
Through the melody通常の歌詞。
例:
Beyond the silent ocean
I remember your voice
一般的な英語歌詞として生成。
結果:
王道的なエレクトロニックサウンド
自然な曲展開
普通の歌モノとして成立
基準パターンとなりました。
Pattern2:擬似英語化
(意味を消して英語風の音だけ残す)
[Verse]
Belonda the silenta oceana
Ai rememora yura voisa
Hidina insidea the darkena
A melodia appeara
Every heartbea
Every momenta
Movesa togethera
Througa the nighta
[Chorus]
Musica remembara
Every dreama
Every voisa
Becomesa one lighta
Belonda the sounda
Belonda the timea
We area connecteda
Througa the melodia英語らしい音を残しながら、意味を崩した歌詞。
例:
Belonda the silenta oceana
Ai rememora yura voisa
結果:
Pattern1との差はほとんど感じられませんでした。
意味を失った文章でも、
音節
母音配置
英語風の発音感
が維持されることで、自然な歌唱になりました。
この結果から、
Sunoは歌詞の意味だけではなく、音としての情報を利用している可能性が考えられます。
Pattern3:母音中心
(意味を完全排除)
[Verse]
Aeou ia eea oaeoa
Ai eeoa uoa oia
Iaia iaia
Aeoea iaea
Every heartbeat
Aa ee ai
Oo ua ei
Ia oea
[Chorus]
Aa ee ai
Oo ua ei
Every dream
Every voice
Ae ae ae
Ooo ia
Music
Remember意味を完全に排除し、母音主体の歌詞。
例:
Aa ee ai
Oo ua ei
結果:
予想していたようなスタッカート感は強くありませんでした。
むしろ、
伸ばし気味
声楽的
滑らかな歌唱
になりました。
母音はメロディラインと相性が良く、音の伸びとして処理されている可能性があります。
Pattern4:音節分割
[Verse]
Be yond the si lent o cean
I re mem ber your voice
Hid den in side the dark ness
A mel o dy ap pears
Ev ery heart beat
Ev ery mo ment
Moves to geth er
Through the night
[Chorus]
Mu sic re mem bers
Ev ery dream
Ev ery voice
Be comes one light
Be yond the sound
Be yond the time単語を細かく分割。
例:
Be
yond
the
sound
結果:
これは明確に変化が感じられました。
発音の区切り
リズム感
歌唱タイミング
に影響しているように感じられました。
前回の改行・ハイフン実験とも一致する結果です。
Sunoでは「区切り情報」が歌唱表現に影響する可能性があります。
Pattern5:子音強調
(母音を減らしリズム成分を見る)
[Verse]
Bnd th slnt cn
Rmmbr yr vcs
Hddn nstd th drk
Mldy prs
Evry hrtbt
Evry mment
Mvs tgethr
Thrgh th nght
[Chorus]
Msc rmmbers
Evry drm
Evry vcs
Bcmes wn lght母音を減らし、子音中心にしたパターン。
例:
Bnd th slnt cn
結果:
意外にも流れるような歌唱になりました。
発音の制約が減り、ボーカルラインを優先した処理になった可能性があります。
Pattern6:人工言語
(完全な未知言語)
[Verse]
Nalera solamai venora
Lunari temora vesai
Kalena moritai
Selora navani
Every rhythm
Every pulse
Nalera flows
Beyond the sky
[Chorus]
Solamai
Nalera
Temora
Vesai
One emotion
Many voices
One melody
Forever意味を持たない完全な造語。
例:
Nalera solamai
Temora vesai
結果:
過去の人工言語実験と同じく、Sunoは違和感なく楽曲化しました。
意味が存在しない言葉でも、
音節
母音
リズム
が成立していれば、歌として処理できる可能性があります。
実験結果まとめ
今回の結果を整理すると、
条件 結果
通常英文 自然な歌唱
擬似英語 ほぼ同じように成立
母音中心 伸びる歌唱
音節分割 区切りが強くなる
子音強調 流れる歌唱
人工言語 違和感なく成立
今回見えてきた仮説
今回の実験だけでSuno内部の処理を断定することはできません。
しかし、今回の条件では、
「文章の意味」
よりも、
音節構造
母音配置
発音リズム
区切り情報
が歌唱生成へ影響している可能性が見えてきました。
仮説モデルとしては、
文字情報
↓
音韻情報へ変換
↓
歌唱リズム生成
↓
メロディ・伴奏と統合
↓
楽曲化という流れが考えられます。
今回の統合曲
最後に今回の実験結果を反映した統合曲を制作しました。
今回の統合曲は、実験結果を反映して、
通常英語
擬似英語
母音
音節分割
人工言語
日本語的な音韻感
を混ぜながら、Sunoが自然な楽曲として統合できるかを見る形にします。
テーマは、
「Different forms, one voice(違う形でも、声はひとつになる)」
です。
Style
Experimental electronic cinematic fusion, futuristic AI music laboratory, intelligent music generation, emotional cyber pop, modern electronic pop, cinematic atmosphere, evolving synthesizers, deep warm analog bass, precise electronic drums, expressive emotional vocals, vocal as rhythmic instrument, phonetic vocal textures, syllable-driven melodic phrasing, dynamic rhythmic articulation, layered electronic percussion, evolving vocal harmonies, immersive stereo production, professional studio quality, progressive musical journey, subtle glitch textures, intelligent arrangement, balanced commercial sound, memorable melodic hooks, hybrid human and artificial creativity, modern J-pop influenced melodic sensibility, polished modern mix, evolving vocal phrasing, expressive legato and subtle staccato contrast, cinematic electronic anthem, futuristic vocal experiment, multilingual phonetic fusion, emotional artificial language performanceLyrics
[Intro]
Before the sound
There is a voice
Before the word
There is a rhythm
Letters can change
Symbols can move
But the melody
Finds a way
[Verse]
Beyond the silent ocean
I remember your voice
Hidden inside the darkness
A melody appears
Belonda the silenta oceana
Ai rememora yura voisa
Every heartbeat
Every moment
Moves together
Through the night
[Pre-Chorus]
Be
yond
the
sound
Be yond the line
Be yond the time
Different letters
Different forms
Still the emotion remains
[Chorus]
Music remembers
Every dream
Every voice
Becomes one light
Mu sic re mem bers
Ev ery voice
Ev ery dream
Aa ee ai
Oo ua ei
Nalera solamai
Temora vesai
One feeling
Many sounds
One melody
Forever
[Instrumental]
Na na na
Ta da ra
Aa ee ai
Oh oh oh
Nalera
Solamai
Beyond
Beyond the sound
[Verse 2]
Space or no space
Line after line
Symbol after symbol
Time after time
Bnd th slnt cn
Rmmbr yr vcs
Even without the words
The song continues
[Bridge]
Oto no kanata
Koe no hikari
Yume no naka de
Oto wa tsunagaru
Aeou ia eea
Nalera moritai
Solamai vesai
Many languages
Many voices
One heartbeat
[Final Chorus]
Beyond the sound
Beyond the letters
Beyond the signs
Different structures
Same emotion
Beyond-the-sound
Beyond_the_sound
Be
yond
the
sound
Music remembers
More than words
More than language
More than form
Every voice
Every dream
Every rhythm
Becomes music
[Outro]
The experiment continues
The melody remains
Different sounds
Different worlds
One voice
One music今回の統合版では、あえて完全な人工言語だけにはせず、
人間が意味を感じられる領域
↓
意味が崩れる領域
↓
純粋な音になる領域
を連続的につなげています。
歌詞には、
通常英文
擬似英語
母音列
音節分割
人工言語
日本語的な音韻
を組み込みました。
目的は、
「意味が変化しても、Sunoはどこまで音楽として統合できるのか」
を見ることです。
結果として、複数の異なる文字・音声要素を混ぜても、Sunoは全体を自然な楽曲として成立させました。
まとめ
今回の実験では、
意味を失っても歌唱は成立する
擬似言語でも自然な楽曲になる
母音は伸びる歌唱へ影響する
音節分割はリズム制御に影響する
人工言語でも音楽として成立する
という傾向が確認できました。
AI音楽生成では、単純な「文章理解」だけではなく、
文字
↓
音
↓
リズム
↓
感情表現
という複数の情報処理が行われているのかもしれません。
今後も条件を一つずつ変えながら、AIがどのように音楽を構築しているのかを実験していきます。
タグ
#Suno #AI音楽 #AI作曲 #生成AI #音楽実験 #AI研究 #プロンプト研究 #音響解析 #人工知能 #音楽生成 #DTM #テクノロジー #実験記録
