简单对比几个模型唱粤语歌

新觉青年,玩物益智

最近打算用模型唱几首粤语歌曲,自己写了几句歌词,丢给模型唱出来,效果总体一般。

试听(播放器没有出来,可以刷新下页面)

模型版本

  • SUNO V4.5-all
  • MiniMax Audio Music-3.0
  • Mureka V7.5-all
  • 网易天音 模型未知
  • SeedMusic 1.0 Preview

歌词带简单编曲提示

纸鹞 --- [Intro](Sparse, warm nylon guitar harmonics, 4 bars) [Verse 1] 怀念细个放纸鹞(Light felt piano enters, soft sustained chords) 看它向着天际飘 甜丝丝感觉微妙(Gentle ambient swell) 快乐仿似七彩桥 [Verse 2] 人大咗少见纸鹞(Minimal brushed snare joins, subtle pulse) 奔波里笑容渐少 但每当记起这份欢笑(Strings gently rise) 心头顿觉暖光照 [Outro](Strings fade, nylon guitar returns alone; final chord rings with long reverb tail, natural decay)

演唱风格提示词

A tender, nostalgic Cantonese female vocal, childlike warmth in tone, breathy head voice with intimate close-mic presence. Crystal-clear Cantonese diction with natural tonal movement, smooth legato phrasing, soft dynamics as if whispering a cherished memory. Minimal vibrato, only a gentle shimmer on final sustained notes. Cantopop ballad delivery, evoking the sweetness and innocence of childhood recollection.

简单对比

模型名称版本咬字准确编曲支持音色音质
SUNOV4.5-all✅✅✅✅✅
MiniMaxAudio Music 3.0✅✅
MurekaV7.5-all
网易天音未知✅✅✅✅✅
SeedMusic1.0 Preview✅✅✅✅✅✅

总结

总体而言,SUNO 效果是比较好的,除了人生毛刺感明显外,其他都还可以。网易天音和 SeedMusic 感觉是同一个模型,出来的人声质感很像,只是采用的编曲样本有些差异。Mureka 总体过于老派,像是那种 Hi-Fi 口水歌训练出来的感觉。MiniMax 不支持歌词混插编曲提示,总体 bpm 节奏好像很局促,听感一般。

粤语模型比普通话模型差距开始拉大了。

BTW: 本人母语是粤语(疍家)

添加新评论