Music Generation
Stories, products, and related signals connected to this tag in Explore.
Stories
Filter storiesPractitioners cite 30-second character-led shots and rap music-video tests, often paired with Midjourney style references. One Spanish-song test reported strong results but weaker distinctions between trap and rap.
Pika Music accepts text, lyrics, voice, and music references separately or together in one diffusion decoder. It is available through Pika API Club, and Pika claims up to 10 times better cost efficiency than other music models.
Cuenca says it rewrote an Alzheimer's-themed music video after an early version resembled Black Mirror's San Junipero in its characters, setting, and story elements. The resulting video, Candela, was released at Cannes Lions.
Mureka V9.5 uses MusiCoT to plan a song’s structure and emotional arc before filling in details. Mureka Co can generate tracks and return beat-aligned stems inside Ableton.
A BytePlus thread says Seed Audio 1.0 can create dialogue, effects, background music, and reference voices for clips up to 120 seconds. The workflow can pass audio timing into Seedance for video generation.
Stages AI previewed Audio Studio and said v1.01 is under checks while v1.02 is slated to add stem separation, alongside voice-design and mobile demos. The rollout points to an in-platform audio workflow for editing, narration, and music tasks that usually live in separate tools.
Bennash released Rendergeist Pulse for free and used it on multi-minute audio-reactive videos including Super Super Saturday and The Longer the Road. The release turns beat-synced motion graphics into a reusable tool instead of a one-off experiment.
Multiple Reddit posts said Suno's cleanup workflow exposed no Audio Influence slider, and separate complaint threads said Google Play refunds were approved after repeated red-error blocks. The complaints point to a gap between Suno's revision workflow and what paying users say they can actually access.
Creators posted finished shorts and ad-style clips built with Midjourney, Seedance, LTX, Suno and Glif. The stacks compress previs, motion and music into days, but the posts still describe manual compositing, editing and local renders.
Reddit posts said v5.5 improved voice tone but still ignores gender-labeled sections, switches singers mid-part, and struggles with detailed instrument instructions. Creators are iterating on renders until the emotion fits, then generating lipsync video to work around the gaps.
Google is rolling out Lyria 3 Pro for full songs and Lyria 3 Clip for 30-second generations in the Gemini API and AI Studio. Musicians can now map intros, verses, choruses and bridges instead of stitching short music clips together.
ElevenLabs launched Flows, a node-based canvas inside ElevenCreative that chains image, video, voice, music, SFX, lip sync, and voice changing in one workspace. Use it to keep context across the pipeline instead of re-exporting between apps.