A user on Reddit's r/StableDiffusion community is inquiring about the capabilities of local music generation models, specifically asking if Yue2 or similar models can perform true reference-to-audio generation. The user is looking for models that can produce audio based on input reference audio and user prompts, contrasting this with Yue2's apparent limitation of only extracting notes and melody. AI
RANK_REASON User-generated question on a subreddit about a specific model's capabilities, not a formal release or announcement.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →