AI
Ask HN: Are there AI models for generating sounds based on a text and reference?
I've been having a hard time finding a solution. Is there really no commercialized model that I can feed in a reference sound and text instruction and get another sound out? Right now having a multimodal inputs to image or text output is a commodotized, solved problem. IE you can put a prompt for some image and use a reference image to guide the model on what you want. After all, a picture is worth a thousand words right? Ive also used a text and audio input in order to get a text description or...
Read the full discussion on HackerNews
This article was aggregated from HackerNews. Click to join the conversation.
View on HackerNews