I had an idea to find a multimodal or model that can listen but not just for SST/transcription, but to sort of describe/understand/interpret what it’s hearing—- and then have it listen to my entire music library, and then have it peg my exact music taste and recommend other music, or to just do something interesting or fun with what it has learned or interprets my library. I don’t think there are any multi modal models that could do this though?
Edit, just found a Reddit thread from 9 mos ago about this concept. Interesting, wonder what the current state of this idea is. Also, Pandora, Google/YT music have been able to do this for years— eg recommend music based off user taste/similar songs etc… I wonder how they are doing it? Obviously some sort of AI/algorithmic thing I would think, but not LLM related I wouldn’t think.
3 comments of 7
[ 6.7 ms ] story [ 31.4 ms ] threadEdit, just found a Reddit thread from 9 mos ago about this concept. Interesting, wonder what the current state of this idea is. Also, Pandora, Google/YT music have been able to do this for years— eg recommend music based off user taste/similar songs etc… I wonder how they are doing it? Obviously some sort of AI/algorithmic thing I would think, but not LLM related I wouldn’t think.
https://www.reddit.com/r/udiomusic/comments/1gzvk70/a_llm_th...
With that, we no longer need humans to listen to music!