I'm sorry if this question is obnoxious but I mean it sincerely, not as a pedantic gotcha.
Is this correct? Do LLMs do images and videos? I thought the models that did those things (Midjourney, DALL-E, Sora, etc) were a different kinds of models, not LLMs