Generative models around images are nothing new and have been around for a while already. But even today, if you really want creative control and expression, you need a designer that's good with Photoshop or Illustrator etc.
This is applicable to LLMs as well. You can get it to write plausible BS but if you really want a rooted in reality, well articulated write up about something, a human has to be taken onboard.
This equally extends to voice over. If you really want expressive and creative control to put some outstanding rendering of something, AI isn't going to cut it.
This is only true if you assume AI isn't going to keep improving. It gets significantly better on a quarterly basis, far faster than the time it takes for an actor to develop their craft and career. The output quality of todays' cutting edge models would have been science fiction only 2-3 years ago.
I'm not so sure about the future. Such models, all the models don't have a well understood input output mapping and that's going to be a problem for a very long time.
This is applicable to LLMs as well. You can get it to write plausible BS but if you really want a rooted in reality, well articulated write up about something, a human has to be taken onboard.
This equally extends to voice over. If you really want expressive and creative control to put some outstanding rendering of something, AI isn't going to cut it.