I'd point out that the lack of copyright on raw AI gens is a feature, not a bug.
But "raw AI gens" is the important part there.
The USPTO has been pretty clear on this for the past two years, the machine-generated parts are public domain, but the parts that aren't can be protected. It all functions exactly like any other public domain material in that regard.
Yeah, you can gank raw gens day in and day out, but you can't really be certain how much work someone's done to any given image unless they tell you.
Maybe that song is 100% AI, but if it's been remixed in post, or if the lyrics are human authored, the human author has a claim.
Maybe that image was 100% generated, or maybe the AI was just used to edit the photo, or the image was composited together from multiple gens. Photoshop works on AI gens just fine.
If that video is more than 10 seconds long, then you should probably err on the side of assuming mixed rights.
And as for the "Learn how to draw" aspect. I did, and how to sew, and how to puppeteer and puppetsmith, and how to paint (on canvas an miniatures), and how to propbuild, (along with a dozen other skills in my jack-of-all-trades-box) and now I can't do any of it because when my hands aren't shaking they are in agony from the arthritis, the pinched shoulder and neck nerves, or both.
I can still work a mouse and type, though not nearly for as long as I used to. I can still edit, composite, and color, sometimes, so that's what I do. With AI, I can parlay those skills into executions of my weird fauxstalgia and unreality projects. I can get things done in the windows of functionality before the pain kills my momentum.
So if you're after inspiration porn, I'm playin' through the pain even when it's AI-assisted work. There's going to come a day when even that will be too much, and I'm enjoying the time I have until then.
On the top, we have "Stunning True-Life Tales of Science Fiction #1 - Robots Ruined the Internet" which was made with public domain comic scans de-colored, composited, re-inked, and re-colored into a new context.
One the bottom we have three pages from "The Secret Origin of Wally Manmoth" which was made with public domain AI gens de-colored, composited, re-inked and recolored into a new context.
The exact same process, except for where the public domain images came from, the first is art by actual human beings that fell into the public domain because the copyrights weren't renewed, back when that was still possible. The second is AI gens made with Midjourney and Dall-E 3, made from math obtained by studying what is essentially the entire human zeitgeist, as manifest in the searchable internet.
I don't really see how one demeans the human spirit while the other does not, or how one is more moral than the other. Further, I've been doing digital photocollage for decades, often with almost entirely copyrighted materials, nary a peep about it.
There's no avenue to hem out AI art that doesn't take out entire mediums of established art. Collage is just the most easy comparison, given how many people mistakenly liken AI to collage.
Also, prompting isn't what it used to be.
Prompting has gotten a lot more complicated as the technology has advanced. For video prompting, for instance, it's rare to just type in an entirely textual request. Most systems use a starting and optional ending frame, several now have multi-keyframe options, and then there's reference-to-video (and image).
For example, with Vidu, a prompt usually winds up being a main prompt establishing the shot direction, and separate reference prompts for the locale, characters, and props.
You can text prompt all you like but you'll never get the above shot with pure text prompting, much less in such a way that allows the characters to be consistent shot-to-shot, scene-to-scene.
Simple prompting couldn't really be considered expression, but at a certain point, you're prompting chunks of script and stage direction.
For the above video, I did the character design work, the concepts, the script, the editing, the sound editing, all the text and graphic design elements (minus the Arkanoid logo), and I voiced the announcer (with Suno filtering).
There's no question that this would qualify as "a significant amount of human expression" were it cut together with appropriated public domain movie and TV clips, or even with unambiguously copyrighted material (as in AMVs). And its hard to get more transformative than 'broken down into math, averaged and cross-reverenced with the math extracted from the entire publicly searchable internet, then reconstituted into new images from that math, prompt guidance, and random jpg noise.
TL:DR: As the technology advanced and demand for consistency, control and coherence increased, the complexity of prompting increased with it, and even if it hasn't reached that point yet, at some point a prompt will get detailed enough that it won't make sense for it not to be considered an expression in and of itself.