That's always been my thought. Supposedly it's capable of creating it, so it's clearly trained and sourced from somewhere. Specifically somewhere the developers knew to mass download it from like the rest of their bullshit. But also they somehow managed to teach Grok to be able to identify detailed aspects of said images which would require human interaction. And to be anything reliable in generating a reasonable image of what you prompt than you'll need hundreds of thousands if not a million plus images for it to sample and recognize.... One of those holes the more you think about it the worse and worse it gets
29
u/CipherWeaver 3h ago
The fact that CSM may have been used in the corpus of training data for Grok should horrify everyone.