Well I can tell you the grok subreddit is MASSIVELY complaining about how everything is awful and it has declined massively, but it’s almost all about moderation – I think they were probably the dudes making some of the problematic NSFW content. Not all NSFW content is problematic, but you probably heard about what grok was happy to do.
I know they just released Imagine 1.5 like… yesterday, and I can’t tell if I have access to it yet. I haven’t noticed a difference but sometimes they roll out updates to different users at different times.
I don’t have enough experience to rate it very well, I’ve only been using it for a couple of days, I’ve mostly been using it to animate already existing images (mostly from midjourney)
There are two modes - speed and quality - other models have something similar like NB2 and NB pro for google but the on grok speed/quality are very different models. Speed tends towards sort of idealized images that look a little fake, too perfect, a little plasticky, but not bad. You see the same faces come up over and over again. Quality is more realistic and has less ideal models that look more like real people. But the interesting thing about speed is that it does infinite scrolling generation so you can crank out dozens of candidate images very quickly whereas quality does a batch of 4 at the same time. The batches of 4 for quality are too close, though, to explore a conceptual space. It’s like NB2 in flow in 4x mode rather than midjourney - 4 slight variations of the same image.
I did find that grok does a decent worldbuilding, like if you put weird elements together it does a decent job of building a world where those two concepts mesh, but I need to spend more time exploring that. I think I posted an example of that earlier, when I used grok through an API, where I had a george washington vs donald trump NBA game on an alien world and the little aliens were wearing “Trump 24” jersies even though I never mentioned that detail.
So… no real opinion yet. the speed generation mode with infinite generation (until you hit your limits) is novel, I haven’t seen image generation that fast / prolific in any other model.
The agentic image / video creation tool is excellent. Flow has something similar now but Grok is probably even better. You create a project and tell the agent (a grok chat bot) what you want to do and it will help you come up with prompts and brain storm and generate the images and videos. Like you can say “I want to do 4 variations on this theme” and it will create them well, then you can say “let’s build off this one but go in this direction” or “take this scene but put it in a city at night” and it’ll do a very competent job of it. You can ask for suggestions or brainstorm ideas for scenes and ask for edits. Before google released an agent in flow I bet this was the best version of this, and I think it’s probably still better than flow because the project grid is a 2d space to organize images and videos you can work with rather than just one timeline. Probably the best feature. If they didn’t have the agent mode when you stopped, you might find that it’s good at helping you create the results you want by acting as essentially an expert prompt interpreter.
Do you have any prompts you want me to test specifically?