One of the things they seem to be emphasizing here is the UX around being able to place specific elements where you want them in an image. If the positions of the components in the overall composition are very important, this seems to make that a lot easier and kind of reminds me of InvokeAI.
Ideogram V4, an open-weight model released back in June can also do this [1], but you have to use a relatively cumbersome JSON structure to describe all the different bounding boxes. So it’s definitely a bit of a hassle.
I'll probably be waiting until it goes open-weight (hopefully soon) like they did with Flux.2 / Klein.
That sort of steering ability that has been possible with the latest Gemini releases has been nice to work with over previous generations. It’s great to see this improve on the platform with declarative controls built into the API and coming soon as an open model.
Ideogram V4, an open-weight model released back in June can also do this [1], but you have to use a relatively cumbersome JSON structure to describe all the different bounding boxes. So it’s definitely a bit of a hassle.
I'll probably be waiting until it goes open-weight (hopefully soon) like they did with Flux.2 / Klein.
[1] - https://docs.ideogram.ai/using-ideogram/getting-started/prom...
Chats can be awful user interfaces.
What's new from the last post? GA?
> We will open up an early access phase for FLUX 3 Image in the following weeks.
Not sure if there was a separate post for early access or if they just skipped to this.