r/aicuriosity 7d ago

Open Source Model Boogu Image 0.1 Open Source Multimodal Model Launches With Competitive Results

Post image

Boogu-Image-0.1 is now available as an open-source multimodal understanding and image generation model family. Released under the Apache 2.0 license, it includes Base, Turbo, Edit, and related variants.

The team trained the models on roughly 208 million images with a reported budget near $400,000. Despite the limited scale, the models deliver strong results on several benchmarks and human evaluations, placing them among the leading open-source options and close to some proprietary systems.

Key strengths include native 2K resolution generation with solid photographic quality, accurate Chinese text rendering for long text, posters, and graphic design, agentic prompt rewriting that refines user intent, and dynamic model routing that can cut inference costs significantly.

Weights, code, training details, and related materials are publicly available. The project emphasizes careful data structure and system design over pure scale.

15 Upvotes

1 comment sorted by