Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingNewsWatching0 independent reports0

Boogu-Image-0.1: Improving Open Agentic Multimodal Generation with Understanding on a Minimal Budget

First seen · 7/16/2026, 12:00 PMLatest activity · 7/16/2026, 12:00 PM

Boogu-Image-0.1 is an open unified multimodal understanding and generation model family with Base, Turbo, Edit, and Edit-Turbo variants. It targets high-quality text-to-image generation, fast inference, instruction-based editing, and Chinese-English text rendering. The authors attribute its performance to a stronger multimodal encoder, agentic prompt rewriting, improved data and training pipelines, and inference-time scaling. They report using 208.62 million unique images, with an estimated theoretical training cost of about $400,000 for the base model. Weights, code, and recipes are released under Apache 2.0.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/16, 12:00 PMnot independentRepresentative
    Boogu-Image-0.1: Improving Open Agentic Multimodal Generation with Understanding on a Minimal Budget