Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
The Bonsai Image Ternary 4B MLX 2-bit is a text-to-image diffusion transformer specifically designed for deployment on Apple Silicon, emphasizing quality in its Bonsai Image variant. This model utilizes ternary weights of {−1, 0, +1} along with FP16 group-wise scaling in its transformer layers, which encompass Q/K/V projections, output projections, and MLP weights. Notably, it reduces the size of the FLUX.2 Klein 4B transformer from 7.75 GB FP16 to just 1.21 GB, achieving a remarkable 6.4× smaller footprint while maintaining visual quality and fidelity to prompts akin to the original model. The deployment package for Apple Silicon is 3.88 GB, which includes the MLX 2-bit diffusion transformer, a 4-bit Qwen3-4B text encoder, and an FP16 Flux2 VAE. After the text encoder handles prompt encoding, it is offloaded to ensure that only the compact transformer and VAE remain in memory during the denoising loop. Furthermore, the model employs a 4-step FlowMatchEuler sampler with guidance set at 1.0 and a shift of 3.0, eliminating the need for CFG and negative prompts, thus streamlining the generation process for enhanced user experience. Overall, this innovation represents a significant advancement in efficient and effective image generation technology.
Description
MAI-Image-2.5 represents the most advanced image model developed by Microsoft AI to date, marking an evolution in the MAI-Image series. Upon its release, it achieved an impressive third place on the Arena text-to-image leaderboard, showcasing its ability to excel in a diverse array of artistic styles. The model adheres closely to user instructions, enhances text rendering capabilities, and generates intricate and coherent images as desired. Compared to its predecessor, MAI-Image-2, this new version offers a significant leap in quality, particularly in areas such as text clarity, stylized illustrations, and commercial imagery enhancements. In addition, it demonstrates a robust capacity for visual reasoning involving objects, scene composition, lighting, scale, and spatial relationships, effectively transforming basic directives into refined images. MAI-Image-2.5 places a strong emphasis on the nuances that elevate creative work to a professional level, resulting in sharper text on promotional materials, cleaner labels for products, improved structuring of product images, more intentional scene compositions, enhanced layouts, and overall more sophisticated visuals that bolster brand identity. This model not only sets a new standard for image generation but also opens up exciting possibilities for creative professionals seeking to elevate their work.
API Access
Has API
API Access
Has API
Integrations
Microsoft Azure
Microsoft Foundry
Pricing Details
No price information available.
Free Trial
Free Version
Pricing Details
No price information available.
Free Trial
Free Version
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Vendor Details
Company Name
PrismML
Founded
2026
Country
United States
Website
prismml.com
Vendor Details
Company Name
Microsoft AI
Country
United States
Website
microsoft.ai/