Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
Karlo serves as an innovative model designed to create images from textual descriptions. It enhances the impressive unCLIP architecture developed by OpenAI by improving the conventional super-resolution model, enabling it to capture complex details at an impressive resolution of 256px, while effectively reducing noise through a limited number of denoising iterations.
In developing Karlo, we undertook a comprehensive training regimen that began from the ground up, leveraging a substantial dataset of 115 million image-text pairs, which included COYO-100M, CC3M, and CC12M. For the Prior and Decoder sections, we utilized the advanced ViT-L/14 text encoder sourced from OpenAI's CLIP library. To boost performance, we implemented a notable alteration to the original unCLIP design; rather than using a trainable transformer in the decoder, we opted to incorporate the text encoder from ViT-L/14, thereby enhancing the model's capability. This strategic choice not only streamlined the architecture but also contributed to improved image quality and fidelity.
Description
Pixray is an innovative system designed for image generation that integrates earlier concepts, including Perception Engines which utilize image augmentation to iteratively refine images through an ensemble of classifiers. This system also incorporates CLIP-guided GAN techniques developed by Ryan Murdoch and Katherine Crowson, along with enhancements like CLIPDraw created by Kevin Frans. Furthermore, it employs effective methods for exploring latent space, derived from Sampling Generative Networks. Users can generate images based on text prompts using Pixray, with predictions executed on Nvidia T4 GPU hardware, typically completed in about seven minutes, although the actual time may fluctuate significantly depending on the specific inputs provided. In addition to its functionality, Pixray is available as both a Python library and a command-line tool, making it accessible for various applications. While Replicate allows users to utilize Pixray for free initially, a credit card is required after a certain period, with charges incurred by the second for the predictions made, and this cost varies according to the hardware used for running different models. As a result, users can select from a range of models, each optimized for distinct types of hardware, allowing for tailored performance based on their specific needs.
API Access
Has API
API Access
Has API
Integrations
B^ DISCOVER
B^ EDIT
Pricing Details
Free
Free Trial
Free Version
Pricing Details
$0.0002 per second
Free Trial
Free Version
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Vendor Details
Company Name
Kakao Brain
Founded
2017
Country
South Korea
Website
github.com/kakaobrain/karlo
Vendor Details
Company Name
Replicate
Founded
2019
Country
United States
Website
replicate.com/pixray/text2image