Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Description

LFM2 represents an advanced series of on-device foundation models designed to provide a remarkably swift generative-AI experience across a diverse array of devices. By utilizing a novel hybrid architecture, it achieves decoding and pre-filling speeds that are up to twice as fast as those of similar models, while also enhancing training efficiency by as much as three times compared to its predecessor. These models offer a perfect equilibrium of quality, latency, and memory utilization suitable for embedded system deployment, facilitating real-time, on-device AI functionality in smartphones, laptops, vehicles, wearables, and various other platforms, which results in millisecond inference, device durability, and complete data sovereignty. LFM2 is offered in three configurations featuring 0.35 billion, 0.7 billion, and 1.2 billion parameters, showcasing benchmark results that surpass similarly scaled models in areas including knowledge recall, mathematics, multilingual instruction adherence, and conversational dialogue assessments. With these capabilities, LFM2 not only enhances user experience but also sets a new standard for on-device AI performance.

Description

Distil Labs enhances AI performance by substituting costly calls to advanced models with tailored small language models designed for specific tasks while ensuring the quality standards are upheld. By monitoring real production traffic and gathering traces from current LLM requests, it constructs an evaluation set to gain insights into actual workload behavior. Following this, the company creates and verifies synthetic training data, aligns the data distribution with the intended workload, and engages in supervised fine-tuning alongside reinforcement learning. The model is then quantized, and an optimized endpoint is established. The outcomes are systematically assessed against the existing model concerning accuracy, latency, and efficiency, providing teams with data to determine when to increase traffic. Ultimately, the OpenAI-compatible endpoint features a specialized small language model, prompt optimization, effective caching, and refined serving tailored for the specific application, ensuring maximum performance. This comprehensive approach allows organizations to maximize the potential of their AI implementations.

API Access

Has API

API Access

Has API

Screenshots View All

Screenshots View All

Integrations

GPT-5.4 nano
Gemini 2.0 Flash-Lite
Hugging Face
OpenAI
OpenRouter
Qwen3
Together AI

Integrations

GPT-5.4 nano
Gemini 2.0 Flash-Lite
Hugging Face
OpenAI
OpenRouter
Qwen3
Together AI

Pricing Details

No price information available.
Free Trial
Free Version

Pricing Details

$0.04 per 1M tokens
Free Trial
Free Version

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Vendor Details

Company Name

Liquid AI

Founded

2023

Country

United States

Website

www.liquid.ai/blog/liquid-foundation-models-v2-our-second-series-of-generative-ai-models

Vendor Details

Company Name

distil labs

Founded

2024

Country

Germany

Website

www.distillabs.ai/

Product Features

Alternatives

Seed2.0 Lite Reviews

Seed2.0 Lite

ByteDance

Alternatives

No Alternatives
Ministral 8B Reviews

Ministral 8B

Mistral AI
Ministral 3B Reviews

Ministral 3B

Mistral AI