Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
Mercury 2 represents a groundbreaking advancement in reasoning models, specifically designed for real-time voice interaction as it can quickly answer phone calls. Unlike traditional autoregressive models that leave callers in silence while generating responses one token at a time, Mercury 2 employs a diffusion large language model architecture capable of producing over 1000 tokens per second with standard NVIDIA GPUs. This remarkable speed allows it to complete a full reasoning process and begin speaking within a timeframe that aligns with natural conversational flow, effectively shortening the typical wait time from several seconds to approximately 300 milliseconds. The operational mechanism of Mercury models involves transforming clear text into noise, after which a conventional Transformer is trained to reverse this transformation and predict the original text across all positions at once. By utilizing a denoising approach that engages multiple tokens simultaneously, generation becomes more efficient, enabling speeds akin to custom silicon on NVIDIA H100s while improving responsiveness in voice applications. As a result, Mercury 2 not only enhances user experience but also sets a new standard for interactive voice technologies.
Description
NVIDIA Alpamayo 2 Super stands as a pioneering open model tailored for robotaxis and autonomous vehicles, designed to navigate rare and intricate driving scenarios while generating decisions that developers can analyze, verify, and rely upon. Utilizing the foundations of NVIDIA Cosmos 3 Super Reasoner and enhanced through reinforcement learning, it merges commercial accessibility with the ability to handle multiple tasks related to autonomous driving. The model comprehensively analyzes full-surround camera input, integrating perspectives from the front, sides, and rear to adeptly manage lane changes, merges, unprotected turns, and complex intersections. In addressing each driving scenario, it can produce a planned trajectory for the vehicle, a chain-of-causation that elucidates the decision-making process, a meta-action such as yielding or stopping, and reasoning auto-labels for both training and validation purposes, along with visual question-answering outputs anchored in specific image regions. These interconnected outputs facilitate the correlation between the model's observations and the actions it undertakes, thereby enhancing transparency in autonomous decision-making. Additionally, this functionality supports developers in refining and optimizing the model's performance in real-world applications.
API Access
Has API
API Access
Has API
Integrations
Cerebras
GPT-4.1
Groq
Inception Labs
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI
Integrations
Cerebras
GPT-4.1
Groq
Inception Labs
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI
Pricing Details
No price information available.
Free Trial
Free Version
Pricing Details
No price information available.
Free Trial
Free Version
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Vendor Details
Company Name
Inception
Country
United States
Website
www.inceptionlabs.ai/blog/mercury-2-the-first-reasoning-model-fast-enough-to-pick-up-the-phone
Vendor Details
Company Name
NVIDIA
Founded
1993
Country
United States
Website
nvidia.com