Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Description

SpeechCAT Professional is an advanced Computer-Aided Transcription (CAT) software created by AudioScribe, specifically designed for voice writers engaged in court reporting, captioning, and Communication Access real-time translation (CART). This software provides real-time speech-to-text functionality along with synchronized audio, accommodating up to five channels of superior digital recording. In addition, it incorporates robust job and case management capabilities, which enhance the organization and consolidation of various assignments. Tailored for official court reporters, SpeechCAT offers specialized features for managing consecutive cases effectively, including a courtroom functionality and a secure case feature that meets the rigorous data protection needs of military courts and grand jury settings. Furthermore, it is compatible with Dragon Professional Individual versions 14 and 15, as well as Dragon NaturallySpeaking Professional or Premium versions 13 and 12, ensuring flawless voice recognition performance. This integration allows users to streamline their workflow and improve transcription accuracy while handling complex cases.

Description

The gpt-4o-mini-realtime-preview model is a streamlined and economical variant of GPT-4o, specifically crafted for real-time interaction in both speech and text formats with minimal delay. It is capable of processing both audio and text inputs and outputs, facilitating “speech in, speech out” dialogue experiences through a consistent WebSocket or WebRTC connection. In contrast to its larger counterparts in the GPT-4o family, this model currently lacks support for image and structured output formats, concentrating solely on immediate voice and text applications. Developers have the ability to initiate a real-time session through the /realtime/sessions endpoint to acquire a temporary key, allowing them to stream user audio or text and receive immediate responses via the same connection. This model belongs to the early preview family (version 2024-12-17) and is primarily designed for testing purposes and gathering feedback, rather than handling extensive production workloads. The usage comes with certain rate limitations and may undergo changes during the preview phase. Its focus on audio and text modalities opens up possibilities for applications like conversational voice assistants, enhancing user interaction in a variety of settings. As technology evolves, further enhancements and features may be introduced to enrich user experiences.

API Access

Has API

API Access

Has API

Screenshots View All

Screenshots View All

Integrations

GPT-4o
OpenAI
WebRTC

Integrations

GPT-4o
OpenAI
WebRTC

Pricing Details

$4,650 one-time payment
Free Trial
Free Version

Pricing Details

$0.60 per input
Free Trial
Free Version

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Vendor Details

Company Name

AudioScribe

Founded

1997

Website

audioscribe.com/portfolio-item/speechcat/

Vendor Details

Company Name

OpenAI

Founded

2015

Country

United States

Website

platform.openai.com/docs/models/gpt-4o-mini-realtime-preview

Product Features

Product Features

Alternatives

MAXScribe Reviews

MAXScribe

Stenograph

Alternatives

Qwen3-Omni Reviews

Qwen3-Omni

Alibaba
InfraWare 360 Reviews

InfraWare 360

InfraWare