An API powered by Google's AI technology allows you to accurately convert speech into text. You can accurately caption your content, provide a better user experience with products using voice commands, and gain insight from customer interactions to improve your service. Google's deep learning neural network algorithms are the most advanced in automatic speech recognition (ASR). Speech-to-Text allows for experimentation, creation, management, and customization of custom resources. You can deploy speech recognition wherever you need it, whether it's in the cloud using the API or on-premises using Speech-to-Text O-Prem. You can customize speech recognition to translate domain-specific terms or rare words. Automated conversion of spoken numbers into addresses, years and currencies. Our user interface makes it easy to experiment with your speech audio.
Learn more

Coevera is an AI-native CRM and sales process platform for B2B sales organizations of 15 to 2,500 quota-carrying sellers — companies with a defined sales process, sales managers accountable for pipeline, and deals complex enough that consultation rather than transaction wins them. Formerly Pipeliner CRM, built since 2011 on one idea: a sales process should be visible, teachable and repeatable. Intelligence is the default state of the system, not a premium add-on. The platform is industry-agnostic; our strongest results have come from manufacturing, financial services, insurance, software, professional services, energy and mining.
Sales leaders buy Coevera for a forecast they can defend. Sales operations teams buy it because it can be configured, changed and reported on without a dedicated administrator, outside consultants or a multi-year build. Reps use it because the pipeline is visual and the next step is obvious.
Voyager AI — predictive deal and pipeline guidance in context, with native MCP so Coevera data is available to the AI tools your team already uses.
Visual pipeline management — drag-and-drop pipelines, buying centers and relationship maps, so who is involved in a deal and what stage it is actually in are both visible.
Guided selling — your stages, required activities and qualification criteria enforced in the flow of work rather than in a document.
Automatizer — no-code workflow automation your own ops team builds and changes.
Reporting and forecasting built in — no BI licence required, with BI Feeder exporting to Tableau, Power BI and others when your analysts want the raw data.
Three editions at three price points — $85, $115 and $150 per user per month. Implementations run in weeks, not quarters. Served markets: United States, Canada, the Un
Learn more
Aiko
Aiko is an AI transcription app from Sindre Sorhus that helps users convert audio into text on macOS, iOS, and visionOS. The app is powered by OpenAI’s Whisper model and runs transcription locally on the user’s device. This on-device approach makes Aiko well suited for private meetings, lectures, interviews, voice notes, and sensitive recordings. On macOS, Aiko uses the Whisper large v2 model, while iOS uses the medium or small model depending on available device memory. The app supports Shortcuts, giving users flexible ways to record, transcribe, copy results, create subtitles, save text, or connect transcripts to other workflows. Users can transcribe from Finder on macOS, record and transcribe from iPhone shortcuts, or build custom workflows that pass transcripts into apps like Notes or ChatGPT. Aiko includes a free 14-day TestFlight trial with full app access and no auto-charges. Older macOS versions are available for users on macOS 13, 14, and 15. By combining local Whisper transcription, Apple platform support, privacy, and Shortcuts automation, Aiko gives users a simple way to turn speech into text.
Learn more
FluidVoice
FluidVoice is a free and open-source dictation application for macOS that combines local speech recognition with an on-device AI model known as Fluid-1, which enhances the quality of dictation. By using a single hotkey, users can dictate text into virtually any input field across various applications such as email, documents, chat interfaces, terminals, code editors, and more, with the text being displayed almost instantaneously. The application relies on local speech models that function offline, allowing for secure dictation without needing an internet connection, while optional AI post-processing can utilize Fluid Intelligence, OpenAI, Groq, or other custom providers. Fluid-1 improves initial dictation by refining rough entries, correcting formatting, capitalization, dates, names, and numbers, and it adjusts the tone according to the currently active application, all while preserving the speaker's intended meaning. Users have the flexibility to develop personalized prompts tailored to different applications, and with modes like Write Mode, Command Mode, and Direct Dictation, transitioning between tasks is seamless. Furthermore, FluidVoice is capable of supporting over 40 languages, leveraging various models such as Nemotron Speech 3.5, Parakeet Flash, Parakeet TDT versions 2 and 3, Cohere Transcribe, Apple Speech, and Whisper, thus catering to a diverse user base and enhancing accessibility in dictation across different linguistic backgrounds. This versatility makes FluidVoice an essential tool for those seeking effective and efficient dictation solutions.
Learn more