Jul 8, 2023 | Read time 4 min

YouTube’s Captions Represent the Direct Need for Speech-to-Text Innovation

YouTube’s automated captioning service is notoriously unreliable and represents the dire need for innovation within the speech-to-text industry. Find out what we’re doing about it.
YouTube’s Captions Represent the Dire Need for Speech-to-Text Innovation
Benedetta Cevoli
Benedetta CevoliSenior Machine Learning Engineer
Carousel slide image
Product

Introducing Agent STT: built for the words that cost you the call

Speechmatics
SpeechmaticsEditorial Team
Carousel slide image
Use Cases

Tresic selects Speechmatics to power the speech layer of its conversation intelligence platform

Speechmatics
SpeechmaticsEditorial Team
[alt: Illustration representing multilingual code-switching for the Speechmatics Melia 1 speech-to-text model.]
Technical

Melia 1 leads speech-to-text code-switching in Arabic, Mandarin and Tamil

Speechmatics
SpeechmaticsEditorial Team
[alt: Dark grid background with a circular symbol on the left and a pixelated "K" on the right connected by a cyan line.]
Product

Speechmatics & LiveKit Inference target the accuracy gap that breaks voice agents in production

Speechmatics
SpeechmaticsEditorial Team
Carousel slide image
Product

Speechmatics on Zapier: No-Code Speech-to-Text Automation

Speechmatics
SpeechmaticsEditorial Team
[alt: Medical model header asset]
Product

Speechmatics launches Medical Model for real-time clinical transcription

Speechmatics
SpeechmaticsEditorial team