Klang Releases Open Swedish AI Model for Real-Time Transcription

Article Featured Image

Klang is releasing Pianissimo, a speech-to-text model trained for Swedish as open source, freely available for others to use and build on.

Pianissimo builds on Nvidia Parakeet, an open speech-to-text model for transcription, which Klang has optimized for Swedish and Swedish dialects.

Klang's transcription model reaches around 3,600 times real time, which means it can process roughly an hour of audio in one second. That speed makes it possible to transcribe speech continuously during a conversation.

"We're building more and more technology that needs real-time transcription ourselves, but we've lacked an open model adapted for Swedish that combines high quality with the speed we need. The fact that Pianissimo runs on ordinary consumer hardware also opens up entirely new kinds of services. Now we're looking forward to seeing what others build on the model, said Mattias Fält, head of AI, co-founder, and lead for Pianissimo at Klang AI, in a statement.

Pianissimo is the first of several AI models that Klang plans to release openly. Klang is also planning versions for more languages, starting with Danish and Norwegian.