Pass IndexThe State of AISign in

Open-weight transcription models

63 in the catalogue today. Every one with what it costs, who sells it and where it stands.

Models that turn recorded speech into text. They are metered by the minute or the second of audio, so cost follows the length of the recording and not the difficulty of it. What separates them is languages covered, whether they mark who is speaking, and whether they run in real time or only on a finished file — a model that is excellent on a podcast may be unusable on a live call.

Weights published under a licence that lets you run them where you like and sell what you build — Apache 2.0, MIT and their kin impose little beyond keeping the notice. This is the list to start from if you need the model on your own hardware, in your own region, or simply want no third party between you and it. The trade is that hosting is now your problem, and the prices shown beside each one are what somebody else charges to do it for you.

Wider

Open-weight modelsTranscription models