polaris-audio/polaris-asr-medium
Speech recognition770M paramssafetensorsLicense: apache-2.0endeid#asr#timestamps
Updated Sep 23, 2026
Part of the Open Pace seed catalog. Weights are listed for the preview and become downloadable at launch.
Overview
Speech recognition with word-level timestamps, robust to background noise and phone-quality audio.
polaris-audio/polaris-asr-medium is a 770M-parameter speech recognition model, published in safetensors under the apache-2.0 license.
Intended use
- Speech recognition in en, de, id.
- Research, prototypes and products that keep a person in the loop.
- Fine-tuning as a starting point for a narrower task.
How to use
# Command-line client (planned; the shape may change)
pace pull polaris-audio/polaris-asr-medium
# Python (planned)
from openpace import load
model = load("polaris-audio/polaris-asr-medium")Limitations
- Heavy accents, overlapping speakers and very noisy audio lower accuracy.
- Never use a synthetic voice to impersonate a real person.
License
Released under apache-2.0. Read the license file before using the weights commercially.