OpenPace

polaris-audio/polaris-asr-medium

Speech recognition770M paramssafetensorsLicense: apache-2.0endeid#asr#timestamps
Updated Sep 23, 2026
Part of the Open Pace seed catalog. Weights are listed for the preview and become downloadable at launch.

Overview

Speech recognition with word-level timestamps, robust to background noise and phone-quality audio.

polaris-audio/polaris-asr-medium is a 770M-parameter speech recognition model, published in safetensors under the apache-2.0 license.

Intended use

  • Speech recognition in en, de, id.
  • Research, prototypes and products that keep a person in the loop.
  • Fine-tuning as a starting point for a narrower task.

How to use

# Command-line client (planned; the shape may change)
pace pull polaris-audio/polaris-asr-medium

# Python (planned)
from openpace import load
model = load("polaris-audio/polaris-asr-medium")

Limitations

  • Heavy accents, overlapping speakers and very noisy audio lower accuracy.
  • Never use a synthetic voice to impersonate a real person.

License

Released under apache-2.0. Read the license file before using the weights commercially.

More speech recognition