prince-canuma's picture
Upload folder using huggingface_hub
5ed3098 verified
|
Raw History Blame Contribute Delete
1.17 kB
metadata
license: apache-2.0
language:
  - multilingual
  - en
  - fr
  - de
  - es
  - pt
  - ja
base_model:
  - ibm-granite/granite-4.0-1b-base
library_name: mlx-audio
tags:
  - mlx
  - speech-to-text
  - speech
  - transcription
  - asr
  - stt
  - mlx-audio

mlx-community/granite-4.0-1b-speech-8bit

This model was converted to MLX format from ibm-granite/granite-4.0-1b-speech using mlx-audio version 0.4.0.

Refer to the original model card for more details on the model.

Use with mlx-audio

pip install -U mlx-audio

CLI Example:

python -m mlx_audio.stt.generate --model mlx-community/granite-4.0-1b-speech-8bit --audio "audio.wav"

Python Example:

from mlx_audio.stt.utils import load_model
from mlx_audio.stt.generate import generate_transcription

model = load_model("mlx-community/granite-4.0-1b-speech-8bit")
transcription = generate_transcription(
    model=model,
    audio_path="path_to_audio.wav",
    output_path="path_to_output.txt",
    format="txt",
    verbose=True,
)
print(transcription.text)