Model profile · live graphModel

Whisper-Lite

A compact speech-to-text model for on-device use.

via Hugging Face
0Graph score
30Connections
11Papers
13Repos
6News

News6

Repos13

Papers11

paperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimod

paperComparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case Study

In our goal to develop personalised dysarthric speech recognition (DSR) models, this study compared the recognition perf

paperFreyaTTS Technical Report

We introduce Freya-TTS, a compact, tokenizer-free, Turkish-first text-to-speech model designed for highly reliable and e

paperLuxEmo: Expressive Text-to-Speech Corpus for Luxembourgish

State-of-the-art speech datasets predominantly focus on widely spoken languages, often overlooking low-resource language

paperWordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTS

While recent Large Language Model (LLM)-based Text-to-Speech (TTS) systems have achieved remarkable naturalness, they pr

paperUnlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction Tuning

Instruction tuning for speech language models (SLMs) is substantially more challenging than for text-based large languag

paperAudio-Native Speech Recognition with a Frozen Discrete-Diffusion Language Model

Automatic speech recognition is dominated by autoregressive decoders that emit one token at a time. We ask whether a dis

paperSIMAX: A Scalable and Interpretable Framework for Multi-Fidelity and Annotated Clinician-Patient Dialogue Simulation

Background. The widespread deployment of ambient digital scribes is driving large-scale capture of clinician-patient dia

Topics