Read original ↗
modelHugging FaceTrust 88 · LabPublished 2mo agoLive · 1mo ago0 graph score

Whisper-Lite

A compact speech-to-text model for on-device use.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

Related to (incoming)

Has model (incoming)

Covers (incoming)

Related across the graph

newsKyutai's Pocket TTS clones a voice from 5 seconds of audio, on CPU, under MIT. Benchmarked against Kokoro, Supertonic, and Inflect-Nano for Eng. TTSreporzru/nightingalerepoattevon-llc/OpenTranscriberepohuggingface/speech-to-speechrepoOpen-Less/openlesspaperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpaperComparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case StudyrepoRYOITABASHI/ShellypaperFreyaTTS Technical Reportrepomirkobozzetto/flowflowpaperLuxEmo: Expressive Text-to-Speech Corpus for LuxembourgishpaperWordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTSrepoamplitudesoldierheed/AI-Voice-Changer-Real-Time-Desktoprepotover0314-w/opentypelessrepoPedal-Intelligence/saypi-userscriptpaperUnlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction TuningrepoNotYuSheng/MeetMemonewsText-to-Speech AI Edits Single Words Mid-Recording: ViiTorVoice Goes Open Source - Tech Timesrepolgy1027/matrix-live-diarizerpaperAudio-Native Speech Recognition with a Frozen Discrete-Diffusion Language ModelrepoBryceWG/BiBi-Keyboardrepoapp-vox/voxpaperSIMAX: A Scalable and Interpretable Framework for Multi-Fidelity and Annotated Clinician-Patient Dialogue SimulationnewsShow HN: Orate – On-device neural text-to-speech queue for MacpaperAdapting Foundation ASR Models to Dysarthric Speech: A Case StudynewsApple's new SpeechAnalyzer API, benchmarked against Whisper and its predecessorpaperNAVER LABS Europe Submission to the Instruction-following 2026 Short TracknewsVocalinux 0.14 Beta Released For Offline Voice Dictation / Speech-To-Text On Linux - PhoronixnewsSeven on-device large language models for mobile phones have received regulatory approval! Apple Intelligence is among them, with support from Alibaba and Baidu. - MoomoopaperA Geometric Perspective on Composable Emotion Steering in Text-to-Speech Models

Topics