modelHugging FaceTrust 88 · LabPublished 2mo agoLive · 1mo ago0 graph score
Whisper-Lite
A compact speech-to-text model for on-device use.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- LinkedLinked via unknownamplitudesoldierheed/AI-Voice-Changer-Real-Time-Desktop →
- LinkedLinked via unknownComparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case Study →
- LinkedLinked via unknownAdapting Foundation ASR Models to Dysarthric Speech: A Case Study →
- LinkedLinked via unknownLuxEmo: Expressive Text-to-Speech Corpus for Luxembourgish →
- LinkedLinked via unknownA Geometric Perspective on Composable Emotion Steering in Text-to-Speech Models →
- PossiblePossibly related (embedding) · 52%attevon-llc/OpenTranscribe →
- PossiblePossibly related (embedding) · 49%Text-to-Speech AI Edits Single Words Mid-Recording: ViiTorVoice Goes Open Source - Tech Times →
Related to (incoming)
repoamplitudesoldierheed/AI-Voice-Changer-Real-Time-Desktoprepoattevon-llc/OpenTranscriberepoPedal-Intelligence/saypi-userscriptrepoRYOITABASHI/Shellyrepolgy1027/matrix-live-diarizerrepomirkobozzetto/flowflowrepohuggingface/speech-to-speechrepoBryceWG/BiBi-Keyboardreporzru/nightingalerepoOpen-Less/openlessrepoapp-vox/voxrepoNotYuSheng/MeetMemorepotover0314-w/opentypeless
Has model (incoming)
paperComparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case StudypaperSIMAX: A Scalable and Interpretable Framework for Multi-Fidelity and Annotated Clinician-Patient Dialogue SimulationpaperAdapting Foundation ASR Models to Dysarthric Speech: A Case StudypaperLuxEmo: Expressive Text-to-Speech Corpus for LuxembourgishpaperA Geometric Perspective on Composable Emotion Steering in Text-to-Speech ModelspaperNAVER LABS Europe Submission to the Instruction-following 2026 Short TrackpaperUnlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction TuningpaperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpaperWordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTSpaperFreyaTTS Technical ReportpaperAudio-Native Speech Recognition with a Frozen Discrete-Diffusion Language Model
Covers (incoming)
newsText-to-Speech AI Edits Single Words Mid-Recording: ViiTorVoice Goes Open Source - Tech TimesnewsKyutai's Pocket TTS clones a voice from 5 seconds of audio, on CPU, under MIT. Benchmarked against Kokoro, Supertonic, and Inflect-Nano for Eng. TTSnewsApple's new SpeechAnalyzer API, benchmarked against Whisper and its predecessornewsSeven on-device large language models for mobile phones have received regulatory approval! Apple Intelligence is among them, with support from Alibaba and Baidu. - MoomoonewsVocalinux 0.14 Beta Released For Offline Voice Dictation / Speech-To-Text On Linux - PhoronixnewsShow HN: Orate – On-device neural text-to-speech queue for Mac
Related across the graph
newsKyutai's Pocket TTS clones a voice from 5 seconds of audio, on CPU, under MIT. Benchmarked against Kokoro, Supertonic, and Inflect-Nano for Eng. TTSreporzru/nightingalerepoattevon-llc/OpenTranscriberepohuggingface/speech-to-speechrepoOpen-Less/openlesspaperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpaperComparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case StudyrepoRYOITABASHI/ShellypaperFreyaTTS Technical Reportrepomirkobozzetto/flowflowpaperLuxEmo: Expressive Text-to-Speech Corpus for LuxembourgishpaperWordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTSrepoamplitudesoldierheed/AI-Voice-Changer-Real-Time-Desktoprepotover0314-w/opentypelessrepoPedal-Intelligence/saypi-userscriptpaperUnlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction TuningrepoNotYuSheng/MeetMemonewsText-to-Speech AI Edits Single Words Mid-Recording: ViiTorVoice Goes Open Source - Tech Timesrepolgy1027/matrix-live-diarizerpaperAudio-Native Speech Recognition with a Frozen Discrete-Diffusion Language ModelrepoBryceWG/BiBi-Keyboardrepoapp-vox/voxpaperSIMAX: A Scalable and Interpretable Framework for Multi-Fidelity and Annotated Clinician-Patient Dialogue SimulationnewsShow HN: Orate – On-device neural text-to-speech queue for MacpaperAdapting Foundation ASR Models to Dysarthric Speech: A Case StudynewsApple's new SpeechAnalyzer API, benchmarked against Whisper and its predecessorpaperNAVER LABS Europe Submission to the Instruction-following 2026 Short TracknewsVocalinux 0.14 Beta Released For Offline Voice Dictation / Speech-To-Text On Linux - PhoronixnewsSeven on-device large language models for mobile phones have received regulatory approval! Apple Intelligence is among them, with support from Alibaba and Baidu. - MoomoopaperA Geometric Perspective on Composable Emotion Steering in Text-to-Speech Models
