Discover and install AI models from our curated collection
Repository: localaiLicense: s1-mini-license

S1-mini by Superwhisper is a 0.6B English text normalizer for raw speech transcripts. It removes fillers and false starts, restores punctuation and capitalization, and formats spoken numbers, dates, currency, and email addresses as written text. This default entry uses the publisher's 462 MB Q4_K_M GGUF and greedy decoding. A higher-fidelity F16 model is available as a variant. Prefix the transcript with the styling, structure, and context control line documented on the model page.
Links
Tags

S1-mini by Superwhisper in the publisher's 1.4 GB F16 GGUF format. This variant preserves full model fidelity for hosts with enough memory.
Links
Tags
Repository: localaiLicense: apache-2.0
Supra2-100M-Instruct is a compact English chat model trained from scratch by SupraLabs on the Qwen3 architecture. It has 100 million parameters, a 2,048-token context window, and is intended for lightweight experiments and constrained edge deployments. This entry uses the publisher's official F16 GGUF build.
Links
Tags
Repository: localaiLicense: apache-2.0
DFM Mimir is an Apache-2.0, instruction-tuned HRM-Text model from Danish Foundation Models. It has about 1 billion parameters and a 4,096-token context window. The model focuses on Danish and English chat, reasoning, mathematics, and code generation, and uses only permissible post-training data. This entry serves the official BF16 safetensors checkpoint with vLLM.
Links
Tags
Repository: localaiLicense: apache-2.0
Mixedbread's mxbai-embed-large-v1 is a 335M-parameter English BERT embedding model for retrieval, semantic search, and RAG. It produces 1,024-dimensional embeddings and supports sequences up to 512 tokens. Prefix retrieval queries with `Represent this sentence for searching relevant passages: `. This entry uses the balanced Q4_K_M GGUF.
Links
Tags
Mixedbread's mxbai-embed-large-v1 in the higher-fidelity Q8_0 GGUF format. This 335M-parameter English BERT model produces 1,024-dimensional embeddings for retrieval, semantic search, and RAG.
Links
Tags
Mixedbread's mxbai-embed-large-v1 in the official full-precision F16 GGUF format. This 335M-parameter English BERT model produces 1,024-dimensional embeddings for retrieval, semantic search, and RAG.
Links
Tags
Repository: localaiLicense: apache-2.0
Streaming English ASR: sherpa-onnx zipformer transducer (int8, chunk-16 left-128). Low-latency real-time transcription with endpoint detection via sherpa-onnx's online recognizer. English-only; for multilingual offline ASR see omnilingual-0.3b-ctc-q8-sherpa.
Links
Tags
VITS-LJS English single-speaker TTS served through the sherpa-onnx backend. Trained on the LJSpeech corpus at 22.05 kHz. Pairs with the sherpa-onnx ASR entries for round-trip audio pipelines.
Links
Tags
English (en_US) single-speaker Piper VITS voice "amy" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
English (en_GB) single-speaker Piper VITS voice "alan" (low quality, 16 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
English (en_GB) single-speaker Piper VITS voice "alan" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
English (en_GB) single-speaker Piper VITS voice "alba" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
Repository: localaiLicense: cc-by-4.0
English (en_GB) multi-speaker (12 voices) Piper VITS voice "aru" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Pick a speaker with the numeric voice/speaker id.
Links
Tags
English (en_GB) single-speaker Piper VITS voice "cori" (high quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
English (en_GB) single-speaker Piper VITS voice "cori" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
Repository: localaiLicense: cc-by-nc-sa-4.0
English (en_GB) single-speaker Piper VITS voice "dii" (high quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Non-commercial use only (CC BY-NC-SA 4.0).
Links
Tags
English (en_GB) single-speaker Piper VITS voice "jenny_dioco" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
Repository: localaiLicense: cc-by-nc-sa-4.0
English (en_GB) single-speaker Piper VITS voice "miro" (high quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Non-commercial use only (CC BY-NC-SA 4.0).
Links
Tags
Repository: localaiLicense: cc-by-sa-4.0
English (en_GB) single-speaker Piper VITS voice "northern_english_male" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.
Links
Tags
Repository: localaiLicense: cc-by-nc-sa-4.0
English (en_GB) multi-speaker (4 voices) Piper VITS voice "semaine" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Pick a speaker with the numeric voice/speaker id. Non-commercial use only (CC BY-NC-SA 4.0).
Links
Tags