Model Gallery

Discover and install AI models from our curated collection

66 models available
1 repositories
Documentation

Find Your Perfect Model

Filter by Model Type

Browse by Tags

s1-mini-q4
S1-mini by Superwhisper is a 0.6B English text normalizer for raw speech transcripts. It removes fillers and false starts, restores punctuation and capitalization, and formats spoken numbers, dates, currency, and email addresses as written text. This default entry uses the publisher's 462 MB Q4_K_M GGUF and greedy decoding. A higher-fidelity F16 model is available as a variant. Prefix the transcript with the styling, structure, and context control line documented on the model page.

Repository: localaiLicense: s1-mini-license

s1-mini-f16
S1-mini by Superwhisper in the publisher's 1.4 GB F16 GGUF format. This variant preserves full model fidelity for hosts with enough memory.

Repository: localaiLicense: s1-mini-license

supra2-100m-instruct
Supra2-100M-Instruct is a compact English chat model trained from scratch by SupraLabs on the Qwen3 architecture. It has 100 million parameters, a 2,048-token context window, and is intended for lightweight experiments and constrained edge deployments. This entry uses the publisher's official F16 GGUF build.

Repository: localaiLicense: apache-2.0

dfm-mimir:vllm
DFM Mimir is an Apache-2.0, instruction-tuned HRM-Text model from Danish Foundation Models. It has about 1 billion parameters and a 4,096-token context window. The model focuses on Danish and English chat, reasoning, mathematics, and code generation, and uses only permissible post-training data. This entry serves the official BF16 safetensors checkpoint with vLLM.

Repository: localaiLicense: apache-2.0

mxbai-embed-large-v1-q4
Mixedbread's mxbai-embed-large-v1 is a 335M-parameter English BERT embedding model for retrieval, semantic search, and RAG. It produces 1,024-dimensional embeddings and supports sequences up to 512 tokens. Prefix retrieval queries with `Represent this sentence for searching relevant passages: `. This entry uses the balanced Q4_K_M GGUF.

Repository: localaiLicense: apache-2.0

mxbai-embed-large-v1-q8
Mixedbread's mxbai-embed-large-v1 in the higher-fidelity Q8_0 GGUF format. This 335M-parameter English BERT model produces 1,024-dimensional embeddings for retrieval, semantic search, and RAG.

Repository: localaiLicense: apache-2.0

mxbai-embed-large-v1-f16
Mixedbread's mxbai-embed-large-v1 in the official full-precision F16 GGUF format. This 335M-parameter English BERT model produces 1,024-dimensional embeddings for retrieval, semantic search, and RAG.

Repository: localaiLicense: apache-2.0

streaming-zipformer-en-sherpa
Streaming English ASR: sherpa-onnx zipformer transducer (int8, chunk-16 left-128). Low-latency real-time transcription with endpoint detection via sherpa-onnx's online recognizer. English-only; for multilingual offline ASR see omnilingual-0.3b-ctc-q8-sherpa.

Repository: localaiLicense: apache-2.0

vits-ljs-sherpa
VITS-LJS English single-speaker TTS served through the sherpa-onnx backend. Trained on the LJSpeech corpus at 22.05 kHz. Pairs with the sherpa-onnx ASR entries for round-trip audio pipelines.

Repository: localaiLicense: apache-2.0

vits-piper-en_US-amy-sherpa
English (en_US) single-speaker Piper VITS voice "amy" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: other

vits-piper-en_GB-alan-low-sherpa
English (en_GB) single-speaker Piper VITS voice "alan" (low quality, 16 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: other

vits-piper-en_GB-alan-medium-sherpa
English (en_GB) single-speaker Piper VITS voice "alan" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: other

vits-piper-en_GB-alba-medium-sherpa
English (en_GB) single-speaker Piper VITS voice "alba" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: cc-by-4.0

vits-piper-en_GB-aru-medium-sherpa
English (en_GB) multi-speaker (12 voices) Piper VITS voice "aru" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Pick a speaker with the numeric voice/speaker id.

Repository: localaiLicense: cc-by-4.0

vits-piper-en_GB-cori-high-sherpa
English (en_GB) single-speaker Piper VITS voice "cori" (high quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: cc0-1.0

vits-piper-en_GB-cori-medium-sherpa
English (en_GB) single-speaker Piper VITS voice "cori" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: cc0-1.0

vits-piper-en_GB-dii-high-sherpa
English (en_GB) single-speaker Piper VITS voice "dii" (high quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Non-commercial use only (CC BY-NC-SA 4.0).

Repository: localaiLicense: cc-by-nc-sa-4.0

vits-piper-en_GB-jenny_dioco-medium-sherpa
English (en_GB) single-speaker Piper VITS voice "jenny_dioco" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: other

vits-piper-en_GB-miro-high-sherpa
English (en_GB) single-speaker Piper VITS voice "miro" (high quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Non-commercial use only (CC BY-NC-SA 4.0).

Repository: localaiLicense: cc-by-nc-sa-4.0

vits-piper-en_GB-northern_english_male-medium-sherpa
English (en_GB) single-speaker Piper VITS voice "northern_english_male" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data.

Repository: localaiLicense: cc-by-sa-4.0

vits-piper-en_GB-semaine-medium-sherpa
English (en_GB) multi-speaker (4 voices) Piper VITS voice "semaine" (medium quality, 22.05 kHz), served through the sherpa-onnx backend with native streaming TTS. Ships espeak-ng phonemization data. Pick a speaker with the numeric voice/speaker id. Non-commercial use only (CC BY-NC-SA 4.0).

Repository: localaiLicense: cc-by-nc-sa-4.0

Page 1 of many