Model Gallery

Discover and install AI models from our curated collection

9 models available
1 repositories
Documentation

Find Your Perfect Model

Filter by Model Type

Browse by Tags

granite-4.2-3b-q4
IBM Granite 4.2 3B is a compact multilingual reasoning model for chat, coding, long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.

Repository: localaiLicense: apache-2.0

granite-4.2-3b-q8
IBM Granite 4.2 3B in the higher-fidelity Q8_0 GGUF format. It is a compact multilingual reasoning model for chat, coding, and tool use.

Repository: localaiLicense: apache-2.0

granite-4.2-8b-q4
IBM Granite 4.2 8B is a multilingual reasoning model for chat, coding, long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.

Repository: localaiLicense: apache-2.0

granite-4.2-8b-q8
IBM Granite 4.2 8B in the higher-fidelity Q8_0 GGUF format. It is a multilingual reasoning model for chat, coding, and tool use.

Repository: localaiLicense: apache-2.0

granite-4.2-30b-q4
IBM Granite 4.2 30B is the family's flagship multilingual reasoning model for chat, coding, long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.

Repository: localaiLicense: apache-2.0

granite-4.2-30b-q8
IBM Granite 4.2 30B in the higher-fidelity Q8_0 GGUF format. It is the family's flagship multilingual reasoning model for chat, coding, and tool use.

Repository: localaiLicense: apache-2.0

granite-4.2-3b:vllm
Granite 4.2 3B is IBM's compact dense reasoning model for code generation, tool calling, agentic workflows, multilingual chat, and long-context tasks. This entry serves the bfloat16 safetensors with vLLM and supports a 128K-token context. It is the smallest fallback in a family that also offers the higher-capacity 8B and 30B checkpoints as variants.

Repository: localaiLicense: apache-2.0

granite-4.2-8b:vllm
Granite 4.2 8B is IBM's mid-sized dense reasoning model for code generation, tool calling, agentic workflows, multilingual chat, and long-context tasks. This entry serves the higher-capacity bfloat16 safetensors with vLLM and supports a 128K-token context.

Repository: localaiLicense: apache-2.0

granite-4.2-30b:vllm
Granite 4.2 30B is IBM's largest dense Granite 4.2 reasoning model for code generation, tool calling, agentic workflows, multilingual chat, and long-context tasks. This entry serves the bfloat16 safetensors with vLLM and supports a 128K-token context.

Repository: localaiLicense: apache-2.0