Discover and install AI models from our curated collection
IBM Granite 4.2 3B is a compact multilingual reasoning model for chat, coding, long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.
Links
Tags
IBM Granite 4.2 3B in the higher-fidelity Q8_0 GGUF format. It is a compact multilingual reasoning model for chat, coding, and tool use.
Links
Tags
IBM Granite 4.2 8B is a multilingual reasoning model for chat, coding, long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.
Links
Tags
IBM Granite 4.2 8B in the higher-fidelity Q8_0 GGUF format. It is a multilingual reasoning model for chat, coding, and tool use.
Links
Tags
IBM Granite 4.2 30B is the family's flagship multilingual reasoning model for chat, coding, long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.
Links
Tags
IBM Granite 4.2 30B in the higher-fidelity Q8_0 GGUF format. It is the family's flagship multilingual reasoning model for chat, coding, and tool use.
Links
Tags
Repository: localaiLicense: apache-2.0
Granite 4.2 3B is IBM's compact dense reasoning model for code generation, tool calling, agentic workflows, multilingual chat, and long-context tasks. This entry serves the bfloat16 safetensors with vLLM and supports a 128K-token context. It is the smallest fallback in a family that also offers the higher-capacity 8B and 30B checkpoints as variants.
Links
Tags
Repository: localaiLicense: apache-2.0
Granite 4.2 8B is IBM's mid-sized dense reasoning model for code generation, tool calling, agentic workflows, multilingual chat, and long-context tasks. This entry serves the higher-capacity bfloat16 safetensors with vLLM and supports a 128K-token context.
Links
Tags
Repository: localaiLicense: apache-2.0
Granite 4.2 30B is IBM's largest dense Granite 4.2 reasoning model for code generation, tool calling, agentic workflows, multilingual chat, and long-context tasks. This entry serves the bfloat16 safetensors with vLLM and supports a 128K-token context.
Links
Tags