IBM Granite 4.2 3B is a compact multilingual reasoning model for chat,
coding, long-context tasks, and tool use. This entry uses the Q4_K_M
GGUF; a higher-fidelity Q8_0 build is available as a variant.
Repository: localaiLicense: apache-2.0
granite-4.2-3b-q4
IBM Granite 4.2 3B is a compact multilingual reasoning model for chat,
coding, long-context tasks, and tool use. This entry uses the Q4_K_M
GGUF; a higher-fidelity Q8_0 build is available as a variant.
IBM Granite 4.2 8B is a multilingual reasoning model for chat, coding,
long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a
higher-fidelity Q8_0 build is available as a variant.
Repository: localaiLicense: apache-2.0
granite-4.2-8b-q4
IBM Granite 4.2 8B is a multilingual reasoning model for chat, coding,
long-context tasks, and tool use. This entry uses the Q4_K_M GGUF; a
higher-fidelity Q8_0 build is available as a variant.
IBM Granite 4.2 30B is the family's flagship multilingual reasoning model
for chat, coding, long-context tasks, and tool use. This entry uses the
Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.
Repository: localaiLicense: apache-2.0
granite-4.2-30b-q4
IBM Granite 4.2 30B is the family's flagship multilingual reasoning model
for chat, coding, long-context tasks, and tool use. This entry uses the
Q4_K_M GGUF; a higher-fidelity Q8_0 build is available as a variant.
Hy-MT2-1.8B is Tencent's compact multilingual translation model. It
follows translation instructions across 33 languages and supports tasks
such as terminology control, style transfer, and structure-preserving
translation.
This default entry uses the 1.1 GB Q4_K_M GGUF. A higher-quality Q8_0
model is available as a variant.
Repository: localaiLicense: apache-2.0
hy-mt2-1.8b-q4
Hy-MT2-1.8B is Tencent's compact multilingual translation model. It
follows translation instructions across 33 languages and supports tasks
such as terminology control, style transfer, and structure-preserving
translation.
This default entry uses the 1.1 GB Q4_K_M GGUF. A higher-quality Q8_0
model is available as a variant.
Ling-3.0-flash is InclusionAI's MIT-licensed hybrid reasoning MoE model
with 124B total parameters and 5.5B active parameters per token. It targets
coding, deep research, instruction following, and agentic workflows with a
native 256K-token context window.
This default entry uses the 36.5 GB AD-IQ1_M GGUF. A higher-quality
44.7 GB AD-IQ2_XS model is available as a variant.
Repository: localaiLicense: mit
ling-3.0-flash-iq1
Ling-3.0-flash is InclusionAI's MIT-licensed hybrid reasoning MoE model
with 124B total parameters and 5.5B active parameters per token. It targets
coding, deep research, instruction following, and agentic workflows with a
native 256K-token context window.
This default entry uses the 36.5 GB AD-IQ1_M GGUF. A higher-quality
44.7 GB AD-IQ2_XS model is available as a variant.
Carbon-3B is Hugging Face's 3B-parameter genomic foundation model for DNA
and RNA sequence generation, recovery, variant-effect prediction, and
motif-perturbation analysis. It supports 32,768 tokens natively and uses a
hybrid tokenizer with 6-mer DNA tokens.
This default entry uses the Q4_K_M GGUF. A higher-quality Q8_0 build is
available as a variant. Prefix DNA sequences with `` and use uppercase
A, C, G, and T characters in groups of six.
Repository: localaiLicense: apache-2.0
carbon-3b-q4
Carbon-3B is Hugging Face's 3B-parameter genomic foundation model for DNA
and RNA sequence generation, recovery, variant-effect prediction, and
motif-perturbation analysis. It supports 32,768 tokens natively and uses a
hybrid tokenizer with 6-mer DNA tokens.
This default entry uses the Q4_K_M GGUF. A higher-quality Q8_0 build is
available as a variant. Prefix DNA sequences with `` and use uppercase
A, C, G, and T characters in groups of six.
Carbon-8B is the largest model in Hugging Face's Carbon family of genomic
foundation models. It targets DNA and RNA sequence generation, recovery,
variant-effect prediction, and motif-perturbation analysis with a native
context length of 32,768 hybrid 6-mer DNA tokens.
This default entry uses the Q4_K_M GGUF. A higher-quality Q8_0 build is
available as a variant. Prefix DNA sequences with `` and use uppercase
A, C, G, and T characters in groups of six.
Repository: localaiLicense: apache-2.0
carbon-8b-q4
Carbon-8B is the largest model in Hugging Face's Carbon family of genomic
foundation models. It targets DNA and RNA sequence generation, recovery,
variant-effect prediction, and motif-perturbation analysis with a native
context length of 32,768 hybrid 6-mer DNA tokens.
This default entry uses the Q4_K_M GGUF. A higher-quality Q8_0 build is
available as a variant. Prefix DNA sequences with `` and use uppercase
A, C, G, and T characters in groups of six.
Ornith-1.0-9B is an MIT-licensed Qwen3.5 model from Ornith AI for
agentic coding, reasoning, repository-level software tasks, and tool use.
It supports text and image input with a context window of 262K tokens.
This default entry uses the Q4_K_M GGUF and F16 vision projector. A
higher-quality Q8_0 model is available as a variant.
Repository: localaiLicense: mit
ornith-1.0-9b-q4
Ornith-1.0-9B is an MIT-licensed Qwen3.5 model from Ornith AI for
agentic coding, reasoning, repository-level software tasks, and tool use.
It supports text and image input with a context window of 262K tokens.
This default entry uses the Q4_K_M GGUF and F16 vision projector. A
higher-quality Q8_0 model is available as a variant.
Ornith-1.5-9B is an MIT-licensed Qwen3.5 model from Ornith AI for
agentic coding, reasoning, repository-level software tasks, and tool use.
It supports text and image input with a context window of 262K tokens.
This default entry uses the Q4_K_M GGUF and BF16 vision projector. A
higher-quality Q8_0 model is available as a variant.
Repository: localaiLicense: mit
ornith-1.5-9b-q4
Ornith-1.5-9B is an MIT-licensed Qwen3.5 model from Ornith AI for
agentic coding, reasoning, repository-level software tasks, and tool use.
It supports text and image input with a context window of 262K tokens.
This default entry uses the Q4_K_M GGUF and BF16 vision projector. A
higher-quality Q8_0 model is available as a variant.
Ornith-1.5-35B-A3B is an MIT-licensed Qwen3.5 mixture-of-experts model
from Ornith AI for agentic coding, reasoning, repository-level software
tasks, and tool use. It activates about 3B parameters per token and
supports text and image input with a context window of 262K tokens.
This default entry uses the Q4_K_M GGUF and BF16 vision projector. A
higher-quality Q8_0 model is available as a variant.
Repository: localaiLicense: mit
ornith-1.5-35b-a3b-q4
Ornith-1.5-35B-A3B is an MIT-licensed Qwen3.5 mixture-of-experts model
from Ornith AI for agentic coding, reasoning, repository-level software
tasks, and tool use. It activates about 3B parameters per token and
supports text and image input with a context window of 262K tokens.
This default entry uses the Q4_K_M GGUF and BF16 vision projector. A
higher-quality Q8_0 model is available as a variant.
Tiel-Coder-35B-A3B is a 35B-parameter mixture-of-experts model for coding,
reasoning, tool use, and vision tasks. This default entry uses the
Q4_K_XL GGUF and BF16 vision projector.
Repository: localaiLicense: mit
tiel-coder-35b-a3b-q4
Tiel-Coder-35B-A3B is a 35B-parameter mixture-of-experts model for coding,
reasoning, tool use, and vision tasks. This default entry uses the
Q4_K_XL GGUF and BF16 vision projector.