Group: Modeljars HuggingFace
Sort by:Popular

1.Qwen3 0.6B GGUF Q4_01 usages

org.modeljars.huggingface » ggml-org.qwen3-0.6b-gguf.q4_0 Apache

Small Apache-2.0 Qwen3 text-generation fixture used by the pure-Java models backend integration tests.
Last Release on Aug 2, 2026
Pinned local-inference GGUF artifact for granite-4.1-3b, verified by immutable revision, byte size, and SHA-256.
Last Release on Sep 17, 2026

3.Qwen3 1.7B GGUF Q8_01 usages

org.modeljars.huggingface » qwen.qwen3-1.7b-gguf.q8_0 Apache

Apache-2.0 Qwen3 model qualified for pure-Java chat and generative tool calling, including complete Spring AI and LangChain4j tool loops.
Last Release on Sep 4, 2026
Pinned local-inference GGUF artifact for DeepSeek-R1-Distill-Qwen-1.5B, verified by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026
Pinned local-inference GGUF artifact for google gemma-3-1b-it, verified by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026
Pinned GGUF embedding artifact for granite-embedding-107m-multilingual, verified by immutable revision, byte size, and SHA-256.
Last Release on Sep 2, 2026
Pinned local-inference GGUF artifact for Llama-3.2-1B-Instruct, verified by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026
Pinned local-inference GGUF artifact for Llama-3.2-3B-Instruct, verified by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026
Apache-2.0 Qwen2.5 math specialist for English and Chinese chain-of-thought and tool-integrated reasoning.
Last Release on Aug 2, 2026
Apache-2.0 Yi-Coder 1.5B Chat imatrix GGUF artifact pinned by immutable revision, byte size, and SHA-256.
Last Release on Aug 5, 2026
Compact 43.6M-parameter Needle 2 model for local tool selection, structured tool calls, argument extraction, and calibrated refusal.
Last Release on Sep 4, 2026
Compact BERT cross-encoder reranker qualified against ONNX logits and an independent implementation of the same quantized artifact.
Last Release on Sep 7, 2026
Meta's 1.26B-parameter, 272M-active on-device mixture-of-experts model, executed directly from its packed group-32 INT4 Safetensors checkpoint by the pure-Java Models backend.
Last Release on Sep 4, 2026
Pinned GGUF embedding artifact for embeddinggemma-300M, verified by immutable revision, byte size, and SHA-256.
Last Release on Aug 11, 2026
Pinned local-inference GGUF artifact for gemma-4-26B-A4B-it, verified by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026

16.SmolLM3 3B GGUF Q4_K_M

org.modeljars.huggingface » ggml-org.smollm3-3b-gguf.q4_k_m Apache

Officially maintained Apache-2.0 SmolLM3 3B artifact for compact long-context reasoning and chat.
Last Release on Sep 4, 2026
Pinned local-inference GGUF artifact for Indian-Legal-Qwen2.5-3B, verified by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026
First-party H2O Danube2 1.8B Chat GGUF artifact pinned by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026
First-party H2O Danube3 500M Chat GGUF artifact pinned by immutable revision, byte size, and SHA-256.
Last Release on Aug 5, 2026
First-party SmolLM2 1.7B Instruct GGUF artifact pinned by immutable revision, byte size, and SHA-256.
Last Release on Aug 2, 2026