togatogah/jinen-v2-small.gguf
togatogah/jinen-v2-small.gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
214,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
togatogah/jinen-v2-small.gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ggml-org/gemma-4-E2B-it-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ai-forever/FRIDA is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
lued/Qwen3.8-27B-INT8-W8A16-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
google/timesfm-3.0-pytorch is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
optimum-intel-internal-testing/tiny-random-qwen3-omni is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
jinaai/jina-embeddings-v5-omni-nano is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
jinaai/jina-embeddings-v5-omni-small-retrieval is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
amd/Qwen3.8-27B-Quark-AWQ-INT4-W4A16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
bartowski/Ornith-1.5-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
rahul7star/gemma-gguf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-E2B-it-unsloth-bnb-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ibm-research/MoLFormer-XL-both-10pct is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
emhltbkars/xxx is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
trl-internal-testing/tiny-MuseGlimmerForConditionalGeneration is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
nvidia/Kimi-K2.7-Code-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Arki05/Qwen3.8-27B-GGUF-shards is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
RedHatAI/Llama-3.2-1B-Instruct-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
Abiray/LTX-2.5-Distilled-GGUF is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
numind/NuExtract3 is a image to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
openbmb/MiniCPM5-1B-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mlx-community/Devstral-Small-2-24B-Instruct-2512-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ornith-ai/Ornith-1.5-35B-A3B-MLX-6bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
orcarouter/Qwen3.8-Flash-Next-Uncensored-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.