AI Models Directory
Filter and search through cloud API and local open-weight AI models.
GPT-4o
Omnimodal LLM
GPT-4o by OpenAI - Omnimodal LLM for enterprise and developer workflows.
OpenAI o3
Reasoning Model
OpenAI o3 by OpenAI - Reasoning Model for enterprise and developer workflows.
Claude 3.5 Sonnet
Frontier LLM
Claude 3.5 Sonnet by Anthropic - Frontier LLM engineered for high safety, reasoning, and long context.
Gemini 2.0 Flash
Omnimodal LLM
Gemini 2.0 Flash - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemini 2.0 Pro
Omnimodal LLM
Gemini 2.0 Pro - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Llama 3.3 70B
Open Weights LLM
Llama 3.3 70B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Mistral Large 2
Open Weights LLM
Mistral Large 2 - European frontier open & commercial model by Mistral AI.
DeepSeek-R1
Reasoning Model
DeepSeek-R1 - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-R1-Zero
Reasoning Model
DeepSeek-R1-Zero - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-V3
MoE LLM
DeepSeek-V3 - DeepSeek open-source frontier model for math, reasoning, and code.
Qwen 2.5 72B
Open Weights LLM
Qwen 2.5 72B - Alibaba Qwen open multilingual model series.
SmolLM2 1.7B
Open Weights LLM
SmolLM2 1.7B - Open-source benchmarked model available on Hugging Face Hub.
BGE-M3
Embedding Model
BGE-M3 - Open-source benchmarked model available on Hugging Face Hub.
BioBERT
Embedding Model
BioBERT - Open-source benchmarked model available on Hugging Face Hub.
StarCoder 2 7B
Code LLM
StarCoder 2 7B - Open-source benchmarked model available on Hugging Face Hub.
InternLM 2.5 7B
Open Weights LLM
InternLM 2.5 7B - Open-source benchmarked model available on Hugging Face Hub.
Solar Pro 22B
Open Weights LLM
Solar Pro 22B - Open-source benchmarked model available on Hugging Face Hub.
Flux 1.1 Pro
Image Generation
Flux 1.1 Pro - Generative image model for digital creative workflows.
Stable Diffusion 3.5 Large
Image Generation
Stable Diffusion 3.5 Large - Generative image model for digital creative workflows.
Gemini 2.0 Flash-Lite
Omnimodal LLM
Gemini 2.0 Flash-Lite - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Llama 3.1 405B
Open Weights LLM
Llama 3.1 405B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Qwen 2.5 Coder 32B
Code LLM
Qwen 2.5 Coder 32B - Alibaba Qwen open multilingual model series.
Flux 1 Dev
Image Generation
Flux 1 Dev - Generative image model for digital creative workflows.
Flux 1 Schnell
Image Generation
Flux 1 Schnell - Generative image model for digital creative workflows.
GPT-4o mini
Lightweight LLM
GPT-4o mini by OpenAI - Lightweight LLM for enterprise and developer workflows.
OpenAI o3-mini
Reasoning Model
OpenAI o3-mini by OpenAI - Reasoning Model for enterprise and developer workflows.
Claude 3.5 Haiku
Lightweight LLM
Claude 3.5 Haiku by Anthropic - Lightweight LLM engineered for high safety, reasoning, and long context.
Gemini 1.5 Pro
Omnimodal LLM
Gemini 1.5 Pro - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemini 1.5 Flash-8B
Omnimodal LLM
Gemini 1.5 Flash-8B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemini 1.0 Ultra
Omnimodal LLM
Gemini 1.0 Ultra - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemma 2 9B
Open Weights LLM
Gemma 2 9B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemma 7B
Open Weights LLM
Gemma 7B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
CodeGemma 7B
Open Weights LLM
CodeGemma 7B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
RecurrentGemma 2B
Open Weights LLM
RecurrentGemma 2B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Imagen 2
Image Generation
Imagen 2 - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Veo 1
Video Generation
Veo 1 - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
AudioLM
Omnimodal LLM
AudioLM - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
PaLM 2 Bison
Omnimodal LLM
PaLM 2 Bison - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Llama 3.2 3B
Open Weights LLM
Llama 3.2 3B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 3.1 70B
Open Weights LLM
Llama 3.1 70B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 3 8B
Open Weights LLM
Llama 3 8B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 2 7B
Open Weights LLM
Llama 2 7B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Code Llama 13B
Code LLM
Code Llama 13B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama Guard 3 1B
Open Weights LLM
Llama Guard 3 1B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
SAM 1
Vision LLM
SAM 1 - Meta AI open source model engineered for high efficiency, vision, and edge performance.
AudioCraft
Audio Model
AudioCraft - Meta AI open source model engineered for high efficiency, vision, and edge performance.
EnCodec
Code LLM
EnCodec - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Mistral Medium
Open Weights LLM
Mistral Medium - European frontier open & commercial model by Mistral AI.
Mistral Small 2
Open Weights LLM
Mistral Small 2 - European frontier open & commercial model by Mistral AI.
Mistral 7B v0.3
Open Weights LLM
Mistral 7B v0.3 - European frontier open & commercial model by Mistral AI.
Mistral 7B v0.1
Open Weights LLM
Mistral 7B v0.1 - European frontier open & commercial model by Mistral AI.
Codestral Mamba
Code LLM
Codestral Mamba - European frontier open & commercial model by Mistral AI.
Pixtral Large
Vision LLM
Pixtral Large - European frontier open & commercial model by Mistral AI.
Mixtral 8x7B
MoE LLM
Mixtral 8x7B - European frontier open & commercial model by Mistral AI.
Mistral Embed 2
Open Weights LLM
Mistral Embed 2 - European frontier open & commercial model by Mistral AI.
DeepSeek-V2.5
MoE LLM
DeepSeek-V2.5 - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-V2
MoE LLM
DeepSeek-V2 - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek Coder V2
Code LLM
DeepSeek Coder V2 - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek Coder V2 Lite
Code LLM
DeepSeek Coder V2 Lite - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek Coder 33B
Code LLM
DeepSeek Coder 33B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek Coder 7B
Code LLM
DeepSeek Coder 7B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek Coder 1.3B
Code LLM
DeepSeek Coder 1.3B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek Math 7B
MoE LLM
DeepSeek Math 7B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek VL2
Vision LLM
DeepSeek VL2 - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek VL 7B
Vision LLM
DeepSeek VL 7B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-R1-Distill-Qwen-32B
Reasoning Model
DeepSeek-R1-Distill-Qwen-32B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-R1-Distill-Qwen-14B
Reasoning Model
DeepSeek-R1-Distill-Qwen-14B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-R1-Distill-Qwen-7B
Reasoning Model
DeepSeek-R1-Distill-Qwen-7B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-R1-Distill-Qwen-1.5B
Reasoning Model
DeepSeek-R1-Distill-Qwen-1.5B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-R1-Distill-Llama-70B
Reasoning Model
DeepSeek-R1-Distill-Llama-70B - DeepSeek open-source frontier model for math, reasoning, and code.
DeepSeek-R1-Distill-Llama-8B
Reasoning Model
DeepSeek-R1-Distill-Llama-8B - DeepSeek open-source frontier model for math, reasoning, and code.
Janus Pro 7B
Vision LLM
Janus Pro 7B - DeepSeek open-source frontier model for math, reasoning, and code.
Janus 1.3B
Vision LLM
Janus 1.3B - DeepSeek open-source frontier model for math, reasoning, and code.
Qwen 2.5 14B
Open Weights LLM
Qwen 2.5 14B - Alibaba Qwen open multilingual model series.
Qwen 2.5 3B
Open Weights LLM
Qwen 2.5 3B - Alibaba Qwen open multilingual model series.
Qwen 2.5 0.5B
Open Weights LLM
Qwen 2.5 0.5B - Alibaba Qwen open multilingual model series.
Qwen 2.5 Coder 14B
Code LLM
Qwen 2.5 Coder 14B - Alibaba Qwen open multilingual model series.
Qwen 2.5 Coder 3B
Code LLM
Qwen 2.5 Coder 3B - Alibaba Qwen open multilingual model series.
Qwen 2.5 Coder 0.5B
Code LLM
Qwen 2.5 Coder 0.5B - Alibaba Qwen open multilingual model series.
Qwen 2.5 Math 7B
Math Model
Qwen 2.5 Math 7B - Alibaba Qwen open multilingual model series.
Qwen 2 VL 7B
Vision LLM
Qwen 2 VL 7B - Alibaba Qwen open multilingual model series.
Qwen 2 72B
Open Weights LLM
Qwen 2 72B - Alibaba Qwen open multilingual model series.
Qwen 1.5 110B
Open Weights LLM
Qwen 1.5 110B - Alibaba Qwen open multilingual model series.
QPad
Open Weights LLM
QPad - Alibaba Qwen open multilingual model series.
Qwen-Agent
Open Weights LLM
Qwen-Agent - Alibaba Qwen open multilingual model series.
Granite 3.0 8B
Open Weights LLM
Granite 3.0 8B - Open-source benchmarked model available on Hugging Face Hub.
Phi-3.5 MoE
Open Weights LLM
Phi-3.5 MoE - Open-source benchmarked model available on Hugging Face Hub.
Nomic Vision v1.5
Vision LLM
Nomic Vision v1.5 - Open-source benchmarked model available on Hugging Face Hub.
All-mpnet-base-v2
Open Weights LLM
All-mpnet-base-v2 - Open-source benchmarked model available on Hugging Face Hub.
Longformer
Open Weights LLM
Longformer - Open-source benchmarked model available on Hugging Face Hub.
EasyOCR
Open Weights LLM
EasyOCR - Open-source benchmarked model available on Hugging Face Hub.
Yi 1.5 6B
Open Weights LLM
Yi 1.5 6B - Open-source benchmarked model available on Hugging Face Hub.
GLM-4 9B
Open Weights LLM
GLM-4 9B - Open-source benchmarked model available on Hugging Face Hub.
Reka Edge
Open Weights LLM
Reka Edge - Open-source benchmarked model available on Hugging Face Hub.
Mamba 2.8B
Open Weights LLM
Mamba 2.8B - Open-source benchmarked model available on Hugging Face Hub.
Command R+
Open Weights LLM
Command R+ - Open-source benchmarked model available on Hugging Face Hub.
SDXL Turbo
Image Generation
SDXL Turbo - Generative image model for digital creative workflows.
Stable Video Diffusion (SVD)
Video Generation
Stable Video Diffusion (SVD) - Generative video model for digital creative workflows.
Midjourney v6
Image Generation
Midjourney v6 - Generative image model for digital creative workflows.
Ideogram 2.0
Image Generation
Ideogram 2.0 - Generative image model for digital creative workflows.
Runway Gen-2
Video Generation
Runway Gen-2 - Generative video model for digital creative workflows.
Kling AI 1.5
Video Generation
Kling AI 1.5 - Generative video model for digital creative workflows.
MiniMax Video-01
Video Generation
MiniMax Video-01 - Generative video model for digital creative workflows.
Hailuo AI
Image Generation
Hailuo AI - Generative image model for digital creative workflows.
Udio v1.5
Audio Synthesis
Udio v1.5 - Generative audio & speech model for digital creative workflows.
Eleven Flash
Audio Synthesis
Eleven Flash - Generative audio & speech model for digital creative workflows.
OpenVoice v2
Audio Synthesis
OpenVoice v2 - Generative audio & speech model for digital creative workflows.
Gaussian Splatting
3D Generation
Gaussian Splatting - Generative 3D mesh model for digital creative workflows.
OpenAI o1
Reasoning Model
OpenAI o1 by OpenAI - Reasoning Model for enterprise and developer workflows.
OpenAI o1-mini
Lightweight Reasoning
OpenAI o1-mini by OpenAI - Lightweight Reasoning for enterprise and developer workflows.
GPT-4 Turbo
Large Language Model
GPT-4 Turbo by OpenAI - Large Language Model for enterprise and developer workflows.
GPT-4
Large Language Model
GPT-4 by OpenAI - Large Language Model for enterprise and developer workflows.
GPT-3.5 Turbo
Lightweight LLM
GPT-3.5 Turbo by OpenAI - Lightweight LLM for enterprise and developer workflows.
DALL-E 3
Image Generation
DALL-E 3 by OpenAI - Image Generation for enterprise and developer workflows.
DALL-E 2
Image Generation
DALL-E 2 by OpenAI - Image Generation for enterprise and developer workflows.
Whisper Large v3
Speech Recognition
Whisper Large v3 by OpenAI - Speech Recognition for enterprise and developer workflows.
Whisper Large v2
Speech Recognition
Whisper Large v2 by OpenAI - Speech Recognition for enterprise and developer workflows.
Whisper Medium
Speech Recognition
Whisper Medium by OpenAI - Speech Recognition for enterprise and developer workflows.
Whisper Small
Speech Recognition
Whisper Small by OpenAI - Speech Recognition for enterprise and developer workflows.
text-embedding-3-large
Embedding Model
text-embedding-3-large by OpenAI - Embedding Model for enterprise and developer workflows.
text-embedding-3-small
Embedding Model
text-embedding-3-small by OpenAI - Embedding Model for enterprise and developer workflows.
text-embedding-ada-002
Embedding Model
text-embedding-ada-002 by OpenAI - Embedding Model for enterprise and developer workflows.
OpenAI TTS-1
Speech Synthesis
OpenAI TTS-1 by OpenAI - Speech Synthesis for enterprise and developer workflows.
OpenAI TTS-1 HD
Speech Synthesis
OpenAI TTS-1 HD by OpenAI - Speech Synthesis for enterprise and developer workflows.
Claude 3 Opus
Frontier LLM
Claude 3 Opus by Anthropic - Frontier LLM engineered for high safety, reasoning, and long context.
Claude 3 Sonnet
Enterprise LLM
Claude 3 Sonnet by Anthropic - Enterprise LLM engineered for high safety, reasoning, and long context.
Claude 3 Haiku
Lightweight LLM
Claude 3 Haiku by Anthropic - Lightweight LLM engineered for high safety, reasoning, and long context.
Claude 2.1
Enterprise LLM
Claude 2.1 by Anthropic - Enterprise LLM engineered for high safety, reasoning, and long context.
Claude 2.0
Large LLM
Claude 2.0 by Anthropic - Large LLM engineered for high safety, reasoning, and long context.
Claude Instant 1.2
Lightweight LLM
Claude Instant 1.2 by Anthropic - Lightweight LLM engineered for high safety, reasoning, and long context.
Gemini 2.0 Flash Thinking
Omnimodal LLM
Gemini 2.0 Flash Thinking - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemini 1.5 Flash
Omnimodal LLM
Gemini 1.5 Flash - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemini 1.0 Pro
Omnimodal LLM
Gemini 1.0 Pro - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemma 2 27B
Open Weights LLM
Gemma 2 27B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemma 2 2B
Open Weights LLM
Gemma 2 2B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Gemma 2B
Open Weights LLM
Gemma 2B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
CodeGemma 2B
Open Weights LLM
CodeGemma 2B - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Imagen 3
Image Generation
Imagen 3 - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Veo 2
Video Generation
Veo 2 - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
MusicLM
Omnimodal LLM
MusicLM - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
PaLM 2
Omnimodal LLM
PaLM 2 - Google AI multimodal model designed for high throughput, reasoning, and synthesis.
Llama 3.2 11B Vision
Vision LLM
Llama 3.2 11B Vision - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 3.2 90B Vision
Vision LLM
Llama 3.2 90B Vision - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 3.2 1B
Open Weights LLM
Llama 3.2 1B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 3.1 8B
Open Weights LLM
Llama 3.1 8B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 3 70B
Open Weights LLM
Llama 3 70B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 2 70B
Open Weights LLM
Llama 2 70B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama 2 13B
Open Weights LLM
Llama 2 13B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Code Llama 70B
Code LLM
Code Llama 70B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Code Llama 34B
Code LLM
Code Llama 34B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Code Llama 7B
Code LLM
Code Llama 7B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Llama Guard 3 8B
Open Weights LLM
Llama Guard 3 8B - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Prompt Guard 86M
Open Weights LLM
Prompt Guard 86M - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Segment Anything 2 (SAM 2)
Vision LLM
Segment Anything 2 (SAM 2) - Meta AI open source model engineered for high efficiency, vision, and edge performance.
SeamlessM4T v2
Open Weights LLM
SeamlessM4T v2 - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Seamless Expressive
Open Weights LLM
Seamless Expressive - Meta AI open source model engineered for high efficiency, vision, and edge performance.
MusicGen
Audio Model
MusicGen - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Bark
Open Weights LLM
Bark - Meta AI open source model engineered for high efficiency, vision, and edge performance.
MovieGen
Open Weights LLM
MovieGen - Meta AI open source model engineered for high efficiency, vision, and edge performance.
Mistral Large
Open Weights LLM
Mistral Large - European frontier open & commercial model by Mistral AI.
Mistral Small 3
Open Weights LLM
Mistral Small 3 - European frontier open & commercial model by Mistral AI.
Mistral NeMo 12B
Open Weights LLM
Mistral NeMo 12B - European frontier open & commercial model by Mistral AI.
Mistral 7B v0.2
Open Weights LLM
Mistral 7B v0.2 - European frontier open & commercial model by Mistral AI.
Codestral 22B
Code LLM
Codestral 22B - European frontier open & commercial model by Mistral AI.
Pixtral 12B
Vision LLM
Pixtral 12B - European frontier open & commercial model by Mistral AI.
Mixtral 8x22B
MoE LLM
Mixtral 8x22B - European frontier open & commercial model by Mistral AI.
Mathstral 7B
Open Weights LLM
Mathstral 7B - European frontier open & commercial model by Mistral AI.
Qwen 2.5 32B
Open Weights LLM
Qwen 2.5 32B - Alibaba Qwen open multilingual model series.
Qwen 2.5 7B
Open Weights LLM
Qwen 2.5 7B - Alibaba Qwen open multilingual model series.
Qwen 2.5 1.5B
Open Weights LLM
Qwen 2.5 1.5B - Alibaba Qwen open multilingual model series.
Qwen 2.5 Coder 7B
Code LLM
Qwen 2.5 Coder 7B - Alibaba Qwen open multilingual model series.
Qwen 2.5 Coder 1.5B
Code LLM
Qwen 2.5 Coder 1.5B - Alibaba Qwen open multilingual model series.
Qwen 2.5 Math 72B
Math Model
Qwen 2.5 Math 72B - Alibaba Qwen open multilingual model series.
Qwen 2 VL 72B
Vision LLM
Qwen 2 VL 72B - Alibaba Qwen open multilingual model series.
Qwen 2 VL 2B
Vision LLM
Qwen 2 VL 2B - Alibaba Qwen open multilingual model series.
Qwen 2 57B A14B
Open Weights LLM
Qwen 2 57B A14B - Alibaba Qwen open multilingual model series.
Qwen 1.5 72B
Open Weights LLM
Qwen 1.5 72B - Alibaba Qwen open multilingual model series.
Qwen-Audio
Open Weights LLM
Qwen-Audio - Alibaba Qwen open multilingual model series.
SmolLM2 360M
Open Weights LLM
SmolLM2 360M - Open-source benchmarked model available on Hugging Face Hub.
SmolLM2 135M
Open Weights LLM
SmolLM2 135M - Open-source benchmarked model available on Hugging Face Hub.
Granite 3.1 8B
Open Weights LLM
Granite 3.1 8B - Open-source benchmarked model available on Hugging Face Hub.
Granite 3.1 2B
Open Weights LLM
Granite 3.1 2B - Open-source benchmarked model available on Hugging Face Hub.
Granite Code 34B
Code LLM
Granite Code 34B - Open-source benchmarked model available on Hugging Face Hub.
Phi-4 14B
Open Weights LLM
Phi-4 14B - Open-source benchmarked model available on Hugging Face Hub.
Phi-3.5 Vision
Vision LLM
Phi-3.5 Vision - Open-source benchmarked model available on Hugging Face Hub.
Phi-3.5 Mini
Open Weights LLM
Phi-3.5 Mini - Open-source benchmarked model available on Hugging Face Hub.
Phi-3 Medium
Open Weights LLM
Phi-3 Medium - Open-source benchmarked model available on Hugging Face Hub.
Phi-3 Small
Open Weights LLM
Phi-3 Small - Open-source benchmarked model available on Hugging Face Hub.
Phi-3 Mini
Open Weights LLM
Phi-3 Mini - Open-source benchmarked model available on Hugging Face Hub.
Phi-2
Open Weights LLM
Phi-2 - Open-source benchmarked model available on Hugging Face Hub.
BGE-Large-EN-v1.5
Embedding Model
BGE-Large-EN-v1.5 - Open-source benchmarked model available on Hugging Face Hub.
BGE-Small-EN-v1.5
Embedding Model
BGE-Small-EN-v1.5 - Open-source benchmarked model available on Hugging Face Hub.
BGE-Reranker-Large
Embedding Model
BGE-Reranker-Large - Open-source benchmarked model available on Hugging Face Hub.
Nomic Embed Text v1.5
Embedding Model
Nomic Embed Text v1.5 - Open-source benchmarked model available on Hugging Face Hub.
GTE-Large-en-v1.5
Embedding Model
GTE-Large-en-v1.5 - Open-source benchmarked model available on Hugging Face Hub.
GTE-Qwen2-7B-instruct
Embedding Model
GTE-Qwen2-7B-instruct - Open-source benchmarked model available on Hugging Face Hub.
Instructor Large
Open Weights LLM
Instructor Large - Open-source benchmarked model available on Hugging Face Hub.
Sentence-Transformers All-MiniLM-L6-v2
Embedding Model
Sentence-Transformers All-MiniLM-L6-v2 - Open-source benchmarked model available on Hugging Face Hub.
ModernBERT Base
Embedding Model
ModernBERT Base - Open-source benchmarked model available on Hugging Face Hub.
ModernBERT Large
Embedding Model
ModernBERT Large - Open-source benchmarked model available on Hugging Face Hub.
BERT Base Uncased
Embedding Model
BERT Base Uncased - Open-source benchmarked model available on Hugging Face Hub.
RoBERTa Large
Embedding Model
RoBERTa Large - Open-source benchmarked model available on Hugging Face Hub.
SciBERT
Embedding Model
SciBERT - Open-source benchmarked model available on Hugging Face Hub.
ClinicalBERT
Embedding Model
ClinicalBERT - Open-source benchmarked model available on Hugging Face Hub.
PubMedBERT
Embedding Model
PubMedBERT - Open-source benchmarked model available on Hugging Face Hub.
FinBERT
Embedding Model
FinBERT - Open-source benchmarked model available on Hugging Face Hub.
BigBird
Open Weights LLM
BigBird - Open-source benchmarked model available on Hugging Face Hub.
LayoutLMv3
Vision LLM
LayoutLMv3 - Open-source benchmarked model available on Hugging Face Hub.
Donut
Vision LLM
Donut - Open-source benchmarked model available on Hugging Face Hub.
TrOCR
Open Weights LLM
TrOCR - Open-source benchmarked model available on Hugging Face Hub.
Kokoro 82M
Audio TTS
Kokoro 82M - Open-source benchmarked model available on Hugging Face Hub.
MiniCPM-V 2.6
Open Weights LLM
MiniCPM-V 2.6 - Open-source benchmarked model available on Hugging Face Hub.
MiniCPM 3 4B
Open Weights LLM
MiniCPM 3 4B - Open-source benchmarked model available on Hugging Face Hub.
StarCoder 2 15B
Code LLM
StarCoder 2 15B - Open-source benchmarked model available on Hugging Face Hub.
StarCoder 2 3B
Code LLM
StarCoder 2 3B - Open-source benchmarked model available on Hugging Face Hub.
CodeGen 16B
Code LLM
CodeGen 16B - Open-source benchmarked model available on Hugging Face Hub.
Yi 1.5 34B
Open Weights LLM
Yi 1.5 34B - Open-source benchmarked model available on Hugging Face Hub.
Yi 1.5 9B
Open Weights LLM
Yi 1.5 9B - Open-source benchmarked model available on Hugging Face Hub.
Yi Lightning
Open Weights LLM
Yi Lightning - Open-source benchmarked model available on Hugging Face Hub.
ERNIE 4.0 Turbo
Open Weights LLM
ERNIE 4.0 Turbo - Open-source benchmarked model available on Hugging Face Hub.
ERNIE 3.5
Open Weights LLM
ERNIE 3.5 - Open-source benchmarked model available on Hugging Face Hub.
Tencent Hunyuan
Open Weights LLM
Tencent Hunyuan - Open-source benchmarked model available on Hugging Face Hub.
GLM-4V 9B
Open Weights LLM
GLM-4V 9B - Open-source benchmarked model available on Hugging Face Hub.
GLM-4 Flash
Open Weights LLM
GLM-4 Flash - Open-source benchmarked model available on Hugging Face Hub.
Baichuan 2 13B
Open Weights LLM
Baichuan 2 13B - Open-source benchmarked model available on Hugging Face Hub.
InternLM 2.5 20B
Open Weights LLM
InternLM 2.5 20B - Open-source benchmarked model available on Hugging Face Hub.
InternVL 2.5 78B
Vision LLM
InternVL 2.5 78B - Open-source benchmarked model available on Hugging Face Hub.
InternVL 2 8B
Vision LLM
InternVL 2 8B - Open-source benchmarked model available on Hugging Face Hub.
Reka Flash
Open Weights LLM
Reka Flash - Open-source benchmarked model available on Hugging Face Hub.
Reka Core
Open Weights LLM
Reka Core - Open-source benchmarked model available on Hugging Face Hub.
DBRX Instruct
Open Weights LLM
DBRX Instruct - Open-source benchmarked model available on Hugging Face Hub.
Snowflake Arctic 480B
Open Weights LLM
Snowflake Arctic 480B - Open-source benchmarked model available on Hugging Face Hub.
DeciLM 7B
Open Weights LLM
DeciLM 7B - Open-source benchmarked model available on Hugging Face Hub.
RWKV 6 7B
Open Weights LLM
RWKV 6 7B - Open-source benchmarked model available on Hugging Face Hub.
Jamba 1.5 Large
Open Weights LLM
Jamba 1.5 Large - Open-source benchmarked model available on Hugging Face Hub.
Jamba 1.5 Mini
Open Weights LLM
Jamba 1.5 Mini - Open-source benchmarked model available on Hugging Face Hub.
Jurassic 2 Ultra
Open Weights LLM
Jurassic 2 Ultra - Open-source benchmarked model available on Hugging Face Hub.
Solar 10.7B
Open Weights LLM
Solar 10.7B - Open-source benchmarked model available on Hugging Face Hub.
Falcon 2 11B
Open Weights LLM
Falcon 2 11B - Open-source benchmarked model available on Hugging Face Hub.
Falcon 180B
Open Weights LLM
Falcon 180B - Open-source benchmarked model available on Hugging Face Hub.
Falcon 40B
Open Weights LLM
Falcon 40B - Open-source benchmarked model available on Hugging Face Hub.
Falcon 7B
Open Weights LLM
Falcon 7B - Open-source benchmarked model available on Hugging Face Hub.
Command R
Open Weights LLM
Command R - Open-source benchmarked model available on Hugging Face Hub.
Cohere Embed v3
Embedding Model
Cohere Embed v3 - Open-source benchmarked model available on Hugging Face Hub.
Cohere Aya 23
Open Weights LLM
Cohere Aya 23 - Open-source benchmarked model available on Hugging Face Hub.
Aya 101
Open Weights LLM
Aya 101 - Open-source benchmarked model available on Hugging Face Hub.
Stable Diffusion 3.5 Medium
Image Generation
Stable Diffusion 3.5 Medium - Generative image model for digital creative workflows.
SDXL 1.0
Image Generation
SDXL 1.0 - Generative image model for digital creative workflows.
SD 1.5
Image Generation
SD 1.5 - Generative image model for digital creative workflows.
SD 2.1
Image Generation
SD 2.1 - Generative image model for digital creative workflows.
Stable Audio 2.0
Audio Synthesis
Stable Audio 2.0 - Generative audio & speech model for digital creative workflows.
Stable Audio Open
Audio Synthesis
Stable Audio Open - Generative audio & speech model for digital creative workflows.
Midjourney v5.2
Image Generation
Midjourney v5.2 - Generative image model for digital creative workflows.
NijiJourney v6
Image Generation
NijiJourney v6 - Generative image model for digital creative workflows.
Ideogram 1.0
Image Generation
Ideogram 1.0 - Generative image model for digital creative workflows.
Runway Gen-3 Alpha
Video Generation
Runway Gen-3 Alpha - Generative video model for digital creative workflows.
Luma Dream Machine 1.5
Video Generation
Luma Dream Machine 1.5 - Generative video model for digital creative workflows.
Luma Ray 2
Video Generation
Luma Ray 2 - Generative video model for digital creative workflows.
Sora (OpenAI)
Video Generation
Sora (OpenAI) - Generative video model for digital creative workflows.
Pika 1.5
Video Generation
Pika 1.5 - Generative video model for digital creative workflows.
Hunyuan Video (Tencent)
Video Generation
Hunyuan Video (Tencent) - Generative video model for digital creative workflows.
CogVideoX 5B
Video Generation
CogVideoX 5B - Generative video model for digital creative workflows.
Suno v4
Audio Synthesis
Suno v4 - Generative audio & speech model for digital creative workflows.
Suno v3.5
Audio Synthesis
Suno v3.5 - Generative audio & speech model for digital creative workflows.
Cartesia Sonic
Audio Synthesis
Cartesia Sonic - Generative audio & speech model for digital creative workflows.
Eleven Multilingual v2
Audio Synthesis
Eleven Multilingual v2 - Generative audio & speech model for digital creative workflows.
Play3.0-mini
Image Generation
Play3.0-mini - Generative image model for digital creative workflows.
XTTS v2
Audio Synthesis
XTTS v2 - Generative audio & speech model for digital creative workflows.
Tripo3D 2.0
3D Generation
Tripo3D 2.0 - Generative 3D mesh model for digital creative workflows.
Meshy 4
3D Generation
Meshy 4 - Generative 3D mesh model for digital creative workflows.
Llama-3.1-Code-GGUF-Q4_K_M
Code LLM
Llama-3.1-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-GGUF-Q4_K_M
Code LLM
Qwen-2.5-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-GGUF-Q4_K_M
Code LLM
DeepSeek-V3-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-GGUF-Q4_K_M
Code LLM
Mistral-7B-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-GGUF-Q4_K_M
Code LLM
Gemma-2-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-GGUF-Q4_K_M
Code LLM
Phi-3.5-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-GGUF-Q4_K_M
Code LLM
Granite-3.1-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-GGUF-Q4_K_M
Code LLM
Yi-1.5-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-GGUF-Q4_K_M
Code LLM
Falcon-2-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-GGUF-Q4_K_M
Code LLM
BGE-M3-Code-GGUF-Q4_K_M - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-GGUF-Q8_0
Code LLM
Llama-3.1-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-GGUF-Q8_0
Code LLM
Qwen-2.5-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-GGUF-Q8_0
Code LLM
DeepSeek-V3-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-GGUF-Q8_0
Code LLM
Mistral-7B-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-GGUF-Q8_0
Code LLM
Gemma-2-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-GGUF-Q8_0
Code LLM
Phi-3.5-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-GGUF-Q8_0
Code LLM
Granite-3.1-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-GGUF-Q8_0
Code LLM
Yi-1.5-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-GGUF-Q8_0
Code LLM
Falcon-2-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-GGUF-Q8_0
Code LLM
BGE-M3-Code-GGUF-Q8_0 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-AWQ-INT4
Code LLM
Llama-3.1-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-AWQ-INT4
Code LLM
Qwen-2.5-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-AWQ-INT4
Code LLM
DeepSeek-V3-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-AWQ-INT4
Code LLM
Mistral-7B-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-AWQ-INT4
Code LLM
Gemma-2-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-AWQ-INT4
Code LLM
Phi-3.5-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-AWQ-INT4
Code LLM
Granite-3.1-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-AWQ-INT4
Code LLM
Yi-1.5-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-AWQ-INT4
Code LLM
Falcon-2-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-AWQ-INT4
Code LLM
BGE-M3-Code-AWQ-INT4 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-GPTQ-4bit
Code LLM
Llama-3.1-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-GPTQ-4bit
Code LLM
Qwen-2.5-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-GPTQ-4bit
Code LLM
DeepSeek-V3-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-GPTQ-4bit
Code LLM
Mistral-7B-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-GPTQ-4bit
Code LLM
Gemma-2-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-GPTQ-4bit
Code LLM
Phi-3.5-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-GPTQ-4bit
Code LLM
Granite-3.1-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-GPTQ-4bit
Code LLM
Yi-1.5-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-GPTQ-4bit
Code LLM
Falcon-2-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-GPTQ-4bit
Code LLM
BGE-M3-Code-GPTQ-4bit - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-EXL2-5.0bpw
Code LLM
Llama-3.1-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-EXL2-5.0bpw
Code LLM
Qwen-2.5-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-EXL2-5.0bpw
Code LLM
DeepSeek-V3-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-EXL2-5.0bpw
Code LLM
Mistral-7B-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-EXL2-5.0bpw
Code LLM
Gemma-2-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-EXL2-5.0bpw
Code LLM
Phi-3.5-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-EXL2-5.0bpw
Code LLM
Granite-3.1-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-EXL2-5.0bpw
Code LLM
Yi-1.5-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-EXL2-5.0bpw
Code LLM
Falcon-2-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-EXL2-5.0bpw
Code LLM
BGE-M3-Code-EXL2-5.0bpw - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-FP8-quant
Code LLM
Llama-3.1-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-FP8-quant
Code LLM
Qwen-2.5-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-FP8-quant
Code LLM
DeepSeek-V3-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-FP8-quant
Code LLM
Mistral-7B-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-FP8-quant
Code LLM
Gemma-2-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-FP8-quant
Code LLM
Phi-3.5-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-FP8-quant
Code LLM
Granite-3.1-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-FP8-quant
Code LLM
Yi-1.5-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-FP8-quant
Code LLM
Falcon-2-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-FP8-quant
Code LLM
BGE-M3-Code-FP8-quant - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-FP16-full
Code LLM
Llama-3.1-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-FP16-full
Code LLM
Qwen-2.5-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-FP16-full
Code LLM
DeepSeek-V3-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-FP16-full
Code LLM
Mistral-7B-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-FP16-full
Code LLM
Gemma-2-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-FP16-full
Code LLM
Phi-3.5-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-FP16-full
Code LLM
Granite-3.1-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-FP16-full
Code LLM
Yi-1.5-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-FP16-full
Code LLM
Falcon-2-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-FP16-full
Code LLM
BGE-M3-Code-FP16-full - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-Uncensored-Instruct
Code LLM
Llama-3.1-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-Uncensored-Instruct
Code LLM
Qwen-2.5-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-Uncensored-Instruct
Code LLM
DeepSeek-V3-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-Uncensored-Instruct
Code LLM
Mistral-7B-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-Uncensored-Instruct
Code LLM
Gemma-2-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-Uncensored-Instruct
Code LLM
Phi-3.5-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-Uncensored-Instruct
Code LLM
Granite-3.1-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-Uncensored-Instruct
Code LLM
Yi-1.5-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-Uncensored-Instruct
Code LLM
Falcon-2-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-Uncensored-Instruct
Code LLM
BGE-M3-Code-Uncensored-Instruct - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-RAG-FineTune
Code LLM
Llama-3.1-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-RAG-FineTune
Code LLM
Qwen-2.5-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-RAG-FineTune
Code LLM
DeepSeek-V3-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-RAG-FineTune
Code LLM
Mistral-7B-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-RAG-FineTune
Code LLM
Gemma-2-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-RAG-FineTune
Code LLM
Phi-3.5-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-RAG-FineTune
Code LLM
Granite-3.1-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-RAG-FineTune
Code LLM
Yi-1.5-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-RAG-FineTune
Code LLM
Falcon-2-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-RAG-FineTune
Code LLM
BGE-M3-Code-RAG-FineTune - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-Hermes-3
Code LLM
Llama-3.1-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-Hermes-3
Code LLM
Qwen-2.5-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-Hermes-3
Code LLM
DeepSeek-V3-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-Hermes-3
Code LLM
Mistral-7B-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-Hermes-3
Code LLM
Gemma-2-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-Hermes-3
Code LLM
Phi-3.5-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-Hermes-3
Code LLM
Granite-3.1-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-Hermes-3
Code LLM
Yi-1.5-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-Hermes-3
Code LLM
Falcon-2-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-Hermes-3
Code LLM
BGE-M3-Code-Hermes-3 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-OpenChat-3.5
Code LLM
Llama-3.1-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-OpenChat-3.5
Code LLM
Qwen-2.5-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-OpenChat-3.5
Code LLM
DeepSeek-V3-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-OpenChat-3.5
Code LLM
Mistral-7B-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-OpenChat-3.5
Code LLM
Gemma-2-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-OpenChat-3.5
Code LLM
Phi-3.5-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-OpenChat-3.5
Code LLM
Granite-3.1-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-OpenChat-3.5
Code LLM
Yi-1.5-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-OpenChat-3.5
Code LLM
Falcon-2-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-OpenChat-3.5
Code LLM
BGE-M3-Code-OpenChat-3.5 - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Code-Nexus-Agent
Code LLM
Llama-3.1-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Qwen-2.5-Code-Nexus-Agent
Code LLM
Qwen-2.5-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
DeepSeek-V3-Code-Nexus-Agent
Code LLM
DeepSeek-V3-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Mistral-7B-Code-Nexus-Agent
Code LLM
Mistral-7B-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Gemma-2-Code-Nexus-Agent
Code LLM
Gemma-2-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Phi-3.5-Code-Nexus-Agent
Code LLM
Phi-3.5-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Granite-3.1-Code-Nexus-Agent
Code LLM
Granite-3.1-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Yi-1.5-Code-Nexus-Agent
Code LLM
Yi-1.5-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Falcon-2-Code-Nexus-Agent
Code LLM
Falcon-2-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
BGE-M3-Code-Nexus-Agent
Code LLM
BGE-M3-Code-Nexus-Agent - Fine-tuned and quantized Code domain model variant optimized for fast local host inference.
Llama-3.1-Math-GGUF-Q4_K_M
Math Model
Llama-3.1-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-GGUF-Q4_K_M
Math Model
Qwen-2.5-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-GGUF-Q4_K_M
Math Model
DeepSeek-V3-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-GGUF-Q4_K_M
Math Model
Mistral-7B-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-GGUF-Q4_K_M
Math Model
Gemma-2-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-GGUF-Q4_K_M
Math Model
Phi-3.5-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-GGUF-Q4_K_M
Math Model
Granite-3.1-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-GGUF-Q4_K_M
Math Model
Yi-1.5-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-GGUF-Q4_K_M
Math Model
Falcon-2-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-GGUF-Q4_K_M
Math Model
BGE-M3-Math-GGUF-Q4_K_M - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-GGUF-Q8_0
Math Model
Llama-3.1-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-GGUF-Q8_0
Math Model
Qwen-2.5-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-GGUF-Q8_0
Math Model
DeepSeek-V3-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-GGUF-Q8_0
Math Model
Mistral-7B-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-GGUF-Q8_0
Math Model
Gemma-2-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-GGUF-Q8_0
Math Model
Phi-3.5-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-GGUF-Q8_0
Math Model
Granite-3.1-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-GGUF-Q8_0
Math Model
Yi-1.5-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-GGUF-Q8_0
Math Model
Falcon-2-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-GGUF-Q8_0
Math Model
BGE-M3-Math-GGUF-Q8_0 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-AWQ-INT4
Math Model
Llama-3.1-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-AWQ-INT4
Math Model
Qwen-2.5-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-AWQ-INT4
Math Model
DeepSeek-V3-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-AWQ-INT4
Math Model
Mistral-7B-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-AWQ-INT4
Math Model
Gemma-2-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-AWQ-INT4
Math Model
Phi-3.5-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-AWQ-INT4
Math Model
Granite-3.1-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-AWQ-INT4
Math Model
Yi-1.5-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-AWQ-INT4
Math Model
Falcon-2-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-AWQ-INT4
Math Model
BGE-M3-Math-AWQ-INT4 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-GPTQ-4bit
Math Model
Llama-3.1-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-GPTQ-4bit
Math Model
Qwen-2.5-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-GPTQ-4bit
Math Model
DeepSeek-V3-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-GPTQ-4bit
Math Model
Mistral-7B-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-GPTQ-4bit
Math Model
Gemma-2-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-GPTQ-4bit
Math Model
Phi-3.5-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-GPTQ-4bit
Math Model
Granite-3.1-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-GPTQ-4bit
Math Model
Yi-1.5-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-GPTQ-4bit
Math Model
Falcon-2-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-GPTQ-4bit
Math Model
BGE-M3-Math-GPTQ-4bit - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-EXL2-5.0bpw
Math Model
Llama-3.1-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-EXL2-5.0bpw
Math Model
Qwen-2.5-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-EXL2-5.0bpw
Math Model
DeepSeek-V3-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-EXL2-5.0bpw
Math Model
Mistral-7B-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-EXL2-5.0bpw
Math Model
Gemma-2-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-EXL2-5.0bpw
Math Model
Phi-3.5-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-EXL2-5.0bpw
Math Model
Granite-3.1-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-EXL2-5.0bpw
Math Model
Yi-1.5-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-EXL2-5.0bpw
Math Model
Falcon-2-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-EXL2-5.0bpw
Math Model
BGE-M3-Math-EXL2-5.0bpw - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-FP8-quant
Math Model
Llama-3.1-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-FP8-quant
Math Model
Qwen-2.5-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-FP8-quant
Math Model
DeepSeek-V3-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-FP8-quant
Math Model
Mistral-7B-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-FP8-quant
Math Model
Gemma-2-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-FP8-quant
Math Model
Phi-3.5-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-FP8-quant
Math Model
Granite-3.1-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-FP8-quant
Math Model
Yi-1.5-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-FP8-quant
Math Model
Falcon-2-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-FP8-quant
Math Model
BGE-M3-Math-FP8-quant - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-FP16-full
Math Model
Llama-3.1-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-FP16-full
Math Model
Qwen-2.5-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-FP16-full
Math Model
DeepSeek-V3-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-FP16-full
Math Model
Mistral-7B-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-FP16-full
Math Model
Gemma-2-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-FP16-full
Math Model
Phi-3.5-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-FP16-full
Math Model
Granite-3.1-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-FP16-full
Math Model
Yi-1.5-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-FP16-full
Math Model
Falcon-2-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-FP16-full
Math Model
BGE-M3-Math-FP16-full - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-Uncensored-Instruct
Math Model
Llama-3.1-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-Uncensored-Instruct
Math Model
Qwen-2.5-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-Uncensored-Instruct
Math Model
DeepSeek-V3-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-Uncensored-Instruct
Math Model
Mistral-7B-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-Uncensored-Instruct
Math Model
Gemma-2-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-Uncensored-Instruct
Math Model
Phi-3.5-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-Uncensored-Instruct
Math Model
Granite-3.1-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-Uncensored-Instruct
Math Model
Yi-1.5-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-Uncensored-Instruct
Math Model
Falcon-2-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-Uncensored-Instruct
Math Model
BGE-M3-Math-Uncensored-Instruct - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-RAG-FineTune
Math Model
Llama-3.1-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-RAG-FineTune
Math Model
Qwen-2.5-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-RAG-FineTune
Math Model
DeepSeek-V3-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-RAG-FineTune
Math Model
Mistral-7B-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-RAG-FineTune
Math Model
Gemma-2-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-RAG-FineTune
Math Model
Phi-3.5-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-RAG-FineTune
Math Model
Granite-3.1-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-RAG-FineTune
Math Model
Yi-1.5-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-RAG-FineTune
Math Model
Falcon-2-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-RAG-FineTune
Math Model
BGE-M3-Math-RAG-FineTune - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-Hermes-3
Math Model
Llama-3.1-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-Hermes-3
Math Model
Qwen-2.5-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-Hermes-3
Math Model
DeepSeek-V3-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-Hermes-3
Math Model
Mistral-7B-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-Hermes-3
Math Model
Gemma-2-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-Hermes-3
Math Model
Phi-3.5-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-Hermes-3
Math Model
Granite-3.1-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-Hermes-3
Math Model
Yi-1.5-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-Hermes-3
Math Model
Falcon-2-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-Hermes-3
Math Model
BGE-M3-Math-Hermes-3 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-OpenChat-3.5
Math Model
Llama-3.1-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-OpenChat-3.5
Math Model
Qwen-2.5-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-OpenChat-3.5
Math Model
DeepSeek-V3-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-OpenChat-3.5
Math Model
Mistral-7B-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-OpenChat-3.5
Math Model
Gemma-2-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-OpenChat-3.5
Math Model
Phi-3.5-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-OpenChat-3.5
Math Model
Granite-3.1-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-OpenChat-3.5
Math Model
Yi-1.5-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-OpenChat-3.5
Math Model
Falcon-2-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-OpenChat-3.5
Math Model
BGE-M3-Math-OpenChat-3.5 - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Math-Nexus-Agent
Math Model
Llama-3.1-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Qwen-2.5-Math-Nexus-Agent
Math Model
Qwen-2.5-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
DeepSeek-V3-Math-Nexus-Agent
Math Model
DeepSeek-V3-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Mistral-7B-Math-Nexus-Agent
Math Model
Mistral-7B-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Gemma-2-Math-Nexus-Agent
Math Model
Gemma-2-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Phi-3.5-Math-Nexus-Agent
Math Model
Phi-3.5-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Granite-3.1-Math-Nexus-Agent
Math Model
Granite-3.1-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Yi-1.5-Math-Nexus-Agent
Math Model
Yi-1.5-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Falcon-2-Math-Nexus-Agent
Math Model
Falcon-2-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
BGE-M3-Math-Nexus-Agent
Math Model
BGE-M3-Math-Nexus-Agent - Fine-tuned and quantized Math domain model variant optimized for fast local host inference.
Llama-3.1-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Llama-3.1-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
Qwen-2.5-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Qwen-2.5-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
DeepSeek-V3-Medical-GGUF-Q4_K_M
Specialized Domain LLM
DeepSeek-V3-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
Mistral-7B-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Mistral-7B-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
Gemma-2-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Gemma-2-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
Phi-3.5-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Phi-3.5-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
Granite-3.1-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Granite-3.1-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
Yi-1.5-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Yi-1.5-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
Falcon-2-Medical-GGUF-Q4_K_M
Specialized Domain LLM
Falcon-2-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.
BGE-M3-Medical-GGUF-Q4_K_M
Specialized Domain LLM
BGE-M3-Medical-GGUF-Q4_K_M - Fine-tuned and quantized Medical domain model variant optimized for fast local host inference.