ModelVaultAI Index
All ModelsLocal RunCloud APIs
AI Cost CalculatorLive
Token CalculatorSoon
Context CalculatorSoon
VRAM CalculatorSoon
Admin PanelDashboardCompare
Compare
Searchable Directory of 500+ Cloud & Local AI Models

Find the Perfect AI Model in Seconds

Discover, compare, and benchmark 500+ cloud APIs and open-weight local models. Filter by modality, context window, VRAM requirements, and pricing.

Explore ModelsCompare Models
Indexed AI Models
520Models

Frontier & open-weight index

AI Labs & Providers
10Labs

OpenAI, Anthropic, Google, etc.

Cloud Endpoints
270APIs

REST APIs & SDK integrations

Local & Ollama
447Models

Self-hostable weights

Open Weights
443Free

Community & commercial open

Taxonomy & Modality

Popular Model Categories

View all categories
156 Models

Text / Chat

General-purpose conversation, writing, translation, and summary models.

133 Models

Reasoning

Advanced step-by-step thinking, complex problem solving, and math.

173 Models

Coding Agents

Code generation, debugging, repository refactoring, and agentic workflows.

23 Models

Image Generation

High-fidelity text-to-image synthesis, artistic styling, and graphics.

1 Models

Image Editing

Inpainting, outpainting, background removal, and image-to-image tasks.

13 Models

Video Generation

Photorealistic text-to-video, image-to-video, and motion synthesis.

🔥 High Demand

Trending & Popular AI Models

Browse all 520 models
OpenAI
Paid API

GPT-4o

Omnimodal LLM

GPT-4o by OpenAI - Omnimodal LLM for enterprise and developer workflows.

Context Window128k
MMLU88.7
Cloud Only
#flagship#omni
OpenAI
Freemium

GPT-4o mini

Lightweight LLM

GPT-4o mini by OpenAI - Lightweight LLM for enterprise and developer workflows.

Context Window128k
MMLU82
Cloud Only
#mini#fast
OpenAI
Paid API

OpenAI o3

Reasoning Model

OpenAI o3 by OpenAI - Reasoning Model for enterprise and developer workflows.

Context Window200k
AIME 202496.7
Cloud Only
#reasoning#math
OpenAI
Paid API

OpenAI o3-mini

Reasoning Model

OpenAI o3-mini by OpenAI - Reasoning Model for enterprise and developer workflows.

Context Window200k
AIME 202487.3
Cloud Only
#reasoning#o3-mini
Anthropic
Paid API

Claude 3.5 Sonnet

Frontier LLM

Claude 3.5 Sonnet by Anthropic - Frontier LLM engineered for high safety, reasoning, and long context.

Context Window200k
SWE-bench49
Cloud Only
#claude#coding
Anthropic
Paid API

Claude 3.5 Haiku

Lightweight LLM

Claude 3.5 Haiku by Anthropic - Lightweight LLM engineered for high safety, reasoning, and long context.

Context Window200k
HumanEval88.1
Cloud Only
#claude#haiku
Curated Recommendations

Best Models for Specific Roles

Best for Coding & Software Engineering

Top-ranked models for repo-level refactoring, debugging, and terminal automation.

View all Coding models
Anthropic
Paid API

Claude 3.5 Sonnet

Frontier LLM

Claude 3.5 Sonnet by Anthropic - Frontier LLM engineered for high safety, reasoning, and long context.

Context Window200k
SWE-bench49
Cloud Only
#claude#coding
DeepSeek
Open Weights

DeepSeek-V3

MoE LLM

DeepSeek-V3 - DeepSeek open-source frontier model for math, reasoning, and code.

Context Window128k
MATH / Code82
Cloud API Local Run
#deepseek#mit-license

Best for Reasoning & Complex Math

Models with deep chain-of-thought verification for scientific and mathematical proofs.

View all Reasoning models
OpenAI
Paid API

OpenAI o3

Reasoning Model

OpenAI o3 by OpenAI - Reasoning Model for enterprise and developer workflows.

Context Window200k
AIME 202496.7
Cloud Only
#reasoning#math
DeepSeek
Open Weights

DeepSeek-R1

Reasoning Model

DeepSeek-R1 - DeepSeek open-source frontier model for math, reasoning, and code.

Context Window128k
MATH / Code80
Cloud API Local Run
#deepseek#mit-license

Best for Local Deployment (Ollama)

High-performing open weights models optimized to run on consumer hardware.

View all Local models
Meta AI
Open Weights

Llama 3.3 70B

Open Weights LLM

Llama 3.3 70B - Meta AI open source model engineered for high efficiency, vision, and edge performance.

Context Window128k
MMLU65
Cloud API Local Run
#meta#llama
Alibaba Cloud
Open Weights

Qwen 2.5 72B

Open Weights LLM

Qwen 2.5 72B - Alibaba Qwen open multilingual model series.

Context Window128k
MMLU / HumanEval72
Cloud API Local Run
#qwen#alibaba
Latest Indexing

Recently Added Models

OpenAI
Paid API

OpenAI o3-mini

Reasoning Model

OpenAI o3-mini by OpenAI - Reasoning Model for enterprise and developer workflows.

Context Window200k
AIME 202487.3
Cloud Only
#reasoning#o3-mini
DeepSeek
Open Weights

Janus 1.3B

Vision LLM

Janus 1.3B - DeepSeek open-source frontier model for math, reasoning, and code.

Context Window128k
MATH / Code82
Cloud API Local Run
#deepseek#mit-license
DeepSeek
Open Weights

Janus Pro 7B

Vision LLM

Janus Pro 7B - DeepSeek open-source frontier model for math, reasoning, and code.

Context Window128k
MATH / Code81
Cloud API Local Run
#deepseek#mit-license
DeepSeek
Open Weights

DeepSeek-R1-Distill-Llama-8B

Reasoning Model

DeepSeek-R1-Distill-Llama-8B - DeepSeek open-source frontier model for math, reasoning, and code.

Context Window128k
MATH / Code80
Cloud API Local Run
#deepseek#mit-license
Platform Architecture

Why Developers Choose ModelVault

Standardized metadata, hardware requirements, and benchmark transparency built for high-performance AI deployment.

520+ AI Models Index

Comprehensive dataset indexing frontier and open-weight models with hardware, benchmark, and context specifications.

Production Verified Specs

Side-by-Side Comparison

Compare up to 4 models simultaneously across context length, benchmark scores, licensing terms, and pricing rates.

Production Verified Specs

Advanced Search & Filtering

Instant real-time search across modalities, providers, tasks, VRAM requirements, and deployment setups.

Production Verified Specs

Verified Benchmark Data

Standardized performance scores including MMLU, HumanEval, SWE-bench, and MATH across all leading architectures.

Production Verified Specs

Transparent API & Hardware Rates

Clear token cost breakdown for commercial APIs and VRAM hardware requirements for local deployments.

Production Verified Specs

Native Local & Ollama Support

Dedicated index of self-hostable open weights optimized for Ollama, vLLM, LM Studio, and local consumer GPUs.

Production Verified Specs
Side-by-Side Matrix Engine

Compare AI Models Side-by-Side Before You Build

Evaluate context windows, benchmark scores (MMLU, HumanEval), pricing rates, and licensing terms across up to 4 models simultaneously.

Launch Comparison Matrix
Matrix Preview4 Models Max
GPT-4o128k ctx • 88.7 MMLU
Claude 3.5 Sonnet200k ctx • 93.7 Code
DeepSeek-R1128k ctx • Open Weights
Ecosystem Roadmap

Upcoming Developer Utilities

In Active Development
Live

AI API Cost Calculator

Estimate monthly API expenditure based on projected input/output token volume across OpenAI, Anthropic, Google, DeepSeek, and Mistral.

Launch Tool
Coming Soon

Context Window Calculator

Calculate document page capacity, RAG chunk limits, and context window fill ratios for long-context LLMs.

In Development
Coming Soon

Token & Pricing Estimator

Analyze raw prompt text to estimate exact token counts and compare generation costs across 20+ models.

In Development
Coming Soon

Local VRAM Hardware Estimator

Determine exact GPU VRAM requirements for quantization levels (GGUF Q4, AWQ INT4, FP16) before downloading weights.

In Development

Indexed AI Research Labs

Leading creators of frontier models, open weights, and API infrastructure.

OpenAI20 Models IndexAnthropic8 Models IndexGoogle AI33 Models IndexMeta AI354 Models IndexMistral AI17 Models IndexDeepSeek21 Models IndexStability AI40 Models IndexCohere2 Models IndexxAI0 Models IndexAlibaba Cloud25 Models Index

Stay Ahead of Frontier & Open-Weight Model Drops

Join 15,000+ AI engineers and researchers. Get weekly breakdowns of new model releases, benchmark shifts, and local LLM quantization tips.

Weekly Release Summaries
Verified Benchmark Data
Zero Spam, Unsubscribe Anytime
ModelVault

The open directory for discovering, comparing, and benchmarking cloud and local AI models. Built for engineers, researchers, and technical leaders.

All 520+ Model Specifications Active
ProductAll AI Models IndexComparison MatrixLocal Models (Ollama)Cloud API ModelsAI Cost Calculator
ResourcesReasoning ModelsCoding AgentsVision-LanguageEmbeddings & RAG
Management & LegalAdmin Console ↗Privacy PolicyTerms of ServiceDisclaimerAbout ModelVaultContact Us
© 2025 ModelVault AI Directory. All model specifications verified.
Built for high-performance AI workflows