Best AI LLM APIs & Models Tools
Foundation models, inference APIs, and model hosting platforms for developers building on LLMs
👑 Premium partnerPaid placement, shown in every category
Build a llm apis & models tool? This is the list AI assistants read.
These are the llm apis & models tools ChatGPT names when someone asks for a recommendation, and Claude Opus 4.8 is the one it names first. If you built one that isn't here, adding it is free — the listing publishes after review. Want it live in minutes with a Verified badge instead? That option is on the form, one-time, no subscription.
All LLM APIs & Models Tools (48)
Claude Opus 4.8
Anthropic's flagship model — stronger coding, agents, and honesty
Mistral Small 4
Mistral's unified open-source model — reasoning + vision + coding, Apache 2.0
Mistral Small 3.1
Mistral's 24B multimodal open-source model — beats GPT-4o Mini, Apache 2.0
Mistral Small 3
Mistral's 24B latency-optimized open model — faster than Llama 3.3 70B, Apache 2.0
Mistral Medium 3.5
Mistral's 128B merged flagship — open weights, coding+reasoning+instructions
Mistral 3
Mistral's MoE flagship + edge model family — Apache 2.0, multimodal, reasoning
North Mini Code
Cohere's open-source agentic coding model — 30B MoE, 3B active, Apache 2.0
Codestral 25.08
Mistral's low-latency code completion model — FIM, 80+ languages, 256k context
Codestral Embed
Mistral's code-specific embedding model — semantic code search and RAG for repos
Mistral OCR 3
Mistral's document OCR model — 74% win rate vs OCR 2, $2/1k pages, forms + handwriting
Devstral 2
Mistral's SOTA open-weight coding model — 72.2% SWE-bench, free API
Magistral
Mistral's first reasoning model — 73.6% AIME2024, native multilingual chain-of-thought, 10x faster via Le Chat
📬 Get the best new AI tools delivered weekly
One concise email with fresh launches, trending picks, and featured standouts.
Mistral Medium 3
Enterprise GPT-4o-class performance at $0.40/$2.00 per million tokens — Mistral's mid-tier powerhouse
Mistral Saba
Mistral's 24B regional LLM — Arabic & South Asian languages, 150+ tok/s, self-hostable
Mistral NeMo
Mistral × NVIDIA 12B open-weight model — 128k context, Tekken tokenizer, FP8 inference, Apache 2.0
Codestral Mamba
Mistral's 7B Mamba-architecture coding model — linear-time inference, 256k context, Apache 2.0
Mathstral 7B
Open-weight 7B math specialist from Mistral AI — STEM reasoning, MATH benchmark SOTA at release
Mixtral 8x22B
Mistral's largest open-weights MoE — 141B total / 39B active, Apache 2.0
Claude Fable 5
Anthropic's most capable model — state-of-the-art coding, vision, and long-horizon tasks
Understudy Labs
Captures traces from your production LLM work, then trains a cheaper open-weight model that beats the eval
Conifer
Local-first least-cost inference router that keeps routine requests on your own hardware
EvoLink
One OpenAI-compatible API key for 75+ LLM, image, video, and audio models with smart routing and failover
Yolo-Auto
Flat-rate, unmetered OpenAI-compatible API serving Qwen3.6-35B, priced by concurrency instead of tokens
Mood Metrics API
Real-time sentiment API with confidence scores, emotions and a financial mode
APIMart
One OpenAI-compatible key for GPT, Claude, Gemini, Sora and Flux, billed pay-as-you-go below list price
Factagora
Fact-verification API returning confidence-scored verdicts with cited sources, for grounding LLM output
Defapi
Credit-based gateway to image, video, text and music models from multiple vendors, from a $5 top-up
LLMAPI
OpenAI-compatible gateway to 400+ models across Anthropic, OpenAI, Google, xAI and DeepSeek with spend analytics and per-key IAM, token pools from $30/month
HiAPI
One API for image, video and audio models with persistent artifact links and MCP support, free API key with up to 2,000 credits then usage-based
Compute Prices
Daily-refreshed price index for 67 GPU providers and 26 LLM inference APIs with an agent-friendly JSON API, free key at 750 requests/day
Supavec
Open-source RAG-as-a-service API built on Supabase row-level security — upload documents, query for cited context, self-host if you want
Apiframe
One credit-based API and MCP endpoint for Midjourney and 70+ other generative media models, with webhooks, a Studio UI and no per-seat fees
Semarize
Conversational intelligence as an API — configurable extraction kits that return structured deal signals from call transcripts
ChatComparison.ai
Runs one prompt against 40+ models side by side with per-response cost and a routing recommendation, on one $15/mo seat
OpenRouter Model Price Snapshot
Export public OpenRouter model names and prompt list prices into Apify dataset records.
AICost
AICost is an independent price desk for AI models, built around one question: what will this actually cost me?
bnrouter
OpenAI-compatible AI API gateway — one endpoint to route, meter, and monetize access to major model providers.
PoYo.ai
One API for AI image, video, music, chat, and 3D generation
Heabsy AI Inference
OpenAI- and Anthropic-compatible EU inference API with zero data retention on the EEA tier.
ToolPilot
Compare AI model API pricing and estimate token costs for developer workloads.
APIClaw
Access multiple AI models with a single API key.
Jev
Typed System One model for decision-making.
ToAPIs
OpenAI-compatible gateway for 50+ text, image, and video AI models.
DiscountedTokens
Prepaid GPT-5.x API, one key, from US$5
ModelRush
One OpenAI-compatible API for text, image, video, and voice models.
XiuRouter
One API for leading AI models, with selected routes saving 90%+ versus provider reference prices.
Standard Compute
Standard Compute — description pending review.
social media chat reconstruct
social media chat reconstruct — description pending review.
Want your tool featured here?
Featured tools appear first on this page and get surfaced to AI search engines like ChatGPT and Perplexity.Every plan includes a permanent backlink to your site. Paid plans make it dofollow.