All API tools (172)
- 9RouterFree MIT AI router and token saver that connects Claude Code, Codex, Cursor, and Cline to 40+ providers with auto-fallback and RTK token compression.
- AnySearchPrivacy-first AI search infrastructure for agents: one API and MCP server unifying web and vertical-domain search with structured results that cut token waste.
- CheaperInferenceCheaperInference resells discounted AI inference through one OpenAI-compatible API, cutting model costs by up to 30% with usage-based billing, no contract.
- Claude AcademyAnthropic's official learning platform. 300+ free courses and tutorials across Claude.ai, Cowork, Claude Code, Tag, and the Platform API, plus an AI Fluency.
- Cloudflare AI GatewayLLM proxy on Cloudflare. Paid logs 10M/gateway; extra requests $0.05/million. Not an unmetered endpoint.
- CozeByteDance agent studio. Coze Studio on GitHub is Apache-2.0 with 21,459 stars. Not a free-for-anyone slogan.
- DifyOpen-source agent workflows. Cloud: Sandbox $0, Professional $59, Team $159 per workspace/month. Community self-host is $0.
- FirecrawlWeb-to-LLM API. AGPL-3.0 self-host. Not the old $16/$83/$333-only table.
- Free Claude CodeMIT local gateway that points Claude Code, Codex, Pi, OpenCode, or Cline at 48 providers.
- GMI CloudAI-native GPU cloud. No invented seat.
- LangChainMIT agent framework. Hosted traces are LangSmith Developer $0 / Plus $39 per seat. Not LangServe-as-the-product.
- LlamaIndexDocument agents plus LlamaParse. 1,000 credits = $1.25. OSS framework is MIT.
- MCP FrameworkTypeScript kit for MCP servers. Software $0. Not a 5-minute Claude Desktop-only tutorial.
- MiniMaxShanghai lab behind Hailuo, M2.5, and the 2026 M3/H3 coding models. TalkAPI M3 is $0.24/$0.98 per 1M tokens.
- mlx-servemlx-serve is a MIT Zig inference server for Apple Silicon that runs MLX and GGUF locally with OpenAI and Anthropic APIs, no Python.
- NInferNInfer is an Apache-2.0 C++/CUDA engine for selected Qwen checkpoints on a single RTX 5090, with OpenAI- and Anthropic-compatible local APIs.
- NotebookLMGoogle's source-grounded research notebook, now presented as Gemini Notebook at notebook.google. Analyzes your files and turns them into audio, video, and more.
- OllamaLocal models plus Ollama cloud. Not a 30-minute Llama3 tutorial.
- oMLXoMLX is an Apache-2.0 Apple Silicon inference server with SSD KV cache and a macOS menu bar app for coding agents.
- OmniRouteFree MIT AI gateway with one endpoint, 339 providers, 90+ free tiers, quota-aware fallback, and token compression for Claude Code, Codex, and Cursor.
- OpenAISan Francisco lab behind ChatGPT and the GPT API. Current flagship is GPT-5.6 Sol; GPT-5.2 remains the previous frontier snapshot.
- OpenGradientDecentralized L1 for verifiable AI: ZKML, TEE, Vanilla, HACA, PIPE, x402, Model Hub. No invented dollar table.
- OpenRouterUnified LLM router. Not a GPT-4o catalog page.
- RAGFlowOpen-source RAG plus agents. Cloud signup exists; no public dollar table. Not a generic document scanner.
- SiliconFlowChina inference API for open models. Official site siliconflow.cn. Qwen3-8B listed at ¥0.27/¥0.88 per 1M tokens. Skip unverified 6M-user claims.
- skills.qiaomu.aiChinese-language Claude Code Skill recommendation site by Yang Xiang Qiaomu (向阳乔木): a curated catalog of popular Skills with one-click install guidance for d.
- Claude Agent SDKAnthropic agent library. Python 7,910 stars, MIT.
- Hugging FaceModels, datasets, and inference. Extra private storage $18/TB. Not a 100k-model slogan.
- Claude 3.5 SonnetPrevious Claude 3.5 Sonnet snapshot. Not the 2026 default.
- Claude 3 HaikuPrevious Claude 3 Haiku snapshot. Not the 2026 fast default.
- Claude 3 OpusPrevious Claude 3 Opus snapshot. Not the 2026 flagship.
- Claude 3 SonnetPrevious Claude 3 Sonnet snapshot. Not the 2026 default.
- Claude Opus 4.5Anthropic's previous Opus snapshot: 200k context, $5/$25 per 1M tokens. Later billed lines include Opus 4.6 and Opus 5.
- Claude Sonnet 4.5Anthropic's previous Sonnet snapshot: 200k context, $3/$15 per 1M tokens. Docs now sell newer Claude 4.6 and Claude 5 lines.
- Claude v1Previous Claude v1 snapshot. Not a 2026 text default.
- BGE-M32024 BAAI multilingual embedding snapshot. Not a 2026 MIRACL or MTEB crown.
- BAAI bge-reranker-v2.5-gemma2-lightweightPrevious BGE reranker v2.5 Gemma2 lightweight snapshot.
- Breeze TTS 2Open-weight multilingual TTS from BreezeBlue, built for real-time interaction. Reference-free voice design and reference-guided voice direction, with sub-40m.
- Claude Fable 5.1Anthropic's generally available Fable-class model for coding and knowledge work. $10/$50 per 1M tokens, cache reads $0.25, API id claude-fable-5-1.
- Claude Fable 5Anthropic's Mythos-class flagship made safe for general use, with 1M-token context, frontier coding and vision, days-long agentic autonomy, and $10/$50 per 1.
- Claude Haiku 4.5Claude Haiku 4.5: live cheap/fast Claude row at $1/$5 per 1M tokens and 200k context. Not leftover 3 Haiku.
- Claude Sonnet 5Claude Sonnet 5: live mid-tier Claude row at $2/$10 per 1M tokens and 1M context. Not leftover Sonnet 4.5.
- cogvlm-base-490-hfPrevious CogVLM snapshot. Not a 2026 VLM default.
- Cohere: Command ACohere Command A: live Command lead on first-party docs. Not leftover Command R.
- Cohere: Command RPrevious Cohere Command R snapshot. Not a 2026 chat / RAG default.
- Cohere Embed v3Previous Cohere embed snapshot. Not a $0.10/1M 2026 default.
- Cohere Rerank 3.5Previous Cohere rerank snapshot.
- OpenAI: dall-e-2Previous OpenAI DALL-E 2 snapshot. DALL-E 3 is already removed from the API.
- OpenAI: dall-e-3Deprecated OpenAI image model from November 2023. Removed from the API; OpenAI now points new work to GPT Image 2.
- Deepgram Nova-2Previous Deepgram ASR. Not the fastest 2026 crown.
- DeepSeek-Coder-V2.5Previous DeepSeek coder snapshot. Not a 90.2% HumanEval census.
- DeepSeek-R1Previous DeepSeek-R1 snapshot. Not an o1 / 2026 reasoning default.
- DeepSeek V2.5Previous DeepSeek V2.5 snapshot. Not a 2026 chat+coder default.
- DeepSeek V3Previous DeepSeek V3 snapshot. Not a GPT-4o / Claude 3.5 census.
- DeepSeek V4 FlashDeepSeek small agent model: 284B/13B MoE, 1M context, MIT weights, $0.14/$0.28 per 1M tokens.
- DeepSeek V4DeepSeek V4 represents the next generation of DeepSeek's flagship AI models, building upon the success of V3 with enhanced capabilities in reasoning, multimodal processing, and agent-based interactions.
- deepseek-vl-7b-basePrevious DeepSeek-VL 7B Base snapshot. Not a 2026 VL default.
- ElevenLabs Turbo v2.5Previous ElevenLabs Turbo v2.5 snapshot. Not the 2026 voice default.
- EmbeddingGemmaLightweight multilingual text embedding model from Google DeepMind, optimized for on-device AI with <200MB RAM usage.
- Flux.1 Dev2024 BFL open-weight stills. Not fully MIT, not the flagship.
- Flux.1 ProPrevious BFL image SKU. Not a 2024 Midjourney/DALL-E 3 crown.
- Gemini 2.0 Flash ThinkingPrevious Gemini 2.0 Flash Thinking snapshot. Not a 2026 o1-mini alternative.
- GLM-4.7An open-source multilingual multimodal chat model from Zhipu AI with advanced thinking capabilities, exceptional coding performance, and enhanced UI generation.
- GLM-5.2Zhipu's June 2026 744B MoE coding model with 1M context and MIT weights. Successor GLM-5.3 is now the flagship.
- GLM-5.3Zhipu coding model on GLM Coding Plan. Token API not on the public table yet. Not $18-only.
- Google: Gemini 1.5 Flash-8BPrevious Gemini 1.5 Flash-8B snapshot. Not the 2026 default.
- Google: Gemini 2.0 FlashPrevious Gemini 2.0 Flash snapshot. Live Flash is Gemini 3.7 Flash.
- Google: Gemini 3.1 Flash-LiteGoogle Gemini 3.1 Flash-Lite: leftover cheap Flash-Lite sibling on the live models page.
- Google: Gemini 3.1 Flash LiveGoogle Gemini 3.1 Flash Live: leftover live/audio sibling on the Gemini models page. Not the 2026 Flash default.
- Google: Gemini 3.1 Flash TTSGoogle Gemini 3.1 Flash TTS: leftover TTS sibling on the live Gemini models page. Not the 2026 Flash default.
- Google: Gemini 3.1 ProGoogle Gemini 3.1 Pro: leftover Pro sibling on the live models page. Not a 2026 Pro default.
- Google: Gemini 3.5 Flash-LiteGoogle Gemini 3.5 Flash-Lite: leftover cheap Flash-Lite row on the live Gemini models page.
- Google: Gemini 3.5 Live TranslateGoogle Gemini 3.5 Live Translate: leftover live-translate sibling on the live Gemini models page. Not the 2026 Flash default. No invent.
- Google: Gemini 3.6 FlashGoogle Gemini 3.6 Flash: previous-generation Flash after 3.5, before live 3.7 Flash.
- Google: Gemini 3 FlashGoogle's latest frontier model delivering breakthrough intelligence at unprecedented speed and cost efficiency.
- Google: Gemini 3 ProPrevious Gemini 3 Pro Preview snapshot. Not a 2026 vision crown.
- Google: Gemini Flash 1.5Previous Gemini 1.5 Flash snapshot. Not the 2026 default.
- Google: Gemini Nano Banana 2 LiteGoogle Gemini Nano Banana 2 Lite: leftover cheaper Nano Banana 2 sibling on the live Gemini image page.
- Google: Gemini Omni FlashGoogle Gemini Omni Flash: leftover Omni Flash sibling on the live Gemini models page. Not the 2026 Flash default.
- Google: Gemini Pro 1.5Previous Gemini 1.5 Pro snapshot. Not the 2026 default.
- Google: Gemma 2 27BPrevious Gemma 2 27B snapshot. Not a 2026 default open model.
- Google: Gemma 2 9BPrevious Gemma 2 9B snapshot. Not a 2026 default small model.
- GPT-5.6 LunaOpenAI GPT-5.6 Luna: fast affordable GPT-5.6 tier, 128k context, $1/$6 per 1M tokens.
- GPT-5.6 TerraOpenAI GPT-5.6 Terra: balanced daily driver in the 5.6 family, 400k context, $2.50/$15 per 1M tokens.
- GPT-5.2OpenAI's previous frontier model: 400k context, 128k max out, $1.75/$14 per 1M tokens. Docs now recommend GPT-5.6 Sol.
- GPT-5.6 SolOpenAI frontier flagship GPT-5.6 Sol: 1.05M context, effort up to max plus ultra, $5/$30 per 1M tokens.
- GPT-6 AstraOpenAI flagship GPT-6 Astra: 1.05M context, computer-use SOTA, $10/$50 per 1M tokens. API id gpt-6-astra.
- GPT Image 2OpenAI's latest image generation model for fast, high-quality create and edit jobs, with flexible sizes up to 4K and token-based API pricing.
- GrokxAI / SpaceXAI Grok family. Current docs flagship is grok-4.6: 500k context, $2/$6 per 1M tokens, plus Imagine image/video and Voice APIs.
- Ideogram 2.0Ideogram Aug 2024 snapshot. No invented later-version or seat table.
- Jina Embeddings v4Previous Jina Embeddings v4 snapshot.
- Jina AI Reranker v3Previous Jina rerank snapshot.
- LFM 2.5-2.6BLiquid AI's open-weight 2.6B dense model trained for on-device agentic workloads, with 128K context and native tool calling.
- Llama 3.1 Euryale 70B v2.2Previous Llama 3.1 Euryale 70B v2.2 snapshot. Not a 2026 Llama default.
- Llama 3 Euryale 70B v2.1Previous Llama 3 Euryale 70B v2.1 snapshot. Not a 2026 Llama default.
- LLaMA Guard 3Previous Llama Guard snapshot. Not the 2026 safety default.
- llava-v1.6-34b-hfPrevious LLaVA-NeXT / LLaVA 1.6 snapshot.
- LongCat 2.0Meituan's open-weight MoE model: 1.6T total / ~48B active params, 1M context, MIT. Strong on coding and agentic tasks (2026-08-25).
- CodeLlama 34B InstructPrevious Code Llama 34B Instruct snapshot. Not a 2026 coding default.
- Llama 3.1 405B InstructPrevious Llama 3.1 405B Instruct snapshot. Not a GPT-4o crown.
- Llama 3.1 70B InstructPrevious Llama 3.1 70B Instruct snapshot. Not the 2026 default.
- Llama 3.1 8B InstructPrevious Llama 3.1 8B Instruct snapshot. Not the 2026 default.
- Llama 3.2 1B InstructPrevious Llama 3.2 1B Instruct snapshot. Not a 1.3B typo.
- Meta Llama 3.2 Vision2024 Llama 3.2 vision snapshot. Not the 2026 multimodal default.
- Llama 3 70B InstructPrevious Llama 3 70B Instruct snapshot. Not the 2026 default.
- Llama 3 8B InstructPrevious Llama 3 8B Instruct snapshot. Not the 2026 default.
- Llama v2 70B ChatPrevious Llama v2 70B Chat snapshot. Not a 2026 chat default.
- Midjourney V6Midjourney V6 snapshot from Dec 2023. Not the current default; no invented seat table.
- MiniMax H3MiniMax open multimodal video model (July 31, 2026): text/image/video/audio in, 768P or 2K out, 4-15s clips, native stereo, pay-as-you-go API.
- MiniMax M2.1Open-source 230B parameter MoE model optimized for multi-language coding, agentic workflows, and real-world development tasks with 74% SWE-bench performance.
- MiniMax-Music3MiniMax's open-weight music generation model with studio-grade sound, natural vocal synthesis, and full-song generation from prompts and lyrics, with official ComfyUI support.
- Mistral: Mistral 7B InstructPrevious Mistral 7B Instruct snapshot.
- Mistral: Mistral NemoPrevious Mistral Nemo 12B snapshot. Live catalog is Medium 3.5 / Small 4 / Large 3.
- Mistral Nemo Inferor 12BPrevious Inferor 12B snapshot. Live catalog is Medium 3.5 / Small 4 / Large 3.
- Mistral OCR 4.1Mistral OCR 4.1: live latest OCR with paragraph boxes, structural labels, and confidence. Not leftover OCR 4.0.
- Mistral Pixtral 12B2024 Mistral 12B vision snapshot. Not the first-and-only multimodal.
- Mistral Shieldstral 1.0Mistral Shieldstral 1.0: live compact multimodal moderation model. Apache 2.0.
- Mistral SmallPrevious Mistral Small v24.09 snapshot.
- Mistral TinyPrevious Mistral Tiny snapshot. Live catalog is Medium 3.5 / Small 4 / Large 3.
- Mistral Voxtral Mini Transcribe 2Mistral Voxtral Mini Transcribe 2: live Premier transcription row, v 26.02. Older 25.07 is deprecated.
- Mistral Voxtral Mini Transcribe RealtimeMistral Voxtral Mini Transcribe Realtime: live realtime transcription sibling. Not leftover Transcribe 25.07.
- Mistral Voxtral TTSMistral Voxtral TTS: live catalog TTS row, v 26.03, CC BY-NC 4.0.
- Muse Spark 1.2Meta's coding-focused flagship model powering Muse Code: 1M context, higher first-attempt accuracy, reliable tool calling, and end-to-end developer workflows.
- mixedbread ai mxbai-rerank-large-v1Previous mxbai-rerank-large-v1 snapshot.
- MythoMax 13BPrevious MythoMax 13B snapshot.
- Nous: Hermes 3 405B InstructPrevious Hermes 3 405B Instruct snapshot. Not a 2026 frontier default.
- NousResearch: Hermes 2 Pro - Llama-3 8BPrevious Hermes 2 Pro Llama-3 8B snapshot.
- NV-Embed-v22024 NVIDIA embedding snapshot.
- NVIDIA nv-rerankqa-mistral-4b-v3Previous NVIDIA nv-rerankqa-mistral-4b-v3 snapshot.
- NVIDIA: Llama 3.1 Nemotron 70B InstructPrevious NVIDIA Llama 3.1 Nemotron 70B Instruct snapshot. Not a 2026 default.
- NVIDIA Nemotron 3.5 Lightning 30B A3BNVIDIA's efficient open-weight 30B MoE hybrid model with 3B active parameters, 1M-token context, and single-GPU deployment for local reasoning and coding.
- omni-moderation-latestHosted OpenAI omni-moderation snapshot. Not the 2026 safety default path.
- OpenAI: GPT-4o-miniPrevious GPT-4o mini snapshot. Not the 2026 latest.
- OpenAI: GPT-4oPrevious GPT-4o snapshot. Not the 2026 latest.
- OpenAI: o1-miniPrevious o1-mini snapshot. Not the 2026 reasoning default.
- OpenAI: o1-previewPrevious o1-preview snapshot. Not a cheaper o1 and not 2026 default.
- Shap-ePrevious OpenAI Shap-E snapshot. Live stills path is GPT Image 2.
- OpenChat 3.5 7BPrevious OpenChat 3.5 7B snapshot.
- Qwen-VLPrevious Alibaba Qwen-VL snapshot. Not a 2026 LVLM default.
- Qwen2.5 72B InstructPrevious Qwen2.5 72B Instruct snapshot. Not a 2026 latest series.
- Qwen2.5-72BPrevious Qwen stills LLM. Not a 2026 Llama-3-405B crown.
- Qwen2.5 Coder 32B InstructPrevious Qwen2.5-Coder 32B Instruct snapshot. Not the latest series.
- Qwen2.5-Coder-32BPrevious Qwen code snapshot.
- Qwen2-VL 72B InstructPrevious Qwen2-VL 72B Instruct snapshot. Not a 2024 VL crown.
- Qwen3-EmbeddingPrevious Qwen3 embedding snapshot.
- Qwen3-VL-EmbeddingPrevious Qwen3-VL embedding snapshot.
- Qwen3-VL-RerankerPrevious Qwen3-VL rerank snapshot.
- QwQ-32B-PreviewPrevious QwQ-32B-Preview snapshot.
- rerank-english-v3.0Previous Cohere English rerank snapshot.
- rerank-multilingual-v3.0Previous Cohere multilingual rerank snapshot.
- Cohere: rerank-v3.5Previous Cohere rerank-v3.5 snapshot.
- Rocinante 12BPrevious Rocinante 12B snapshot.
- Seedance 2.5ByteDance Seed's audio-video model for 30-second storytelling, precise reference control, extend-twice generation, and pro tools like green-screen editing.
- Stable Diffusion 3.5Stability 2024 stills family. Homepage also leads with audio and Brand Studio.
- text-embedding-3-large2024 OpenAI embedding snapshot.
- text-embedding-3-smallPrevious OpenAI embedding snapshot. Not a 2026 embedding crown.
- text-embedding-ada-002Previous OpenAI embedding snapshot. Not the 2026 embedding default.
- text-moderation-latestPrevious OpenAI text-moderation snapshot. Not the 2026 safety default.
- Tiel-Coder 35B A3BAgentic coding MoE. Not a lab release.
- OpenAI: tts-1-hdHosted OpenAI tts-1-hd snapshot. Not the 2026 default TTS path.
- OpenAI: tts-1Hosted OpenAI tts-1 snapshot. Not the 2026 default TTS path.
- voyage-3-largeVoyage AI's latest SOTA general-purpose embedding model, ranking first across 8 evaluation domains spanning 100 datasets, outperforming OpenAI and Cohere by 9.74% and 20.71% on average.
- Voyage AI Rerank 2Previous Voyage rerank snapshot.
- WeLM-617BWeChat AI's 617B MoE model with 23B active parameters, demonstrating the first sequence-length scaling at frontier scale via Hidden Decoding, powering WeChat.
- WeMM-EmbeddingTencent WeChat Vision's universal multimodal embedding family (2B/4B/9B) built on Qwen3.5. Text, image, video, and visual-document inputs map to one L2-normalized embedding space.
- OpenAI: whisper-1Hosted OpenAI whisper-1 snapshot. Not the 2026 default STT path.
- Whisper V3Open-weight Whisper large-v3 family. Not OpenAI's current hosted STT default.
- WizardLM-2 7BPrevious WizardLM-2 7B snapshot. Do not keep leftover 10x-larger crown.
- WizardLM-2 8x22BPrevious WizardLM-2 8x22B snapshot.
- xAI: Grok BetaPrevious xAI Grok Beta snapshot. Not a 2026 reasoning default.
- Documentation Search SkillProactively search auto-generated documentation before implementing - function signatures, API docs, class definitions.
- MCP Builder SkillAnthropic's official guide for creating high-quality MCP (Model Context Protocol) servers in Python (FastMCP) or Node/TypeScript to enable LLMs to interact w.
Items tagged with API (172)
Month Visit15000000DR: 95AS: 95
Month Visit45000000DR: 95AS: 88
Month Visit15000000DR: 95AS: 95
Month Visit15000000DR: 95AS: 95
Month Visit15000000DR: 95AS: 95
Month Visit12000000DR: 88AS: 90
Month Visit5000000DR: 95AS: 85
Month Visit5000000DR: 88AS: 90
Month Visit5000000DR: 88AS: 90
Month Visit4800000DR: 95AS: 80
Month Visit3000000DR: 92AS: 95
Month Visit2800000DR: 148AS: 162
Month Visit2500000DR: 92AS: 95
Month Visit2500000DR: 92AS: 95
Month Visit2500000DR: 92AS: 95
Month Visit2500000DR: 92AS: 95
Month Visit2000000DR: 95AS: 95
Month Visit1000000DR: 85AS: 85
Month Visit1000000DR: 45AS: 38
Month Visit800000DR: 90AS: 88
Month Visit800000DR: 90AS: 88
Month Visit800000DR: 90AS: 88
Month Visit800000DR: 90AS: 88
Month Visit800000DR: 90AS: 88