{"providers":[{"vendor":"Hugging Face","category":"open-source-host","description":"ML model hub — $0.10/month free inference credits, 200+ models via Inference Providers, unlimited model hosting on Hub","tier":"Free","url":"https://huggingface.co/docs/inference-providers/pricing","tags":["ai","ml","inference","models","open-source","free tier"],"verifiedDate":"2026-08-09","vendor_page":"/vendor/hugging-face"},{"vendor":"Groq","category":"inference","description":"Ultra-fast LLM inference on LPU hardware — free tier: 30 RPM, 100K-500K tokens/day depending on model. Supports Llama 4 Scout 17B, Llama 3.3 70B, Qwen3 32B, Whisper, and more. No credit card required","tier":"Free","url":"https://groq.com/pricing","tags":["ai","ml","llm","inference","fast","free tier"],"verifiedDate":"2026-08-15","vendor_page":"/vendor/groq"},{"vendor":"Google Gemini API","category":"frontier","description":"Free tier is Flash-tier only: Gemini 2.5 Flash (10 RPM), Gemini 2.5 Flash-Lite (15 RPM), Gemini 3.0 Flash Preview, Gemini 3.1 Flash-Lite Preview, Gemini Embedding, and Gemma 4. 2.5 Pro and 3.1 Pro Preview are paid-only. Per-model paid pricing: Gemini 3.1 Pro Preview $2/$12 per MTok (≤200K ctx, doubles above), Gemini 3.0 Flash Preview $0.50/$3, Gemini 3.1 Flash-Lite Preview $0.25/$1.50, Gemini 2.5 Pro $1.25/$10 (≤200K, doubles above), Gemini 2.5 Flash $0.30/$2.50. Gemini 2.0 Flash and 2.0 Flash-Lite deprecated June 1, 2026 — migrate to 2.5 Flash or 3.x Flash. All models support Batch/Flex at 50% discount. Mandatory spend caps enforced since April 1, 2026.","tier":"Free (Reduced)","url":"https://ai.google.dev/pricing","tags":["ai","ml","llm","inference","multimodal","free tier","deal-change","per-model-pricing"],"verifiedDate":"2026-08-18","vendor_page":"/vendor/google-gemini-api"},{"vendor":"Mistral AI","category":"frontier","description":"Access to all Mistral models including Large, Codestral, Pixtral — 2 RPM, 1B tokens/month. No credit card required","tier":"Experiment","url":"https://mistral.ai/pricing","tags":["ai","ml","llm","inference","code","free tier"],"verifiedDate":"2026-08-20","vendor_page":"/vendor/mistral-ai"},{"vendor":"OpenRouter","category":"inference","description":"AI model router — ~30 free models (DeepSeek R1, Llama 3.3, Qwen3, Gemma 3), OpenAI-compatible API, ~20 RPM per model","tier":"Free","url":"https://openrouter.ai/pricing","tags":["ai","ml","llm","inference","api","free tier"],"verifiedDate":"2026-08-25","vendor_page":"/vendor/openrouter"},{"vendor":"Cloudflare Workers AI","category":"open-source-host","description":"AI inference at the edge — 10,000 neurons/day free across text generation, image classification, translation, speech-to-text models","tier":"Free","url":"https://developers.cloudflare.com/workers-ai/platform/pricing/","tags":["ai","ml","inference","edge","serverless","free tier"],"verifiedDate":"2026-08-09","vendor_page":"/vendor/cloudflare-workers-ai"},{"vendor":"Cerebras","category":"inference","description":"Ultra-fast LLM inference API. Free tier: 1M tokens/day, 10-30 requests/min (varies by model). Models include Llama 3.1 8B, Qwen 3 235B, GPT-OSS 120B. Multi-thousand tokens/sec inference speed","tier":"Free","url":"https://cerebras.ai/","tags":["ai","llm","inference","api"],"verifiedDate":"2026-08-15","vendor_page":"/vendor/cerebras"},{"vendor":"Cohere","category":"frontier","description":"AI model API. Trial key: 1,000 API calls/month across all endpoints (Chat, Embed, Rerank). Access to Command R+, Rerank 3.5, Embed 4. Non-commercial use only","tier":"Trial Key","url":"https://cohere.com/pricing","tags":["ai","llm","embeddings","reranking","api"],"verifiedDate":"2026-08-16","vendor_page":"/vendor/cohere"},{"vendor":"OpenAI","category":"frontier","description":"AI API platform — free tier limited to GPT-3.5 Turbo model only, 3 requests/minute rate limit. No free trial credits for new accounts (discontinued mid-2025). Paid tiers use token-based pricing: GPT-4o from $2.50/1M input tokens. Batch processing at 50% discount available","tier":"Free","url":"https://openai.com/api/pricing/","tags":["ai","ml","llm","gpt","api","nlp","generative-ai"],"verifiedDate":"2026-06-30","vendor_page":"/vendor/openai"},{"vendor":"Pinecone","category":"specialized","description":"Vector database — 2 GB storage, 2M write units/month, 1M read units/month, 5 indexes, 5M embedding tokens/month. Pinecone Assistant: 100 docs / 1 GB","tier":"Starter","url":"https://pinecone.io/pricing","tags":["ai","ml","vector-database","embeddings","search"],"verifiedDate":"2026-07-31","vendor_page":"/vendor/pinecone"},{"vendor":"Qdrant","category":"specialized","description":"Vector database — 1 GB free forever cluster, fully managed on AWS/GCP/Azure. Unlimited users, backups included. No credit card required","tier":"Free Forever","url":"https://qdrant.tech/pricing/","tags":["ai","ml","vector-database","embeddings","search"],"verifiedDate":"2026-07-30","vendor_page":"/vendor/qdrant"},{"vendor":"Deepgram","category":"specialized","description":"Speech-to-text and text-to-speech API — $200 free credits on signup (no credit card required), ~43K minutes transcription with Nova model, credits never expire. Access to all model endpoints","tier":"Free Credits","url":"https://deepgram.com/pricing","tags":["speech-to-text","transcription","voice","ai","nlp"],"verifiedDate":"2026-07-30","vendor_page":"/vendor/deepgram"},{"vendor":"Kaggle","category":"specialized","description":"Data science platform (Google) — entirely free: 30 hrs/week GPU compute (Tesla T4, 16 GB VRAM), 20 hrs/week TPU, 20 GB working disk per session. Unlimited public notebooks, datasets, and competitions","tier":"Free","url":"https://www.kaggle.com","tags":["ml","data-science","notebooks","gpu","competitions","datasets"],"verifiedDate":"2026-08-24","vendor_page":"/vendor/kaggle"},{"vendor":"Roboflow","category":"specialized","description":"Computer vision platform — Public plan: 250,000 images, 10 projects, 2 users, $60/month free inference credits. Includes AI-assisted labeling, model training, cloud deployment. All data publicly shared","tier":"Public (Free)","url":"https://roboflow.com/pricing","tags":["ml","computer-vision","image-labeling","object-detection","model-training"],"verifiedDate":"2026-07-28","vendor_page":"/vendor/roboflow"},{"vendor":"Scale AI","category":"specialized","description":"Data labeling and AI data platform — free tier: first 1,000 annotation units and 10,000 images uploaded free. Nucleus free tier for individuals and academia. Pay-as-you-go after free allocation","tier":"Free","url":"https://scale.com/pricing","tags":["ml","data-labeling","annotation","ai-data","training-data"],"verifiedDate":"2026-07-31","vendor_page":"/vendor/scale-ai"},{"vendor":"AssemblyAI","category":"specialized","description":"Speech-to-text and audio intelligence API — free: $50 in credits (~185 hours pre-recorded transcription). Up to 5 concurrent streams. Core STT and Audio Intelligence models. Pay-as-you-go at $0.15/hr after","tier":"Free","url":"https://www.assemblyai.com/pricing","tags":["ai","speech-to-text","transcription","audio","nlp","api"],"verifiedDate":"2026-07-30","vendor_page":"/vendor/assemblyai"},{"vendor":"Replicate","category":"open-source-host","description":"ML model hosting and inference platform — free runs on curated model collection without billing. Pay-per-second billing by hardware type (CPU/GPU) after free allowance. No credit card required to start","tier":"Free","url":"https://replicate.com/pricing","tags":["ai","ml","model-hosting","inference","gpu"],"verifiedDate":"2026-08-15","vendor_page":"/vendor/replicate"},{"vendor":"Baseten","category":"open-source-host","description":"ML model deployment platform — $30 in free credits for new accounts. Basic plan is $0/month with pay-as-you-go billing after credits. Per-minute GPU/CPU billing for custom deployments, per-token for Model APIs","tier":"Basic (Free Credits)","url":"https://www.baseten.co/pricing/","tags":["ai","ml","model-deployment","inference","gpu","serverless"],"verifiedDate":"2026-08-01","vendor_page":"/vendor/baseten"},{"vendor":"Weights & Biases","category":"specialized","description":"ML experiment tracking and model registry — free tier: experiment tracking, unlimited projects, 5 GB cloud storage, up to 5 Model seats. Community support. Model registry and dataset versioning included. Non-commercial use only","tier":"Free","url":"https://wandb.ai/site/pricing/","tags":["ml","experiment-tracking","model-registry","mlops","datasets"],"verifiedDate":"2026-08-17","vendor_page":"/vendor/weights-biases"},{"vendor":"Comet ML","category":"specialized","description":"ML experiment tracking and LLM evaluation — free tier: 1 user, 100 GB data storage, experiment tracking and comparison, dataset versioning, model registry, LLM evaluation. Free Pro plan for academics","tier":"Free","url":"https://www.comet.com/site/pricing/","tags":["ml","experiment-tracking","model-registry","mlops","llm-evaluation"],"verifiedDate":"2026-07-31","vendor_page":"/vendor/comet-ml"},{"vendor":"Clarifai","category":"specialized","description":"Full-stack AI platform (computer vision, NLP, audio) — Community tier: 1,000 API calls/month. Access to pre-trained models for image recognition, NLP, and audio. No credit card required","tier":"Community (Free)","url":"https://www.clarifai.com/pricing","tags":["ai","computer-vision","nlp","image-recognition","api"],"verifiedDate":"2026-07-15","vendor_page":"/vendor/clarifai"},{"vendor":"Neptune.ai","category":"specialized","description":"ML experiment tracking — free for individuals/researchers. 200 GB metadata storage, 50 GB file storage, 100K tracking calls/hour, 20 active runs, unlimited archived experiments, Jupyter/Colab integration","tier":"Free (Individual)","url":"https://neptune.ai/pricing","tags":["ml","experiment-tracking","mlops","foundation-models","research"],"verifiedDate":"2026-05-13","vendor_page":"/vendor/neptune-ai"},{"vendor":"Labelbox","category":"specialized","description":"Data labeling and annotation platform — free tier: 500 Labelbox Units (LBUs)/month. Image, video, text, and geospatial annotation. Free for qualified educational institutions","tier":"Free","url":"https://labelbox.com/pricing/","tags":["ai","data-labeling","annotation","training-data","computer-vision"],"verifiedDate":"2026-06-09","vendor_page":"/vendor/labelbox"},{"vendor":"Arize AI","category":"specialized","description":"ML observability platform (now Arize AX). AX Free plan: 25K trace spans/month, 1 GB ingestion/month, 15 days data retention. Includes online evals, product observability, and community support.","tier":"Free","url":"https://arize.com/pricing","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-08-26","vendor_page":"/vendor/arize-ai"},{"vendor":"Composio","category":"specialized","description":"Integration platform for AI Agents and LLMs. Integrate over 200+ tools across the agentic internet.","tier":"Free","url":"https://composio.dev/pricing","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-08-26","vendor_page":"/vendor/composio"},{"vendor":"DeepAR","category":"specialized","description":"Augmented reality face filters for any platform with one SDK. The free plan provides up to 10 monthly active users (MAU) and tracks up to 4 faces","tier":"Free","url":"https://developer.deepar.ai","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-07-31","vendor_page":"/vendor/deepar"},{"vendor":"Maxim AI","category":"specialized","description":"Simulate, evaluate, and observe your AI agents. Maxim is an end-to-end evaluation and observability platform, helping teams ship their AI agents reliably and >5x faster. Free forever for indie developers and small teams (3 seats).","tier":"Free","url":"https://getmaxim.ai/","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-07-30","vendor_page":"/vendor/maxim-ai"},{"vendor":"OCR.Space","category":"specialized","description":"An OCR API parses image and pdf files that return the text results in JSON format. 25,000 requests per month are free and a 1MB file size limit.","tier":"Free","url":"https://ocr.space/","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-08-01","vendor_page":"/vendor/ocr-space"},{"vendor":"Parseur","category":"specialized","description":"20 free pages/month: Extract data from PDFs, emails. AI powered. Full API access.","tier":"Free","url":"https://parseur.com","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-08-28","vendor_page":"/vendor/parseur"},{"vendor":"Reducto","category":"specialized","description":"Turn any unstructured documents (PDF, XLSX, JPG, PPTX, etc.) into structured JSON data. Parse, extract data, and edit PDF forms. Free tier with 15k free credits and pay-as-you-go.","tier":"Free","url":"https://reducto.ai","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-08-01","vendor_page":"/vendor/reducto"},{"vendor":"Tavily AI","category":"specialized","description":"API for online search and rapid insights and comprehensive research, with the capability of organization of research results. 1000 request/month for the Free tier with No credit card required.","tier":"Free","url":"https://www.tavily.com/pricing","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-08-01","vendor_page":"/vendor/tavily-ai"},{"vendor":"paperspace","category":"specialized","description":"ML platform (now part of DigitalOcean). Free Gradient plan: public projects and 5 GB storage only. No free compute — all GPU and CPU instances are billed hourly. Paid plans from $8/month.","tier":"Free","url":"https://www.paperspace.com/pricing","tags":["developer-tools","free-for-dev"],"verifiedDate":"2026-08-22","vendor_page":"/vendor/paperspace"},{"vendor":"Arize AX","category":"specialized","description":"AI engineering platform for evaluating and observing AI applications and agents. AX Free plan: 25K trace spans/month, 1 GB ingestion/month, 15 days data retention. Includes online evals, product observability, built-in Alyx agent, and community support.","tier":"Free","url":"https://arize.com/pricing","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-20","vendor_page":"/vendor/arize-ax"},{"vendor":"Audio Enhancer","category":"specialized","description":"AI-powered audio enhancer SaaS that removes noise and echo while preserving natural vocal clarity. totally Free: unlimited one-click enhancements, no login required, supports MP3/WAV/FLAC","tier":"Free","url":"https://voice-clone.org/tools/audio-enhancer","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-06-16","vendor_page":"/vendor/audio-enhancer"},{"vendor":"Braintrust","category":"specialized","description":"Evals, prompt playground, and data management for Gen AI. Starter plan (free): 1 GB processed data/month, 10k scores/month, 14 days data retention. Unlimited users, projects, datasets, playgrounds, and experiments. Overage: $4/GB data, $2.50 per 1k scores.","tier":"Free","url":"https://www.braintrust.dev/pricing","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-20","vendor_page":"/vendor/braintrust"},{"vendor":"Clair","category":"specialized","description":"Clinical AI Reference. Students have free access to the professional tool suite, which includes Open Search, Clinical Summary, Med Review, Drug Interactions, ICD-10 Codes, and Stewardship. Additionally, a free trial for the professional suite is available.","tier":"Free","url":"https://askclair.ai/","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-10","vendor_page":"/vendor/clair"},{"vendor":"Keywords AI","category":"specialized","description":"The best LLM monitoring platform. One format to call 200+ LLMs with 2 lines of code. 10,000 free requests every month and $0 for platform features!","tier":"Free","url":"https://keywordsai.co","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-08","vendor_page":"/vendor/keywords-ai"},{"vendor":"Langfuse","category":"specialized","description":"Open-source LLM engineering platform that helps teams collaboratively debug, analyze, and iterate on their LLM applications. Free forever plan includes 50k observations per month and all platform features. [#opensource](https://github.com/langfuse/langfuse)","tier":"Free","url":"https://langfuse.com/pricing","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-09-01","vendor_page":"/vendor/langfuse"},{"vendor":"LangSmith","category":"specialized","description":"LLM observability and evaluation platform by LangChain. Developer (Free) tier includes 5,000 traces/month, 14-day data retention, 1 seat, 1 workspace. Provides prompt playground, dataset management, annotation queues, and monitoring dashboards. Plus plan at $39/seat/month adds 10,000 traces, 3 workspaces, and unlimited Agent Builder agents. Overage at $0.50 per 1,000 traces","tier":"Free","url":"https://www.langchain.com/pricing","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-08","vendor_page":"/vendor/langsmith"},{"vendor":"Langtrace","category":"specialized","description":"enables developers to trace, evaluate, manage prompts and datasets, and debug issues related to an LLM application’s performance. It creates open telemetry standard traces for any LLM which helps with observability and works with any observability client. Free plan offers 50K traces/month.","tier":"Free","url":"https://langtrace.ai","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-05-07","vendor_page":"/vendor/langtrace"},{"vendor":"LangWatch","category":"specialized","description":"A LLMOps platform helping AI teams measure, monitor, and optimize LLM applications for reliability, cost-efficiency, and performance. With a powerful DSPy component, we enable seamless collaboration between engineers and non-technical teams to fine-tune and productionize GenAI products. Free plan includes all platform features, 1k traces/month and 1 workflow DSPy optimizers. [#opensource](https://github.com/langwatch/langwatch)","tier":"Free","url":"https://langwatch.ai/pricing","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-10","vendor_page":"/vendor/langwatch"},{"vendor":"Lumenfall.ai","category":"specialized","description":"AI media gateway providing unified access to leading image generation models via an OpenAI-compatible API. The platform itself is free to use with zero markup and no subscription fee. Inference costs for most models are billed at provider price, but FLUX.1 [schnell] FP8 is offered free forever with unlimited usage for registered users. Built-in failover and provider resilience included.","tier":"Free","url":"https://lumenfall.ai/pricing","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-09-01","vendor_page":"/vendor/lumenfall-ai"},{"vendor":"Mediaworkbench.ai","category":"specialized","description":"MediaWorkbench.ai offers 100,000 free words for Azure OpenAI, DeepSeek, and Google Gemini models, enabling users to access powerful tools for code generation, deep research, and image creation.","tier":"Free","url":"https://mediaworkbench.ai","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-10","vendor_page":"/vendor/mediaworkbench-ai"},{"vendor":"Othor AI","category":"specialized","description":"An AI-native fast, simple, and secure alternative to popular business intelligence solutions like Tableau, Power BI, and Looker. Othor utilizes large language models (LLMs) to deliver custom business intelligence solutions in minutes. The Free Forever plan provides one workspace with five datasource connections for one user, with no limits on analytics. [#opensource](https://github.com/othorai/othor.ai)","tier":"Free","url":"https://othor.ai/pricing","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-09","vendor_page":"/vendor/othor-ai"},{"vendor":"Pollinations.AI","category":"specialized","description":"easy-to-use, free image generation AI with free API available. No signups or API keys required, and several option for integrating into a website or workflow. [#opensource](https://github.com/pollinations/pollinations)","tier":"Free","url":"https://pollinations.ai/","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-07","vendor_page":"/vendor/pollinations-ai"},{"vendor":"Portkey","category":"specialized","description":"Control panel for Gen AI apps featuring an observability suite & an AI gateway. Send & log up to 10,000 requests for free every month.","tier":"Free","url":"https://portkey.ai/","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-12","vendor_page":"/vendor/portkey"},{"vendor":"ReportGPT","category":"specialized","description":"AI Powered Writing Assistant. The entire platform is free as long as you bring your own API key.","tier":"Free","url":"https://ReportGPT.app","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-06","vendor_page":"/vendor/reportgpt"},{"vendor":"Zenable","category":"specialized","description":"Instantly auto-fix outputs from tools like Cursor, Windsurf, and Copilot to meet your company's quality and compliance standards using guardrails built with Policy as Code. The free tier includes 100 tools calls per day to the MCP server and 25 free automated pull request reviews per day via the GitHub App.","tier":"Free","url":"https://zenable.io","tags":["ai","ml","generative-ai","free-for-dev"],"verifiedDate":"2026-08-08","vendor_page":"/vendor/zenable"},{"vendor":"Vast.ai","category":"specialized","description":"GPU marketplace — startup program offers $2,500 in free GPU credits. Access to A100s, H200s, and consumer cards. Up to 81% savings vs. major cloud GPU providers. 24/7 priority support during credit period. No long-term contracts","tier":"Startup Program","url":"https://vast.ai/startup","tags":["gpu","ai","ml","compute","startup credits","infrastructure"],"verifiedDate":"2026-08-14","vendor_page":"/vendor/vast-ai"},{"vendor":"xAI","category":"frontier","description":"Sign-up gives $25 in free API credits. Additional $150/month via data sharing program (opt-in, requires $5 minimum spend first). Access to Grok models including Grok 4.1 series. Starting at $0.20/M input tokens, $0.50/M output tokens for Grok 4.1 Fast.","tier":"Free Credits","url":"https://docs.x.ai/developers/models","tags":["ai","ml","llm","inference","free credits","api"],"verifiedDate":"2026-08-14","vendor_page":"/vendor/xai"},{"vendor":"Anthropic API","category":"frontier","description":"Claude API access with usage-based pricing. Opus 4.6: $5/$25 per MTok (input/output), 67% below previous Opus pricing. Sonnet 4.6: $3/$15 per MTok. Haiku 4.5: $0.80/$4 per MTok. Batch API at 50% discount. Free tier: limited access via console with rate limits.","tier":"Pay-as-you-go","url":"https://docs.anthropic.com/en/docs/about-claude/models","tags":["ai","ml","api","llm","claude","deal-change"],"verifiedDate":"2026-08-15","vendor_page":"/vendor/anthropic-api"},{"vendor":"GitHub Models","category":"specialized","description":"GitHub retired GitHub Models on 2026-07-30, so there is no free tier. GitHub's own documentation states \"GitHub Models has been retired.\" The former offer was free access to 100+ models via GitHub Marketplace at 10-15 RPM and 50-150 requests/day.","tier":"Retired","url":"https://docs.github.com/en/github-models/about-github-models","tags":["ai","ml","llm","inference","github","free tier","openai-compatible"],"verifiedDate":"2026-08-20","vendor_page":"/vendor/github-models"},{"vendor":"NVIDIA NIM","category":"inference","description":"Free serverless APIs for LLM inference — access Llama 3.1, Mistral, and NVIDIA models. Free tier: ~40 RPM, 1,000 free API credits. No credit card required for development","tier":"Free","url":"https://build.nvidia.com","tags":["ai","ml","llm","inference","nvidia","gpu","free tier"],"verifiedDate":"2026-08-15","vendor_page":"/vendor/nvidia-nim"},{"vendor":"Ollama Cloud","category":"specialized","description":"Cloud-hosted Ollama for running open-source LLMs — free tier for light usage with 1 concurrent model. Access Llama, Mistral, Gemma, and other open models via API","tier":"Free","url":"https://ollama.com/pricing","tags":["ai","ml","llm","inference","open-source","self-hosted","free tier"],"verifiedDate":"2026-08-19","vendor_page":"/vendor/ollama-cloud"},{"vendor":"LLM7.io","category":"specialized","description":"UK-based free LLM inference gateway. Free tier supported by donors — access to 30+ models including DeepSeek R1, Qwen2.5 Coder, text, image, and speech-to-text models. No published rate limits on free tier.","tier":"Free","url":"https://llm7.io","tags":["ai","ml","llm","inference","gateway","free tier"],"verifiedDate":"2026-08-17","vendor_page":"/vendor/llm7-io"},{"vendor":"SiliconFlow","category":"inference","description":"Free inference for open-source models with 100 requests/day limit and $1 free credits. Supports DeepSeek-R1, DeepSeek-V3, QwQ-32B, and other open-source models. China-based provider.","tier":"Free (Limited)","url":"https://siliconflow.cn/pricing","tags":["ai","ml","llm","inference","open-source","free tier"],"verifiedDate":"2026-08-20","vendor_page":"/vendor/siliconflow"},{"vendor":"Zhipu AI","category":"specialized","description":"Free tier for GLM-4 series models. 20 million tokens welcome package plus permanently free Flash models (GLM-4.7-Flash, GLM-4.5-Flash, GLM-4.6V-Flash) with no rate limits or expiration. China-based, function calling support.","tier":"Free","url":"https://open.bigmodel.cn","tags":["ai","ml","llm","inference","glm","free tier"],"verifiedDate":"2026-08-20","vendor_page":"/vendor/zhipu-ai"},{"vendor":"DeepSeek API","category":"specialized","description":"Among the cheapest LLM APIs available. DeepSeek V3.2 (deepseek-chat): $0.28/M input, $0.42/M output (1M context). DeepSeek V3.2 Thinking (deepseek-reasoner): $0.28/M input, $0.42/M output. Cache hits 90% cheaper. Off-peak discounts available. China-based.","tier":"Free Credits + Pay-as-you-go","url":"https://api-docs.deepseek.com/quick_start/pricing","tags":["ai","ml","llm","inference","api","deepseek","reasoning"],"verifiedDate":"2026-08-18","vendor_page":"/vendor/deepseek-api"},{"vendor":"MiniMax","category":"specialized","description":"Chinese AI lab (稀宇科技, Shanghai) offering multimodal APIs via platform.minimax.io. Pay-as-you-go: MiniMax-M2.7 text $0.3/$1.2 per Mtok (highspeed $0.6/$2.4), prompt-cache read $0.06/Mtok; speech/TTS $60/M chars (standard) or $100/M chars (HD), rapid voice clone $1.50/voice; MiniMax-Hailuo-02/2.3 video $0.10-$0.56/video depending on resolution and duration; image-01 $0.0035/image; Music-2.6 $0.15 per ~5 min with a 2-week free trial. Alternative Token Plan subscriptions from $10/mo (Starter: 1,500 M2.7 requests/5hr) to $150/mo (Ultra-Highspeed: 30K requests/5hr, 5 Hailuo videos/day). No documented permanent API free tier. Also maker of Hailuo AI (consumer video) and Talkie.","tier":"Pay-as-you-go","url":"https://platform.minimax.io/docs/guides/pricing-paygo","tags":["ai","ml","llm","api","inference","video","speech","tts","image-generation","music","multimodal","china"],"verifiedDate":"2026-07-09","vendor_page":"/vendor/minimax"},{"vendor":"Exa","category":"specialized","description":"AI-native search engine API — x402 pay-per-query search, semantic and keyword search, auto-generated summaries, contents retrieval. No API key required with x402. 1,000 free searches/month with API key.","tier":"Free","url":"https://exa.ai/pricing","tags":["ai","search","api","semantic-search","x402","agent-payments","free tier"],"verifiedDate":"2026-08-16","vendor_page":"/vendor/exa"},{"vendor":"GPU-Bridge","category":"specialized","description":"Multi-provider GPU inference API — 30+ AI services (LLM, image generation, embeddings, speech) accessible via x402 micropayments. Unified API across providers, no accounts needed. Pay per inference call.","tier":"Pay-per-use","url":"https://gpu-bridge.com","tags":["gpu","inference","llm","ai","x402","agent-payments","machine-learning"],"verifiedDate":"2026-04-14","vendor_page":"/vendor/gpu-bridge"},{"vendor":"Google Vertex AI Agent Engine","category":"specialized","description":"Vertex AI Agent Builder with sessions and memory now GA. Build, deploy, and manage AI agents with persistent memory, multi-turn sessions, and tool governance. Part of the Vertex AI platform.","tier":"Pay-as-you-go","url":"https://cloud.google.com/vertex-ai/docs/agent-builder","tags":["ai","agents","google cloud","vertex-ai","agentic"],"verifiedDate":"2026-04-15","vendor_page":"/vendor/google-vertex-ai-agent-engine"},{"vendor":"Google Vector Search 2.0","category":"specialized","description":"Vector Search 2.0 now GA on Vertex AI. High-performance approximate nearest neighbor (ANN) vector similarity search for AI/RAG applications. Supports billion-scale indexes.","tier":"Pay-as-you-go","url":"https://cloud.google.com/vertex-ai/docs/vector-search/overview","tags":["ai","vectors","search","google cloud","vertex-ai","rag"],"verifiedDate":"2026-08-15","vendor_page":"/vendor/google-vector-search-2-0"},{"vendor":"Google GLM 5","category":"specialized","description":"Experimental model optimized for complex systems engineering and agentic tasks. Designed for multi-step reasoning, code generation, and autonomous agent workflows. Preview access via Vertex AI.","tier":"Experimental Preview","url":"https://cloud.google.com/vertex-ai/docs/generative-ai/model-reference/overview","tags":["ai","llm","google cloud","vertex-ai","agentic","experimental"],"verifiedDate":"2026-08-19","vendor_page":"/vendor/google-glm-5"},{"vendor":"Google Gemini Embedding 2","category":"frontier","description":"First natively multimodal embedding model — text, images, video, audio, and documents in a single embedding space. Enables cross-modal similarity search and retrieval. Available via Vertex AI and Gemini API.","tier":"Pay-as-you-go","url":"https://cloud.google.com/vertex-ai/docs/generative-ai/embeddings/get-text-embeddings","tags":["ai","embeddings","multimodal","google cloud","vertex-ai","rag"],"verifiedDate":"2026-08-17","vendor_page":"/vendor/google-gemini-embedding-2"},{"vendor":"Alibaba Cloud Qwen Code","category":"specialized","description":"Qwen Code free tier reduced to 100 requests/day (from 1,000). Full access requires Coding Plan Pro ($50/month). AI coding assistant based on Qwen LLM family.","tier":"Free (Reduced)","url":"https://qwen.ai","tags":["ai","ml","llm","coding","free tier","deal-change"],"verifiedDate":"2026-08-18","vendor_page":"/vendor/alibaba-cloud-qwen-code"},{"vendor":"E2B","category":"specialized","description":"Cloud sandboxes for AI agents. Secure code execution environments with filesystem, process management, and network access. Free tier: 100 sandbox hours/month","tier":"Free (100 sandbox hrs/mo)","url":"https://e2b.dev/pricing","tags":["ai","sandboxes","code-execution","agents","free tier"],"verifiedDate":"2026-08-22","vendor_page":"/vendor/e2b"},{"vendor":"Together AI","category":"specialized","description":"Fast inference API for open-source LLMs (Llama, Mixtral, Code Llama). Free tier: $1 free credits on signup. Pay-per-token after","tier":"Free ($1 credits)","url":"https://www.together.ai/pricing","tags":["ai","inference","llm","open-source","free tier"],"verifiedDate":"2026-08-24","vendor_page":"/vendor/together-ai"},{"vendor":"Fireworks AI","category":"specialized","description":"Fast inference platform for LLMs and image models. Free tier: $1 free credits. Serverless and on-demand deployment options","tier":"Free ($1 credits)","url":"https://fireworks.ai/pricing","tags":["ai","inference","llm","image-generation","free tier"],"verifiedDate":"2026-08-25","vendor_page":"/vendor/fireworks-ai"},{"vendor":"Modal","category":"specialized","description":"Serverless cloud for AI/ML. Run code on GPUs without managing infrastructure. Free tier: $30/month free compute credits","tier":"Free ($30/mo credits)","url":"https://modal.com/pricing","tags":["ai","gpu","serverless","inference","free tier"],"verifiedDate":"2026-08-21","vendor_page":"/vendor/modal"}],"changes":[{"vendor":"Cloudflare Workers AI","change_type":"pricing_restructured","date":"2026-09-01","date_source":"discovered","summary":"The pricing has been restructured. While a free tier of 10,000 Neurons per day still exists, the pricing is now more granular and based on per-model unit pricing, billed in Neurons at $0.011 / 1,000 Neurons. Some models now require a paid plan or AI Gateway credits.","previous_state":"AI inference at the edge — 10,000 neurons/day free across text generation, image classification, translation, speech-to-text models","current_state":"Workers AI has a free tier with 10,000 Neurons per day. Usage above this is $0.011 / 1,000 Neurons. Some models require a paid plan or AI Gateway credits.","impact":"high","source_url":"https://developers.cloudflare.com/workers-ai/platform/pricing/","category":"AI / ML","alternatives":[],"detected_by":"reverify-ai","recorded_date":"2026-09-01"},{"vendor":"OpenAI","change_type":"free_tier_removed","date":"2026-08-26","summary":"Assistants API deprecated, full shutdown August 26, 2026. Developers must migrate to Responses API + Conversations API","previous_state":"Assistants API available with code interpreter, file search, function calling, and Threads for conversation management","current_state":"Deprecated. Full shutdown Aug 26, 2026. Migration required: Assistants → Prompts + Responses API, Threads → Conversations API. No automated migration tool provided","impact":"high","source_url":"https://community.openai.com/t/assistants-api-beta-deprecation-august-26-2026-sunset/1354666","category":"AI/ML","alternatives":["Anthropic Claude API","Google Gemini API","Cohere API"],"recorded_date":"2026-02-26","date_source":"hand_written"},{"vendor":"Google Gemini API","change_type":"product_deprecated","date":"2026-06-01","summary":"Gemini 2.0 Flash and 2.0 Flash-Lite will be deprecated June 1, 2026. Developers must migrate to Gemini 2.5 Flash or 3.x Flash models. Google is consolidating the model lineup — 2.0 generation reaching end of life as 2.5 and 3.x become production-ready.","previous_state":"Gemini 2.0 Flash and 2.0 Flash-Lite available on free and paid tiers","current_state":"Gemini 2.0 Flash and 2.0 Flash-Lite deprecated June 1, 2026. Migrate to 2.5 Flash or 3.x Flash","impact":"high","source_url":"https://ai.google.dev/gemini-api/docs/deprecations","category":"AI / ML","alternatives":["Gemini 2.5 Flash","Gemini 3.0 Flash","OpenRouter","Groq"],"recorded_date":"2026-04-17","date_source":"hand_written"},{"vendor":"OpenAI","change_type":"free_tier_removed","date":"2026-05-12","summary":"DALL-E 2 and DALL-E 3 API access discontinued. Developers must migrate to gpt-image-1 (different pricing model, quality tiers changed from standard/hd to low/medium/high) or switch to free alternatives like Pollinations.AI or Lumenfall.ai","previous_state":"DALL-E 3 API: $0.040-$0.120/image (1024x1024), paid via API credits","current_state":"DALL-E 2 & 3 API shut down. Replacement: gpt-image-1 at $0.011-$0.167/image with new quality tiers","impact":"high","source_url":"https://platform.openai.com/docs/deprecations","category":"AI / ML","alternatives":["gpt-image-1","Pollinations.AI","Lumenfall.ai","Cloudflare Workers AI","Stability AI","Replicate"],"recorded_date":"2026-04-10","date_source":"hand_written"},{"vendor":"OpenAI","change_type":"product_deprecated","date":"2026-05-07","summary":"Realtime API beta endpoints deprecated. Developers must remove OpenAI-Beta header, use new client_secrets endpoint, specify session_type, and update event names. GA Realtime API is the direct replacement.","previous_state":"Realtime API beta: required OpenAI-Beta: realtime=v1 header, single session type, beta event names","current_state":"Realtime API GA: no beta header, POST /v1/realtime/client_secrets for ephemeral keys, session_type required (speech-to-speech or transcription), updated event names","impact":"high","source_url":"https://platform.openai.com/docs/deprecations","category":"AI / ML","alternatives":["OpenAI Realtime API (GA)","Deepgram","AssemblyAI","Azure OpenAI Realtime","ElevenLabs","Google Cloud Speech-to-Text"],"recorded_date":"2026-04-10","date_source":"hand_written"},{"vendor":"DeepSeek","change_type":"pricing_restructured","date":"2026-04-17","summary":"DeepSeek V3.2 replaces V4 branding. Pricing dropped: chat model $0.30→$0.28/M input, $0.50→$0.42/M output. Reasoner model pricing unified with chat at $0.28/$0.42 (was $0.55/$2.19). Free token welcome package appears removed.","previous_state":"DeepSeek V4: $0.30/M input, $0.50/M output. R1: $0.55/M input, $2.19/M output. 5M free tokens for new accounts.","current_state":"DeepSeek V3.2 (chat + reasoner): $0.28/M input, $0.42/M output. Cache hits 90% cheaper. No free token mention on pricing page.","impact":"medium","source_url":"https://api-docs.deepseek.com/quick_start/pricing","category":"AI / ML","alternatives":["OpenAI API","Anthropic Claude API","Google Gemini API"],"recorded_date":"2026-04-16","date_source":"hand_written"},{"vendor":"xAI","change_type":"free_tier_removed","date":"2026-04-13","summary":"$25/month free API credits no longer offered — Grok API paid-only","previous_state":"","current_state":"$25/month free API credits no longer offered — Grok API paid-only","impact":"low","source_url":"","category":"AI / ML","alternatives":[],"recorded_date":"2026-04-13","date_source":"hand_written"},{"vendor":"Gem","change_type":"product_deprecated","date":"2026-04-12","summary":"Removed: source page no longer accessible or deal program discontinued","previous_state":"3 months free. Access via: Free for Stripe Corporate Credit card users","current_state":"Removed from index","impact":"low","source_url":"https://stripe.com/en-de/corporate-card","category":"Startup Perks","alternatives":[],"recorded_date":"2026-04-12","date_source":"hand_written"},{"vendor":"Mode","change_type":"product_deprecated","date":"2026-04-12","summary":"Removed: source page no longer accessible or deal program discontinued","previous_state":"$12,000 credits for 12 months, free White Label Embeds, then 50% graduation discount in year one.. A","current_state":"Removed from index","impact":"low","source_url":"https://segment.com/industry/startups/","category":"Startup Perks","alternatives":[],"recorded_date":"2026-04-12","date_source":"hand_written"},{"vendor":"Google Gemini API","change_type":"restriction","date":"2026-04-08","summary":"Gemini API free tier restricted to Flash and Flash-Lite models only (April 2026). Gemini 2.5 Pro and other Pro models now require a paid billing account. Previously free users could access Pro models with rate limits. Combined with April 1 spend cap enforcement, this significantly narrows what is available at $0.","previous_state":"Free tier included access to both Flash and Pro models with rate limits. Gemini 2.5 Pro accessible to free users at reduced rates","current_state":"Free tier restricted to Flash and Flash-Lite models only. Pro models (Gemini 2.5 Pro) require paid billing account. Free users limited to lighter-weight models","impact":"high","source_url":"https://ai.google.dev/gemini-api/docs/pricing","category":"AI / ML","alternatives":["OpenRouter","Groq","Together AI","Anthropic Claude API"],"recorded_date":"2026-04-08","date_source":"hand_written"},{"vendor":"OpenAI Codex","change_type":"pricing_restructured","date":"2026-04-03","summary":"Switched from per-seat subscription to pay-as-you-go token-based pricing. Teams can add Codex-only seats billed on token consumption with no rate limits. ChatGPT Business price cut from $25 to $20/month (annual). $100 credit per new Codex team member (up to $500/team, limited time)","previous_state":"Per-seat subscription model, Business plan $25/user/mo","current_state":"Pay-as-you-go token-based pricing, Business $20/user/mo (annual), Codex-only seats with usage-based billing, $100 new member credits","impact":"medium","source_url":"https://openai.com/index/codex-flexible-pricing-for-teams/","category":"AI Coding","alternatives":["GitHub Copilot","Claude Code","Devin"],"recorded_date":"2026-04-15","date_source":"hand_written"},{"vendor":"Google Gemini API","change_type":"restriction","date":"2026-04-01","summary":"Billing-account-level spend caps enforced starting April 1, 2026. When a tier's spend cap is reached, API requests pause until the next billing month. Affects pay-as-you-go developers who may see unexpected request pausing.","previous_state":"No hard spend caps — usage billed without automatic pausing","current_state":"Tier-level spend caps enforced at billing account level. Requests pause when cap is hit until next month","impact":"medium","source_url":"https://ai.google.dev/gemini-api/docs/billing","category":"AI / ML","alternatives":["OpenRouter","Anthropic Claude API","OpenAI API"],"recorded_date":"2026-03-26","date_source":"hand_written"},{"vendor":"Google","change_type":"pricing_restructured","date":"2026-03-30","summary":"Google Developer Program annual subscription ending March 30. Replaced by Google AI Pro ($10/mo) and AI Ultra ($100/mo) tiers with bundled cloud credits and AI model access.","previous_state":"Google Developer Program: annual subscription with standalone cloud credits, developer tools access, and learning resources","current_state":"Google AI Pro ($10/mo) and AI Ultra ($100/mo) tiers bundle cloud credits with AI model access (Gemini, etc). Developer benefits folded into AI-centric subscriptions","impact":"high","source_url":"https://developers.google.com/profile/help/benefits","category":"Cloud IaaS","alternatives":["AWS Activate","Microsoft for Startups","Google for Startups Cloud"],"verified_date":"2026-03-04","recorded_date":"2026-03-11","date_source":"hand_written"},{"vendor":"xAI (Grok)","change_type":"free_tier_removed","date":"2026-03-19","summary":"Grok Imagine free image/video generation locked behind SuperGrok subscription ($30/mo). Over 1B generations in first month overwhelmed capacity","previous_state":"Free image and video generation for all Grok users","current_state":"Image/video generation requires SuperGrok subscription ($30/month). Text-only Grok remains free with limits","impact":"medium","source_url":"https://x.com/xai","category":"AI / ML","alternatives":["DALL-E (credits)","Stable Diffusion (self-hosted)","Pollinations.AI"],"recorded_date":"2026-03-24","date_source":"hand_written"},{"vendor":"Anthropic Claude","change_type":"limits_increased","date":"2026-03-13","summary":"Temporary usage promotion: double five-hour usage limits during off-peak hours, March 13-27, 2026. Applies to Claude Pro and Team subscribers","previous_state":"Standard five-hour usage limits for Pro and Team subscribers","current_state":"2x usage limits during off-peak hours (March 13-27, 2026 only). Temporary promotion to showcase increased capacity","impact":"low","source_url":"https://www.anthropic.com/news","category":"AI/ML","alternatives":["OpenAI ChatGPT Plus","Google Gemini Advanced"],"recorded_date":"2026-03-15","date_source":"hand_written"},{"vendor":"Google Gemini 2.0 Flash","change_type":"product_deprecated","date":"2026-03-03","summary":"Gemini 2.0 Flash and 2.0 Flash-Lite models deprecated, scheduled for retirement. Free tier continues with Gemini 2.5 models. Developers on 2.0 must migrate","previous_state":"Gemini 2.0 Flash and Flash-Lite available as free-tier models with standard rate limits","current_state":"2.0 Flash and Flash-Lite deprecated. Developers must migrate to Gemini 2.5 Flash or 2.5 Pro. Free tier preserved for 2.5 series","impact":"medium","source_url":"https://www.isumsoft.com/internet/gemini-2-flash-deprecation-migration-guide.html","category":"AI/ML","alternatives":["Google Gemini 2.5 Flash","Google Gemini 2.5 Pro","OpenRouter","Groq"],"recorded_date":"2026-02-26","date_source":"hand_written"},{"vendor":"OpenAI","change_type":"limits_reduced","date":"2026-02-09","summary":"Ads launched in ChatGPT Free and Go ($8/mo) tiers. Sponsored units from major brands appear below responses on first prompt. $60 CPM, $200K minimum ad commitment","previous_state":"Ad-free experience across all ChatGPT tiers including Free","current_state":"Free and Go tiers show contextual ads (labeled 'sponsored') below responses. Plus/Pro/Business/Enterprise/Education tiers remain ad-free. Users can opt out at cost of reduced daily message limits","impact":"high","source_url":"https://openai.com/index/testing-ads-in-chatgpt/","category":"AI/ML","alternatives":["Anthropic Claude","Google Gemini","Perplexity"],"recorded_date":"2026-02-26","date_source":"hand_written"},{"vendor":"Anthropic","change_type":"limits_increased","date":"2026-02-05","summary":"Claude Opus 4.6 API pricing at $5/$25 per MTok (input/output) — 67% below previous Opus 4/4.1 pricing of $15/$75. Frontier AI model at mid-tier prices","previous_state":"Opus 4/4.1 priced at $15/$75 per million tokens (input/output)","current_state":"Opus 4.6 at $5/$25 per MTok. Extended context (>200K tokens): $10/$37.50. Batch API: $2.50/$12.50. Price reduction began with Opus 4.5 generation","impact":"high","source_url":"https://docs.anthropic.com/en/docs/about-claude/models","category":"AI/ML APIs","alternatives":["OpenAI GPT-4o","Google Gemini","Mistral"],"recorded_date":"2026-02-26","date_source":"hand_written"},{"vendor":"Cloudflare","change_type":"new_free_tier","date":"2026-02-04","summary":"Queues added to Workers free plan — message queuing now free","previous_state":"Queues was paid-only","current_state":"Queues included in free Workers plan","impact":"medium","source_url":"https://developers.cloudflare.com/queues/","category":"Cloud IaaS","alternatives":[],"recorded_date":"2026-02-25","date_source":"hand_written"},{"vendor":"Google","change_type":"pricing_restructured","date":"2026-01-27","summary":"Google Developer Program Premium merged into Google One AI Pro ($19.99/mo, includes $10 Cloud credits) and AI Ultra ($100 Cloud credits). Developer benefits now bundled into consumer AI subscriptions","previous_state":"Separate Developer Program Premium subscription ($299/year) with dev tools and perks","current_state":"Developer benefits folded into AI Pro ($19.99/mo, $10 Cloud credits for Vertex AI/Cloud Run/Gemini API) and AI Ultra ($100 Cloud credits). Developer Program Premium being phased out for @gmail.com accounts","impact":"medium","source_url":"https://blog.google/innovation-and-ai/technology/developers-tools/gdp-premium-ai-pro-ultra/","category":"Developer Tools","alternatives":["GitHub Copilot","JetBrains AI"],"recorded_date":"2026-02-26","date_source":"hand_written"},{"vendor":"Google Gemini","change_type":"limits_reduced","date":"2025-12-15","summary":"Google quietly slashed Gemini API free tier rate limits by 50-80% in late 2025. Gemini 2.5 Flash went from ~250 requests/day to 20-50. Pro model free tier removed entirely. A Google PM admitted generous limits were only supposed to be available for a single weekend","previous_state":"Pro model available on free tier, Flash ~250 RPD, Flash-Lite generous limits","current_state":"Pro free tier removed. Flash reduced to 10 RPM (~20-50 RPD). Flash-Lite 15 RPM. 50-80% reduction across the board. Gemini 2.0 Flash retiring March 2026","impact":"high","source_url":"https://ai.google.dev/gemini-api/docs/pricing","category":"AI / ML","alternatives":["OpenRouter","Groq","Together AI"],"recorded_date":"2026-03-09","date_source":"hand_written"},{"vendor":"OpenAI","change_type":"limits_reduced","date":"2025-06-01","summary":"Free trial credits ($5-$18 for new accounts) completely discontinued. Free tier now limited to GPT-3.5 Turbo at 3 RPM with no credits. All advanced models (GPT-4, DALL-E, Whisper) require paid access","previous_state":"$5-$18 in free trial credits for new accounts, 3-month expiry, access to all models","current_state":"No free credits. Free tier: GPT-3.5 Turbo only, 3 requests/minute. Must add payment method for any other model access","impact":"high","source_url":"https://openai.com/api/pricing/","category":"AI / ML","alternatives":["Google Gemini API","Anthropic Claude API","Mistral API","Groq"],"recorded_date":"2026-03-11","date_source":"hand_written"}],"count":70,"categories":["frontier","inference","open-source-host","specialized"]}