Compare 65+ free AI/ML tools — Groq, Cerebras, OpenAI, Hugging Face, GitHub Copilot, Cursor, Langfuse, and more. Exact free tier limits by AI domain. Verified April to September 2026.
AI infrastructure is evolving faster than any other developer tooling category. The good news: competition has driven generous free tiers. Groq offers blazing-fast Llama 3.3 70B inference at ~30 RPM free. Cerebras gives 1M tokens/day free. Mistral offers access to all models including Large and Codestral at 1B tokens/month. And open-source tools like Cline, Aider, and Gemini CLI are completely free — just bring your own API key.
This page compares every free AI and ML tool in our index — 82 tools across LLM APIs, AI coding assistants, ML platforms, observability, and specialized services. Whether you need an OpenAI alternative or a free AI coding assistant, we have the comparison with exact free tier limits.
Large language model APIs for text generation, embeddings, and more. These are the foundational building blocks for AI applications — from GPT-4 and Claude to open-source models like Llama and Mistral.
Ultra-fast LLM inference on LPU hardware — free tier: 30 RPM, 100K-500K tokens/day depending on model. Supports Llama 4 Scout 17B, Llama 3.3 70B, Qwen3 32B, Whisper, and more. No credit card required
Free tier covers Gemini 2.5 Pro plus the Flash-tier models: Gemini 2.5 Flash (10 RPM), Gemini 2.5 Flash-Lite (15 RPM), Gemini 3.0 Flash Preview, Gemini 3.1 Flash-Lite Preview, Gemini Embedding, and Gemma 4. 3.1 Pro Preview is paid-only. Per-model paid pricing: Gemini 3.1 Pro Preview $2/$12 per MTok (≤200K ctx, doubles above), Gemini 3.0 Flash Preview $0.50/$3, Gemini 3.1 Flash-Lite Preview $0.25/$1.50, Gemini 2.5 Pro $1.25/$10 (≤200K, doubles above), Gemini 2.5 Flash $0.30/$2.50. Gemini 2.0 Flash and 2.0 Flash-Lite deprecated June 1, 2026 — migrate to 2.5 Flash or 3.x Flash. All models support Batch/Flex at 50% discount. Mandatory spend caps enforced since April 1, 2026.
Free plan includes $10/mo in API credits and access to Mistral models in Studio, alongside limited messages, web searches and coding sessions. Paid API rates start at $0.5/M input and $1.5/M output tokens for Mistral Large; batch processing halves the price and cached input tokens cost up to 90% less.
Superseded: As of 2026-09-07, openrouter.ai/pricing reads: Free tier includes 25+ free models, 4 free providers, 50 reqs/day rate limit, $25,000 of list price inference / month with no fees, 5% fee after that. We are not publishing our stored OpenRouter terms beside it — our own pricing change record, discovered 2026-09-07, names them as the previous ones. Read what we recorded ↓
Superseded: As of 2026-09-01, developers.cloudflare.com/workers-ai/platform/pricing reads: Workers AI has a free tier with 10,000 Neurons per day. Usage above this is $0.011 / 1,000 Neurons. Some models require a paid plan or AI Gateway credits. We are not publishing our stored Cloudflare Workers AI terms beside it — our own pricing change record, discovered 2026-09-01, names them as the previous ones. Read what we recorded ↓
Ultra-fast LLM inference API. Free tier: 1M tokens/day, 10-30 requests/min (varies by model). Models include Llama 3.1 8B, Qwen 3 235B, GPT-OSS 120B. Multi-thousand tokens/sec inference speed
AI model API. Trial key: 1,000 API calls/month across all endpoints (Chat, Embed, Rerank). Access to Command R+, Rerank 3.5, Embed 4. Non-commercial use only
AI API platform. One model is priced Free in OpenAI's own table — the moderation model omni-moderation-latest. Everything else is per-token: embeddings from $0.02/1M, chat-latest $5.00/1M input and $30.00/1M output. Three further free amounts are sub-quotas inside paid tools: 1 GB per day of File search storage, 1 GB per account per month of ChatKit upload storage, and web-search content tokens on non-reasoning models. No free token allowance for the flagship models; trial credits for new accounts were discontinued in mid-2025.
AI media gateway providing unified access to leading image generation models via an OpenAI-compatible API. The platform itself is free to use with zero markup and no subscription fee. Inference costs for most models are billed at provider price, but FLUX.1 [schnell] FP8 is offered free forever with unlimited usage for registered users. Built-in failover and provider resilience included.
Social media scheduling, publishing and analytics. Free plan at $0/month. The pricing page states no per-plan entitlements: the feature list shown is identical for all four plans.
easy-to-use, free image generation AI with free API available. No signups or API keys required, and several option for integrating into a website or workflow. [#opensource](https://github.com/pollinations/pollinations)
Sign-up gives $25 in free API credits. Additional $150/month via data sharing program (opt-in, requires $5 minimum spend first). Access to Grok models including Grok 4.1 series. Starting at $0.20/M input tokens, $0.50/M output tokens for Grok 4.1 Fast.
Claude API access with usage-based pricing. Fable 5.1: $10/$50 per MTok (input/output). Opus 5: $5/$25 per MTok. Sonnet 5: $2/$10 per MTok. Haiku 4.5: $1/$5 per MTok. Batch API at 50% discount. Free tier: limited access via console with rate limits.
AI-powered coding assistants, agents, and app builders. From inline autocomplete (GitHub Copilot, Cursor) to fully autonomous agents (Devin, Claude Code) and no-code app builders (Bolt.new, Lovable).
AI-powered code editor built on VS Code. Hobby is the free plan: no credit card required, limited Agent requests, access to Composer — cursor.com/pricing states no completion or request figure for it. Individual paid plans: Pro $20/month, Pro+ $60/month (3x Pro's Agent limits), Ultra $200/month (20x). Teams $40/user/month (Standard; Premium at 5x Standard's Agent limits). Enterprise is custom-priced with pooled usage, invoice/PO billing and SCIM. Cursor renders only the selected tier's price in the page body; the full ladder is published in the page's JSON-LD Offer list.
Autonomous AI software engineer by Cognition Labs. Core plan pay-as-you-go starting at $20 (ACUs at $2.25 each). Team plan $500/mo (250 ACUs at $2.00 each). Enterprise with custom pricing, VPC deployment, SSO. Previously $500-only pricing before January 2026 restructuring
AI app builder by StackBlitz — generates frontend, backend, and database code in a live runtime environment. Free tier: 1M tokens/month (300K daily cap), unlimited databases, public + private projects, native hosting with Bolt branding, 10MB file upload limit. Paid plans from $20/mo for higher token limits.
AI app builder (formerly GPT Engineer) — generates full-stack apps from natural language. Free tier: 5 daily credits (up to 30/month), public projects only (private projects require paid plan), cloud hosting on lovable.app, Lovable branding badge. Credits consumed at variable rates (simple styling ~0.50, auth setup ~1.20, full landing page ~1.70). Pro $25/mo (100 credits + 5 daily). Business $50/mo.
Agentic coding tool by Anthropic — terminal-based AI coding agent that reads/writes files, runs commands, and manages git workflows. Available to Claude Pro ($20/mo), Team ($25/seat/mo), and Max ($100-200/mo) subscribers. Uses Claude Sonnet/Opus models. Also available via Anthropic API with usage-based pricing. Free tier via Claude.ai (limited Sonnet usage, Projects, Artifacts). April 2026 changes: (1) Third-party tool restrictions — external tools like OpenClaw now billed separately on pay-as-you-go basis, can no longer apply standard usage limits. (2) Extra usage option — all paid plans (Pro, Max 5x, Max 20x) now have pay-as-you-go overage at standard API rates beyond included usage. (3) Peak-hour throttling — 5-hour session limits reduced during weekdays 5 AM–11 AM Pacific.
AI coding assistant by GitHub/Microsoft. Works in VS Code/JetBrains/GitHub.com. Copilot Free: 2,000 completions/month, an allowance of GitHub AI Credits, limited agents, auto model selection only. Copilot Student is free to verified students. Pro $10/mo (base 1,000 AI credits), Pro+ $39/mo (3,900), Max $100/mo (10,000), Business $19/granted seat/mo (1,900 per user), Enterprise $39/granted seat/mo (3,900 per user). Usage beyond the allowance is billed at $0.01/AI credit; code completions and next edit suggestions are unlimited on paid plans and not billed in credits. The April 2026 pause on new Pro, Pro+ and Student signups has been lifted — all individual plans accept new signups today. New self-serve purchase of Copilot Business and Copilot Enterprise remains paused (2026-04-22), with GitHub saying sign-ups reopen soon for card and PayPal customers.
AI coding assistant by AWS. Free tier: inline code suggestions, chat, 1,000 lines code transformation/month, 50 agentic requests/month, access to latest Claude models. Supports VS Code, JetBrains, CLI, and AWS Console. Pro plan $19/user/month (higher limits, admin controls, IP indemnity)
Open-source autonomous AI coding agent for VS Code. Fully free — users provide their own API keys (OpenRouter, Anthropic, OpenAI, etc.). Supports file editing, terminal commands, browser interaction. No usage limits beyond API provider costs. MIT licensed.
Open-source AI pair programming CLI tool. Fully free — users provide their own LLM API keys (supports GPT-4o, Claude, Gemini, DeepSeek, Ollama local models). Git-aware editing, multi-file changes, voice coding, in-chat image support. Apache 2.0 licensed.
AI-powered IDE by Codeium (formerly Codeium Editor). Free tier: limited daily quotas for Cascade AI flows, code completions, AI chat (credit system replaced by quotas Mar 2026). Pro $20/mo, Teams $40/mo, Max $200/mo. SWE-1.5 Fast Agent model for faster iteration. Grandfathered Pro subscribers keep $15/mo indefinitely. All tiers include premium models
AI coding assistant with deep codebase understanding. No free tier. Standard $20/month flat and Business $100/month flat, each covering up to 50 seats with $20 and $100 of included monthly usage respectively across LLM, Context Engine and compute; top-ups are pay-as-you-go. Enterprise is custom-priced. Flat team pricing with no per-seat charge. Supports VS Code, JetBrains and CLI. SOC 2 Type II certified
Agent-first IDE by Google, powered by Gemini 3. 100% free during public preview — no paid tiers yet. Built-in browser automation, multi-agent orchestration, cross-platform (Mac/Windows/Linux). Announced November 2025.
Open-source terminal AI coding agent by Google. Free tier with personal Google account: 60 requests/minute, 1,000 requests/day, 1M token context window. Powered by Gemini 3. Supports MCP tools, shell commands, file editing. Apache 2.0 licensed
Cloud-native coding agent by OpenAI. Available with a ChatGPT subscription and, separately, with an API key billed on token use; both routes are current. ChatGPT Business has two seat types, both carrying Codex: Standard at $20/user/mo billed annually ($25 monthly) and Premium at $100/user/mo billed annually ($125 monthly), where Premium buys 5x more usage than Standard and removes the five-hour usage limit. Also in ChatGPT Plus ($20/mo) and Pro (from $100/mo). Codex-only pay-as-you-go seats closed to new Business workspaces in June 2026; existing seats continue. The Codex pricing page we cite lists only the Standard seat — the Premium seat is published on OpenAI's business pricing page. $100 credit per new Codex team member (up to $500/team, limited time). 2M+ weekly users. API via codex-mini at $1.50/1M input, $6/1M output tokens.
Model hosting, training, experiment tracking, and GPU compute. The infrastructure layer for building, training, and deploying custom ML models.
ML model hub — $0.10/month free inference credits, 200+ models via Inference Providers, unlimited model hosting on Hub
Data science platform (Google) — entirely free: 30 hrs/week GPU compute (NVIDIA Tesla P100), 20 hrs/week TPU with a 9-hour cap per session, 20 GB working disk per session. Unlimited public notebooks, datasets, and competitions
ML model hosting and inference platform — free runs on curated model collection without billing. Pay-per-second billing by hardware type (CPU/GPU) after free allowance. No credit card required to start
ML model deployment platform — $30 in free credits for new accounts. Basic plan is $0/month with pay-as-you-go billing after credits. Per-minute GPU/CPU billing for custom deployments, per-token for Model APIs
Superseded: As of 2026-09-03, wandb.ai/site/pricing reads: The free tier is $0/mo and includes 1 user seat, experiment tracking, registry & lineage tracking, and is for personal projects only. Corporate use is not allowed. We are not publishing our stored Weights & Biases terms beside it — our own pricing change record, discovered 2026-09-03, names them as the previous ones. Read what we recorded ↓
Superseded: As of 2026-08-28, comet.com/site/pricing reads: There are multiple free options: 'Free' (Open Source GitHub), 'Free Cloud' (up to 10 team members, 25k spans per month, 60-day data retention), and a 'Free' plan for academics. The 'Free Cloud' plan includes Agent tracing & analysis, Test Suites & assertions, and Agent Playground. We are not publishing our stored Comet ML terms beside it — our own pricing change record, discovered 2026-08-28, names them as the previous ones. Read what we recorded ↓
ML experiment tracking — free for individuals/researchers. 200 GB metadata storage, 50 GB file storage, 100K tracking calls/hour, 20 active runs, unlimited archived experiments, Jupyter/Colab integration
ML platform (now part of DigitalOcean). Free Gradient plan: public projects and 5 GB storage only. No free compute — all GPU and CPU instances are billed hourly. Paid plans from $8/month.
LLM monitoring, prompt engineering, evaluation, and debugging tools. Essential for production AI applications — track costs, latency, quality, and catch regressions.
Integration platform for AI Agents and LLMs. Integrate over 200+ tools across the agentic internet.
Simulate, evaluate, and observe your AI agents. Maxim is an end-to-end evaluation and observability platform, helping teams ship their AI agents reliably and >5x faster. Free forever for indie developers and small teams (3 seats).
AI engineering platform for evaluating and observing AI applications and agents. AX Free plan: 25K trace spans/month, 1 GB ingestion/month, 15 days data retention. Includes online evals, product observability, built-in Alyx agent, and community support.
Evals, prompt playground, and data management for Gen AI. Starter plan (free): 1 GB processed data/month, 10k scores/month, 14 days data retention. Unlimited users, projects, datasets, playgrounds, and experiments. Overage: $4/GB data, $2.50 per 1k scores.
Rebranded to Respan; keywordsai.co redirects to respan.ai. Free plan: full platform, 100k logs, 1k scores, 5 datasets, 2 evaluators, 5 prompts. No credit card required.
Open-source LLM engineering platform that helps teams collaboratively debug, analyze, and iterate on their LLM applications. Free forever plan includes 50k observations per month and all platform features. [#opensource](https://github.com/langfuse/langfuse)
enables developers to trace, evaluate, manage prompts and datasets, and debug issues related to an LLM application’s performance. It creates open telemetry standard traces for any LLM which helps with observability and works with any observability client. Free plan offers 50K traces/month.
A LLMOps platform helping AI teams measure, monitor, and optimize LLM applications for reliability, cost-efficiency, and performance. With a powerful DSPy component, we enable seamless collaboration between engineers and non-technical teams to fine-tune and productionize GenAI products. Free plan includes all platform features, 1k traces/month and 1 workflow DSPy optimizers. [#opensource](https://github.com/langwatch/langwatch)
Control panel for Gen AI apps featuring an observability suite & an AI gateway. Send & log up to 10,000 requests for free every month.
Superseded: As of 2026-09-01, zenable.io reads: Free accounts get 15 PR reviews/week and 100 agentic reviews/day. We are not publishing our stored Zenable terms beside it — our own pricing change record, discovered 2026-09-01, names them as the previous ones. Read what we recorded ↓
Speech-to-text, computer vision, vector databases, document parsing, and other domain-specific AI tools. These services handle specific AI tasks that general-purpose LLMs don't cover well.
Vector database — 2 GB storage, 2M write units/month, 1M read units/month, 5 indexes, 5M embedding tokens/month. Pinecone Assistant: 100 docs / 1 GB
Vector database — 1 GB free forever cluster, fully managed on AWS/GCP/Azure. Unlimited users, backups included. No credit card required
Speech-to-text and text-to-speech API — $200 free credits on signup (no credit card required), ~43K minutes transcription with Nova model, credits never expire. Access to all model endpoints
Superseded: As of 2026-08-28, roboflow.com/pricing reads: Public plan is free, requires no credit card, and includes 15 credits / month, 2 users, and Community Support. Data and models are open source on Roboflow Universe. We are not publishing our stored Roboflow terms beside it — our own pricing change record, discovered 2026-08-28, names them as the previous ones. Read what we recorded ↓
Data labeling and AI data platform — free tier: first 1,000 annotation units and 10,000 images uploaded free. Nucleus free tier for individuals and academia. Pay-as-you-go after free allocation
Speech-to-text and audio intelligence API — free: $50 in credits (~185 hours pre-recorded transcription). Up to 5 concurrent streams. Core STT and Audio Intelligence models. Pay-as-you-go at $0.15/hr after
Full-stack AI platform (computer vision, NLP, audio) — Community tier: 1,000 API calls/month. Access to pre-trained models for image recognition, NLP, and audio. No credit card required
Data labeling and annotation platform — free tier: 500 Labelbox Units (LBUs)/month. Image, video, text, and geospatial annotation. Free for qualified educational institutions
Augmented reality face filters for any platform with one SDK. The free plan provides up to 10 monthly active users (MAU) and tracks up to 4 faces
An OCR API parses image and pdf files that return the text results in JSON format. 25,000 requests per month are free and a 1MB file size limit.
20 free pages/month: Extract data from PDFs, emails. AI powered. Full API access.
Turn any unstructured documents (PDF, XLSX, JPG, PPTX, etc.) into structured JSON data. Parse, extract data, and edit PDF forms. Free tier with 15k free credits and pay-as-you-go.
API for online search and rapid insights and comprehensive research, with the capability of organization of research results. 1000 request/month for the Free tier with No credit card required.
AI-powered audio enhancer SaaS that removes noise and echo while preserving natural vocal clarity. totally Free: unlimited one-click enhancements, no login required, supports MP3/WAV/FLAC
Clinical AI Reference. Students have free access to the professional tool suite, which includes Open Search, Clinical Summary, Med Review, Drug Interactions, ICD-10 Codes, and Stewardship. Additionally, a free trial for the professional suite is available.
An AI-native fast, simple, and secure alternative to popular business intelligence solutions like Tableau, Power BI, and Looker. Othor utilizes large language models (LLMs) to deliver custom business intelligence solutions in minutes. The Free Forever plan provides one workspace with five datasource connections for one user, with no limits on analytics. [#opensource](https://github.com/othorai/othor.ai)
AI Powered Writing Assistant. The entire platform is free as long as you bring your own API key.
Additional AI and ML tools with free tiers.
Superseded: As of 2026-09-01, langchain.com/pricing reads: Developer plan is $0/seat/month with up to 5k base traces/month, then pay-as-you-go. LCU is $1.50/unit and LSU is $1.00/unit. Plus plan is $39/seat/month with up to 10k base traces/month, then pay-as-you-go. We are not publishing our stored LangSmith terms beside it — our own pricing change record, discovered 2026-09-01, names them as the previous ones. Read what we recorded ↓
GitHub retired GitHub Models on 2026-07-30, so there is no free tier. GitHub's own documentation states "GitHub Models has been retired." The former offer was free access to 100+ models via GitHub Marketplace at 10-15 RPM and 50-150 requests/day.
Free serverless APIs for LLM inference — access Llama 3.1, Mistral, and NVIDIA models. Free tier: ~40 RPM, 1,000 free API credits. No credit card required for development
Cloud-hosted Ollama for running open-source LLMs — free tier for light usage with 1 concurrent model. Access Llama, Mistral, Gemma, and other open models via API
UK-based free LLM inference gateway. Free tier supported by donors — access to 30+ models including DeepSeek R1, Qwen2.5 Coder, text, image, and speech-to-text models. No published rate limits on free tier.
Free inference for open-source models with 100 requests/day limit and $1 free credits. Supports DeepSeek-R1, DeepSeek-V3, QwQ-32B, and other open-source models. China-based provider.
Free tier for GLM-4 series models. 20 million tokens welcome package plus permanently free Flash models (GLM-4.7-Flash, GLM-4.5-Flash, GLM-4.6V-Flash) with no rate limits or expiration. China-based, function calling support.
Superseded: As of 2026-09-03, api-docs.deepseek.com/quick_start/pricing reads: Pricing is now per 1M tokens with different rates for input and output tokens, peak and off-peak hours, and cache hits/misses. For example, deepseek-v4-flash has input token prices ranging from $0.007 to $0.44 depending on these factors, and output tokens cost $0.66 to $1.32 per 1M tokens. We are not publishing our stored DeepSeek API terms beside it — our own pricing change record, discovered 2026-09-03, names them as the previous ones. Read what we recorded ↓
Superseded: As of 2026-08-28, platform.minimax.io/docs/guides/pricing-paygo reads: Music generation APIs are being discontinued for new users. Hailuo video pricing varies by resolution and duration, with models like MiniMax-Hailuo-2.3-Fast and MiniMax-H3. MiniMax-M3 is available with pay-as-you-go pricing. We are not publishing our stored MiniMax terms beside it — our own pricing change record, on 2026-08-20, names them as the previous ones. Read what we recorded ↓
Superseded: As of 2026-09-03, kiro.dev/pricing reads: KIRO FREE $0 per month 50 credits. KIRO PRO $20 per user / month 1,000 credits. KIRO PRO+ $40 per user / month 2,000 credits. KIRO PRO MAX $100 per user / month 5,000 credits. KIRO POWER $200 per user / month 10,000 credits. Add-on credits are $0.04/credit. $20 credit towards first paid plan upgrade. We are not publishing our stored Amazon Kiro terms beside it — our own pricing change record, discovered 2026-09-03, names them as the previous ones. Read what we recorded ↓
Superseded: As of 2026-09-03, exa.ai/pricing reads: Starter Free $20 credits on sign-up with $10 credits every month. No payment method required Includes: Free $10 credits per month MCP server access Claude Connector ChatGPT plugin 50+ integrations Access to all endpoints 5 Search Queries Per Second (QPS). We are not publishing our stored Exa terms beside it — our own pricing change record, discovered 2026-09-03, names them as the previous ones. Read what we recorded ↓
Multi-provider GPU inference API — 30+ AI services (LLM, image generation, embeddings, speech) accessible via x402 micropayments. Unified API across providers, no accounts needed. Pay per inference call.
Vertex AI Agent Builder with sessions and memory now GA. Build, deploy, and manage AI agents with persistent memory, multi-turn sessions, and tool governance. Part of the Vertex AI platform.
Vector Search 2.0 now GA on Vertex AI. High-performance approximate nearest neighbor (ANN) vector similarity search for AI/RAG applications. Supports billion-scale indexes.
First natively multimodal embedding model — text, images, video, audio, and documents in a single embedding space. Enables cross-modal similarity search and retrieval. Available via Vertex AI and Gemini API.
Qwen Code free tier reduced to 100 requests/day (from 1,000). Full access requires Coding Plan Pro ($50/month). AI coding assistant based on Qwen LLM family.
Superseded: As of 2026-09-05, e2b.dev/pricing reads: Hobby Free + Usage Costs One-time $100 of usage in credits Community support Up to 1-hour sandbox session length Up to 20 concurrently running sandboxes. We are not publishing our stored E2B terms beside it — our own pricing change record, discovered 2026-09-05, names them as the previous ones. Read what we recorded ↓
Fast inference API for open-source LLMs (Llama, Mixtral, Code Llama). Free tier: $1 free credits on signup. Pay-per-token after
Fast inference platform for LLMs and image models. Free tier: $1 free credits. Serverless and on-demand deployment options
Serverless cloud for AI/ML. Run code on GPUs without managing infrastructure. Free tier: $30/month free compute credits
Top free AI/ML tools compared by domain, free tier limits, and best use case.
| Service | Domain | Free Tier | OSS | Best For |
|---|---|---|---|---|
| Groq | LLM API | ~30 RPM, Llama 3.3 70B | No | Fastest free LLM inference (LPU hardware) |
| Cerebras | LLM API | 1M tokens/day, 30 RPM | No | High-volume free inference, Llama & Qwen |
| Mistral AI | LLM API | 1B tokens/mo, 2 RPM | No | Access to all Mistral models including Codestral |
| OpenRouter | LLM API | ~30 free models, ~20 RPM | No | Multi-model router, OpenAI-compatible API |
| GitHub Copilot | AI Coding | 2,000 completions/mo, 50 chats | No | Inline code completion in any IDE |
| Cursor | AI Coding | Limited Agent requests | No | AI-native code editor with Composer |
| Gemini CLI | AI Coding | 1,000 req/day, 60 RPM | Yes | Free terminal AI agent, open-source |
| Hugging Face | ML Platform | Free inference, unlimited hosting | Yes | Model hub, 200+ inference providers |
| Kaggle | ML Platform | 30 hrs/week GPU, 20 hrs TPU | No | Free GPU compute for ML training |
| Langfuse | Observability | 50K observations/mo | Yes | Open-source LLM observability & tracing |
| Deepgram | Speech AI | $200 free credits (~43K min) | No | Speech-to-text and text-to-speech API |
| Pinecone | Vector DB | 2 GB storage, 5 indexes | No | Managed vector database for RAG/search |
Groq and Cerebras lead on free LLM inference volume — Groq for speed (LPU), Cerebras for daily token quota (1M/day). Mistral offers the broadest model access on free tier (all models including Large). For AI coding, GitHub Copilot and Cursor both offer 2,000 free completions/month, while Gemini CLI is completely free and open-source. Langfuse is the standout for LLM observability (open-source, 50K observations free). Verified April to September 2026.
Looking for more? Browse all AI / ML and AI Coding tools in our full index of 1,547+ developer deals.
Get AI tool recommendations from your AI assistant. Compare LLM APIs, coding tools, and ML platforms — directly in your editor.
claude mcp add agentdeals -- npx -y agentdealsWorks with Claude Desktop, Cursor, Cline, Windsurf → Full setup guide