Most AI tool lists are: - ❌ Outdated (prices/limits from 2023) - ❌ Filled with affiliate links and sponsored placements - ❌ General-purpose directories with no developer focus - ❌ Missing production-critical details (rate limits, commercial use, architecture patterns)
Free AI Tools
Snapshot 2026-08-03 23:56:39 UTC · version 1
Research document
Free AI Tools
Curated list of free LLM APIs, coding copilots, AI IDEs, agents, and infrastructure tools for building real AI applications.
What's Inside
- ✅ Free GPT-5.5 / Claude Fable 5 / Gemini API access
- 🤖 Coding copilots and AI-native IDEs (Cursor, Trae, Windsurf)
- 💰 Cheapest AI APIs ($0.08-0.50 per 1M tokens)
- 📚 RAG stack tools (vector DBs, embeddings, frameworks)
- 🎯 Agent frameworks and automation tools
- 🔒 Local models for privacy (Ollama, Llama, Qwen)
- 🏗️ Production-ready stack configurations
- 🆕 Claude Fable 5, Claude Opus 4.8, Sonnet 4.6, Haiku 4.5 — GPT-5.5 (Instant/Thinking) — GitHub Copilot AI Credits — Windsurf Max — Trae Ultra — OpenCode 167k⭐ — Kiro Cloud Agent — Xiaomi MiMo V2.5 Pro
Goal: Help developers build AI apps without paying $200/month.
[!NOTE]
Please don't abuse these services, else we might lose them for everyone. The number becomes 550+ when you add all the models and sub services of all the tools provided. When raising issues or pull requests please dont add your own paid, expensive personal projects.
[!WARNING]
Early 2026 Model Tier Changes: Major providers (OpenAI, Anthropic, Google) have restricted flagship reasoning and pro models (GPT-5.5 Pro, Claude Fable 5, Gemini 3.1 Pro) to paid tiers. Free tiers now get highly optimized or lighter versions (GPT-5.5 Instant, Claude Sonnet/Haiku, Gemini Flash). Entries marked with[verify]need confirmation.June 2026 Pricing & Billing Updates: Windsurf switched to a quota-based model (Pro $20, Teams $40, new Max $200) on Mar 18. Trae moved to a 5-tier token system (Lite $3, Pro $10, Pro+ $30, Ultra $100) on Feb 24. Qoder's 50% launch promo ended Apr 30 — standard pricing is now Pro $20, Pro+ $60, Ultra $200. GitHub Copilot moved to usage-based billing (GitHub AI Credits) on Jun 1, with a new Max tier at $100. Anthropic added Claude Fable 5, Opus 4.8, Sonnet 4.6, and Haiku 4.5. Xiaomi MiMo V2.5 Pro API permanently cut 99% (May 26) — $0.435/$0.87 with $0.0036 cache.
🎯 Why This Repo Exists
Most AI tool lists are:
- ❌ Outdated (prices/limits from 2023)
- ❌ Filled with affiliate links and sponsored placements
- ❌ General-purpose directories with no developer focus
- ❌ Missing production-critical details (rate limits, commercial use, architecture patterns)
This repo focuses only on:
- ✅ Tools developers actually use in production
- ✅ Generous free tiers (no "5 requests then paywall")
- ✅ Production-capable models (SWE-bench verified, not toys)
- ✅ Real infrastructure (APIs, hosting, vector DBs, not just chatbots)
- ✅ Minimal fluff, maximum utility
Unlike: awesome-ai (general list), ai-collection (marketing focus), toolify (affiliate-heavy)
This is for: Builders who want to ship AI features this week.
⭐ Support This Project
If this repo helped you build something or saved you money:
⭐ Star this repo — it helps more builders discover free AI resources.
[🔄 Share with your team] — spread the knowledge.
📝 Contribute — found a new free tier? Updated pricing? PRs welcome!
📅 Updates
2026-06-25
- 🔄 Major model verification and name alignment: Migrated old placeholders to official Claude Fable 5, Claude Opus 4.8, and GPT-5.5 (Instant/Thinking/Pro) architectures.
2026-06-16
- 🆕 Added OpenCode (167k⭐ OSS CLI), AWS Kiro (full spec-driven family), Xiaomi MiMo Token Plan (Chinese coding subscription)
- 🧹 Removed weak/no-longer-free items from Free LLM providers: Cohere (non-commercial only), GitHub Models (Copilot-required), SambaNova/Hyperbolic (trial-only), HuggingFace (~$0.10/mo), Vercel ($5/mo), Mistral Codestral, Together AI, iFlow (7-day key), Perplexity API
- 🔄 Updated Gemini CLI entry: 3.1 Pro is paid-only; 3 Flash is the free tier (1,500 req/day)
- 🔄 Pricing refresh: Windsurf (Mar 18), Trae (Feb 24), Qoder (Apr 30), GitHub Copilot (Jun 1) billing changes
- ➕ Added GitHub Copilot Max tier ($100/mo, $200 AI Credits) and Claude Haiku 4.5
- 🐛 Fixed stale Cursor / Qoder / Windsurf / GitHub Copilot pricing throughout
2026-05-18
- ✨ added github PR review tools
2026-04-12
- ✨ added a website for easy navigation
2026-04-11
- ✨ Initial release
Table of Contents
- Quick Comparison
- Free LLM API Providers
- AI-Powered IDEs
- CLI Coding Tools
- API Providers for AI Coding Tools
- Paid Tiers Comparison
- Local Models
- free-coding-models CLI
- Additional 2026 AI Tools
- 🏗️ Recommended Stacks
- ⚡ Realtime & Streaming APIs
- 🎙️ Speech Models
- 🎨 Image Generation Models
- 🎬 Video Generation APIs
- 🌐 AI Browser Automation
- 💾 Cheap Vector DB Hosting
- 🏛️ Common AI Architecture Patterns
- 💵 Model Price Comparison
- 🎯 Best Models by Use Case
- ⏱️ Rate Limit Comparison
- ✅ Commercial Use Summary
- 🧩 RAG Stack Tools
- 🔢 Best Free Embedding APIs
- 🖥️ AI Hosting & GPU Providers
- 📊 AI Evaluation Tools
- 📐 Structured Output Tools
- 🏷️ Legend
- Contributing
- License
Quick Comparison
Free LLM API Providers Summary
| Provider | Models | Free Tier | Credit Card |
|---|---|---|---|
| NVIDIA NIM | 46 | 40 req/min | No |
| OpenRouter | 25 | 50/day (1K/day with $10) | No |
| Groq | 20+ | 1K-14.4K req/day | No |
| Google AI Studio | 9 | 5-500 req/day | No |
| Cloudflare Workers AI | 47+ | 10K neurons/day | No |
| Cerebras | 4 | 1M tokens/day | No |
| Mistral La Plateforme | 10+ | 1B tokens/month | No |
AI-Powered IDEs with Free Pro-Grade Access
| IDE | Pro-grade Models | Free Tier Limit | Credit Card |
|---|---|---|---|
| Cursor | GPT-5.5-Instant / Custom | Limited free tier (Hobby) | No |
| Trae | DeepSeek V4, GPT-5.5-Instant (Claude removed) | 5,000 auto-completions/month | No |
| Windsurf | OpenAI, Anthropic, Google, xAI | Light quota (daily/weekly) | No |
| Qoder | Qwen3.6-Plus, Qwen3-Coder-480B, GPT-5.5-Instant | Unlimited completions + limited chat | No |
AI GitHub PR Review Tools
| Tool | Starting Price | Free Tier | Features | Credit Card |
|---|---|---|---|---|
| PrixAI | Free / $10 paid plan | Free trial available | Unlimited reviews Auto-fix PRs, issue planning | No |
| Bito | Free / $25 paid plans | Free trial available | AI PR reviews/Unlimited reviews | No |
| Sourcery | ~$12/month | Free trial available | Code quality reviews | No |
CLI Coding Tools with Free Pro-Grade Access
| Tool | Pro-grade Models | Free Tier Limit | Credit Card |
|---|---|---|---|
| Gemini CLI | Gemini 3 Flash | 1,500 req/day | No |
| Rovo Dev CLI | Claude Sonnet 4.6, GPT-5.5-Instant | 5M tokens/day | No |
| Warp | GPT-5.5-Instant, Claude Sonnet 4.6 | 150 credits/mo (first 2 mo), 75/mo after | No |
| GitHub Copilot | GPT-5.5-Instant, Sonnet 4.6, Gemini Flash | 50 chat + 2K completions/month | No |
| Jules | Gemini 2.5 Pro | 15 tasks/day | No |
| AWS Kiro | Claude Fable 5 [verify], Opus 4.8, Sonnet 4.6 | 50 credits/month + 500 bonus | No |
| OpenCode | 75+ providers (BYOK) + Go bundle | Free (Zen) / Go $10/mo | No |
| Xiaomi MiMo | MiMo-V2.5-Pro, MiMo-V2.5, MiMo-V2-Omni | Free API credits | No |
| ForgeCode | 300+ models via OpenRouter | 10K tokens/day | No |
| Amazon Q Developer | Claude Sonnet 4.6 | 50 agentic req/month (Deprecated) | Required |
| RooCode | Bring your own keys | Unlimited (BYOK) | No |
| Goose | Bring your own keys | Unlimited (BYOK) | No |
| OhMyPi | Bring your own keys | Unlimited (BYOK) | No |
| Antigravity | BYOK / Local model fallback | Unlimited (BYOK / 100% Offline) | No |
What Qualifies as "Pro-Grade"?
Models achieving ≥60% on SWE-bench Verified / Pro:
| Model | SWE-bench Pro / Verified | Provider | Status |
|---|---|---|---|
| Claude Fable 5 | ~85% / S-Tier State of the Art | Anthropic | Flagship Reasoning |
| Claude Opus 4.8 | 69.2% (SWE-Bench Pro) | Anthropic | Flagship General |
| GPT-5.5 Pro | ~81% [verify] | OpenAI | Research-grade |
| GPT-5.5 Thinking | ~78.5% [verify] | OpenAI | Deep Reasoning |
| Claude Sonnet 4.6 | 79.3% | Anthropic | Premium Speed |
| Gemini 3.1 Pro | 77.4% | Premium Context | |
| Qwen3.6-Plus | 71.2% | Alibaba | Premium Open-weight |
Note:
[verify]indicates scores need verification from official sources. Always check current benchmarks before making decisions.
🏗️ Recommended Stacks
Ready-made combinations for different use cases. Copy-paste these configurations.
🟢 Fully Free Coding Stack (No Credit Card)
| Layer | Tool | Why |
|---|---|---|
| IDE | Cursor Hobby / Qoder | Limited completions + GPT-5.5-Instant chat |
| CLI | Gemini CLI (3 Flash) / Rovo Dev | 1,500 req/day Flash, 5M tokens/day Rovo |
| API | OpenRouter + Groq | 50 req/day + 14.4K req/day combo |
| Local | Ollama + Qwen3.6-Plus | Unlimited offline |
| Automation | n8n Self-hosted | Unlimited workflows |
| Vector DB | ChromaDB / LanceDB | Free local storage |
Total Cost: $0/month
⚡ Fastest Stack (Low Latency)
| Layer | Tool | Speed |
|---|---|---|
| Inference | Groq / Cerebras | 2,000 tokens/sec (Cerebras) |
| Coding | Qwen3.6-Plus via Groq | 1,000 req/day (71.2% SWE) |
| Agent | OpenCode Zen | Big Pickle (72.0%), MiniMax M2.5 (80.2%) |
| Cache | DeepSeek V4 | $0.30/$0.50 per 1M, 90% cache discount |
| Edge | Cloudflare Workers AI | Global CDN |
Best for: Real-time apps, trading bots, live coding assistants
💰 Cheapest Pro Stack (<$10/month)
| Layer | Tool | Cost |
|---|---|---|
| IDE | Trae Lite | $3/mo ($5 basic usage + bonus) |
| IDE | Trae Pro | $10/mo ($20 basic usage + bonus, SOLO mode) |
| API | OpenRouter $10 | 1K req/day + BYOK 1M/month free |
| CLI | OpenCode | Free (BYOK) or Go $10/mo |
| CLI | Xiaomi MiMo Lite | $6/mo (60M credits, ~120 tasks) |
| CLI | Gemini CLI | v0.37.1 (Gemini 3.1 Pro/Flash) |
| Local | Ollama | Free |
| Embeddings | Jina AI | Free tier |
Total Cost: ~$10/month for pro-grade everything
🔒 Local Privacy Stack (100% Offline)
| Layer | Tool | Privacy |
|---|---|---|
| Models | Ollama + Llama 3.3 / Qwen3-Coder | Runs locally |
| IDE | Continue.dev + VS Code | BYO local models |
| CLI | Aider + local Ollama | Git-integrated, offline |
| Chat UI | Open WebUI | Self-hosted ChatGPT alternative |
| Vector DB | ChromaDB / LanceDB | Local embeddings storage |
| Speech | Whisper (local) | Offline transcription |
Best for: Healthcare, legal, finance - any sensitive data
🤖 Agentic AI Stack (Autonomous Workflows)
| Component | Tool | Role |
|---|---|---|
| Orchestrator | n8n / Gumloop | Workflow automation |
| Reasoning | DeepSeek R1 / DeepSeek V4 | Complex decision making |
| Execution | Qwen3.6-Plus | Code generation |
| Memory | ChromaDB / Supabase Vector | Long-term context |
| Embeddings | Jina Embeddings v3 (1M tokens/day free) | Semantic search |
| Monitoring | LangSmith | Trace agent steps |
Best for: Autonomous research assistants, code review bots, data processing pipelines
📊 RAG Stack (Document Q&A)
| Component | Tool | Purpose |
|---|---|---|
| Framework | LlamaIndex / LangChain | RAG orchestration |
| Vector DB | ChromaDB / Weaviate / Supabase | Document storage |
| Embeddings | E5-Mistral-7B (best accuracy) | Text vectorization |
| Chunking | LlamaIndex | Smart document splitting |
| Reranking | Cohere Rerank | Improve retrieval accuracy |
| LLM | Claude Sonnet 4.6 (79.3%) / GPT-5.5 | Answer generation |
| Eval | RAGAS | Measure RAG performance |
Best for: ExamAi, legal document analysis, knowledge bases
Free LLM API Providers
Fully Free Providers
OpenRouter
Limits: 20 RPM, 29 free models (262K context max, March 2026), models share quota
- Llama 3.3 70B ✅
- NEW: Nemotron 3 Super (262K context)
- NEW: MiniMax M2.5
- NEW: Devstral 2 (Apache 2.0)
- NEW: Gemma 3n family (mobile-optimized)
- qwen/qwen3.6-plus:free ✅
- Hermes 3 Llama 3.1 405B
- Llama 3.2 3B Instruct
- Mistral Small 3.1 24B
- Full list
OfoxAI
Unified API gateway for 100+ LLMs. OpenAI and Anthropic SDK-compatible. China-friendly with Hong Kong direct access (100-300ms latency). No monthly fees, pay per token.
Limits: Not published | 1 free model
- GLM-4.7-Flash (200K context, 128K output, $0/M input, $0/M output)
Google AI Studio
Data is used for training when used outside UK/CH/EEA/EU.
Rate limits: Tier 1 (default): 250 RPD | Tier 2: Requires $250 spend + 30 days
| Model | Free Tier Limits |
|---|---|
| Gemini 3.1 Pro [verify: now paid] | 250 RPD (Tier 1) |
| Gemini 3 Flash | 1,500 RPD |
| All others | Check console |
Note: Data training outside UK/CH/EEA/EU still applies.
NVIDIA NIM
Phone number verification required. Models tend to be context window limited.
Limits: 1K credits signup, up to 5K total, 40 RPM (phone verify required)
- 46+ models including Llama 3.3 70B, Llama 4 Scout, Mistral Large, Qwen3 235B
Mistral (La Plateforme)
Free tier requires opting into data training; phone verification required
Limits (per-model): 1 req/s, 500K tokens/min, 1B tokens/month
- Open and Proprietary Mistral models (Mistral Large 3, Small 3.1, etc.)
Vercel AI Gateway
Routes to various supported providers.
Limits: $5/month
OpenCode Zen
AI gateway with curated models. Free models may use data for improvement.
- Big Pickle Stealth (S+, 72.0% SWE-bench)
- MiniMax M2.5 Free (S+, 80.2% SWE-bench)
- MiMo V2 Pro/Omni/Flash Free
- Nemotron 3 Super Free
- GPT 5 Nano
- Trinity Large Preview Free
Cerebras
| Model | Limits |
|---|---|
| GPT-OSS 120B | 30 req/min, 60K tokens/min, 900 req/hour, 1M tokens/day |
| Llama 3.1 8B | Same limits as above |
| Qwen3-235B | Available via API |
Groq
| Model | Limits |
|---|---|
| Llama 3.1 8B | 14,400 req/day, 6K tokens/min |
| Llama 3.3 70B | 1,000 req/day, 12K tokens/min |
| Llama 4 Maverick/Scout | 1,000 req/day |
| Whisper Large v3/v3 Turbo | 7,200 audio-sec/min, 2,000 req/day |
| Qwen3-32B | 1,000 req/day, 6K tokens/min |
| Kimi K2 Instruct | 1,000 req/day, 10K tokens/min |
| GPT-OSS 20B/120B | 1,000 req/day, 8K tokens/min |
| And 15+ more |
Cloudflare Workers AI
Limits: 10,000 neurons/day
- @cf/aisingapore/gemma-sea-lion-v4-27b-it
- @cf/ibm-granite/granite-4.0-h-micro
- @cf/openai/gpt-oss-120b, @cf/openai/gpt-oss-20b
- @cf/qwen/qwen3-30b-a3b-fp8
- @cf/zai-org/glm-4.7-flash
- DeepSeek R1 Distill Qwen 32B
- Deepseek Coder 6.7B Base/Instruct (AWQ)
- Deepseek Math 7B Instruct
- Gemma 2B/3 12B/7B Instruct (LoRA)
- Hermes 2 Pro Mistral 7B
- Llama 2 7B/13B Chat (FP16/INT8/AWQ/LoRA)
- Llama 3 8B Instruct, Llama 3.1 8B Instruct (AWQ/FP8)
- Llama 3.2 1B/3B/11B Vision Instruct
- Llama 3.3 70B Instruct (FP8), Llama 4 Scout Instruct
- Mistral 7B Instruct v0.1/v0.2 (AWQ/LoRA)
- Mistral Small 3.1 24B Instruct
- Qwen 1.5 0.5B/1.8B/7B/14B Chat (AWQ)
- Qwen 2.5 Coder 32B Instruct, Qwen QwQ 32B
- Phi-2, SQLCoder 7B 2
- And more...
Providers with Trial Credits
| Provider | Credits | Duration | Notes |
|---|---|---|---|
| Fireworks | $1 | Permanent | Various open models |
| Baseten | $30 | Permanent | Pay by compute time |
| Nebius | $1 | Permanent | Various open models |
| Novita | $0.50 | 1 year | Various open models |
| AI21 | $10 | 3 months | Jamba family |
| Upstage | $10 | 3 months | Solar Pro/Mini |
| NLP Cloud | $15 | Permanent | Phone verification required |
| Alibaba Cloud | 1M tokens/model | 90 days | Qwen models |
| Modal | $5-30/month | Monthly | Pay by compute time |
| Inference.net | $1 (+$25 on survey) | Permanent | Various open models |
| Hyperbolic | $1 | Permanent | DeepSeek, Llama, Qwen, GPT-OSS |
| SambaNova Cloud | $5 | 3 months | Llama, Qwen, DeepSeek |
| Scaleway | 1M tokens | Permanent | DeepSeek, Llama, Mistral, Gemma |
Additional Free API Providers
| Provider | Models | Free Tier | Environment Variable |
|---|---|---|---|
| ZAI | 7 | Free tier (generous quota) | ZAI_API_KEY |
| SiliconFlow | 6 | 1K RPM, 50K TPM | SILICONFLOW_API_KEY |
| OVHcloud AI Endpoints | 8 | 2 req/min (no key), 400 RPM with key | OVH_AI_ENDPOINTS_ACCESS_TOKEN |
| Chutes AI | 4 | Free community GPU-powered | CHUTES_API_KEY |
| DeepInfra | 4 | 200 concurrent requests | DEEPINFRA_API_KEY |
| Replicate | 2 | 6 req/min (no payment), up to 3K RPM with payment | REPLICATE_API_TOKEN |
AI-Powered IDEs
Full-featured integrated development environments with built-in AI assistance.
IDEs with Pro-Grade Models
Cursor
Model: GPT-5.5-Instant (Default Adaptive routing)
- Free tier (Hobby): Limited Agent requests + Limited Tab completions/month + 1-week Pro trial
- Free models: Cursor Small, Deepseek v3, Gemini 2.5 Flash, GPT-5.5-Instant (Limited access)
- Premium tiers required for manual model selections like GPT-5.5 Pro or Claude Fable 5
- Credit-based billing since Jun 2025: each paid plan includes a credit pool equal to its price; Tab completions unlimited, Auto mode effectively unlimited, credits only deplete when you manually pick a premium model
- AI-powered code editor with autonomous coding capabilities
- Pro ($20/mo or $16/mo annually): $20/mo credit pool + Unlimited Tab completions + Auto mode
- Pro+ ($60/mo or $48/mo annually): $60/mo credit pool + 3x Pro usage + Background Agents
- Ultra ($200/mo or $160/mo annually): $400/mo credit pool (20x Pro) + Priority access
- Teams ($40/user/mo or $32/user/mo annually): Pro-equivalent per seat + Centralized billing + Usage analytics + SAML/OIDC SSO
- Enterprise (Custom): Everything in Teams + Pooled usage + SCIM + AI code tracking API + Audit logs
- Bugbot add-on: $40/user/month (Pro/Teams) — automated PR review
Trae
Models: DeepSeek V4, GPT-5.5-Instant, Gemini 2.5 Pro (Claude models removed)
- New token-based pricing (effective Feb 24, 2026) — replaced the legacy "fast/slow request" model
- Free: Limited usage, 5,000 auto-completions/month, Standard queue
- Lite ($3/mo): $5 basic usage + bonus, Unlimited auto-completions
- Pro ($10/mo): $20 basic usage + bonus, Unlimited auto-completions, SOLO mode included, 10 concurrent cloud tasks
- Pro+ ($30/mo): $90 basic usage + bonus (4.5x Pro), 15 concurrent cloud tasks
- Ultra ($100/mo): $400 basic usage + bonus, Model early access, 20 concurrent cloud tasks
- 7-day free Pro trial (replaces the legacy $3 first-month deal)
- Annual: Pro $90/yr (
$7.5/mo), Pro+ $270/yr ($22.5/mo), Ultra $900/yr (~$75/mo) - On-Demand Usage: pay-as-you-go at API rates after basic + bonus usage is exhausted
- Migration bonus: $20 in dollar usage for current Pro users who manually switch (valid 90 days)
Windsurf
Models: OpenAI, Anthropic, Google, xAI model access
- New quota-based pricing (effective Mar 19, 2026) — replaced the legacy "prompt credits" model
- Daily + weekly usage allowance instead of monthly credit pool
- Existing paid subscribers are grandfathered at the old price but moved to the new quota system (with a free extra week to try it)
- Free ($0): Light quota + Unlimited Tab completions + 1 app deploy/day
- Pro ($20/mo): Standard quota + Full model access (Claude Fable 5, GPT-5.5 Thinking, Sonnet 4.6) + Purchase extra usage at API price
- ~7-27 messages/day on Premium Plus models (Fable 5, GPT-5.5 Thinking)
- ~8-101 messages/day on Premium models (Sonnet 4.6, Gemini Pro)
- Max ($200/mo) — NEW Mar 2026: Heavy quota (~6x Pro) + Priority support
- ~42-170 messages/day on Premium Plus models
- ~291-1,190 messages/day on Lightweight models (Haiku, Flash)
- Teams ($40/user/mo): Standard quota per seat + Centralized billing + Admin dashboard + Priority support
- Enterprise ($60+/user/mo): Custom volume + SSO + Audit logs
Pricing | Pricing Announcement (Mar 18, 2026)
Void IDE
Models: Multi-agent (frontend/backend/testing agents)
- Agent-first IDE - new 2026 category
- Multiple specialized agents coordinate across codebase
- Free preview tier with high usage limits
- VS Code-based
Best for: Full-stack development with natural language direction
Qoder
Models: Qwen3.6-Plus (71.2% SWE), Qwen-Coder-Qoder, GPT-5.5-Instant
- Free tier: Unlimited completions + limited chat/agent (basic models) + 2-week Pro trial (1,000 credits)
- Experts Mode: Multi-agent collaboration (new Mar 2026)
- Quest Mode: Fully autonomous app building
- Nextnew: Tab predictions
- Windows/macOS, VS Code-based
- 50% launch promo ended Apr 30, 2026 — now back to standard pricing
Pricing (standard, post-promo — effective Apr 30, 2026):
- Free: Basic models, limited messages
- Pro: $20/mo — 2,000 credits
- Pro+: $60/mo — 6,000 credits
- Ultra: $200/mo — 20,000 credits
- Teams: $40/seat/mo — 3,000 credits/seat
- Personal Add-on Credits: $20 for 1,000 credits
- Credits: $0.02/credit, expire 1mo
- Teams new capabilities (rollout): BYOK, Security controls over MCP/Skills, Plugin management, Knowledge Engine
Docs | Pricing | Adjustment Notice
RooCode
Models: Bring your own API keys (any provider)
- Open-source AI-powered coding assistant for VS Code
- Whole dev team of AI agents in your editor
- No subscription required - pay-as-you-go with your own keys
- Custom modes for different coding tasks
IDEs with Basic Models
Codeium
Model: Base model (Llama 3.3 70B), pro-grade models require subscription
- Individual plan: Free forever with unlimited code completions, AI chat, commands
- 70+ programming languages supported
- IDE integrations: VS Code, JetBrains, Vim/Neovim, Jupyter
- No credit card required
- Limited context awareness (expanded in paid tiers)
- Pro ($10/mo): Unlimited usage with advanced context awareness, Claude Sonnet 4.6, GPT-5.5 access
- Teams ($12/user/mo): Pro features + team management
- Enterprise (Custom): On-premise deployment, custom models
JetBrains AI Assistant
Models: Local models + cloud models with limited quota
- AI Free tier included with IDEs
- Unlimited code completion and local model support
- Limited quota for cloud-based features
- 30-day AI Pro trial included
- Offline mode with local models via Ollama/LM Studio
- AI Pro ($15/mo): Increased cloud quota + unlimited local models
- AI Ultimate ($25/mo): Maximum cloud quota + advanced features
Tabnine
Models: Claude Sonnet 4.6, GPT-5.5, Llama 3.3 70B, proprietary models
- Free tier with limited features
- Basic AI code completions and chat (limited)
- Local processing available
- Context heavily limited in free tier
- 600+ programming languages supported
- Pro ($12/mo): Enhanced AI completions and chat
- Enterprise ($39/user/mo): Multiple LLMs, private deployment, on-premises and air-gapped options
Bolt.new
Models: Unspecified models
- $1 credit/mo = ~100K tokens (reduced Mar 2026)
- Specific model not publicly specified
- Credit card required
- $20/mo: 20M tokens/month
- $200/mo: 200M tokens/month
Lovable
Models: Unspecified models
- 5 daily credits, max 30 per month (free)
- Models not publicly enumerated
- Credit card required
- Pro ($25/mo): 150 credits/month (5 daily credits)
- Teams ($30/mo): Higher limits (undisclosed)
v0.dev
Models: Proprietary models (not frontier)
- $5 in credits/month limit
- Uses proprietary models with varied routing
- Credit card required
- GPT-5.5 access requires v0 Premium subscription
Additional 2026 AI Chat Platforms
General-purpose chat interfaces with free tiers.
| Platform | Free Model | Key Capabilities | Limitations |
|---|---|---|---|
| ChatGPT | GPT-5.5 Instant | Sora 3, DALL-E 4, GPT Store | ~20 msgs/5hr |
| Gemini | Gemini 3.1 Flash | 2M Context, 20 Deep Research/mo | Research quota |
| Claude | Claude Sonnet/Haiku | Technical reasoning | ~30 msgs/5h |
| Grok | Grok 4.2 | Aurora 2 images, voice | 15 msgs/12hr |
| Mistral Le Chat | Mistral Medium 3 | Structured output | Fewer integrations |
CLI Coding Tools
Command-line tools for AI-assisted coding in your terminal.
CLI Tools with Pro-Grade Models
Gemini CLI
Models: Gemini 3.1 Flash, Gemini 2.5 Pro
- Gemini 3.1 Pro latest version (v0.37.1 April 2026 is paid-only tier fallback)
- 100 requests/day for Gemini 2.5 Pro (free tier fallback)
- 1,500 requests/day for Gemini 3 Flash
- No credit card required for free tier
- MCP server support, Google Search grounding
- Install:
npm install -g @google/gemini-cli
Rovo Dev CLI
[!IMPORTANT]
Rovo Dev CLI isn’t available during a Rovo Dev Standard trial. To use this feature, you need a paid Rovo Dev Standard subscription.
Models: Claude Sonnet 4.6, GPT-5.5 Instant
- 5M tokens/day free tier
- No credit card required during beta
- Token limits reset at midnight UTC
- Jira/Confluence integration, MCP server support
- Requires Atlassian account
- Pro ($19.99/mo): 100 tasks/day, 5x higher limits
- Ultra: 300 tasks/day, 20x higher limits, priority access to latest models (GPT-5.5 Thinking)
Warp
Models: GPT-5.5 Instant, Claude Sonnet 4.6, Gemini 2.5 Pro
- 150 AI credits/month (first 2 months), then 75 AI credits/month
- No credit card required for basic signup
- AI-powered terminal with code generation
- Build ($20/mo): 1,500 AI credits/month
- Bring Your Own API Key (BYOK) option available
OpenCode
167k+ GitHub stars • 850+ contributors • 6.5M monthly users • Apache 2.0
Models: 75+ providers via BYOK — Anthropic, OpenAI, Google, Groq, AWS Bedrock, Azure, OpenRouter, local Ollama
- MIT/Apache 2.0 licensed — fork, customize, self-host
- Five agent modes (Tab-switchable): Build (full tools), Plan (read-only), Debug, Review, Docs
- LSP-driven self-correction — auto-spawns Language Server Protocol servers and feeds compiler diagnostics back to the model
- Multi-agent support: up to 10 parallel agents per workspace
- Local inference via Ollama: $0 — no data leaves your machine
OpenCode Go (recommended for getting started): Subscription bundle of curated open-weight models
- $5 first month, then $10/mo (beta)
- Models included: GLM-5.1, Kimi K2.5, MiniMax M2.5, DeepSeek V4 Pro/Flash, Qwen3.7 Max, MiMo-V2.5-Pro
- Usage limits: $12/5h, $30/week, $60/month
- "Use balance" option falls back to your Zen credits when limits are hit
OpenCode Zen: Pay-per-request credits (PAYG from $20)
Install: curl -fsSL https://opencode.ai/install | bash • brew install opencode • npm install -g opencode-ai
GitHub Copilot
Models: GPT-5.5 Instant, Claude Sonnet 4.6, Gemini Flash, Grok Code Fast 1 (Free tier); Claude Fable 5 & GPT-5.5 Thinking available in Pro/Pro+/Max/Business/Enterprise only
- MAJOR: Usage-based billing effective Jun 1, 2026 — premium request units (PRUs) replaced by GitHub AI Credits (token-based)
- 50 agent mode or chat requests + 2,000 completions/month (Free tier)
- Agent Mode with autonomous multi-step coding
- No credit card required for Free
- Free Copilot Pro for students/educators (GitHub Student Pack)
- Code completions and Next Edit suggestions remain included on all plans and do not consume AI Credits
- Pro ($10/mo): $15 monthly AI Credits + unlimited completions + cloud agent
- Pro+ ($39/mo): $70 monthly AI Credits + 1,500 premium req equivalent + Fable 5 access
- Max ($100/mo) — NEW Jun 2026: $200 monthly AI Credits + Priority access to new models + 2.9x Pro+ usage
- Business ($19/user/mo): $19 in AI Credits (promo: $30 in Jun/Jul/Aug 2026) + unlimited completions
- Enterprise ($39/user/mo): $39 in AI Credits (promo: $60 in Jun/Jul/Aug 2026) + unlimited completions
Plans Details | Usage-Based Billing Announcement (Apr 27, 2026)
Jules
Model: Gemini 2.5 Pro
- 15 tasks/day free tier
- 3 concurrent tasks
- Rolling 24-hour window reset
- Pro ($19.99/mo): 100 tasks/day, 5x higher limits
- Ultra (via Google AI Ultra): 300 tasks/day, 20x higher limits, 60 concurrent tasks, priority access to latest models
AWS Kiro
AWS's spec-driven agentic IDE and CLI — official replacement for Amazon Q Developer (EOL Apr 30, 2027; new signups stopped May 15, 2026)
Models (all AWS Bedrock-hosted): Claude Fable 5 [verify], Claude Opus 4.8, Claude Sonnet 4.6, Claude Haiku 4.5
- 50 credits/month (Free tier)
- 14-day welcome bonus: 500 credits
- No credit card required for Free
- Pro ($20/mo): 1,000 credits
- Pro+ ($40/mo): 2,000 credits
- Power ($200/mo): 10,000 credits
- $0.04/credit overage rate
- Spec-driven development:
requirements.md→design.md→tasks.mdin.kiro/specs/ - IAM Policy Autopilot + native AWS MCP Server integration
Xiaomi MiMo Token Plan
Xiaomi's subscription plan for AI coding scenarios — bundled access to MiMo flagship models Compatible with OpenCode, OpenClaw, Claude Code, and other mainstream toolchains
Models: MiMo-V2.5-Pro, MiMo-V2.5, MiMo-V2.5-TTS, MiMo-V2-Omni
- No context-length multiplier — same rate for 10K or 500K context (big deal for agentic workflows)
- 1:2 credit ratio for Pro vs Omni models (consumed in parallel, not independently)
- Night discount: 0.8x consumption (00:00–08:00 Beijing Time)
Monthly Pricing:
| Tier | Price (USD) | Price (CNY) | Monthly Credits | ~Tasks/mo |
|---|---|---|---|---|
| Lite | $6/mo | ¥39/mo | 60M | ~120 medium-complexity |
| Standard | $16/mo | ¥99/mo | 200M | ~400 |
| Pro | $50/mo | ¥329/mo | 700M | ~1,400 |
| Max | $100/mo | ¥659/mo | 82B Credits | ~160,000+ (Upgraded 51x May 26) |
API Pricing (permanently reduced 99% on May 26, 2026):
| Model | Input (per 1M) | Output (per 1M) | Cache Hit (per 1M) |
|---|---|---|---|
| MiMo V2.5 Pro | $0.435 | $0.87 | $0.0036 |
| MiMo V2.5 Standard | $0.20 | $0.60 | $0.002 |
Claude Code
Models: Claude Sonnet 4.6, Claude Opus 4.8 [verify: paid-only], Haiku 4.5
- Free tier available with limited usage
- Pro ($20/mo): Sonnet 4.6 access with extended usage
- Max 5x ($100/mo): ~225 messages/5 hours
- Max 20x ($200/mo): ~900 messages/5 hours
- Extended thinking modes: "think" (
4K tokens), "megathink" (10K), "ultrathink" (~32K)
OpenAI Codex CLI
Model: GPT-5.5 (Custom dynamic endpoints)
- Free with ChatGPT Plus ($20/mo): 30–150 messages/5 hours
- ChatGPT Pro ($200/mo): 300–1,500 messages/5 hours with GPT-5.5 Pro
- Pay-as-you-go API: $1.25/$10 per million tokens (input/output)
- First model with session "compaction" for multi-million token deep sessions
Antigravity
Models: Bring your own API keys (OpenRouter, Groq) + local models (qwen2.5-0.5b-embedded fallback)
- Offline-first design: Runs completely local or connects to cloud endpoints (such as OpenRouter or Groq).
- AI-Powered Data Sorting: Auto-ingests messy, raw, lowercase brain dumps and sorts them into logical categories, subcategories, and structured spreadsheet columns/rows.
- Self-Correction & Lint loops: Features tight compilation/lint integration to verify code sanity and fix compiler/linter issues iteratively.
- Open-source & Free: Developed by the Advanced Agentic Coding team.
API Providers for AI Coding Tools
These services provide API access to coding-optimized models for tools like Cursor, Continue.dev, Cline, etc.
OpenRouter
- 50 requests/day free tier (1,000/day with $10+ credits)
- Qwen3-Coder-480B, Qwen3-30B-A3B, Qwen3-235B-A22B, Gemini Flash
- OpenAI-compatible API
Cerebras
- 1.5M tokens/day free tier (expanded Feb 2026)
- 30 req/min, 8,192 token context
- Models: Qwen3.6-Plus-480B, Llama 3.1 70B
- Ultra-fast: 2,400 t/s (Qwen3.6)
- OpenAI-compatible API (works with Cursor, Continue.dev, Cline, RooCode, etc.)
Paid Tiers Comparison
AI-Powered IDEs - Paid Plans
| IDE | Entry Tier | Credits/Requests | Key Features |
|---|---|---|---|
| Cursor | Pro ($20/mo) | $20/mo credit pool | Unlimited completions, Auto mode |
| Trae | Lite ($3/mo) / Pro ($10/mo) | $5 / $20 basic usage + bonus | SOLO mode, 5-tier token system |
| Windsurf | Pro ($20/mo) | Standard quota (daily/weekly) | Multi-provider, Claude Fable 5, Max $200 tier |
| Qoder | Pro ($20/mo) | 2,000 credits | Quest Mode, Experts Mode |
| Codeium | Pro ($10/mo) | Unlimited | Claude Sonnet 4.6, GPT-5.5 access |
CLI Tools - Paid Plans
| Tool | Entry Tier | Credits/Requests | Key Features |
|---|---|---|---|
| Claude Code | Pro ($20/mo) | ~225 messages/5h | Sonnet 4.6 + Opus 4.8 [verify] |
| Warp | Build ($20/mo) | 1,500 credits/month | BYOK available |
| GitHub Copilot | Pro ($10/mo) | $15 monthly AI Credits | Usage-based token billing since Jun 1, 2026 |
| OpenCode | Go ($10/mo) | $12/5h, $30/wk, $60/mo | Apache 2.0, 75+ providers, BYOK |
| AWS Kiro | Pro ($20/mo) | 1,000 credits | Spec-driven dev, replaces Q Developer |
| Xiaomi MiMo | Lite ($6/mo) | 60M credits | OpenCode/Claude Code compatible |
Local Models
Running open-weight frontier models locally provides unlimited coding assistance without API costs.
Notable Local Models (2026):
- Qwen3.6-Plus-480B (71.2% SWE, ~150GB VRAM)
- Gemma 4 [verify] (Google, Apache 2.0, fully open-source flagship)
- GLM-5.1 / GLM-5V-Turbo [verify] (Zhipu MoE-based SOTA coders)
- Devstral 2 (24B, Apache 2.0, agent-optimized)
- DeepSeek Coder V4 (lite version ~18GB)
free-coding-models CLI
Find the fastest free coding model in seconds. Ping 238 models across 25 providers in real-time.
npm install -g free-coding-models
free-coding-models
Features
- Parallel pings — all 238 models tested simultaneously
- Stability Score (0-100) — composite score from p95 latency, jitter, spike rate, uptime
- Smart ranking — top 3 highlighted 🥇🥈🥉
- Favorites — star models with
F, persisted across sessions - Tool Integration — auto-configure OpenCode, Goose, Aider, Continue, Cline, etc.
- OpenCode Zen Models — 8 exclusive free models (Big Pickle, MiniMax M2.5 Free, MiMo V2, etc.)
Quick Usage
# Most reliable model right now
free-coding-models --fiable
# Configure Goose with S-tier model
free-coding-models --goose --tier S
# NVIDIA top models only
free-coding-models --origin nvidia --tier S
# JSON output for scripting
free-coding-models --tier S --json | jq -r '.[0].modelId'
Tool Launcher Flags
| Flag | Launches |
|---|---|
--opencode |
📦 OpenCode CLI |
--openclaw |
🦞 OpenClaw |
--goose |
🪿 Goose |
--aider |
🛠 Aider |
--qwen |
🐉 Qwen Code |
--continue |
▶️ Continue CLI |
--cline |
🧠 Cline |
--gemini |
♊ Gemini CLI |
--rovo |
🦘 Rovo Dev CLI |
| And 8 more... |
Tier Scale
| Tier | SWE-bench | Best For |
|---|---|---|
| S+ | ≥75% | Claude Opus 4.6 [verify], GPT-5.4 [verify] |
| S | 65-75% | Qwen3.6-Plus (71.2%), Claude Sonnet 4.6 [verify] |
| A+/A | 40–60% | Solid alternatives |
| A-/B+ | 30–40% | Smaller tasks |
| B/C | < 30% | Code completion |
License Summary
All 238 models allow commercial use of generated output. You own what the models generate.
| License | Models | Commercial |
|---|---|---|
| Apache 2.0 | Qwen3/Qwen2.5 Coder, GPT-OSS 120B/20B, Devstral Small 2, Gemma 4, MiMo V2 Flash | ✅ Unrestricted |
| MIT | GLM 4.5/4.6/4.7/5, MiniMax M2.1, Devstral 2 | ✅ Unrestricted |
| Llama Community License | Llama 3.3 70B, Llama 4 Scout/Maverick | ✅ Attribution required. >700M MAU → separate Meta license |
| DeepSeek License | DeepSeek V3/V3.1/V3.2, R1 | ✅ Use restrictions on model (no military, no harm) — output is yours |
| NVIDIA Nemotron License | Nemotron Super/Ultra/Nano | ✅ Updated Mar 2026, now near-Apache 2.0 permissive |
| MiniMax Model License | MiniMax M2, M2.5 | ✅ Royalty-free, non-exclusive. Prohibited uses policy applies to model |
| Proprietary (API) | Claude (Rovo), Gemini (CLI), Perplexity Sonar, Mistral Large, Codestral | ✅ You own outputs per provider ToS |
| OpenCode Zen | Big Pickle, MiMo V2 Pro/Flash/Omni Free, GPT 5 Nano, MiniMax M2.5 Free, Nemotron 3 Super Free | ✅ Per OpenCode Zen ToS |
Key Points:
- Generated code is yours — no model claims ownership of your output
- Apache 2.0 / MIT models (Qwen, GLM, GPT-OSS, MiMo, Devstral Small) are the most permissive — no strings attached
- Llama requires "Built with Llama" attribution; >700M MAU needs a Meta license
- DeepSeek / MiniMax have use-restriction policies (no military use) that govern the model, not your generated code
- API-served models (Claude, Gemini, Perplexity) grant full output ownership under their terms of service
⚠️ Disclaimer: This is a summary, not legal advice. License terms can change. Always verify the current license on the model's official page before making legal decisions.
Comparison Notes
- Goal: Compare AI coding tools by their access to pro-grade models and free tier limits.
- What qualifies a model as "pro-grade"? Models must achieve ≥60% on SWE-bench Verified, demonstrating real-world software engineering capability. Current qualifying models: Claude Opus 4.5 (80.9% [verify]), GPT-5.1-Codex-Max (77.9% [verify]), Claude Sonnet 4.5 (77.2% [verify]), Gemini 3 Pro (76.2% [verify]), GPT-5 (74.9% [verify]), Claude Opus 4.1 (74.5% [verify]), Claude Sonnet 4 (72.7% [verify]), GPT-5 mini (71.0% [verify]), Qwen3-Coder-480B (69.6% [verify]), and Gemini 2.5 Pro (63.2% [verify]).
[verify]tag: Indicates information needs verification from official sources. Pricing, limits, and model availability change frequently.- Different limit types: Tools use various quota systems - requests, tokens, credits, chats - making direct comparison challenging. Check documentation for specifics.
- Real-world usage: Actual consumption varies dramatically based on coding style, task complexity, and tool implementation.
Education & Student Programs
| Program | What You Get | Requirements |
|---|---|---|
| GitHub Student Pack | Free Copilot Pro for students | Verify with .edu email |
| GitHub Copilot Free | 50 chat + 2,000 completions/month | VS Code users |
| Copilot Pro for Teachers/Maintainers | Free Copilot Pro | Open source maintainers & educators |
Additional 2026 AI Tools
Agentic Workflow Platforms
Visual orchestration tools for building autonomous AI agents without coding.
| Platform | Free Tier | Best For | Key Features |
|---|---|---|---|
| Make (Integromat) | 1,000 ops/month | Visual builders | Drag-and-drop AI Agents, 3,000+ app integrations |
| n8n | Unlimited (self-hosted) | Technical teams | Self-hosted RAG systems, private data automation |
| Gumloop | 2,000 credits/month | No-code agents | Natural-language builder, "Gummie" troubleshooting agent |
| Relay.app | Generous free plan | Beginners | Simple agentic workflows |
| Activepieces | 1,000 tasks/month | Open-source | Flat pricing, self-hostable |
| Podium | Entry-level tiers | Sales/communication | 24/7 lead response AI agents |
| QuantFlow Pilot | Free | Autonomous development | #1 Terminal Benchmark 2.0 — AI that ships your tickets |
Data Visualization & Analysis
AI-powered tools for conversational data analysis and narrative visualization.
| Tool | Function | Free Tier Detail | Key Feature |
|---|---|---|---|
| Julius | Chat-with-data | Upload spreadsheets, generate instant visualizations | |
| Anomaly AI | AI Dashboards | Generate interactive dashboards from natural language | |
| Flourish | Data Storytelling | No-code interactive maps, "scrollytelling" features | |
| Datawrapper | Publishing | Publish-ready charts in seconds, journalism-focused | |
| Looker Studio | Marketing Data | Seamless Google Analytics/Ads integration | |
| Power BI Desktop | Microsoft reports | Copilot recommendations, local report building | |
| AI for Database | Natural language DB queries | Freemium - free tier available | Connect any DB (PostgreSQL, MySQL, MongoDB) and query in plain English — no SQL needed, with self-refreshing dashboards and workflow automation |
Creative & Multimedia Tools
Professional-grade content creation with generous free tiers.
| Tool | Output | Free Tier | Key Capability |
|---|---|---|---|
| Veo | Video | Basic Free | Cinematic clips with realistic motion and sound |
| Sora 2 (via ChatGPT) | Video | Limited free tier | Deep ChatGPT integration, high-quality video |
| DALL-E 4 (via ChatGPT) | Image | Limited free tier | Latest OpenAI image model |
| Synthesia | Video Avatars | Free individual plan | "Video Agents" in 120+ languages |
| 1 More Shot | Music Videos | Free plan | Advanced lip-sync, frame-by-frame control |
| Leonardo.Ai | Images | 150 tokens/day (~70 images) | Commercial use allowed |
| Recraft AI | Vector/SVG | 30 credits/day | Infinitely scalable icons and logos |
| Ideogram | Images | 10-20 prompts/day | Perfect text rendering, "Magic Prompt" |
| Suno AI | Music | 50 credits/day (~10 tracks) | Complete songs with vocals and instruments |
| ElevenLabs | Voice | Basic Free | Realistic voice cloning |
| Canva AI | Design | Robust free tier | AI design assets, brochures, short videos |
Productivity & Research Tools
| Tool | Function | Free Tier Detail | Key Feature |
|---|---|---|---|
| Grammarly | Writing | 100 AI prompts/month | Rewrites and tone detection |
| LanguageTool | Grammar | 10,000 characters/text | 25+ languages, open-source |
| Fathom | Meetings | Forever Free | Records/transcribes Zoom/Teams, auto-sync to CRM |
| NotebookLM | Research | Free | Audio Overview podcasts, grounded in your documents |
| Humata | PDF Analysis | 60 pages/month | Clickable source citations |
| QuillBot | Rewriting | 125 words/time | Fluency & Standard modes |
| DeepL | Translation | Basic Free | Incognito sensitive mode |
| MemoryPalace | AI Memory | Free, open source | 96.6% LongMemEval — memory palace technique for AI |
Vertical AI (Specialized Domains)
Medical AI:
| Tool | Pricing | Key Value |
|---|---|---|
| iatroX | Free | Adaptive Q-Bank, NICE/BNF clinical reference |
| DxGPT | Free | Diagnostic assistant (500K+ users, 6K doctors) |
| OpenEvidence | Free (US verified) | Evidence-grounded search, ambient note generation |
Legal AI:
| Tool | Pricing | Key Value |
|---|---|---|
| DocLegal.Ai | $10/month | Clause suggestion, risk detection |
| Doculex.ai | Varies | Case-data-driven drafting from medical records |
| Spellbook | 7-day trial | In-editor contract analysis |
| Harvey AI | Enterprise | Regulatory matters, high security |
Marketing & SEO Tools
| Tool | Function |
|---|---|
| Wellows | AI Visibility Score tracking across ChatGPT, Gemini, Perplexity |
| Google SGE Labs | See how AI Overviews interpret target keywords |
| NeuronWriter | AI content scoring |
| Surfer SEO | Content optimization |
| Jasper | AI copywriting with brand voice |
| Writesonic | Scalable copywriting |
Open Source & Local Tools
| Tool | Function | Description |
|---|---|---|
| Open WebUI | Local Chat Interface | ChatGPT-like experience running entirely offline with Ollama |
| Whisper (OpenAI) | Speech-to-Text | Most accurate open-source transcription |
| Piper | Text-to-Speech | High-quality offline audio generation |
| ComfyUI | Image Generation | Node-based interface for Stable Diffusion |
| Zed | AI IDE | 50 AI prompts/month, native performance, high speed |
| Void IDE | Agent-first IDE | Multi-agent frontend/backend/testing |
| MemoryPalace | AI Memory System | 96.6% LongMemEval — memory palace technique for AI conversations |
⚡ Realtime & Streaming APIs
Low-latency APIs for voice assistants, live coding copilots, trading tools, and realtime chat.
Streaming LLM APIs
| Provider | Latency | Best For | Free Tier |
|---|---|---|---|
| Groq Streaming | ~50-150ms (0.4ms/token) | Live coding, chat | 14.4K req/day |
| OpenAI Realtime API | Low | Voice assistants, agents | No free tier (pay-per-use only, trial credits new accounts) |
| Gemini Live API | Low | Multimodal streaming | Dynamic caps (varies by prompt complexity) |
| Cerebras | 2,400 tok/sec (Qwen3.6) | Batch + streaming | 1.5M tokens/day |
| Cloudflare Workers AI | Edge | Global low-latency | 10K neurons/day |
Speech Streaming APIs
| Provider | Type | Latency | Free Tier |
|---|---|---|---|
| Deepgram | STT streaming | ~300ms | $200 credits |
| AssemblyAI Streaming | Realtime STT | ~400ms | 50 hours/month |
| Groq Whisper | STT fast | ~200ms | 2,000 req/day |
| ElevenLabs Streaming | TTS streaming | ~100ms | 10K chars/month |
| OpenAI Realtime | STT + LLM + TTS | ~200ms | Limited |
Best for:
- Trading bots: Groq streaming (fastest)
- Voice assistants: OpenAI Realtime API (end-to-end)
- Live captions: AssemblyAI or Deepgram
- Realtime chat: Gemini Live API
🎙️ Speech Models
Speech-to-text and text-to-speech models comparison.
Speech-to-Text (STT)
| Model | Provider | Accuracy | Speed | Free Tier | Best For |
|---|---|---|---|---|---|
| Whisper Large v3 | OpenAI/Groq/Local | Excellent | Fast | 2,000 req/day (Groq) | General purpose, local |
| Deepgram Nova | Deepgram | Superior | Very Fast | $200 credits | Production, enterprise |
| AssemblyAI | AssemblyAI | Excellent | Fast | 50 hours/month | Streaming, diarization |
| Whisper API | OpenAI | Excellent | Medium | Pay-per-use | Reliable, consistent |
| Google Speech | Google Cloud | Good | Fast | 60 min/month | Google ecosystem |
| Whisper (local) | OpenAI/Ollama | Excellent | GPU-dependent | Unlimited offline | Privacy, cost control |
Text-to-Speech (TTS)
| Model | Provider | Quality | Speed | Free Tier | Best For |
|---|---|---|---|---|---|
| ElevenLabs | ElevenLabs | 🏆 Best | Fast | 10K chars/month | Voice cloning, pro voice |
| OpenAI TTS | OpenAI | Excellent | Fast | Pay-per-use | Reliable, cheap |
| Piper | Local | Good | Very Fast | Unlimited offline | Privacy, self-hosted |
| Bark | Suno/Local | Good | Medium | Free (local) | Expressive, local |
| Google TTS | Google Cloud | Good | Fast | 1M chars/month | Google ecosystem |
| WhisperSpeech | Local | Good | Fast | Unlimited | Whisper-based TTS |
All-in-One Voice APIs
| API | Input | Output | Latency | Use Case |
|---|---|---|---|---|
| OpenAI Realtime | Audio | Audio | ~200ms | Voice agents |
| Deepgram Voice | Audio | Text/Audio | ~300ms | Voice bots |
| AssemblyAI LeMUR | Audio | LLM response | ~1s | Voice RAG |
🎨 Image Generation Models
Comparison of image generation models and APIs.
| Model | Provider | Quality | Speed | Free Tier | Best For |
|---|---|---|---|---|---|
| FLUX.2 | Black Forest Labs | 🏆 Excellent | Fast | Local/Replicate | High quality, open |
| DALL-E 4 | OpenAI | 🏆 Best | Medium | ChatGPT Plus | Latest OpenAI |
| Ideogram 2.0 | Ideogram | Excellent | Fast | 20 prompts/day | Text in images |
| Recraft V4 | Recraft | Excellent | Fast | 50 credits/day | Vector/SVG output |
| Stable Diffusion XL | Stability AI | Good | Fast | Local/DreamStudio | Flexibility, local |
| Midjourney v6 | Midjourney | 🏆 Excellent | Slow | None (paid only) | Artistic, Discord |
| Leonardo.ai | Leonardo | Very Good | Fast | 150 tokens/day | Commercial use, gaming |
| Adobe Firefly | Adobe | Good | Fast | 25 credits/month | Safe, commercial |
| Imagen 3 | Excellent | Medium | Vertex AI trial | Photorealistic | |
| DiffusionBee | Local | Good | Fast | Local unlimited | Easy setup, open-source |
| ComfyUI | Local | Good | Fast | Local unlimited | Advanced, node-based |
Free Image Model APIs
| Provider | Model | Free Tier | Notes |
|---|---|---|---|
| Replicate | FLUX.1-schnell | Free tier | Fast inference |
| Pollinations | Various | Unlimited | No signup |
| HuggingFace | SDXL/FLUX | $0.10 credits | Inference API |
| Leonardo | Phoenix | 150 tokens/day | Commercial OK |
🎬 Video Generation APIs
Text-to-video and image-to-video generation. Hot area in 2026.
| Model | Provider | Quality | Duration | Free Tier | Best For |
|---|---|---|---|---|---|
| Veo 3 | 🏆 Excellent | 1080p, 60s clips | Limited preview | Cinematic, realistic | |
| Sora 3 | OpenAI | 🏆 Excellent | 120s | ChatGPT Plus | High quality, physics |
| Runway Gen-3 | Runway | Excellent | 10 seconds | 3 free credits | Creative, filmmaking |
| Pika 3.0 | Pika | Very Good | 3-5 seconds | Free tier | Lip-sync improved |
| Luma Dream Machine | Luma | Very Good | 5 seconds | 30 generations/mo | Fast, realistic |
| Kling | Kuaishou | Excellent | 2-10 minutes | Limited | Long-form, Chinese |
| Hailuo AI | MiniMax | Good | 6 seconds | Free tier | Character consistency |
| Stable Video Diffusion | Stability | Good | 4 seconds | Local | Open, flexible |
Video API Pricing (approximate)
| Provider | Cost per video | Generation time |
|---|---|---|
| Runway | ~$0.20-0.50 | 1-5 min |
| Pika | ~$0.10-0.30 | 30s-2 min |
| Luma | ~$0.30-0.60 | 2-5 min |
| Kling | ~$0.05-0.20 | 1-10 min |
🌐 AI Browser Automation
Tools for AI agents to control browsers - web scraping, form filling, testing.
| Tool | Type | Pricing | Best For |
|---|---|---|---|
| Browserbase | Managed browsers | $5 free tier | Production agents |
| Steel.dev | Browser API | Free tier | AI-native browser control |
| Stagehand | AI browser framework | Open source | Next-gen Playwright |
| Playwright | Browser automation | Free | Reliable, well-documented |
This HTML preview is truncated for page performance. The canonical Markdown file contains the complete snapshot.
Why MDRSS assigned this score
- Production catalog audit 2026-08-04
- Taxonomy classified from title, annotation, source and Markdown signals
- Agent usefulness evaluated from structure, procedures, examples, evidence and retrieval value
Discussion 0
Sign in to join the discussion.