66 tools Avg score 8.5/10 56 with a free tier

AI agent tools

"Agent" is the least stable word in this index. Three unrelated products answer to it.

Data last refreshed 2026-08-28

Three things called the same thing

Sorting this list starts with admitting the label covers three different products. There are frontier chat models that can call tools, which are agents in the sense that they take actions on your behalf inside a conversation. There are coding agents that operate on a repository across many steps without a human in the loop for each one. And there is workflow automation — the connective software that fires model calls on triggers and moves results between systems.

A buyer for one of those is almost never a buyer for the others. Nothing about evaluating an automation platform prepares you to evaluate a repo agent, and the pricing shapes are unrelated: per seat, per run, per token, per resolved outcome.

This tag also overlaps development heavily by construction, because the most mature agents in the world right now write code. If you came here to shop for a coding agent specifically, the developer page is the better-filtered version of this list.

What carries this tag

  • Assistants that take multi-step actions rather than answering one prompt.
  • Coding agents that write, run and revise across a project.
  • Workflow and automation platforms that orchestrate model calls between apps.
  • Research agents that plan and execute their own search strategy.
  • A plain chatbot with no tool use does not carry this tag.

Which categories these come from

54 of 66 are free or freemium, 12 are paid only.

The top 12, ranked

Ranked on AI Score, then adjusted for how recently the tool shipped, whether it is an editor's pick, and what readers actually open — so a strong recent release can edge out a slightly higher score. Every entry shows what it is good for and what it costs you, including the parts the vendor leads away from.

1 Grok Bot logo
Grok Bot 9.6/10 Included on eligible Cursor ($20+/mo), SuperGrok ($30+/mo), and Cursor Teams ($40+/seat/mo) plans. Product page also lists a Get started for free path. ★ Pick Productivity

Always-on AI teammates with their own cloud computer. They sign into the apps you already use, finish the job, and only come back when something needs approval.

Best for Anyone who wants an always-on teammate that signs into real apps, finishes work end to end, and only pings you for approval.

  • Each bot has its own cloud computer and keeps working after you close the laptop
  • Now included on regular Cursor and SuperGrok plans, not just the old $200–$300 tiers
  • Multi-bot group chat and learn-by-demonstration routines are first-class
  • Not Grok chat and not the Grok Build CLI — a separate teammate product
Free tier

Product page lists a Get started for free path. Cursor plans from $20/mo include weekly Grok Bot usage. Eligible Cursor, SuperGrok, and Teams plans include access.

Trade-offs
  • Still early beta — no published SLA or independent evals
  • Cheaper Cursor plans list weekly usage, so heavy days can hit a cap
2 Cursor logo
Cursor 9.5/10 Free Hobby + Individual $20/mo + Teams $40/user/mo + Enterprise Custom ★ Pick Coding

AI-first code editor built on VS Code. Multi-file editing, intelligent refactoring, and a built-in chat that understands your entire codebase. A favorite among developers for complex projects.

Best for Professional developers handling complex, multi-file refactors who want AI built into a familiar VS Code-based editor.

  • Composer edits multiple files at once, handling imports and cross-file references automatically
  • Indexes the entire codebase for context-aware suggestions
  • Cmd+K inline editing and Tab completion
  • Agent mode for autonomous multi-step tasks
Free tier

Limited Agent requests

Trade-offs
  • Can be resource-heavy on older machines
  • Free tier has limited premium model access
3
Claude Code 9.3/10 Included with Claude Pro $17/mo, Max 5x $100/mo, Max 20x $200/mo, Team and Enterprise plans ★ Pick Coding

Anthropic's coding agent, running in the terminal, IDEs, Slack, the web and mobile, with parallel subagents and scheduled routines.

Best for Developers who want an agent that works inside an existing repo and toolchain rather than in a hosted editor, and who already pay for a Claude plan.

  • Runs locally and talks directly to model APIs, with no backend server and no remote code index
  • Asks permission before editing files or running commands, so the blast radius stays under the developer's control
  • Fans work out across what Anthropic describes as 10s to 100s of parallel subagents
  • Routines run the same configured job on a schedule, from an API call, or in response to an event
Free tier

Not established for Claude Code. The $0 Free plan exists ("Free for everyone"), but its listed features do not include Claude Code; Pro is listed as "Everything in Free, plus: ... Includes Claude Code", and the Claude Code page's individual options start at Pro.

Trade-offs
  • Anthropic publishes no Claude Code specific usage limit, only the line "Usage limits apply", so the practical ceiling is unknown until you hit it
  • The pricing page shows "From $100 per month" for both Max 5x and Max 20x, which makes the step between them impossible to read from the page
4 Claude logo
Claude 9.5/10 Free tier + Pro $20/mo + Team $30/mo/user ★ Pick Chatbots

Anthropic's AI assistant with Opus 4.6, 1M-token context, computer use, agentic coding via Claude Code, and MCP integrations.

Best for Developers and professionals who need agentic coding, computer control, and large-document or codebase analysis in one assistant.

  • 1M-token context window for entire codebases or book-length documents in one session
  • Computer use lets Claude click, type, and navigate apps directly
  • Claude Code CLI writes, tests, and iterates on code autonomously in a repo
  • MCP ecosystem connects Claude to third-party databases, APIs, and creative tools
Free tier

Free access to Claude Sonnet 4.6 with daily usage limits

Trade-offs
  • No image or video generation — strictly text and code output
  • Free tier usage limits are restrictive, especially during peak hours
5 ChatGPT logo
ChatGPT 9.5/10 Free tier + Plus $20/mo + Pro $200/mo ★ Pick Chatbots

OpenAI's flagship AI assistant with o3/o4-mini reasoning, GPT-4o, Advanced Voice, Sora video gen, Operator agent, and Deep Research — the most feature-packed chatbot available.

Best for Users who want one subscription covering reasoning, voice, vision, image and video generation, and agentic browsing in a single app.

  • Broadest feature set: voice, vision, image gen, Sora video, browsing, and agents together
  • o3 and o4-mini reasoning models for complex math, science, and coding
  • Operator agent and Deep Research add autonomous multi-step task completion
  • Free tier now includes limited GPT-4o access, not just the mini model
Free tier

Free access to GPT-4o mini and limited GPT-4o with basic features

Trade-offs
  • Pro plan at $200/mo is hard to justify unless you need heavy o3 or Operator usage
  • Writing quality has fallen behind Claude for nuanced, long-form content
6 GPT-5.5 logo
GPT-5.5 9.4/10 API: $5/$30 per 1M tokens (in/out). ChatGPT Plus $20/mo, Pro $200/mo ★ Pick Coding

OpenAI's most capable frontier model, built for complex multi-step reasoning, agentic tool use, and deep coding tasks. Powers ChatGPT and Codex with up to 272K context.

Best for Developers and teams needing a frontier reasoning model for agentic coding workflows and large-codebase context handling.

  • 272K context window ingests entire codebases or long documents without chunking workarounds
  • Agentic tool use with self-correction across multi-step workflows, not just single-turn answers
  • Powers the Codex agent for autonomous pull-request writing, testing, and iteration inside ChatGPT
  • Specialized GPT-5.5-Cyber variant tuned for cybersecurity analysis
Free tier

No free API tier. Free ChatGPT users get GPT-4o, not GPT-5.5.

Trade-offs
  • API pricing is premium — $30/M output tokens adds up fast for heavy usage
  • No free tier for the API; ChatGPT Free users are stuck on GPT-4o
7 Gemini logo
Gemini 9.2/10 Free tier + Advanced $19.99/mo ★ Pick Chatbots

Google's multimodal AI assistant with Gemini 2.5 Pro/Flash, 2M-token context, Workspace integration, Gems agents, and native image generation.

Best for Google Workspace users who want an assistant that can read their Gmail, Drive, and Calendar while reasoning across huge documents or videos in one pass.

  • 2M-token context window, largest among major chatbots
  • Deep Google Workspace integration reads actual Gmail, Drive, and Calendar data
  • Native multimodal understanding across text, images, audio, and video
  • $19.99/mo tier bundles full 2.5 Pro access with 2TB of storage
Free tier

Gemini 2.5 Flash with image generation and basic multimodal features

Trade-offs
  • Creative writing and nuanced tone trail Claude and ChatGPT
  • Best features require Google ecosystem buy-in — less useful if you're not on Workspace
8 Grok logo
Grok 8.7/10 Free + SuperGrok $30/mo + SuperGrok Plus $100/mo + Heavy $300/mo ★ Pick Chatbots

xAI's flagship chatbot on Grok 4.6. Consumer plans now run Free, SuperGrok $30, SuperGrok Plus $100, and Heavy $300, plus a separate API.

Best for People who want Grok 4.6 in chat and need a clear map of Free, SuperGrok, SuperGrok Plus, Heavy, and the API.

  • Current flagship is Grok 4.6, not a later unreleased number
  • SuperGrok Plus at $100/mo sits between SuperGrok and Heavy and is the first listed consumer plan for 1080p video
  • Native X search on the consumer app and, with tools enabled, on the API
  • Grok Build and Grok Bot are separate products, not chat toggles
Free tier

Yes — a limited grok.com tier with web/X search, voice, and connectors. Caps change; check x.ai/pricing.

Trade-offs
  • Grok Bot and in-chat Build Mode are still Heavy (or top Cursor for Bot), not SuperGrok or SuperGrok Plus
  • API realtime answers require search tools; the raw model cutoff is February 1, 2026
9 GPT-5.6 Sol logo
GPT-5.6 Sol 9.1/10 API usage-based: Sol $5/$30, Terra $2.50/$15, Luna $1/$6 per 1M tokens ★ Pick Chatbots

OpenAI's flagship GPT-5.6 family (Sol, Terra, Luna) targeting frontier coding, agentic tasks, and cybersecurity — currently in limited preview.

Best for Development and agentic-workflow teams evaluating a frontier model family they can route by task difficulty and token cost.

  • Three-tier family (Sol, Terra, Luna) trades capability against cost for the same generation
  • Claimed state-of-the-art results on Terminal-Bench, a real-world coding and agentic benchmark
  • Major reported cybersecurity capability gains, released via a gov-coordinated limited preview
  • Luna tier priced at $1/$6 per 1M tokens, cheap for a current-generation model
Pricing

API usage-based: Sol $5/$30, Terra $2.50/$15, Luna $1/$6 per 1M tokens

Trade-offs
  • Launched in limited preview — not broadly available, so you can't reliably build production on it yet
  • Sol's $5/$30 pricing is premium; costs add up fast on token-heavy agentic loops
10 CodeRabbit logo
CodeRabbit 9/10 Free tier + Pro $24/user/mo, Pro Plus $48/user/mo (annual) ★ Pick Coding

AI code reviewer that posts inline, context-aware feedback on every pull request across GitHub, GitLab, Azure DevOps, and Bitbucket.

Best for Engineering teams that want automatic, context-aware review posted on every pull request across their git provider.

  • Builds full-repository context so feedback reflects how a change interacts with the rest of the codebase
  • Chains 40+ open-source linters and SAST tools into the same review pass
  • Works across GitHub, GitLab, Azure DevOps, and Bitbucket, plus IDE and CLI reviews
  • Interactive review chat lets you question, request fixes, or dismiss suggestions in the PR thread
Free tier

Permanent free tier with PR summaries and IDE/CLI reviews, plus a 14-day Pro Plus trial that needs no card.

Trade-offs
  • Per-PR-author billing at $24–$48/user/mo adds up fast for larger engineering teams
  • AI review comments can still be noisy or surface false positives that reviewers must triage
11
Ollama 8.8/10 Free, with Pro at $20/mo and Team at $25/seat/mo for cloud usage ★ Pick Coding

Runs open-weight models on your own machine with one command, and optionally spills over to Ollama's cloud when the model is too big for local hardware.

Best for Anyone who wants open-weight models running on their own hardware for privacy, offline work or zero marginal cost, with a cloud fallback for models that will not fit.

  • Local-first by default, so prompts and files stay on the machine unless a cloud model is chosen explicitly
  • Positions itself as the runtime other agents plug into, naming Claude Code, OpenClaw and Codex on its own front page
  • Cloud tier is an extension of the same tool rather than a separate product, with servers in the US, Europe and Singapore
  • States that prompt and response data is never logged or trained on
Free tier

Yes, and it is the main event. Local inference is free forever on your own hardware. The paid tiers exist to buy cloud capacity, and their allowances are published only as multipliers of an unstated base.

Trade-offs
  • Cloud allowances are published only as multipliers ("50x more than Free", "5x more than Pro") against a base quantity Ollama never states
  • The $100 Max tier is currently listed as paused for new sign-ups, so the top individual plan may not be available
12 Zapier logo
Zapier 8.9/10 Free tier + Professional from $19.99/mo + Team from $69/mo + Enterprise contact Productivity

Connect 7,000+ apps with AI-powered automation. Zapier's AI features help build workflows from natural language, auto-generate Zaps, and add intelligent decision-making to your automations.

Best for Teams that need to connect many SaaS apps and build automations from a plain-English description instead of manual setup.

  • Largest app ecosystem in the category at 7,000+ integrations
  • AI Builder generates multi-step Zaps from a natural-language description
  • AI also auto-suggests optimizations to existing workflows
  • Pricing is task-based, which can scale unpredictably with volume
Free tier

Free forever with 100 tasks per month

Trade-offs
  • Gets expensive with high task volumes
  • Free tier is very limited

The other 54 in the index

Same tag, lower down the ranking. Scores, pricing and full write-ups behind each name.

Retell AI Chatbots · Free $10 credit, then ~$0.07/min usage-based 8.9/10 Gemini 3.7 Flash Chatbots · Intro $0.75 input / $3.75 output per 1M tokens through Dec 31, 2026; then $1.50 / $7.50. Half the original 3.6 Flash token price during the intro window. 8.9/10 Make Productivity · Free tier with 1,000 credits/mo; paid from $9/mo for 10,000 credits 8.7/10 Zed Coding · Free (open source) — BYOK for AI features 8.8/10 Replit Agent Coding · Free Starter tier + Core $25/mo + Pro $100/mo + Enterprise custom 8.8/10 HeyGen Video · Free tier + Creator $29/mo 8.8/10 GitHub Copilot Coding · Free tier + Pro $10/mo 8.8/10 Fathom Productivity · Freemium — Free core; Team $19/user/mo 8.8/10 Vapi Chatbots · Free tier, then ~$0.05/min + provider costs 8.8/10 Windsurf Coding · Freemium 9.1/10 GLM-5.2 Coding · Free 20M tokens on signup + open weights; paid API & coding plans (check site) 8.7/10 Writer Writing · From $18/user/mo, Enterprise custom 8.7/10 Fireflies.ai Productivity · Freemium — Pro $10/user/mo, Business $19/user/mo 8.7/10 Bland AI Chatbots · From $0.09/min or plans from ~$499/mo 8.6/10 OpenAI Codex Coding · Included with ChatGPT Free, Go, Plus, Pro, Business, Edu and Enterprise plans 8.8/10 Bonsai 27B Research · Free — open-source weights under Apache 2.0 8.5/10 WorkBeaver Productivity · Free (25 tasks one-time) + Priority $23.95/mo + Founder $108/mo 8.5/10 Glean Productivity · Custom enterprise pricing 8.8/10 Otter.ai Productivity · Freemium — Free (300 min/mo), Pro $8.33/user/mo, Business $20/user/mo 8.5/10 Kimi K3 Chatbots · Free tier + paid from $19/mo; API $3/$15 per M tokens 8.5/10 Sierra Chatbots · Enterprise only — outcome-based, contact sales 8.7/10 n8n Productivity · Freemium 8.7/10 Grok Build Coding · Product page says you can try it free; SuperGrok and X Premium+ still listed as paid paths 8.2/10 DeepSeek Harness Coding · Free (MIT). Inference is whatever model you plug in — DeepSeek V4-Flash and V4-Pro are the vendor pairing. 8.4/10 LM Studio Coding · Free desktop app; cloud inference is pay as you go per token 8.5/10 OpenClaw Productivity · Free (Open Source) 8.6/10 Laguna XS.2 Coding · Free (Apache 2.0) 8.5/10 Command A+ Coding · Free API tier + open weights (Apache 2.0) 8.5/10 MiniMax M3 Coding · Freemium — Plus $20/mo, API $0.30/M input / $1.20/M output 8.5/10 Comet Productivity · Free to download; Perplexity Max at $200/mo unlocks the highest limits 8.4/10 Devin Coding · Teams ~$500/user/mo, Enterprise custom 8.5/10 Hume AI Chatbots · Free credits + usage-based from $0.07/min 8.4/10 xAI Voice Agent Builder Chatbots · Usage-based from $0.05/min 8.2/10 Qwen App Chatbots · Free; Alibaba publishes no paid consumer tier 8.2/10 Siri AI Chatbots · Free with iOS 27 / macOS 27 8.3/10 Fugu-Cyber Research · Token plan only: $6–$12/M input, $36–$54/M output (application required, no free tier) 8.2/10 Copy.ai Writing · Freemium 8.3/10 Goose Coding · Free (open-source, Apache 2.0) 8.2/10 Fundraisly Productivity · Custom plans (contact for pricing) 8.2/10 Muse Spark 1.1 Coding · Free in Meta AI app; API in public preview (pricing not yet disclosed) 8.2/10 OpenScience Research · Free — open-source (self-hosted; you pay your own model API costs) 8.2/10 ZONOS2 Music · Free (open-source) + paid cloud tiers 8.2/10 North Mini Code Coding · Free (open-source, Apache 2.0) 8.2/10 Parallel Search Turbo Research · API usage-based, Turbo from $1 per 1,000 requests 8.2/10 Co-Scientist Research · Free (experimental access via registration) 8.2/10 Cofounder 2 Productivity · Free (open beta — pricing TBA) 8/10 You.com Research · Free tier + YouPro $20/mo 8/10 Jobbie Productivity · Free to start (no credit card) + paid plans for higher volume 7.8/10 T3MP3ST Coding · Free, open source (AGPL-3.0) — you pay for the underlying AI agent's API usage 7.8/10 SellerClaw Productivity · Freemium (public beta) 7.8/10 OpenYabby Productivity · Free — open-source, self-hosted (bring your own model API keys) 7.7/10 Osaurus Coding · Free (open-source) 7.5/10 Microsoft Scout Productivity · Private preview (Frontier program + GitHub Copilot license required) 7.5/10 Overtone Chatbots · Freemium — public tiers not yet announced 7.3/10

Open all 66 in the filterable directory →

Where to start without paying

56 of the 66 tools here publish a free tier. These are the strongest of them, with what the free tier actually gets you.

Grok Bot Free tier

Product page lists a Get started for free path. Cursor plans from $20/mo include weekly Grok Bot usage. Eligible Cursor, SuperGrok, and Teams plans include access.

Cursor Free tier

Limited Agent requests

Claude Code Free tier

Not established for Claude Code. The $0 Free plan exists ("Free for everyone"), but its listed features do not include Claude Code; Pro is listed as "Everything in Free, plus: ... Includes Claude Code", and the Claude Code page's individual options start at Pro.

Claude Free tier

Free access to Claude Sonnet 4.6 with daily usage limits

ChatGPT Free tier

Free access to GPT-4o mini and limited GPT-4o with basic features

GPT-5.5 Free tier

No free API tier. Free ChatGPT users get GPT-4o, not GPT-5.5.

See all 56 free and freemium options →

The questions autonomy raises

What can it actually change

Read-only agents are a different risk class to agents with write access to a repo, an inbox, a CRM or a payment system. Establish the blast radius before the capability list.

Where credentials live

Agents need keys to be useful. Whether those keys sit in a vendor database, in your own infrastructure, or scoped per action is a security decision, not a feature preference.

Is there a stop and an undo

Long chains fail in the middle. Look for step-level visibility, an interrupt, and a way to roll back partial work. Plenty of tools have none of the three.

Cost per run at real volume

Multi-step agents multiply token spend by the number of steps and the number of retries. A cheap-looking per-call price becomes a different number once a run is forty calls.

Where autonomy actually breaks

  • Reliability compounds in the wrong direction. A step that succeeds 95% of the time succeeds about 60% of the time across ten steps, and most real workflows are longer than ten steps.
  • Demos are single-happy-path. The interesting question is what happens on the branch the vendor did not record, and the answer is usually that the agent confidently does the wrong thing.
  • Silent partial completion again — the most expensive failure mode in the category, because it looks exactly like success until somebody checks.
Comparison explorer Put the top three side by side Opens with Grok Bot, Cursor, Claude Code already loaded. Swap any of them out and read pricing plans, features, pros and cons in one table.

Related reading

Browse another job

Weekly issue

The 5 AI tools that mattered this week.

One email, Fridays. No spam, unsubscribe anytime.