The datasheet
Guides
Practical, sourced guides to choosing, building and running AI agents: what works, what it costs, and when to trust it.
Explore this section
On this page
How to Audit a Coding-Agent Benchmark Claim
Check the task set, harness, scaffold, model and attempt budget behind a coding-agent score before deciding what to evaluate on your repository.
Read the guideScope · 7 topics
- Reconstructing a vendor claim and marking missing methodology
- Distinguishing benchmark, variant, harness, scaffold and model
- Recording model versions, evaluation dates and execution budgets
- Separating attempts per task, turns, batch selection and aggregation
- Exposing a hypothetical full-set versus filtered-subset comparison
- Assessing repository relevance without forecasting productivity
- Documentation limitations and monthly canonical maintenance
Agent Autonomy: Approvals, Sandboxes and Stop Controls
Read autonomy claims critically: separate approval gates, allowlists, execution boundaries, output review and stopping controls for the exact product.
Read the guideScope · 9 topics
- Autonomy in deciding versus autonomy in acting
- Human approval, classifier review and reusable permissions
- Filesystem, network and credential boundaries
- Input, output, tool and logging-only guardrails
- Product-surface scope and undocumented defaults
- Output destinations and review gates
- Mid-flight stopping and execution limits
- Hypothetical proposal-only and isolated pull-request agents
- Five-part buyer control record
Audit an MCP Server Before You Install It
Map declared tools, credential permissions, host approvals and untrusted inputs, then record an approve, restrict or refuse decision.
Read the guideScope · 9 topics
- Pre-install tool inventory and parameter review
- Declared behavior versus credential entitlement
- Local runtime and remote discovery reach
- Host tool filtering and confirmation policy
- Untrusted-content influence on tool arguments
- Hypothetical ticket-server blast-radius audit
- Confused-deputy and token-passthrough review questions
- Approve, restrict or refuse decision record
- Weekly canonical-guide maintenance and evidence limitations
Does Claude Code Support AGENTS.md?
No. Claude Code reads CLAUDE.md natively. Here is the official symlink and import workaround, and which coding agents in our index already read AGENTS.md on their own.
Read the guideScope · 5 topics
- Claude Code
- AGENTS.md
- CLAUDE.md
- Coding agents
- Compatibility
Framework, Platform, or Ready-Made AI Agent?
The decision that comes before any tool comparison: what you are actually buying at three levels of the stack, and the tradeoff at each one.
Read the guideScope · 4 topics
- Frameworks
- Agent platforms
- Decision guide
- Build vs buy
n8n vs Lindy vs Dify: 16 AI Agent Platforms Compared
All 16 agent platforms in our index, split on what you are really choosing: open-source developer tool, no-code builder, managed enterprise, or personal assistant.
Read the guideScope · 18 topics
- Agent platforms
- Decision guide
- n8n
- Lindy
- Relevance AI
- Dify
- Botpress
- Sierra
- Salesforce Agentforce
- Moveworks
- Glean
- Botsify
- Vecbase
- OpenClaw
- MindStudio
- Sim
- Vellum
- ElevenLabs
How to Choose an AI Coding Agent: 25 Tools, 4 Jobs
Coding agents is our biggest category at 25 tools, and they split into four genuinely different jobs, so “which is best” is the wrong question.
Read the guideScope · 10 topics
- Coding agents
- Decision guide
- Cursor
- Claude Code
- Devin
- OpenCode
- CodeRabbit
- Zed
- Goose
- Kilo Code
AI Support Agents: Decagon vs Sierra vs Intercom Fin
Every customer-support agent we index, compared: six genuinely different ways to buy one, and the right pick is whichever you actually want.
Read the guideScope · 10 topics
- Support agents
- Decision guide
- Decagon
- Sierra
- Intercom Fin
- Zendesk AI Agents
- Crescendo
- Cresta
- Ada
- Botsify
How to Choose an AI Research Agent: 11 Compared
All 11 research agents in our index look like rivals, but each is built for a different job: web answers, literature, legal, raw web data, or your own data.
Read the guideScope · 13 topics
- Research agents
- Decision guide
- Perplexity
- Manus
- Genspark
- GPT Researcher
- Webhound
- Elicit
- Undermind
- GC AI
- Microsoft Fabric Data Agent
- Julius AI
- TinyFish
11x vs Artisan vs Lead Scorer: 8 AI Sales Agents Compared
All 8 sales-and-marketing agents in our index: eight tools solving eight different jobs at eight different budgets, not eight competing chatbots.
Read the guideScope · 10 topics
- Sales agents
- Decision guide
- 11x
- Artisan
- Qualified
- Rox
- Unify
- Warmly
- Amplemarket
- Lead Scorer
AI Voice Agents: Vapi vs Retell AI vs ElevenLabs
Every voice agent we index, compared by the layer of the stack you own: raw model, developer platform, no-code builder, or managed incumbent.
Read the guideScope · 13 topics
- Voice agents
- Decision guide
- OpenAI Realtime API
- Vapi
- Retell AI
- Bland AI
- ElevenLabs Conversational AI
- Synthflow AI
- LiveKit Agents
- PolyAI
- Deepgram Voice Agent API
- Cartesia
- VoiceAgent
LangChain vs CrewAI vs 12 More Agent Frameworks
All 14 agent frameworks in our index side by side: what each is actually for, live GitHub stars, and how to pick one for what you’re building.
Read the guideScope · 14 topics
- Frameworks
- Decision guide
- LangChain
- LangGraph
- CrewAI
- LlamaIndex
- AutoGen
- Microsoft Agent Framework
- Mastra
- Letta
- Google ADK
- Claude Agent SDK
- Agno
- AgentsKit
AI Agent Pricing in 2026: Every Tool in Our Index
The pricing model, free tier and API access for every AI agent we index. Freemium leads, but usage-based billing is nowhere near an edge case.
Read the guideScope · 3 topics
- Pricing
- Market data
- Directory census
New guides appear here when they are published. Follow the RSS feed.
Advertise here
Reach buyers mid-decision. Reach builders choosing their next agent. Promote your brand with a display placement or bring your listing into focus with Featured.
Explore owner options →Advertise on this page →