AI Developer Tools
Practical tools for working with AI APIs and models every day. Compare pricing and specs across providers, count tokens before you hit send, test prompt variations, validate function schemas, and build function-calling contracts — all without leaving your browser. Everything runs client-side, so your data stays private.
72 tools
Agent Skill Validator
Validate skill definitions across OpenClaw, Claude, Codex, and MCP with portability scoring and exact fixes
SkillSpec Converter
Convert one canonical skill definition into OpenClaw SKILL.md, Claude blocks, Codex scaffolds, and MCP manifest snippets
Skill Regression Suite Builder
Build deterministic regression suites for skill updates with risk-weighted pass-rate gates and CI-ready case definitions
Skill Scope Collision Detector
Detect cross-scope skill version and enabled-state collisions across global, user, project, and local configuration layers
Skill Payload Budget Optimizer
Plan which agent skills to keep, compress or defer when their combined token payload exceeds the context window minus reserved tokens.
Tool Approval Matrix Compiler
Compile cross-platform allow, ask, and deny decisions for tool capabilities across Codex, Claude, and managed MCP policies
Skill Release Canary Planner
Generate staged canary rollout plans for skill updates with deterministic stop conditions and rollback checklists
Trace Failure Classifier
Classify failed trace events into root-cause buckets and output deterministic remediation guidance for agent incident triage
LLM Crawl Policy Validator
Validate robots.txt and llms.txt files, detect conflicts, simulate AI bot access, and export corrected policies
MCP Governance Composer
Compose managed MCP governance packs with allow/deny lists, approval boundaries, and operator rollout checklists
MCP Tool Search Budget Simulator
Simulate context-window usage for full MCP tool injection versus search-first retrieval strategies
Claude Settings Scope Diff
Diff managed, user, project, and local settings scopes and compute the effective merged Claude configuration
Claude Hook Policy Simulator
Simulate Claude hook decisions for pre/post tool events and validate policy rule coverage before rollout
OpenClaw Skill Trust Scanner
Scan SKILL.md instructions for destructive command patterns, missing safety boundaries, and trust posture
Agent Tool Blast Radius Mapper
Map tool capability blast radius, score operational risk, and produce least-privilege policy buckets
AI Model Comparison
Compare AI models side by side: pricing, context windows, max output, and release dates
MCP Server Directory
Browse and search Model Context Protocol (MCP) servers by category, with GitHub stars, transport type and copyable install commands.
AI Agent Framework Comparison
Compare AI agent frameworks such as LangChain, CrewAI, AutoGen and Mastra by language, GitHub stars, multi-agent, tool, RAG and MCP support.
AI Pricing Calculator
Calculate daily, monthly and yearly LLM API costs per model from your request volume and average input and output tokens.
LLM Token Counter
APICount tokens for GPT, Claude, Gemini, Grok, DeepSeek, Mistral, Qwen and Muse models and estimate API cost. Also available as a REST API.
CLAUDE.md / Rules File Generator
Generate CLAUDE.md, .cursorrules, copilot-instructions.md, .windsurfrules and .clinerules files from one form or a stack template.
AI Model Picker Quiz
Answer 7 questions and get personalized AI model recommendations across GPT, Claude, Gemini, Grok, DeepSeek, and open-weight models
System Prompt Library
Browse and copy curated system prompts for coding, writing, analysis, business, education and creative tasks, searchable by keyword.
MCP Server Config Generator
Generate MCP server configurations for Claude Desktop, Cursor, and Windsurf with visual editor and presets
JSON Schema Generator
APIGenerate Draft-07 or 2020-12 JSON Schemas from sample JSON for function calling, structured outputs and validation. Also available as a REST API.
Prompt Template Builder
Build AI prompt templates with variables, live preview, and export to JSON/YAML
AI Cost Estimator
Estimate monthly or annual LLM API costs per model from workload presets, request volume, token counts and prompt-cache hit rate.
System Prompt Editor
Write and analyze AI system prompts with live token counting, variable detection, and XML highlighting
LLM Output Diff Tool
Compare two to four LLM outputs side by side, diff any pair by line or by word, and compare response length in characters and words.
AI Context Window Visualizer
Visualize how your AI model's context window is allocated across system prompt, tools, conversation, and RAG
AI Prompt Tester & Comparator
Compare two to four prompt variants side by side with word-level diffs, token estimates and the {{variables}} each one uses.
Prompt Token Budget Planner
Split a model's context window across system prompt, tools, memory, retrieved documents and history, with a warning when allocation passes 80%.
AI Tool Schema Builder
Build AI tool definitions in a visual editor with typed, enum and nested parameters, and export them as OpenAI, Anthropic, MCP or JSON Schema.
LLM Structured Output Validator
APIValidate LLM JSON output against a JSON Schema, with OpenAI, Anthropic and MCP presets, per-field fixes and sample output. Also has a REST API.
Markdown Memory File Builder
Build markdown memory files for AI agents from guided forms: SOUL.md, USER.md, AGENTS.md, daily logs, decision records, project status and lessons.
Embedding Similarity Calculator
Compare two embedding vectors by cosine similarity, dot product, Euclidean and Manhattan distance, with dimension hints for common embedding models.
RAG Chunk Size Calculator
Recommend a RAG chunk size, overlap and chunking strategy from document type, query type and embedding model, with a context-window token budget.
Agent Trace Viewer
Inspect AI agent traces from LangChain runs, OpenAI Assistants run steps or a JSON step array, with LLM calls, tool calls, errors, timings and tokens.
Prompt Version Diff
Diff two prompt versions line by line, with markers for template variables, XML tags and instruction keywords, plus added variables and token delta.
AI Guardrail Rule Tester
Write keyword and regex guardrail rules that block, flag or redact text, test them against sample inputs and outputs, and export them as JSON.
Function Call Flow Simulator
Script a multi-step tool-calling conversation with mock tool results and errors, then export it as an OpenAI or Anthropic request body.
AI API Key Tester
Identify an AI API key's provider from its prefix (Anthropic, OpenAI, OpenRouter, Groq, Google AI, xAI and more) and check its format locally.
AI Token + Pricing Calculator
Paste text to estimate its tokens and compare input and output cost across GPT, Claude, Gemini and other catalog models, with CSV export.
WebMCP Playground
Validate WebMCP tool definitions or a webmcp.json manifest against lint rules, test them with a form built from the schema and preview discovery.
MCP Server Starter Generator
Generate a runnable MCP server project in TypeScript, Python or Go with tools, resources, prompts, stdio or HTTP transport and optional auth.
Cursor Rules Generator
Generate .cursorrules, .windsurfrules, .clinerules, CLAUDE.md or Copilot instructions from your stack, conventions, testing approach and rules.
AI Response Comparator
Compare up to four LLM responses side by side, with line or word diffs and an analysis of length, shared sentences and content unique to each.
System Prompt Analyzer
Estimate a system prompt's tokens for Claude, GPT, Llama and Gemini, its share of a 4K to 200K context window, and its template variables.
MCP Tool Tester
Validate a WebMCP tool definition's name, description and inputSchema, generate a test form from the schema, and mock a call locally.
Web-to-Markdown Converter
Convert pasted HTML or a fetched URL to Markdown for LLM input, dropping scripts, navigation and footers, and estimate the token savings.
AI Code Smell Detector
Scan code for AI-generated anti-patterns: hallucinated imports, over-abstraction, verbose error handling, redundancy, security smells and AI tells.
LLM Workflow Cost Calculator
Price a multi-step LLM pipeline (embed, retrieve, generate, validate) per run, day and month, with cached-input rates and a provider comparison.
Codebase Context Packer
Pack code files into an LLM prompt under a token budget with per-file truncation (full, signatures, first N lines) and XML, Markdown or plain output.
AI Prompt Injection Tester
Check a system prompt's defenses against known injection patterns: role hijacking, instruction override, delimiter abuse, encoding tricks, jailbreaks.
AI API Error Decoder
Decode an error response from the OpenAI, Anthropic, Gemini, Mistral or Cohere API into its cause, a Python and Node.js fix and a retry strategy.
AI Agent Cost Simulator
Simulate multi-agent LLM costs with per-turn context growth and tool-call overhead, compare against a single agent and project monthly spend.
AI Rules Linter
Lint CLAUDE.md, .cursorrules, and copilot-instructions files for redundancy, conflicting instructions, missing sections, and token efficiency.
Git Diff Token Counter
Paste a git diff and see token counts per file, cost across AI models for code review, and chunking suggestions when diffs exceed context limits.
LLM Latency Estimator
Estimate time to first token, generation time and total latency for current LLMs from your token counts, with a UX pattern for each wait.
Prompt A/B Test Designer
Plan a prompt A/B test: sample size per variant from a two-proportion z-test with Bonferroni correction, cost, duration and a Markdown plan.
MCP Permission Auditor
Audit an MCP server config for risky capabilities, dangerous combinations such as file read plus network, risk scores and least-privilege fixes.
AI Doc Readability Scorer
Score Markdown docs on structure, code examples, API discoverability, schema coverage, LLM parseability and readability, with fixes ranked by impact.
AI Model Sunset Tracker
Track deprecation and shutdown dates for API models from OpenAI, Anthropic, Google, xAI and others, with each replacement and its breaking changes.
Fine-Tuning JSONL Validator
Validate OpenAI or Anthropic fine-tuning JSONL row by row for schema, role order, token counts and duplicates, with per-row issues and summary stats.
Prompt Cache ROI Calculator
Compute prompt-caching breakeven and monthly savings from your static prefix, query tokens and call volume for every current model with cache pricing.
LLM Judge Rubric Builder
Build an LLM-as-judge rubric with weighted criteria on Likert, pass/fail, percentage or 0/1 scales, then export YAML, JSON and a judge prompt.
Eval Dataset Builder
Build a golden eval dataset of prompt, expected, category and tag rows, remove duplicates, check category balance, and export to four eval formats.
Anthropic Stream Event Viewer
Decode an Anthropic Messages SSE stream into text, thinking and tool_use blocks, stop_reason and token usage including cache writes and reads.
AGENTS.md Generator
Generate an AGENTS.md for coding agents from Node, Python, Go or monorepo presets, with setup, build, test and lint commands and done criteria.
llms.txt Generator
Generate an llms.txt file with an H1, summary blockquote and Markdown link sections, or validate an existing one for structure and absolute URLs.
Token & Whitespace Inspector
Show a prompt's token boundaries as a heatmap and flag zero-width, non-breaking and bidirectional control characters that change how a model reads it.
JSON Schema Example Generator
Generate a valid example JSON payload from a JSON Schema, honoring types, enums, required fields, formats, bounds and local $ref references.
The same list is available as JSON at /api/tools-by-category.json.