Skip to content

AI Developer Tools

Practical tools for working with AI APIs and models every day. Compare pricing and specs across providers, count tokens before you hit send, test prompt variations, validate function schemas, and build function-calling contracts — all without leaving your browser. Everything runs client-side, so your data stays private.

72 tools

Agent Skill Validator

Validate skill definitions across OpenClaw, Claude, Codex, and MCP with portability scoring and exact fixes

SkillSpec Converter

Convert one canonical skill definition into OpenClaw SKILL.md, Claude blocks, Codex scaffolds, and MCP manifest snippets

Skill Regression Suite Builder

Build deterministic regression suites for skill updates with risk-weighted pass-rate gates and CI-ready case definitions

Skill Scope Collision Detector

Detect cross-scope skill version and enabled-state collisions across global, user, project, and local configuration layers

Skill Payload Budget Optimizer

Plan which agent skills to keep, compress or defer when their combined token payload exceeds the context window minus reserved tokens.

Tool Approval Matrix Compiler

Compile cross-platform allow, ask, and deny decisions for tool capabilities across Codex, Claude, and managed MCP policies

Skill Release Canary Planner

Generate staged canary rollout plans for skill updates with deterministic stop conditions and rollback checklists

Trace Failure Classifier

Classify failed trace events into root-cause buckets and output deterministic remediation guidance for agent incident triage

LLM Crawl Policy Validator

Validate robots.txt and llms.txt files, detect conflicts, simulate AI bot access, and export corrected policies

MCP Governance Composer

Compose managed MCP governance packs with allow/deny lists, approval boundaries, and operator rollout checklists

MCP Tool Search Budget Simulator

Simulate context-window usage for full MCP tool injection versus search-first retrieval strategies

Claude Settings Scope Diff

Diff managed, user, project, and local settings scopes and compute the effective merged Claude configuration

Claude Hook Policy Simulator

Simulate Claude hook decisions for pre/post tool events and validate policy rule coverage before rollout

OpenClaw Skill Trust Scanner

Scan SKILL.md instructions for destructive command patterns, missing safety boundaries, and trust posture

Agent Tool Blast Radius Mapper

Map tool capability blast radius, score operational risk, and produce least-privilege policy buckets

AI Model Comparison

Compare AI models side by side: pricing, context windows, max output, and release dates

MCP Server Directory

Browse and search Model Context Protocol (MCP) servers by category, with GitHub stars, transport type and copyable install commands.

AI Agent Framework Comparison

Compare AI agent frameworks such as LangChain, CrewAI, AutoGen and Mastra by language, GitHub stars, multi-agent, tool, RAG and MCP support.

AI Pricing Calculator

Calculate daily, monthly and yearly LLM API costs per model from your request volume and average input and output tokens.

LLM Token Counter

API

Count tokens for GPT, Claude, Gemini, Grok, DeepSeek, Mistral, Qwen and Muse models and estimate API cost. Also available as a REST API.

CLAUDE.md / Rules File Generator

Generate CLAUDE.md, .cursorrules, copilot-instructions.md, .windsurfrules and .clinerules files from one form or a stack template.

AI Model Picker Quiz

Answer 7 questions and get personalized AI model recommendations across GPT, Claude, Gemini, Grok, DeepSeek, and open-weight models

System Prompt Library

Browse and copy curated system prompts for coding, writing, analysis, business, education and creative tasks, searchable by keyword.

MCP Server Config Generator

Generate MCP server configurations for Claude Desktop, Cursor, and Windsurf with visual editor and presets

JSON Schema Generator

API

Generate Draft-07 or 2020-12 JSON Schemas from sample JSON for function calling, structured outputs and validation. Also available as a REST API.

Prompt Template Builder

Build AI prompt templates with variables, live preview, and export to JSON/YAML

AI Cost Estimator

Estimate monthly or annual LLM API costs per model from workload presets, request volume, token counts and prompt-cache hit rate.

System Prompt Editor

Write and analyze AI system prompts with live token counting, variable detection, and XML highlighting

LLM Output Diff Tool

Compare two to four LLM outputs side by side, diff any pair by line or by word, and compare response length in characters and words.

AI Context Window Visualizer

Visualize how your AI model's context window is allocated across system prompt, tools, conversation, and RAG

AI Prompt Tester & Comparator

Compare two to four prompt variants side by side with word-level diffs, token estimates and the {{variables}} each one uses.

Prompt Token Budget Planner

Split a model's context window across system prompt, tools, memory, retrieved documents and history, with a warning when allocation passes 80%.

AI Tool Schema Builder

Build AI tool definitions in a visual editor with typed, enum and nested parameters, and export them as OpenAI, Anthropic, MCP or JSON Schema.

LLM Structured Output Validator

API

Validate LLM JSON output against a JSON Schema, with OpenAI, Anthropic and MCP presets, per-field fixes and sample output. Also has a REST API.

Markdown Memory File Builder

Build markdown memory files for AI agents from guided forms: SOUL.md, USER.md, AGENTS.md, daily logs, decision records, project status and lessons.

Embedding Similarity Calculator

Compare two embedding vectors by cosine similarity, dot product, Euclidean and Manhattan distance, with dimension hints for common embedding models.

RAG Chunk Size Calculator

Recommend a RAG chunk size, overlap and chunking strategy from document type, query type and embedding model, with a context-window token budget.

Agent Trace Viewer

Inspect AI agent traces from LangChain runs, OpenAI Assistants run steps or a JSON step array, with LLM calls, tool calls, errors, timings and tokens.

Prompt Version Diff

Diff two prompt versions line by line, with markers for template variables, XML tags and instruction keywords, plus added variables and token delta.

AI Guardrail Rule Tester

Write keyword and regex guardrail rules that block, flag or redact text, test them against sample inputs and outputs, and export them as JSON.

Function Call Flow Simulator

Script a multi-step tool-calling conversation with mock tool results and errors, then export it as an OpenAI or Anthropic request body.

AI API Key Tester

Identify an AI API key's provider from its prefix (Anthropic, OpenAI, OpenRouter, Groq, Google AI, xAI and more) and check its format locally.

AI Token + Pricing Calculator

Paste text to estimate its tokens and compare input and output cost across GPT, Claude, Gemini and other catalog models, with CSV export.

WebMCP Playground

Validate WebMCP tool definitions or a webmcp.json manifest against lint rules, test them with a form built from the schema and preview discovery.

MCP Server Starter Generator

Generate a runnable MCP server project in TypeScript, Python or Go with tools, resources, prompts, stdio or HTTP transport and optional auth.

Cursor Rules Generator

Generate .cursorrules, .windsurfrules, .clinerules, CLAUDE.md or Copilot instructions from your stack, conventions, testing approach and rules.

AI Response Comparator

Compare up to four LLM responses side by side, with line or word diffs and an analysis of length, shared sentences and content unique to each.

System Prompt Analyzer

Estimate a system prompt's tokens for Claude, GPT, Llama and Gemini, its share of a 4K to 200K context window, and its template variables.

MCP Tool Tester

Validate a WebMCP tool definition's name, description and inputSchema, generate a test form from the schema, and mock a call locally.

Web-to-Markdown Converter

Convert pasted HTML or a fetched URL to Markdown for LLM input, dropping scripts, navigation and footers, and estimate the token savings.

AI Code Smell Detector

Scan code for AI-generated anti-patterns: hallucinated imports, over-abstraction, verbose error handling, redundancy, security smells and AI tells.

LLM Workflow Cost Calculator

Price a multi-step LLM pipeline (embed, retrieve, generate, validate) per run, day and month, with cached-input rates and a provider comparison.

Codebase Context Packer

Pack code files into an LLM prompt under a token budget with per-file truncation (full, signatures, first N lines) and XML, Markdown or plain output.

AI Prompt Injection Tester

Check a system prompt's defenses against known injection patterns: role hijacking, instruction override, delimiter abuse, encoding tricks, jailbreaks.

AI API Error Decoder

Decode an error response from the OpenAI, Anthropic, Gemini, Mistral or Cohere API into its cause, a Python and Node.js fix and a retry strategy.

AI Agent Cost Simulator

Simulate multi-agent LLM costs with per-turn context growth and tool-call overhead, compare against a single agent and project monthly spend.

AI Rules Linter

Lint CLAUDE.md, .cursorrules, and copilot-instructions files for redundancy, conflicting instructions, missing sections, and token efficiency.

Git Diff Token Counter

Paste a git diff and see token counts per file, cost across AI models for code review, and chunking suggestions when diffs exceed context limits.

LLM Latency Estimator

Estimate time to first token, generation time and total latency for current LLMs from your token counts, with a UX pattern for each wait.

Prompt A/B Test Designer

Plan a prompt A/B test: sample size per variant from a two-proportion z-test with Bonferroni correction, cost, duration and a Markdown plan.

MCP Permission Auditor

Audit an MCP server config for risky capabilities, dangerous combinations such as file read plus network, risk scores and least-privilege fixes.

AI Doc Readability Scorer

Score Markdown docs on structure, code examples, API discoverability, schema coverage, LLM parseability and readability, with fixes ranked by impact.

AI Model Sunset Tracker

Track deprecation and shutdown dates for API models from OpenAI, Anthropic, Google, xAI and others, with each replacement and its breaking changes.

Fine-Tuning JSONL Validator

Validate OpenAI or Anthropic fine-tuning JSONL row by row for schema, role order, token counts and duplicates, with per-row issues and summary stats.

Prompt Cache ROI Calculator

Compute prompt-caching breakeven and monthly savings from your static prefix, query tokens and call volume for every current model with cache pricing.

LLM Judge Rubric Builder

Build an LLM-as-judge rubric with weighted criteria on Likert, pass/fail, percentage or 0/1 scales, then export YAML, JSON and a judge prompt.

Eval Dataset Builder

Build a golden eval dataset of prompt, expected, category and tag rows, remove duplicates, check category balance, and export to four eval formats.

Anthropic Stream Event Viewer

Decode an Anthropic Messages SSE stream into text, thinking and tool_use blocks, stop_reason and token usage including cache writes and reads.

AGENTS.md Generator

Generate an AGENTS.md for coding agents from Node, Python, Go or monorepo presets, with setup, build, test and lint commands and done criteria.

llms.txt Generator

Generate an llms.txt file with an H1, summary blockquote and Markdown link sections, or validate an existing one for structure and absolute URLs.

Token & Whitespace Inspector

Show a prompt's token boundaries as a heatmap and flag zero-width, non-breaking and bidirectional control characters that change how a model reads it.

JSON Schema Example Generator

Generate a valid example JSON payload from a JSON Schema, honoring types, enums, required fields, formats, bounds and local $ref references.

The same list is available as JSON at /api/tools-by-category.json.