benchmark-model-providerBenchmark and rank AI providers/models against a user-specific prompt suite derived from the user's purpose, domain, and usage frequency. Use when users ask...
Install via ClawdBot CLI:
clawdbot install tankisstank/benchmark-model-providerRequires:
Grade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://vercel.com/docs/deployments/static-sitesUses known external API (expected, informational)
api.openai.comAudited Apr 17, 2026 · audit v1.0
Generated May 23, 2026
A user wants to determine which AI model is most cost-effective and efficient for their daily tasks like email drafting, summarization, and data entry. The benchmark builds a custom prompt suite based on their actual workflow, compares models on speed and cost, and recommends the best fit.
A researcher needs a model that excels at deep reasoning and accurate citations for literature reviews. The skill creates domain-specific questions, runs benchmarks on depth and quality, and outputs a detailed report with raw outputs for auditability.
A software developer seeks a model that provides high-quality code suggestions with minimal errors. The benchmark tests multiple models on programming tasks, scoring them on correctness, efficiency, and cost, then generates a ranking report.
A privacy-conscious user evaluates whether to run a local model or use a cloud API. The skill benchmarks both options on latency, throughput, and quality, helping them decide based on their use case and available compute resources.
A content creator working in multiple languages wants a model that handles Vietnamese, Chinese, and English reliably. The benchmark includes language-specific prompts and evaluates output quality and rendering, producing a bilingual report.
Offer a tiered subscription service where users can run custom benchmarks on a schedule, access historical reports, and get priority support. Revenue comes from monthly or annual plans based on benchmark count and report export options.
Provide enterprise consulting to build tailored benchmark suites for organizations, including integration with their existing AI pipelines. Revenue is generated per project or through retainers for ongoing model evaluation.
Allow free generation of basic markdown reports, with premium features like HTML/PDF exports, custom branding, and web publishing (e.g., via Vercel). Revenue comes from one-time payments or subscriptions for advanced reporting.
💬 Integration Tip
Ensure the BENCHMARK_API_KEY environment variable is set, and verify network access to target model endpoints before running benchmarks. The generated HTML reports can be deployed to static hosting like Vercel for sharing.
Scored Jun 19, 2026
Use this skill when users need to search academic papers, download research documents, extract citations, or gather scholarly information. Triggers include: requests to "find papers on", "search research about", "download academic articles", "get citations for", or any request involving academic databases like arXiv, PubMed, Semantic Scholar, or Google Scholar. Also use for literature reviews, bibliography generation, and research discovery. Requires OpenClawCLI installation from clawhub.ai.
Get institutional-grade CEO performance analytics for S&P 500 companies. Proprietary scores: CEORaterScore (composite), AlphaScore (market outperformance), R...
Creates formal academic research papers following IEEE/ACM formatting standards with proper structure, citations, and scholarly writing style. Use when the user asks to write a research paper, academic paper, or conference paper on any topic.
Search, download, and summarize academic papers from arXiv. Built for AI/ML researchers.
Suggest optimal academic venues for paper submission based on paper content, novelty level, and author's goals. Activated when Chopin asks 'which journal sho...
Baidu Scholar Search - Search Chinese and English academic literature (journals, conferences, papers, etc.)