openclaw-smartness-evalOpenClaw 智能度综合评伌技能。围绕 14 个维度(含规划能力、幻觉控制)输出综合评分、证据、风险与趋势。对齐 CLEAR/T-Eval/Anthropic 行业标准。
Install via ClawdBot CLI:
clawdbot install yh22e/openclaw-smartness-evalGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Potentially destructive shell commands in tool definitions
rm -rf /Uses known external API (expected, informational)
api.deepseek.comAudited Apr 18, 2026 · audit v1.0
Generated Mar 21, 2026
After deploying a new version of an AI agent, this skill runs automated evaluations to verify performance improvements or detect regressions across multiple dimensions like reasoning and latency. It ensures updates genuinely enhance intelligence rather than just superficial changes.
Organizations using AI agents for customer support or data analysis can schedule weekly assessments to generate structured reports on capability trends. This helps identify degradation risks early and maintain consistent service quality for operational reliability.
Research labs conducting AI evaluations leverage this skill to aggregate benchmark results and analyze trends over time. It supports academic studies by providing detailed scores, evidence, and risk flags for comparative analysis of AI models.
Before showcasing an AI agent to clients or investors, teams use this skill to produce unified assessment reports. It standardizes evaluation outputs, highlighting strengths and areas for improvement to build confidence in the product's capabilities.
Companies in regulated industries employ this skill to conduct capability audits and self-evaluations, ensuring AI systems meet internal standards and external requirements. It generates evidence-based reports for transparency and compliance documentation.
Offer this skill as a cloud-based service where AI developers subscribe to access automated evaluation tools and detailed reports. Revenue comes from monthly or annual subscriptions based on usage tiers and advanced features like trend analysis.
Provide professional services to integrate this skill into clients' AI systems, offering customization, training, and ongoing support. Revenue is generated through project-based fees and retainer contracts for continuous monitoring and optimization.
Release the core skill as open-source to build community adoption, while monetizing premium features such as advanced analytics, LLM judge integrations, and priority support. Revenue streams include licensing fees for enterprise add-ons and support packages.
💬 Integration Tip
Ensure all required data sources like response-latency-metrics.json are accessible and properly configured before running evaluations to avoid errors and obtain accurate results.
Scored Jun 19, 2026
Transform AI agents from task-followers into proactive partners that anticipate needs and continuously improve. Now with WAL Protocol, Working Buffer, Autonomous Crons, and battle-tested patterns. Part of the Hal Stack 🦞
Use the ClawdHub CLI to search, install, update, and publish agent skills from clawdhub.com. Use when you need to fetch new skills on the fly, sync installed skills to latest or a specific version, or publish new/updated skill folders with the npm-installed clawdhub CLI.
Mission control dashboard for OpenClaw - real-time session monitoring, LLM usage tracking, cost intelligence, and system vitals. View all your AI agents in o...
Transcribe YouTube videos to text by extracting captions and subtitles directly from the video URL using yt-dlp without audio processing.
Manage a self-hosted Trello-like board via `wekancli`. Create, move and archive cards, lists and boards on a WeKan server. Use when user asks about task boar...
Proactive security monitoring, threat scanning, and auto-remediation for OpenClaw deployments