smartness-eval-open-sourceOpenClaw 智能度综合评伌技能。围绕 14 个维度(含规划能力、幻觉控制)输出综合评分、证据、风险与趋势。对齐 CLEAR/T-Eval/Anthropic 行业标准。
Install via ClawdBot CLI:
clawdbot install yh22e/smartness-eval-open-sourceGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Potentially destructive shell commands in tool definitions
rm -rf /Calls external URL not in known-safe list
https://github.com/xyva-yuangui/smartness-evalUses known external API (expected, informational)
api.deepseek.comAI Analysis
The skill is primarily a read-only evaluation tool that analyzes local system logs and benchmark results. The external API call to DeepSeek is optional, requires explicit user flag (`--llm-judge`), and is consistent with the skill's stated purpose of providing LLM-judged subjective scores. The 'UNSAFE_SHELL' signal appears to be from a rule scanning for the pattern 'rm -rf /' within documentation or code comments, not an active command, as the skill's declared behavior is read-only.
Usage Guide
Loading usage data… refresh in a few seconds.
Scored Jun 29, 2026
Audited Apr 17, 2026 · audit v1.0
Transform AI agents from task-followers into proactive partners that anticipate needs and continuously improve. Now with WAL Protocol, Working Buffer, Autonomous Crons, and battle-tested patterns. Part of the Hal Stack 🦞
Use the ClawdHub CLI to search, install, update, and publish agent skills from clawdhub.com. Use when you need to fetch new skills on the fly, sync installed skills to latest or a specific version, or publish new/updated skill folders with the npm-installed clawdhub CLI.
Mission control dashboard for OpenClaw - real-time session monitoring, LLM usage tracking, cost intelligence, and system vitals. View all your AI agents in o...
Transcribe YouTube videos to text by extracting captions and subtitles directly from the video URL using yt-dlp without audio processing.
Manage a self-hosted Trello-like board via `wekancli`. Create, move and archive cards, lists and boards on a WeKan server. Use when user asks about task boar...
Proactive security monitoring, threat scanning, and auto-remediation for OpenClaw deployments