llm-deploy在 GPU 服务器上部署 LLM 模型服务(vLLM)。支持多服务器配置,自动检查 GPU 和端口占用,一键部署流行的开源大语言模型。
Install via ClawdBot CLI:
clawdbot install wang-junjian/llm-deployGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/vllm-project/vllmAudited Apr 18, 2026 · audit v1.0
Generated Mar 22, 2026
Research teams can quickly deploy multiple LLMs on GPU clusters for experimentation and benchmarking. This skill automates server checks and model launches, reducing setup time from hours to minutes.
Companies can deploy proprietary or open-source LLMs as internal APIs for chatbots, document analysis, or code generation. It supports multi-server configurations for scaling across GPU resources.
Providers can use this skill to manage LLM deployments for customers, offering on-demand model hosting with automated resource monitoring and port management.
Startups can rapidly prototype AI products by deploying models like Llama 3 or Mistral on available GPU servers, enabling quick iteration without deep DevOps expertise.
Institutions can set up hands-on workshops where students deploy and interact with LLMs, using the skill's simple commands to manage models and server states.
Offer a subscription-based service where clients pay for hosted LLM instances on GPU servers. Revenue comes from monthly fees based on model size and usage hours, with automated deployment reducing operational costs.
Provide consulting services to help enterprises integrate LLMs into their workflows, using this skill for setup and optimization. Revenue is generated through project-based contracts and ongoing support retainers.
License the skill as part of a white-label AI platform for other businesses to resell. Revenue comes from licensing fees and a percentage of customer sales, leveraging the skill's ease of use for quick market entry.
💬 Integration Tip
Integrate with existing CI/CD pipelines by automating server checks before deployment, and use the custom model configuration to align with internal model repositories.
Scored Jun 19, 2026
Parse, search, and analyze application logs across formats. Use when debugging from log files, setting up structured logging, analyzing error patterns, correlating events across services, parsing stack traces, or monitoring log output in real time.
Control remote Windows machines via SSH. Use when executing commands on Windows, checking GPU status (nvidia-smi), running scripts, or managing remote Windows systems. Triggers on "run on Windows", "execute on remote", "check GPU", "nvidia-smi", "远程执行", "Windows 命令".
Perform reverse lookup of gTLD domains hosted on a specified nameserver with optional filters by TLD and domain prefix length.
Configure OpenClaw installations with optimized settings, channel setup, security hardening, and production recommendations.
Essential curl commands for HTTP requests, API testing, and file transfers.
Connect to remote desktops via RDP, VNC, and SSH X11 with secure tunneling and troubleshooting.