model-deployUse this skill when users request to deploy LLMs (Qwen, DeepSeek, etc.) on specified GPU servers and start the model service. This skill can Download models...
Install via ClawdBot CLI:
clawdbot install wangwei1237/model-deployGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → http://127.0.0.1:8001/v1/chat/completionsCalls external URL not in known-safe list
http://127.0.0.1:8001/v1/chat/completionsAudited Apr 18, 2026 · audit v1.0
Generated Mar 21, 2026
Researchers need to deploy custom LLMs like Qwen or DeepSeek on GPU clusters for experimentation and benchmarking. This skill automates the setup of vLLM services, allowing quick iteration on model testing without manual server configuration.
Companies deploying AI-powered chatbots can use this skill to scale model services across multiple GPU servers. It ensures efficient deployment of models like Qwen for customer support, handling increased user loads with tensor parallelism.
Providers offering LLM-as-a-Service can deploy models for clients on dedicated GPU infrastructure. This skill streamlines the process using ModelScope and vLLM, enabling rapid provisioning of models like Qwen with configurable ports and GPU allocation.
Data science teams prototyping AI applications require quick deployment of models for internal testing. This skill facilitates setting up vLLM services on shared GPU servers, allowing teams to experiment with different models and parameters efficiently.
Offer GPU server rental with pre-deployed LLM services using this skill. Clients pay for compute resources and model access, generating revenue through subscription or usage-based pricing for scalable AI inference.
Provide end-to-end deployment and maintenance of LLM services for enterprises. Use this skill to automate setup and updates, charging clients for ongoing support, monitoring, and optimization of model performance on their servers.
Consult firms assist clients in deploying specific models like Qwen for tailored applications. Revenue comes from project-based fees for integration, troubleshooting, and training teams to use the skill effectively in production environments.
💬 Integration Tip
Ensure SSH key-based authentication is set up between servers to automate deployments without manual password entry, and verify GPU compatibility with vLLM to avoid memory issues.
Scored Apr 19, 2026
Control remote Windows machines via SSH. Use when executing commands on Windows, checking GPU status (nvidia-smi), running scripts, or managing remote Windows systems. Triggers on "run on Windows", "execute on remote", "check GPU", "nvidia-smi", "远程执行", "Windows 命令".
Perform reverse lookup of gTLD domains hosted on a specified nameserver with optional filters by TLD and domain prefix length.
Configure OpenClaw installations with optimized settings, channel setup, security hardening, and production recommendations.
Connect to remote desktops via RDP, VNC, and SSH X11 with secure tunneling and troubleshooting.
Essential curl commands for HTTP requests, API testing, and file transfers.
Deploy and manage Vercel projects. Use when deploying applications to Vercel, managing environment variables, checking deployment status, viewing logs, or performing Vercel operations. Supports production and preview deployments. Practical infrastructure operations - no "AI will build your app" magic.