mlx-local-inferenceUse when calling local AI on this Mac — text generation, embeddings, speech-to-text, OCR, or image understanding. LLM/VLM via oMLX gateway at localhost:8000/...
Install via ClawdBot CLI:
clawdbot install bendusy/mlx-local-inferenceGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
http://localhost:8000/v1Audited Apr 17, 2026 · audit v1.0
Generated Mar 20, 2026
A clinic uses the MLX Local Inference Stack to process patient notes and medical images locally, ensuring HIPAA compliance by avoiding cloud APIs. The LLM summarizes patient histories, while the VLM and OCR extract information from scanned documents and X-rays, enabling quick, private diagnostic support.
A bank employs this skill to analyze financial reports and contracts on-premises, maintaining data privacy for sensitive client information. The OCR extracts text from PDFs, and the LLM identifies key clauses or risks, streamlining compliance checks without external data exposure.
An e-learning platform uses local inference to generate personalized study materials and transcribe lecture audio offline. The ASR converts spoken content to text, and the LLM creates summaries or quizzes, reducing latency for real-time student interactions in remote areas.
A video production studio leverages the stack to transcribe interviews and analyze visual content without internet dependency. The ASR handles speech-to-text for subtitles, while the VLM assists in scene description or object detection, speeding up post-production workflows.
A law firm implements local AI to search through case files and legal precedents securely on Mac devices. The embedding models index document content for retrieval, and the LLM answers queries or drafts briefs, enhancing research efficiency while protecting confidential data.
Offer a subscription service where businesses pay a monthly fee to access and use the MLX Local Inference Stack on their Apple Silicon Macs. This includes updates, model management, and technical support, targeting organizations needing private, low-latency AI without cloud costs.
Provide consulting services to integrate this skill into existing workflows, such as setting up oMLX servers or optimizing Python scripts for specific use cases. Charge per project or hourly, helping clients deploy and maintain local inference solutions tailored to their needs.
Sell pre-configured model packages or offer premium support for downloading and managing models like Qwen3.5-35B or PaddleOCR-VL. Generate revenue through one-time purchases or support contracts, assisting users with storage and performance optimization on their Macs.
💬 Integration Tip
Ensure all models are downloaded to ~/models/ before use and run ASR/OCR with Python 3.11 via uv to avoid threading issues.
Scored Jun 17, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
Sync OpenRouter models used by OpenClaw into this installation's config. Fetches the OpenClaw app leaderboard from OpenRouter, verifies model IDs against the...