afrexai-sre-platformComprehensive SRE platform enabling SLO definition, reliability assessment, incident response, chaos engineering, and error budget management without externa...
Install via ClawdBot CLI:
clawdbot install 1kalin/afrexai-sre-platformGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://afrexai-cto.github.io/context-packs/Audited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
An online retailer experiencing seasonal traffic spikes uses the skill to define SLOs for its payment and inventory APIs, implement error budget policies to manage deployment risks during peak sales, and establish incident response runbooks to reduce downtime during critical shopping events.
A financial technology startup handling sensitive transactions applies the skill to assess maturity across monitoring and automation, set high-availability SLOs (e.g., 99.99%) for core services, and create structured postmortem processes to meet regulatory requirements and maintain customer trust.
A software-as-a-service company with microservices architecture uses the skill to catalog services by tier, implement chaos engineering to test resilience in production, and automate capacity planning to support growing user bases while minimizing on-call burnout.
A healthcare organization running batch jobs for patient data processing employs the skill to define SLIs for freshness and correctness, set error budget alerts to catch issues early, and establish incident escalation paths to ensure data integrity and compliance with health regulations.
A streaming platform with diurnal traffic patterns leverages the skill to assess latency and availability SLIs, implement burn rate alerts for slow degradation, and use maturity scoring to prioritize automation and toil reduction for improved viewer experience.
Offer tailored SRE assessments and implementation workshops using the skill's templates and maturity framework, charging per engagement or subscription for ongoing support to help clients build reliability practices from scratch.
Develop a cloud-based tool that automates the skill's phases—such as SLO tracking, incident management, and chaos experiments—and sell it as a subscription service to DevOps teams, with tiered pricing based on service count or usage.
Create online courses and certification programs based on the skill's structured approach, targeting engineers and managers seeking to upskill in SRE practices, with revenue from course sales and exam fees.
💬 Integration Tip
Start by completing the maturity assessment to identify gaps, then use the provided YAML templates to document services and SLOs incrementally, integrating with existing monitoring tools like Prometheus or Datadog for data collection.
Scored Jun 19, 2026
Automatically update Clawdbot and all installed skills once daily. Runs via cron, checks for updates, applies them, and messages the user with a summary of what changed.
Location awareness via privacy-friendly GPS tracking (Home Assistant, OwnTracks, GPS Logger). Set location-based reminders and ask about movement history, travel time, and nearby POIs.
Query real-time road conditions, closures, and traffic issues in Norway. Use when the user asks about road status, closed roads, traffic conditions, weather on roads, or planning a route in Norway. Handles queries like "Are there road closures between Oslo and Bergen?", "What's the road condition on E6?", "Any issues driving to Trondheim today?", or general road condition checks for Norwegian roads.
Set up and manage Xian blockchain nodes. Use when deploying a Xian node to join mainnet/testnet, creating a new Xian network, or managing running nodes. Covers Docker-based setup via xian-stack, CometBFT configuration, and node monitoring.
Monitor and manage Dell PowerEdge servers via iDRAC Redfish API (iDRAC 8/9). Use when asked to: - Check server hardware status, health, or temperatures - Que...
Prometheus monitoring — scrape configuration, service discovery, recording rules, alert rules, and production deployment for infrastructure and application metrics.