cn-text-encoding-detectorDetect text file encoding (UTF-8, GBK, Latin-1, etc). Auto-detect BOM markers. Pure Python standard library, no API key required.
Install via ClawdBot CLI:
clawdbot install freedompixels/cn-text-encoding-detectorGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://clawhub.ai/user/freedompixelsAudited Jun 5, 2026 · audit v1.0
Generated Aug 12, 2026
Users with old files that appear as mojibake can use the detector to identify the encoding, then convert to a readable format. This is common in archival and document migration projects, especially with Chinese texts.
Data engineering teams processing files from various sources need to know the encoding for correct parsing. This tool can be integrated into batch jobs to detect and convert files before further processing.
Researchers and developers building NLP models for multilingual text need clean, correctly encoded data. The detector ensures text files are in a consistent encoding (like UTF-8) before feeding into models.
Online retailers importing product descriptions from suppliers may receive files in various encodings. Using the detector helps standardize data to avoid display issues on websites and catalogs.
Developers sometimes encounter files with mixed encodings causing compilation errors or display issues. Detecting and normalizing encoding in source code files aids in consistent project maintenance.
Offer the detector as a free CLI tool to build user base, with premium features like batch processing, a GUI, or API access for a subscription fee.
Provide consulting services to enterprises for integrating the detector into their workflows, customizing it for specific encoding detection needs, and training staff.
Release the tool as open-source, but offer official support, certification, and managed services for businesses that need reliability and scalability.
💬 Integration Tip
The tool is a simple Python script, so it can be easily invoked from shell scripts or wrapped in a small API. Integrate it into your data processing pipeline by calling it before file parsing to ensure correct encoding.
Scored Aug 12, 2026
Fetch and read transcripts from YouTube videos. Use when you need to summarize a video, answer questions about its content, or extract information from it.
Monitor RSS and Atom feeds for content research. Track blogs, news sites, newsletters, and any feed source. Use when monitoring competitors, tracking industr...
用 MinerU API 解析 PDF/Word/PPT/图片为 Markdown,支持公式、表格、OCR。适用于论文解析、文档提取。
Provides a personalized morning report with today's reminders, undone Notion tasks, and vault storage summary for daily planning.
Extract text from PDFs with OCR support. Perfect for digitizing documents, processing invoices, or analyzing content. Zero dependencies required.
Fetch scheduled economic events and data releases from the FMP API for specified dates, filtering by impact, country, and type, and output a chronological ma...