Document Keyword Occurrence Count
Count keyword occurrences in PDF/DOCX/PPTX and other documents, with support for batch report export
Browser execution mode: Your data is processed in your browser and is not uploaded to the server.
Speed and Stability: Processing speed depends on your device and browser. For large batch work, the desktop version may be more stable.
Loading tool, please wait...
If the online tool fails to load or run, try the desktop tool.https://tools.yikeaigc.com/
Tool Usage
Return to old versionCount Settings
Enter one keyword per line; in regex mode, enter one expression per line.
Count multiple keywords at once. Empty lines are ignored.
Set snippet max to 0 to skip context extraction.
Set end page to 0 to count to the last page.
After saving, keywords, match mode, PDF pages, and export format will be restored next time.
Add content
Paste text directly, or choose one or more document files.
Supports PDF, DOCX, PPTX, XLSX, TXT, MD, HTML, CSV, JSON, SRT.
Download stats
After starting stats, you can preview results and download the report.
Instructions
Software Usage Instructions
- Set keywords: Enter the keywords to count line by line in the “Keyword List”; in regular expression mode, enter one expression per line.
- Select a matching method:
- Normal match: Matches characters based on the entered content, suitable for Chinese words and fixed phrases.
- Whole-word match: Suitable for English and numeric terms, reducing false matches when contained within longer words.
- Regular expression: Suitable for counting patterned content such as IDs, dates, and formatted codes.
- Configure counting parameters: Enable case sensitivity, overlapping counts, whitespace merging, line-break hyphen merging, hit line counts, and PDF page-by-page counting as needed.
- Set snippets and page numbers: Set the number of hit snippets, snippet length, and the PDF start and end pages; entering 0 for the end page means counting to the last page of the document.
- Add content: You can paste text or select files such as PDF, DOCX, PPTX, XLSX, TXT, Markdown, HTML, CSV, JSON, and SRT. Multiple files will all be processed, and the interface displays only the first 20.
- Save settings: Click “Save Settings” to keep the current keywords, matching method, page range, and export format for future use.
- Start counting: After entering the “Count & Download” step, click “Start Counting” and wait for processing to finish to view keyword occurrences, hit line counts, PDF page-by-page results, and hit snippets.
- Download results: For a single content item, you can download a TXT, CSV, or JSON report; batch files will generate a ZIP containing the report for each file and an optional summary CSV. If result file names are duplicated, numbers will be appended automatically to distinguish them.
FAQ
INV-\d+.
Related Tools
Markdown to WeChat Official Account Formatting Tool
Convert Markdown articles into WeChat Official Account fo...
Extract Specified PDF Pages as Images
Extract specified pages from a PDF and convert them to im...
Batch Document Page Splitting Tool
Batch split PDF, DOCX, and PPTX document pages, with supp...
Article Keyword Density Processor
Analyze and adjust article keyword density. Automatically...