mcpbeat Sign in

File Organization Agent Skill

File organization toolkit. Provides duplicate detection, file merging/splitting, pattern matching, text processing (like case conversion, word counting), and file classification by size or time.

21k tokens
context cost
the whole folder, loaded on every use
10
files
ships runnable scripts
0
copies elsewhere
how many repositories repackaged it
133
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/MassLab-SII/open-agent-skills --skill file_organization

The instruction itself

15 sections, as written by the author

File Organization Skill

This skill provides file analysis, manipulation, and organization:

  • Duplicate detection: Find and organize files with identical content
  • File merging: Combine multiple small files into one
  • File splitting: Split large files into smaller equal-sized parts
  • Pattern matching: Find files containing specific substrings
  • Text transformation: Convert files to uppercase and count words
  • Size classification: Organize files by size thresholds
  • Time classification: Organize files by creation time

1. Duplicate File Detection

Scans all files in a directory, identifies files with duplicate content, and moves them to a separate subdirectory.

Example

# Find and organize duplicate files (default directory name: 'duplicates')
python find_duplicates.py /path/to/directory

# Use a custom directory name for duplicates
python find_duplicates.py /path/to/directory --duplicates-dir my_duplicates

2. File Merging

Identifies the N smallest .txt files, sorts them alphabetically, and merges their content into a single file.

Example

# Merge the 10 smallest .txt files (default)
python merge_files.py /path/to/directory

# Merge the 5 smallest files
python merge_files.py /path/to/directory --count 5

# Use a custom output filename
python merge_files.py /path/to/directory --output merged_content.txt

3. File Splitting

Splits a large text file into multiple smaller files with equal character counts.

Example

# Split large_file.txt into 10 equal parts (default)
python split_file.py /path/to/directory large_file.txt

# Split into 5 parts
python split_file.py /path/to/directory large_file.txt --parts 5

4. Pattern Matching

Finds all files that contain a substring of N or more characters that also appears in a reference file.

Example

# Find files with 30+ character matches (default)
python pattern_matching.py /path/to/directory large_file.txt

# Use a custom minimum length
python pattern_matching.py /path/to/directory large_file.txt --min-length 50

5. Text Transformation (Uppercase)

Converts text files to uppercase and counts words in each file.

Example

# Convert specific files to uppercase
python convert_uppercase.py /path/to/directory --files file_01.txt file_02.txt file_03.txt

6. File Size Classification

Classifies files into different subdirectories based on their file sizes.

Example

# Using default thresholds (300 and 700 bytes)
python classify_files_by_size.py /path/to/directory

# Custom thresholds
python classify_files_by_size.py /path/to/directory --small 1024 --large 10240

# Custom category names
python classify_files_by_size.py /path/to/directory --small-category tiny --medium-category normal --large-category huge

7. File Time Classification

Classifies files into MM/DD directory structure based on their creation time.

Example

# Classify files by creation time
python classify_files_by_time.py /path/to/directory

Other skills for the same job

different authors, same section of the catalogue
Internal Comms
by anthropics
vendor ×13

A set of resources to help me write all kinds of internal communications, using the formats that my company likes to use. Claude should use this skill whenever asked to write some sort of internal communications (status reports, leadership updates, 3P updates, company newsletters, FAQs, incident reports, project updates, etc.).

6k tokens
Competitive Ads Extractor
by frostant
×10

Extracts and analyzes competitors' ads from ad libraries (Facebook, LinkedIn, etc.) to understand what messaging, problems, and creative approaches are working. Helps inspire and improve your own ad campaigns.

2k tokens
Lead Research Assistant
by frostant
×8

Identifies high-quality leads for your product or service by analyzing your business, searching for target companies, and providing actionable contact strategies. Perfect for sales, business development, and marketing professionals.

2k tokens
Developer Growth Analysis
by frostant
×6

Analyzes your recent Claude Code chat history to identify coding patterns, development gaps, and areas for improvement, curates relevant learning resources from HackerNews, and automatically sends a personalized growth report to your Slack DMs.

4k tokens
App Store Optimization
by alirezarezvani
×3

Complete App Store Optimization (ASO) toolkit for researching, optimizing, and tracking mobile app performance on Apple App Store and Google Play Store

55k tokens scripts
Deeptools
by christophacham
×3

NGS analysis toolkit. BAM to bigWig conversion, QC (correlation, PCA, fingerprints), heatmaps/profiles (TSS, peaks), for ChIP-seq, RNA-seq, ATAC-seq visualization.

21k tokens scripts
Pymatgen
by christophacham
×3

Materials science toolkit. Crystal structures (CIF, POSCAR), phase diagrams, band structure, DOS, Materials Project integration, format conversion, for computational materials science.

26k tokens scripts
Enhance Prompt
by google-labs-code
vendor ×2

Transforms vague UI ideas into polished, Stitch-optimized prompts. Enhances specificity, adds UI/UX keywords, injects design system context, and structures output for better generation results.

3k tokens

How to use it

Copy the folder

Take masslab-sii/file_organization from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.