mcpbeat Sign in

Eyeball Agent Skill

Document analysis with inline source screenshots. When you ask Copilot to analyze a document, Eyeball generates a Word doc where every factual claim includes a highlighted screenshot from the source material so you can verify it with your own eyes.

9k tokens
context cost
the whole folder, loaded on every use
2
files
ships runnable scripts
0
copies elsewhere
how many repositories repackaged it
37394
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/github/awesome-copilot --skill eyeball

What comes with it

28 494 bytes besides the instruction
tools/eyeball.py

The instruction itself

13 sections, as written by the author

Eyeball

Analyze documents with visual proof. When activated, Eyeball produces a Word document on the user's Desktop where every factual assertion includes an inline screenshot from the source material with the cited text highlighted in yellow.

Activation

When the user invokes this skill (e.g., "use eyeball", "run eyeball on this", "eyeball this document"), respond with:

> Eyeball is active. I'll analyze the document and produce a Word doc with inline source screenshots so you can verify every claim with your own eyes.

Then follow the workflow below.

Supported Sources

  • Local files: Word documents (.docx, .doc), PDFs (.pdf), RTF files
  • Web URLs: Any publicly accessible web page

Tool Location

The Eyeball Python utility is located at:

<plugin_dir>/skills/eyeball/tools/eyeball.py

To find the actual path, run:

find ~/.copilot/installed-plugins -name "eyeball.py" -path "*/eyeball/*" 2>/dev/null

If not found there, check the project directory or the user's home directory for the eyeball repo.

First-Run Setup

Before first use, check that dependencies are installed:

python3 <path-to>/eyeball.py setup-check

If anything is missing, install the required dependencies:

pip3 install pymupdf pillow python-docx playwright
python3 -m playwright install chromium

On Windows, also install pywin32 for Word automation:

pip install pywin32

Workflow

Follow these steps exactly. The order matters.

Step 1: Read the source text

Before writing any analysis, extract and read the full text of the source document:

python3 <path-to>/eyeball.py extract-text --source "<path-or-url>"

Read the output carefully. Identify actual section numbers, headings, page numbers, and key language.

CRITICAL: Do not skip this step. Do not write analysis based on assumptions about how the document is structured. Read the actual text.

Step 2: Write analysis with exact citations

For each point in your analysis, you must:

  • Reference the correct section number as it appears in the document (e.g., "Section 9" not "Section 8" because you assumed the numbering).
  • Reference the correct page number where the section appears in the extracted text.
  • Select anchors that are verbatim phrases from the source that directly support your claim.

Step 3: Select anchors correctly

This is the most important step. Anchors determine what gets highlighted in the screenshots.

DO:

  • Use verbatim phrases from the source text that directly support your assertion
  • Use multiple anchors to span the full range of text the reader should see
  • Use specific, uncommon phrases that appear only where you intend

DO NOT:

  • Use generic topic labels (e.g., "Confidentiality") that appear throughout the document
  • Use section titles alone when they appear as cross-references elsewhere
  • Use single common words that match in many places

Examples:

WRONG -- uses a generic topic label that matches everywhere:

{"anchors": ["User-Generated Content"], "target_page": 8}

RIGHT -- uses the specific language that supports the claim:

{"anchors": ["retain ownership", "Ownership of Content, Right to Post"], "target_page": 8}

WRONG -- section title appears as a cross-reference on earlier pages:

{"anchors": ["LIMITATION OF LIABILITY"]}

RIGHT -- includes the section number for precision, targets the correct page:

{"anchors": ["12. LIMITATION OF LIABILITY", "INDIRECT", "CONSEQUENTIAL"], "target_page": 13}

Step 4: Build the analysis document

Construct a JSON array of sections and call the build command:

python3 <path-to>/eyeball.py build \
  --source "<path-or-url>" \
  --output ~/Desktop/<title>.docx \
  --title "Analysis Title" \
  --subtitle "Source description" \
  --sections '[
    {
      "heading": "1. Section Title",
      "analysis": "Your analysis text here. Reference Section X on page Y...",
      "anchors": ["verbatim phrase 1", "verbatim phrase 2"],
      "target_page": 5,
      "context_padding": 40
    },
    {
      "heading": "2. Another Section",
      "analysis": "More analysis...",
      "anchors": ["exact quote from source"],
      "target_pages": [10, 11],
      "context_padding": 50
    }
  ]'

Section object fields:

  • heading (required): Section heading in the output document
  • analysis (required): Your analysis text
  • anchors (required): List of verbatim phrases from the source to search for and highlight
  • target_page (optional): Single page number (1-indexed) to search on
  • target_pages (optional): List of page numbers to search across (screenshots stitched vertically)
  • context_padding (optional): Padding in PDF points above/below the anchor region (default: 40). Increase for more context.

Step 5: Deliver the output

Save the output to the user's Desktop. Tell the user the filename and that they can open it to verify each claim against the highlighted source screenshots.

Self-Check Before Delivery

Before saving the final document, mentally verify:

  • Does each section's analysis text reference the correct section number from the source?
  • Are the anchors verbatim phrases that appear on the target page?
  • Does each anchor directly support the claim in the analysis, not just relate to the same topic?
  • If the screenshot doesn't match the analysis, is the analysis wrong or is the anchor wrong? Fix whichever is incorrect.

Notes

  • The output document includes highlighted screenshots that are dynamically sized. If you provide multiple anchors, the screenshot expands to cover all of them.
  • When a search term is not found, the output document will note this. If this happens, the anchor was likely not verbatim enough. Adjust and rebuild.
  • For web pages, Playwright renders the page to PDF first. The resulting page numbers may differ from what you see in a browser. Use the extracted text output (step 1) to determine correct page numbers.
  • If the user has already provided the source text or you have already read it in the current conversation, you can skip step 1. But always verify section numbers and page references against the actual text before writing analysis.

Other skills for the same job

different authors, same section of the catalogue
DOCX
by anthropics
vendor ×16

Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks

7k tokens
PDF
by anthropics
vendor ×16

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

13k tokens scripts
PPTX
by JayZeeDesign
×15

Presentation creation, editing, and analysis. When Claude needs to work with presentations (.pptx files) for: (1) Creating new presentations, (2) Modifying or editing content, (3) Working with layouts, (4) Adding comments or speaker notes, or any other presentation tasks

308k tokens scripts
Canvas Design
by anthropics
vendor ×13

Create beautiful visual art in .png and .pdf documents using design philosophy. You should use this skill when the user asks to create a poster, piece of art, design, or other static piece. Create original visual designs, never copying existing artists' work to avoid copyright violations.

1388k tokens
PDF
by anthropics
vendor ×10

Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multiple PDFs into one, splitting PDFs apart, rotating pages, adding watermarks, creating new PDFs, filling PDF forms, encrypting/decrypting PDFs, extracting images, and OCR on scanned PDFs to make them searchable. If the user mentions a .pdf file or asks to produce one, use this skill.

15k tokens scripts
DOCX
by w95
×6

Use this skill whenever the user wants to create, read, edit, or manipulate Word documents (.docx files). Triggers include: any mention of 'Word doc', 'word document', '.docx', or requests to produce professional documents with formatting like tables of contents, headings, page numbers, or letterheads. Also use when extracting or reorganizing content from .docx files, inserting or replacing images in documents, performing find-and-replace in Word files, working with tracked changes or comments, or converting content into a polished Word document. If the user asks for a 'report', 'memo', 'letter', 'template', or similar deliverable as a Word or .docx file, use this skill. Do NOT use for PDFs, spreadsheets, Google Docs, or general coding tasks unrelated to document generation.

5k tokens
PPTX
by w95
×4

Use this skill any time a .pptx file is involved in any way — as input, output, or both. This includes: creating slide decks, pitch decks, or presentations; reading, parsing, or extracting text from any .pptx file (even if the extracted content will be used elsewhere, like in an email or summary); editing, modifying, or updating existing presentations; combining or splitting slide files; working with templates, layouts, speaker notes, or comments. Trigger whenever the user mentions \"deck,\" \"slides,\" \"presentation,\" or references a .pptx filename, regardless of what they plan to do with the content afterward. If a .pptx file needs to be opened, created, or touched, use this skill.

2k tokens
Obsidian Markdown
by ZhanlinCui
×3

Create and edit Obsidian Flavored Markdown with wikilinks, embeds, callouts, properties, and other Obsidian-specific syntax. Use when working with .md files in Obsidian, or when the user mentions wikilinks, callouts, frontmatter, tags, embeds, or Obsidian notes.

3k tokens

How to use it

Copy the folder

Take github/eyeball from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.

Install what it needs

The instructions reference pip. Without those the skill loads but fails at the first command.