mcpbeat Sign in

Taobao Product Reviews Agent Skill

Fetch customer reviews for a Taobao or Tmall product by itemId, returning reviewer name, date, purchased variant, review text, and photo URLs. Use when user asks to get product reviews from Taobao, scrape Taobao customer feedback, extract buyer reviews by item ID, collect Tmall ratings and comments, 采集淘宝商品评价, 抓取淘宝买家评论, 获取淘宝商品评论, 天猫商品评价抓取, 按商品ID获取评价. Also applies to sentiment analysis of product reviews, building review datasets, and monitoring product rating changes.

3k tokens
context cost
the whole folder, loaded on every use
3
files
ships runnable scripts
0
copies elsewhere
how many repositories repackaged it
5133
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/browser-act/skills --skill taobao-product-reviews

What comes with it

4 744 bytes besides the instruction
scripts/extract-reviews.py
scripts/next-review-page.py

What it tells the agent to use

found in the instruction text
Bash runs shell commands — read the instruction before connecting

The instruction itself

16 sections, as written by the author

Taobao — Product Reviews

> itemId → paginated customer reviews (reviewer, date, purchased SKU, text, photos)

Language

All process output to user (progress updates, process notifications) follows the user's language.

Objective

Navigate to a Taobao/Tmall product page, load the reviews section, and extract customer review content.

Prerequisites

  • Target page is already open in the browser: https://item.taobao.com/item.htm?id={itemId}
  • User is logged in to Taobao (user avatar or nickname visible in the page header)

Pre-execution Checks

1. Tool Readiness

If browser-act has been confirmed available in the current session → skip this step.

Invoke browser-act via Skill tool to load usage. If installation or configuration issues arise, follow its guidance to resolve then retry.

2. Login Verification

If login status for Taobao has been confirmed in the current session → skip this step.

Otherwise: open https://www.taobao.com and observe the page header:

  • User nickname visible → logged in, continue execution
  • Login button visible → not logged in, inform the user that Taobao login is needed first, assist the user in completing the login flow

User refuses or cannot log in → terminate execution.

Capability Components

> This Skill's operational boundary = what the user can manually do in their browser. It only reads data already displayed to the user on the page, never bypassing authentication or access controls. JS code is encapsulated in Python files under the scripts/ directory, invoked via eval "$(python scripts/xxx.py {params})". $(...) is bash syntax; it is recommended to use the bash tool for execution.

DOM: product reviews (data extraction)

The reviews section is lazy-loaded below the main product area. Follow these steps to load and extract reviews:

  • navigate "https://item.taobao.com/item.htm?id={itemId}"
  • wait stable
  • Close any popup: look for buttons with text "开心收下", "不了", "关闭" and click to dismiss
  • Scroll to trigger lazy loading of the tabs/reviews section:

scroll down --amount 8000

  • wait --selector "[class*='tabTitleItem--']" --state attached --timeout 10000
  • If timeout: scroll down --amount 8000 again and retry wait once more
  • If still no tabs after 2 attempts: take screenshot to confirm page state; the product page may be rendering in a condensed mode — check Known Limitations below
  • eval "$(python scripts/extract-reviews.py '{itemId}')"

Output example:

[
  {
    "username": "一笑奈何",
    "date": "2026-06-03",
    "purchasedSku": "轻巧白|英转中转换器【适用国内电器】适用马来西亚/新加坡等国家",
    "content": "商品非常好,造工很用心!,还会再回购!",
    "photos": [
      "https://gw.alicdn.com/bao/uploaded/i1/O1CN015Cyg4b2FPR2YNq3PD_!!4611686018427383816-0-rate.jpg"
    ],
    "rating": null
  }
]

Notes:

  • purchasedSku: the specific variant the reviewer purchased (extracted from "已购:{sku}" prefix in review header)
  • content: review text body; may be empty if reviewer submitted only photos
  • photos: review photo URLs; empty array if no photos
  • rating: star rating; not always visible in current page layout (null is common)
  • Reviews shown are the default sort (most recent or most helpful as determined by Taobao)

Error handling: if result count = 0 after scroll attempts, the reviews section may not have loaded in the current browser rendering environment. Try navigating to the product page fresh (navigate again) and repeating the scroll sequence. If still failing, this is a known rendering limitation — see Known Limitations below.

DOM: paginate to next review page

After extracting current page reviews:

  • eval "$(python scripts/next-review-page.py)"
  • Returns {"hasNext": true, "buttonText": "下一页"} if next page exists, or {"hasNext": false} if on last page
  • If hasNext is true: state to find the "下一页" button index → click <index>
  • wait stable
  • Re-run eval "$(python scripts/extract-reviews.py '{itemId}')"

Enum Parameters

[collection failed] Sort/filter options for reviews (e.g., newest, most helpful): these controls exist in the reviews section UI but require the tabs section to be loaded; their URL parameters are not exposed and must be set via UI clicks on the sort tabs within the reviews section.

Pagination

DOM Pagination: Click the "下一页" button in the reviews section footer. Each page shows ~10 reviews. Termination: "下一页" button is absent or hasNext returns false.

Success Criteria

result count >= 1 and username non-null rate = 100%

Known Limitations

  • Tab section lazy-loading: The reviews section (along with all tabs: specs, images, recommendations) is lazy-loaded and requires scrolling past the main product area to appear. In some browser sessions or rendering environments, the tabs section does not load even after multiple scroll attempts. This is an intermittent behavior of the Taobao product page rendering engine and does not indicate a site change. Workaround: close and reopen the browser session, then navigate fresh.
  • Requires Taobao login; unauthenticated sessions redirect to login page
  • Review content is only visible on the product page; there is no standalone reviews URL for Taobao/Tmall products
  • Only shows positive buyer reviews by default; negative reviews may require clicking a filter tab within the reviews section (if visible)

Execution Efficiency

  • Batch orchestration: Write a bash script to loop through itemIds serially within a single session; add 3–5 second intervals to allow the lazy-loaded reviews section to render.
  • Test before batch execution: After writing a batch script, you must first test with 1–2 items to verify the reviews section loads correctly; only then run the full batch. Never skip testing and execute in batch directly.
  • Reduce redundant pre-operations: When collecting multiple pages of reviews for one product, stay on the same page and paginate via button click rather than re-navigating.
  • Error resumption: Save results page by page; on failure, resume from the last successful page.

Experience Notes

Path: {working-directory}/browser-act-skill-forge-memories/taobao-product-reviews.memory.md

Before execution: If the file exists, read it first — it records unexpected situations encountered during past executions (e.g., a strategy has become ineffective); adjust strategy order accordingly.

After execution: If an unexpected situation is encountered (strategy became ineffective, page redesigned, anti-scraping upgraded, better path discovered), append a line:

{YYYY-MM-DD}: {what happened} → {conclusion}

Normal execution does not write to the file. Do not record what keywords were used or how many results were returned — those are task outputs, not experience.

Other skills for the same job

different authors, same section of the catalogue
Product Reel Generator
by gooseworks-ai

Generates Instagram-ready product reels from any e-commerce product page URL. Scrapes product images, classifies by type, generates AI-animated clips via Higgsfield API, creates text overlays with style presets, and composes a 15-20 second reel with music. Supports model-based and product-only reels.

4k tokens scripts
Podcast Transcript Fetcher
by Varnan-Tech

Use when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast. Tier 2 (RSS+Groq Whisper) is the recommended approach -- fast, free, and most reliable. Also use when asked to "get transcript", "find episode", "summarize podcast", or "search podcast content". Do not use for general web scraping or non-podcast audio transcription.

20k tokens scripts
Comfyui Launch Flags
by artokun

Pick the right ComfyUI startup flags for VRAM, attention, caching, and speed — the full decision matrix for OOM (--novram / --cache-none / --disable-smart-memory), shared-VRAM creep on Windows (--reserve-vram N), model-switching with big text encoders (--cache-none), high-VRAM throughput (--gpu-only / --highvram), and attention-backend selection (--use-sage-attention for speed, --use-pytorch-cross-attention as the highest-quality / Z-Image-safe fallback). Also the acceleration-stack + Blackwell/RTX 5000 (sm_120) notes. Use when a graph OOMs (especially long video like LTX 2 / WAN), when the GPU spills into shared VRAM and slows to a crawl, when switching between models eats all RAM, when Z-Image produces black/garbled output under Sage, or when deciding which attention backend to launch with. Flag names verified against upstream comfy/cli_args.py — see Sources.

3k tokens
Gallery Scraper
by jdrhyne

Bulk download images from login-protected gallery websites using an attached browser session. Use when asked to scrape, download, or save images from authenticated gallery pages, extract full-size images from thumbnails, or batch download from multi-page galleries.

3k tokens scripts
Yandex Webmaster
by artwist-polyakov

| сайтмапы, переобход, ссылки, фиды, диагностика. Плюс scraping раздела Alice / Share of Voice (нет публичного API). вебмастер индексация, вебмастер запросы, вебмастер переобход, share of voice, sov, алиса, alice efficiency, конкуренты в алисе.

34k tokens scripts ru
Image Scraper
by aAAaqwq

Scrape and download all images from a given URL. Takes a URL, extracts image URLs from the page, and downloads them. Uses python3/curl as primary method, falls back to browser automation if needed. Use when user provides a URL and wants to download images from that page.

2k tokens scripts
Rehab Estimator
by miron-tech

Generate a photo-based rehab estimate for any property. Accepts photos from listing sites (Redfin/Zillow via Chrome), a local folder on your computer, or a shared Google Drive link. Use when a wholesaler needs repair cost estimates before making an offer, building a deal package, or validating their numbers. Grades property condition across 6 zones using the R.E.H.A.B.+F scoring framework and produces three-scenario rehab budgets (rental-ready, mid-range flip, full worst-case). Uses Chrome MCP for Redfin photo browsing, Perplexity for local contractor costs, and Firecrawl for finding listing URLs.

11k tokens
Canvas Design
by anthropics
vendor ×13

Create beautiful visual art in .png and .pdf documents using design philosophy. You should use this skill when the user asks to create a poster, piece of art, design, or other static piece. Create original visual designs, never copying existing artists' work to avoid copyright violations.

1388k tokens

How to use it

Copy the folder

Take browser-act/taobao-product-reviews from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.