Deep research over the Semantic Scholar Graph API. Covers endpoints missing from allenai's lookup skill — paper references (backward citations), recommendations, batch paper lookup (up to 500 IDs), snippet search, and multi-hop citation graph traversal (BFS forward/backward). Use when the user asks to build a citation graph, expand a literature seed, find related work, run a reference network traversal, explore what a paper cites or what cites it beyond simple lookup, or batch-resolve many DOI/arXiv/S2 IDs. For multi-step research questions, delegate to the deep-paper-researcher subagent to keep the main context clean. Not for single paper-by-ID lookups (use semantic-scholar-lookup) or topical discovery (use web_search_advanced_exa).
npx skills add https://github.com/CodeAlive-AI/ai-driven-development --skill semantic-scholar-deep
Purpose: fill the gaps that semantic-scholar-lookup (allenai) leaves — references, recommendations, batch, and multi-hop citation-graph traversal.
ss_client.py + citation_graph.pyTwo execution modes:
Use when the user asks for one specific endpoint:
ss_client.py references <id>ss_client.py recommendations <id>ss_client.py batch ...ss_client.py snippets "..."Fast, cheap, no orchestration overhead.
deep-paper-researcher subagentUse when the task is multi-step or would otherwise flood the context:
Mandatory prompt contents. The subagent runs in isolated context with no access to this conversation's system reminders. Include exactly these two things:
Today is YYYY-MM-DD. Pull from the currentDate system-reminder field, or run date -I via Bash before delegating if it's missing. Never rely on training-data intuitions about the current year.Do NOT do any of these:
Call:
Agent(
subagent_type="deep-paper-researcher",
description="<3–5 word task>",
prompt="Today is 2026-04-22.\n\nUser's request: найди современные 10 статей про AI Code Review на arXiv.\n\n<optional: output format hints, language preference>"
# model: "opus" ← add only when the user opts in (see below)
)
The subagent's Freshness Mode section handles classification; keep this layer thin.
The subagent's model frontmatter is sonnet — that's the default.
Override to Opus by passing model: "opus" to the Agent tool only if the user explicitly requests deeper reasoning. Triggers (any of):
Never auto-upgrade to Opus without a user signal — Sonnet handles the default literature-review workflow fine and costs less.
Trigger this skill for:
Do NOT use for:
semantic-scholar-lookup (faster, no Python)web_search_advanced_exa with category: "research paper" (Exa MCP)deep-paper-researcher subagent, which orchestrates all three toolsLocated under ${SKILL_DIR}/scripts/.
ss_client.py — raw API clientSubcommands (all output JSON on stdout):
| Command | Endpoint | Notes |
|---------|----------|-------|
| search <query> | /graph/v1/paper/search | --bulk switches to /search/bulk (up to 1000/page) |
| paper <id> | /graph/v1/paper/{id} | ID forms: raw, DOI:, ARXIV:, CorpusId:, PMID:, URL: |
| citations <id> | /graph/v1/paper/{id}/citations | paginated; up to 1000 per page |
| references <id> | /graph/v1/paper/{id}/references | paginated; up to 1000 per page |
| recommendations <id> | /recommendations/v1/papers/forpaper/{id} | --pool recent|all-cs |
| batch <id1> <id2> ... | POST /graph/v1/paper/batch | up to 500 IDs |
| author-search <query> | /graph/v1/author/search | |
| author <id> | /graph/v1/author/{id} | |
| author-papers <id> | /graph/v1/author/{id}/papers | |
| snippets <query> | /graph/v1/snippet/search | Full-text snippets |
Common flags: --limit, --offset, --fields, --year, --fields-of-study, --venue, --min-citation-count.
citation_graph.py — BFS traversalpython3 ${SKILL_DIR}/scripts/citation_graph.py <paperId> \
--direction both \
--depth 2 \
--max-nodes 200 \
--per-hop-limit 50 \
--output graph.json
Directions: forward (citations), backward (references), both. Output schema described in the script docstring — nodes: {paperId → metadata+depth}, edges: [{src, dst, direction}].
SEMANTIC_SCHOLAR_API_KEY env var: much higher limits.Retry-After.references/endpoints.md — complete field list per endpoint + query examplesreferences/workflows.md — lit-review, novelty-check, seed-expansion patternsScripts emit raw JSON — redirect to files for anything beyond ~20 results. For graphs >50 nodes always pass --output graph.json to avoid flooding the conversation context.
Typical pipeline inside the deep-paper-researcher subagent:
mcp__exa__web_search_advanced_exa (neural + multi-source)ss_client.py search / batch to get paperId from titles or DOIscitation_graph.py with the top 3-5 seedsA paired subagent definition ships alongside the skill at agents/deep-paper-researcher.md. It orchestrates Exa MCP + allenai semantic-scholar-lookup + this skill's scripts into a token-isolated research agent with:
Anchor date / Mode / Window headerTo install for Claude Code (manual, one-time):
cp ~/.agents/skills/semantic-scholar-deep/agents/deep-paper-researcher.md ~/.claude/agents/
(Path may differ on other agents — copy to the agent's subagents directory, then restart the session.)
Prerequisites for full pipeline: Exa MCP connected, allenai/asta-plugins@"Semantic Scholar Lookup" skill installed.
Assists in writing high-quality content by conducting research, adding citations, improving hooks, iterating on outlines, and providing real-time feedback on each section. Transforms your writing process from solo effort to collaborative partnership.
Identifies high-quality leads for your product or service by analyzing your business, searching for target companies, and providing actionable contact strategies. Perfect for sales, business development, and marketing professionals.
Use this skill to query your Google NotebookLM notebooks directly from Claude Code for source-grounded, citation-backed answers from Gemini. Browser automation, library management, persistent auth. Drastically reduced hallucinations through document-only responses.
Efficient database search tool for bioRxiv preprint server. Use this skill when searching for life sciences preprints by keywords, authors, date ranges, or categories, retrieving paper metadata, downloading PDFs, or conducting literature reviews.
Query and analyze scholarly literature using the OpenAlex database. This skill should be used when searching for academic papers, analyzing research trends, finding works by authors or institutions, tracking citations, discovering open access publications, or conducting bibliometric analysis across 240M+ scholarly works. Use for literature searches, research output analysis, citation analysis, and academic database queries.
Access USPTO APIs for patent/trademark searches, examination history (PEDS), assignments, citations, office actions, TSDR, for IP analysis and prior art searches.
Multiagent AI system for scientific research assistance that automates research workflows from data analysis to publication. This skill should be used when generating research ideas from datasets, developing research methodologies, executing computational experiments, performing literature searches, or generating publication-ready papers in LaTeX format. Supports end-to-end research pipelines with customizable agent orchestration.
Automated LLM-driven hypothesis generation and testing on tabular datasets. Use when you want to systematically explore hypotheses about patterns in empirical data (e.g., deception detection, content analysis). Combines literature insights with data-driven hypothesis testing. For manual hypothesis formulation use hypothesis-generation; for creative ideation use scientific-brainstorming.
Take codealive-ai/semantic-scholar-deep from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.