mcpbeat Sign in

Diffdock Agent Skill

> Predict small-molecule binding poses with DiffDock-L (Corso et al. 2023/2024, github.com/gcorso/DiffDock) — blind diffusion docking that places a ligand into a protein pocket without a predefined search box and ranks the samples with a learned confidence model. Reach for this skill to dock a SMILES or SDF against a PDB, to generate ranked 3D poses for a small fragment library, or to get a starting pose for downstream rescoring. DiffDock predicts geometry, not affinity.

2k tokens
context cost
the whole folder, loaded on every use
2
files
instructions only
0
copies elsewhere
how many repositories repackaged it
859
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/xuzhougeng/wisp-science --skill diffdock

The instruction itself

7 sections, as written by the author

DiffDock-L

DiffDock-L is a blind pose predictor: given a protein structure and a ligand,

it samples ligand placements over the whole surface with a diffusion model and

ranks them with a separately trained confidence head. The confidence score

correlates with pose correctness, not with binding free energy — DiffDock does

not predict whether or how tightly the ligand binds, so for hit triage you

still pair it with a scorer (GNINA, MM-GBSA) or with boltz's affinity head.

For protein–protein and nucleic-acid co-folding, route to boltz or chai1.

Code and weights are MIT (github.com/gcorso/DiffDock).

Running it

cd $DIFFDOCK_REPO   # a clone of github.com/gcorso/DiffDock
python3 -m inference \
  --config default_inference_args.yaml \
  --protein_path target.pdb \
  --ligand_description "COc1ccc(C#N)cc1" \
  --out_dir out

For more than one complex, give --protein_ligand_csv batch.csv instead of the

two single-complex flags; the CSV has four columns — complex_name,

protein_path, ligand_description (SMILES or an .sdf/.mol2 path), and

protein_sequence. Leave protein_path empty and fill protein_sequence to

have DiffDock fold the receptor with ESMFold first; that path and a

larger-library screening recipe are in references/workflows.md.

Under --out_dir/<complex_name>/ each sample is written as

rank{N}_confidence{score}.sdf, plus a copy of rank1.sdf for convenience.

The confidence value in the filename is a logit, so it is unbounded and can be

negative; among samples for the *same* complex higher is better, but values are

not comparable across different complexes or ligands.

The YAML config overwrites your CLI flags

inference.py loads --config default_inference_args.yaml *after* argparse

and replaces every key it finds, so passing --samples_per_complex 40 or

--model_dir ... on the command line is silently ignored if the same key sits

in the YAML. To change sampling depth or any other key the YAML defines, copy

the YAML, edit the copy, and point --config at it.

The first run is silent for ~11 minutes and needs ≥32 GB host RAM

Before the first complex, DiffDock precomputes SO(3) and torus lookup tables.

That step is silent on stderr, takes ~11 minutes, and may exhaust a small machine.

Use a probed SSH context with at least 64 GiB RAM and precompute the tables while

building the environment; do not assume a quiet Run has crashed.

The README's --ligand works on the CLI by accident — use --ligand_description

The upstream README shows --ligand, which only works because argparse

prefix-matches it to the real flag --ligand_description. That shortcut is

CLI-only: as a CSV column header or YAML key, ligand matches nothing and the

row is silently treated as having no ligand. Spell the flag and the column

header out in full.

Wisp execution

Use python only for bounded interactive checks. For a long or GPU-backed

workload, require a selected and probed ssh:<alias> context and load

remote-compute-ssh. Put the documented invocation in a self-contained project

script, activate the remote environment explicitly, stage only small files with

input_paths, and make the command write to a known absolute remote result

path. Submit it with run_in_context and register that exact ssh:// path in

output_specs. Call monitor_run once when waiting is needed, get_run once

for a snapshot, or cancel_run to stop. Do not send a scheduler submission

through the SSH-direct runner.

Errors worth recognizing

| You see | It means / do this |

|---|---|

| ValueError: not allowed to raise maximum limit at startup | setrlimit(NOFILE, 64000) exceeds the sandbox hard limit — sed the constant in inference.py to min(64000, rlimit[1]). |

| Silent SIGKILL a few minutes into the SO(3) precompute | Host RAM exhausted — see the gotcha above. |

| python3: not found | You are on the upstream rbgcsail/diffdock image — that one runs from /home/appuser/DiffDock under micromamba. |


Next: rescore the rank1.sdf poses before ranking ligands against each

other — boltz's affinity head is the in-tree option — since the DiffDock

confidence head alone is not an affinity predictor.

Other skills for the same job

different authors, same section of the catalogue
MCP Builder
by anthropics
vendor ×13

Guide for creating high-quality MCP (Model Context Protocol) servers that enable LLMs to interact with external services through well-designed tools. Use when building MCP servers to integrate external APIs or services, whether in Python (FastMCP) or Node/TypeScript (MCP SDK).

30k tokens scripts
Changelog Generator
by frostant
×9

Automatically creates user-facing changelogs from git commits by analyzing commit history, categorizing changes, and transforming technical commits into clear, customer-friendly release notes. Turns hours of manual changelog writing into minutes of automated generation.

774 tokens
Finishing A Development Branch
by ZhanlinCui
×7

Use when implementation is complete, all tests pass, and you need to decide how to integrate the work - guides completion of development work by presenting structured options for merge, PR, or cleanup

1k tokens
MCP Builder
by JayZeeDesign
×7

Guide for creating high-quality MCP (Model Context Protocol) servers that enable LLMs to interact with external services through well-designed tools. Use when building MCP servers to integrate external APIs or services, whether in Python (FastMCP) or Node/TypeScript (MCP SDK).

37k tokens scripts
Vercel React Native Skills
by vercel-labs
vendor ×6

React Native and Expo best practices for building performant mobile apps. Use when building React Native components, optimizing list performance, implementing animations, or working with native modules. Triggers on tasks involving React Native, Expo, mobile performance, or native platform APIs.

39k tokens
Vercel React Best Practices
by ratacat
×5

React and Next.js performance optimization guidelines from Vercel Engineering. This skill should be used when writing, reviewing, or refactoring React/Next.js code to ensure optimal performance patterns. Triggers on tasks involving React components, Next.js pages, data fetching, bundle optimization, or performance improvements.

34k tokens
Next Best Practices
by vercel-labs
vendor ×4

Next.js best practices - file conventions, RSC boundaries, data patterns, async APIs, metadata, error handling, route handlers, image/font optimization, bundling

20k tokens
Using Git Worktrees
by ZhanlinCui
×4

Use when starting feature work that needs isolation from current workspace or before executing implementation plans - creates isolated git worktrees with smart directory selection and safety verification

1k tokens

How to use it

Copy the folder

Take xuzhougeng/diffdock from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.