nvidia/generate-sandbox-policy
Generate sandbox security policies from plain-language requirements and optional REST API documentation. Produces L4 or fine-grained L7 network policies and ordered network middleware configuration. Use for API access rules, middleware host selection, failure behavior, or built-in and operator-run middleware attachment. Trigger keywords - generate policy, create policy, update policy, change policy, sandbox policy, network policy, API policy, security policy, allow API, restrict API, network middleware, supervisor middleware.
npx skills add https://github.com/NVIDIA/OpenShell --skill generate-sandbox-policy
Generate YAML sandbox network policies and network middleware configuration from API documentation and natural-language user requirements.
This skill translates a user's plain-language policy intent into a valid sandbox policy. The amount of detail the user provides determines the granularity of the generated policy — from broad L4 or preset-based policies (just a host:port) up to fine-grained per-endpoint L7 rules (full API docs).
The output is a network_policies YAML block, an optional network_middlewares block, and optionally a full policy file that conforms to the sandbox policy schema.
The user's input falls into one of three tiers. Work with whatever the user provides — do not require a higher tier than needed.
| Tier | User provides | What you can generate |
|------|--------------|----------------------|
| Minimal | Host(s) and plain-language intent | L4-only policies, or L7 with access presets (read-only, read-write, full) |
| Moderate | Host(s) + some known URL paths or resources | L7 with targeted glob rules for known paths, presets for the rest |
| Full | Complete API docs (OpenAPI, Swagger, markdown, URL) | Fine-grained per-endpoint L7 rules with specific method+path combinations |
The user provides API endpoints and a broad intent. No API docs needed.
Examples:
This is sufficient for:
read-only, read-write, full on all paths)For this tier, default to:
access: read-only when the user says "read", "browse", "view", "query", "fetch"access: read-write when the user says "read-write", "create", "update" (but not "delete")access: full when the user says "full access", "everything", "unrestricted"protocol) when the user says "just allow it", "pass through", "no inspection"The user knows some API paths but doesn't have full docs.
Examples:
Generate explicit rules for the known paths. If the user also wants broader access beyond the specific paths, combine with a catch-all rule or suggest a preset instead.
The user provides full API documentation. Accepted formats:
| Format | How to consume |
|--------|----------------|
| URL | Fetch with WebFetch and parse the endpoint list |
| File path | Read the file (OpenAPI JSON/YAML, markdown, etc.) |
| Pasted text | Parse inline from the conversation |
| OpenAPI/Swagger spec | Extract paths object for all method+path combinations |
From the API docs, build an endpoint inventory — a list of (method, path, description) tuples. Group them logically (e.g., by resource or tag). Then generate precise rules that allow only what the user's intent requires.
Regardless of tier, extract (or infer) these from the user's description:
| Aspect | What to identify | Required? |
|--------|-----------------|-----------|
| Scope | Which API host(s) and port(s) | Yes — always needed |
| Access level | Broad intent: read-only, read-write, full, or custom | Yes — ask if unclear |
| Methods | Specific HTTP methods to allow | Only for custom/fine-grained |
| Paths | Specific URL paths or patterns | Only for custom/fine-grained |
| Enforcement | enforce or audit? Default to enforce. | No — has a default |
| Binary | Which binary/process should have access | Yes — ask if not stated |
| Middleware | Whether admitted HTTP requests need an ordered built-in or operator-run processing stage | No |
If the host and access level are clear but binaries are not specified, ask the user which binary or process will be making the requests. Suggest common defaults like /usr/bin/curl, /usr/local/bin/claude, etc.
Before generating the policy, proactively ask clarifying questions to help the user scope the policy down as narrowly as possible. The goal is the most restrictive policy that still satisfies the user's needs.
Always ask about these if the user hasn't already specified them:
| Missing info | Question to ask |
|-------------|----------------|
| Binary not specified | "Which binary or process will make these requests? (e.g., /usr/bin/curl, /usr/local/bin/claude)" |
| Port not specified | "Which port does this API use? (443 for HTTPS is typical)" |
| Enforcement not stated | "Should policy violations be blocked (enforce) or just logged for review (audit)? I'll default to enforce if you're not sure." |
Ask these when the user's intent is broad and more specificity is possible:
| User says | Ask to narrow |
|-----------|--------------|
| "Full access" / "allow everything" | "Do you actually need DELETE access, or would read-write (everything except DELETE) be enough?" |
| "Allow access to api.example.com" (no method/path detail) | "Do you know which specific API paths or operations you need? If so, I can lock the policy down to just those. Otherwise I'll use a broad preset." |
| L4-only / "just pass it through" | "L4-only means the proxy won't inspect HTTP traffic at all — any method and path will be allowed. Are you sure you don't want at least read-only or read-write restriction?" |
| Wildcard binary (/usr/bin/*) | "A wildcard binary pattern means any binary in that directory can use this policy. Can you narrow it to specific binaries?" |
| Multiple hosts in one policy | "Do all of these hosts need the same access level? If some need tighter restrictions, I can split them into separate policies." |
| access: full with enforcement: audit | "Full access in audit mode means nothing is actually restricted — all traffic flows through and violations are only logged. Is that intentional, or did you want to enforce restrictions?" |
| path glob on all rules | "Using on all paths allows any URL path. Do you know the specific API path prefixes you need (e.g., /api/v1/)?" |
| Private/internal IP destination | "Does this service resolve to a private IP (10.x, 172.16.x, 192.168.x)? If so, you'll need allowed_ips to permit access — what CIDR range should be allowed?" |
When the user mentions a recognizable API host but hasn't provided docs, and the current tier is Minimal, attempt to upgrade to Full by searching for the API documentation online.
When to trigger:
api.github.com, api.anthropic.com, api.openai.com, integrate.api.nvidia.com, api.stripe.com, api.slack.com, api.gitlab.com)How to do it:
WebSearch with a query like "[service name] REST API documentation endpoints" or "[service name] OpenAPI spec"WebFetch and extract the endpoint inventory (method + path pairs)When to skip:
Graceful fallback: If the search doesn't return usable API docs (results are irrelevant, docs are behind authentication, the page is too large to parse), fall back to the current tier without stalling. Say: "I couldn't find usable API docs for [host], so I'll generate the policy using a [preset/L4] approach. You can always provide docs later to tighten it."
If the user confirms the policy must stay broad (they don't know the paths, need genuinely broad access, etc.), accept it but flag the breadth. Do not block policy generation — just make sure the warnings are visible in the output (see Step 6).
You may need to go back and forth a few times. Keep the loop tight:
Do not over-interrogate. If the user has given a clear, specific request, skip clarification and go straight to generation. Only ask when there is genuine ambiguity or an opportunity to meaningfully reduce the attack surface.
Read the full policy schema reference:
Read docs/reference/policy-schema.mdx
Key sections to reference:
network_policies — rule structureNetworkEndpoint fields — host, port, protocol, tls, enforcement, access, rules, allowed_ipsL7Rule / L7Allow — method + path matchingread-only, read-write, fullallowed_ips — CIDR allowlist for private IP spaceWhen middleware is requested, also read the full operational reference:
Read docs/extensibility/supervisor-middleware.mdx
Also read the architecture overview for enforcement context. The default policy is baked into the community base image (ghcr.io/nvidia/openshell-community/sandboxes/base:latest). For reference, consult:
Read architecture/security-policy.md
Follow this decision tree based on the detail tier and user intent:
Is L7 inspection needed?
├─ No (user wants pass-through / "just allow it")
│ └─ Generate L4-only policy (no protocol, no tls, no rules/access)
│
└─ Yes (user wants method/path control)
│
├─ Does a preset match the intent exactly?
│ ├─ Read-only (GET, HEAD, OPTIONS) → access: read-only
│ ├─ Read-write (no DELETE) → access: read-write
│ └─ Everything → access: full
│
└─ No preset fits (specific paths, mixed broad+narrow, exclude certain paths)
└─ Build explicit rules list
└─ Requires either known paths from the user or full API docs
Principle: always choose the simplest representation that satisfies the intent. A preset is preferable to explicit rules when it covers the use case.
| API host port | TLS setting |
|--------------|-------------|
| Port 443 (HTTPS) and L7 rules/preset needed | tls: terminate (required for inspection) |
| Port 443 (HTTPS) and L4-only | Omit tls (passthrough, no L7) |
| Non-443 (HTTP) | Omit tls |
Critical: protocol: rest on port 443 without tls: terminate will not work — the proxy cannot inspect encrypted traffic. Always set tls: terminate when combining port 443 with L7 rules.
Add network_middlewares only when the user asks to inspect, transform, redact, or independently authorize admitted HTTP requests. Middleware runs after network and L7 policy admission and before provider credential injection.
openshell/regex without gateway registration.[[openshell.supervisor.middleware]] and reachable from both the gateway and sandbox supervisors.on_error to fail_closed. Use fail_open only when bypassing the stage preserves the user's stated security requirement.order values across the complete policy. Lower values run first, and at most 10 configs may be selected.endpoints.include; use exclude when a broad selector has trusted exceptions.tls: skip endpoints because the supervisor cannot inspect that traffic.Only needed for the Moderate and Full tiers. Translate API path parameters to glob patterns:
| API path | Glob pattern |
|----------|-------------|
| /repos/{owner}/{repo} | /repos/*/* |
| /repos/{owner}/{repo}/issues | /repos/*/issues |
| /repos/{owner}/{repo}/issues/{id} | /repos/*/issues/* |
| /api/v1/models/{model_id}/versions/{version} | /api/v1/models/*/versions/* |
| All sub-paths under /api/v1/ | /api/v1/** |
Path matching uses the runtime glob engine. Both * and ** may cross /
boundaries; ? matches one character, and bracket classes such as [0-9] and
[!0] are supported. Prefer segment-shaped patterns such as
/repos/*/issues for readability, but do not rely on * to stop at /.
For each allowed operation, create an allow entry:
rules:
- allow:
method: GET
path: "/api/v1/models/*"
- allow:
method: POST
path: "/api/v1/completions"
Use the most specific pattern that covers the intent. Prefer narrow globs over ** when the API structure is known.
Generate a complete network_policies entry. Use this template:
network_policies:
<policy_key>:
name: <policy_key>
endpoints:
- host: <api_host>
port: <port>
protocol: rest # Required for L7 inspection
tls: terminate # Required for HTTPS + L7
enforcement: enforce # or audit
# Use ONE of: access OR rules (never both)
access: <preset> # read-only | read-write | full
# OR
rules:
- allow:
method: <METHOD>
path: "<glob_pattern>"
# Optional: allow private IP destinations (CIDR or exact IP)
# allowed_ips:
# - "10.0.5.0/24"
binaries:
- { path: <binary_path> }
When middleware is requested, add it as a separate top-level map rather than nesting it under a network policy:
network_middlewares:
<config_key>:
name: <human_readable_name>
middleware: <built_in_or_registered_name>
order: 10
config: {}
on_error: fail_closed
endpoints:
include: ["<api_host>"]
# exclude: ["<trusted_host>"]
The map key is the stable policy-local identity. Middleware selection is independent of the network policy entry that admitted the request.
Use deny_rules to block specific dangerous operations while allowing broad access. Deny rules are evaluated after allow rules and take precedence. This is the inverse of the rules approach — instead of enumerating every allowed operation, you grant broad access and block a small set of dangerous ones.
# Example: Allow full access to GitHub but block admin operations
github_api:
name: github_api
endpoints:
- host: api.github.com
port: 443
protocol: rest
enforcement: enforce
access: read-write
deny_rules:
- method: POST
path: "/repos/*/pulls/*/reviews"
- method: PUT
path: "/repos/*/branches/*/protection"
- method: "*"
path: "/repos/*/rulesets"
binaries:
- { path: /usr/bin/curl }
Deny rules support the same matching capabilities as allow rules: method, path, command (SQL), and query parameter matchers. When generating policies, prefer deny rules when the user needs broad access with a small set of blocked operations — it produces a shorter, more maintainable policy than enumerating 60+ allow rules.
When the endpoint resolves to a private IP (RFC 1918), the proxy's SSRF protection blocks the connection by default. Use allowed_ips to selectively allow specific private IP ranges:
host + allowed_ips — domain must resolve to an IP in the allowlistallowed_ips only (no host) — any domain on the port is allowed if it resolves to an IP in the allowlistLoopback (127.0.0.0/8) and link-local (169.254.0.0/16) are always blocked regardless of allowed_ips.
# Example: Allow access to internal service at a known private IP range
internal_api:
name: internal_api
endpoints:
- host: api.internal.corp
port: 8080
allowed_ips:
- "10.0.5.0/24"
binaries:
- { path: /usr/bin/curl }
Use descriptive snake_case keys: github_api, nvidia_inference, internal_service_readonly.
If the user needs access to multiple hosts or the same host with different rules, either:
endpoints if the binary set is the sameBefore presenting the policy to the user, verify correctness and flag breadth concerns.
rules and access are NOT both present on the same endpointprotocol is set, either rules or access is also presenttls: terminate is set, protocol is also setrules list is not empty when presentprotocol: sql, enforcement is not enforcemiddleware name and non-empty endpoints.includeorder values are unique and no selected chain exceeds 10 stagestls: skip endpointprotocol: rest on port 443 should have tls: terminate*name, endpoints, and binarieshost and portpathname fieldinclude and exclude patternsEvaluate the generated policy for overly broad access and include warnings in the output to the user. These do not block generation, but the user must see them.
| Condition | Warning to show |
|-----------|----------------|
| L4-only (no protocol) | "This policy allows all HTTP methods and paths without inspection. The proxy will only check host:port and binary identity. Consider adding protocol: rest with a preset if you want method-level control." |
| access: full | "This policy allows all HTTP methods (including DELETE) on all paths. If you don't need DELETE, read-write is safer. If you only need to read, read-only is the most restrictive option." |
| access: full + enforcement: audit | "Full access in audit mode provides no actual restriction — all traffic flows through. This is effectively a monitoring-only policy." |
| access: read-write when user hasn't confirmed write need | "This policy allows POST, PUT, and PATCH on all paths. If you only need to read data, read-only is more restrictive." |
| Wildcard binary (* or ** in binary path) | "This policy allows any binary matching the glob pattern. A compromised or unexpected binary in that directory could use this policy. Consider listing specific binary paths." |
| path glob on all explicit rules | "All rules use path patterns, which match any URL path. This is equivalent to a preset — consider using access: read-only (or similar) for clarity, or narrowing paths if you know the API structure." |
| Multiple broad endpoints in one policy | "This policy grants the same broad access to N different hosts. If any of these hosts needs tighter restrictions later, you'll need to split the policy." |
| Hostless allowed_ips (no host field) | "This endpoint has no host — any domain resolving to the allowed IP range on this port will be permitted. Consider adding a host field to restrict which domains can use this allowlist." |
| Broad CIDR in allowed_ips (e.g., 10.0.0.0/8) | "This allowed_ips entry covers a very broad range. Consider narrowing to a specific subnet (e.g., 10.0.5.0/24) to minimize exposure." |
| on_error: fail_open | "This middleware can be bypassed when it is unavailable, rejects configuration, returns an invalid result, or exceeds its body limit. Use fail_closed unless availability is more important than this control." |
| Broad middleware host selector | "This middleware applies independently of the admitting network rule to every matching HTTP destination. Narrow endpoints.include or add exclusions if the stage is not required for every matching host." |
Format breadth warnings clearly in the output, e.g.:
⚠️ Breadth warning: This policy uses `access: full`, which allows all HTTP
methods (including DELETE) on all paths. If you don't need DELETE, consider
using `read-write` instead.
If there are no breadth warnings, say so explicitly: "No breadth concerns — this policy is well-scoped."
The policy needs to go somewhere. Determine which mode applies:
| Signal | Mode |
|--------|------|
| User names an existing policy file (e.g., "add to my-sandbox-policy.yaml") | Update existing file |
| User says "update my policy", "add this to my policy file" | Update existing file — ask which file to update |
| User asks to modify an existing policy rule by name | Update existing file — edit the named policy in place |
| User says "create a new policy file" or names a file that doesn't exist | Create new file |
| No file context given | Present only — show the YAML and ask if the user wants it written to a file |
network_policiesfilesystem_policy, landlock, and process sections look like{ host: ..., port: ... }) or expanded YAML stylenetwork_policies, maintaining the file's existing indentation and style.filesystem_policy, landlock, process, or other policies unless the user explicitly asks.Generate a complete, standalone policy file. Use the full schema scaffolding:
version: 1
filesystem_policy:
include_workdir: true
read_only:
- /usr
- /lib
- /proc
- /dev/urandom
- /app
- /etc
- /var/log
read_write:
- /sandbox
- /tmp
- /dev/null
landlock:
compatibility: best_effort
network_policies:
# <generated policies go here>
The filesystem_policy and landlock sections above are sensible defaults.
Process identity is omitted so the selected compute driver can choose it. For
Docker and Podman, each omitted identity field falls back to the image's OCI
USER. Tell the user these are defaults and may need adjustment for their
environment. Gateway inference is configured separately through `openshell
inference set/get. The generated network_policies` block is the primary
output.
If the user provides a file path, write to it. Otherwise, ask where to place it. A common convention is a project-local policy file (e.g., sandbox-policy.yaml) passed to openshell sandbox create --policy <path> or set via the OPENSHELL_SANDBOX_POLICY env var.
Show the generated policy YAML with:
network_policies block, ready to pasteopenshell sandbox create --policy <path> or set OPENSHELL_SANDBOX_POLICY=<path>After presenting or applying the policy, ask if the user wants to:
my_api:
name: my_api
endpoints:
- { host: api.example.com, port: 443 }
binaries:
- { path: /usr/bin/curl }
my_api_readonly:
name: my_api_readonly
endpoints:
- host: api.example.com
port: 443
protocol: rest
tls: terminate
enforcement: enforce
access: read-only
binaries:
- { path: /usr/bin/curl }
my_api_custom:
name: my_api_custom
endpoints:
- host: api.example.com
port: 443
protocol: rest
tls: terminate
enforcement: enforce
rules:
- allow:
method: GET
path: "/api/v1/**"
- allow:
method: POST
path: "/api/v1/data"
binaries:
- { path: /usr/bin/curl }
- { path: /usr/local/bin/myapp }
internal_svc:
name: internal_svc
endpoints:
- host: api.internal.svc
port: 8080
protocol: rest
enforcement: enforce
rules:
- allow:
method: GET
path: "/health"
- allow:
method: POST
path: "/api/v1/jobs"
binaries:
- { path: /usr/bin/curl }
internal_db:
name: internal_db
endpoints:
- host: db.internal.corp
port: 5432
allowed_ips:
- "10.0.5.0/24"
binaries:
- { path: /usr/bin/curl }
private_services:
name: private_services
endpoints:
- port: 8080
allowed_ips:
- "10.0.5.0/24"
- "10.0.6.0/24"
binaries:
- { path: /usr/bin/curl }
ghcr.io/nvidia/openshell-community/sandboxes/base:latest)Take nvidia/generate-sandbox-policy from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.