mcpbeat Sign in

LLM Council Skill for Claude

Multi-LLM collaborative brainstorming and planning. Use when user explicitly requests consultation with multiple AI models (ChatGPT, Gemini, other LLMs) before presenting an implementation plan, or asks to "consult the council", "ask other models", or "get perspectives from other AIs". Queries external LLM APIs, synthesizes their perspectives, and presents an adapted implementation plan.

4k tokens
context cost
the whole folder, loaded on every use
4
files
ships runnable scripts
0
copies elsewhere
how many repositories repackaged it
394
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/gcpdev/llm-council-skill --skill llm-council

The instruction itself

6 sections, as written by the author

LLM Council

Consult multiple AI models (ChatGPT and Gemini) for their perspectives before presenting implementation plans to users.

Workflow

When user requests consultation with other AI models, use phrases like:

  • "Consult with ChatGPT and Gemini about..."
  • "Ask other AI models what they think about..."
  • "Get perspectives from the council on..."
  • "Consult the LLM council: [your question]"

Process:

  • Query external LLMs: Run scripts/query_llms.py with the user's prompt to get perspectives from both ChatGPT and Gemini
  • Analyze responses: Review what each model suggests, identifying valuable insights, alternative approaches, and potential concerns
  • Synthesize plan: Create an implementation plan that incorporates the best ideas from all three models (Claude's own analysis + ChatGPT + Gemini)
  • Present to user: Show the final plan along with a brief summary of key contributions from each model

Setup Requirements

The skill requires API keys and optional model configuration stored in a .env file in the working directory:

OPENAI_API_KEY=sk-...
GEMINI_API_KEY=...

# Optional: Specify which models to use (defaults shown below)
OPENAI_MODEL=gpt-5-nano
GEMINI_MODEL=gemini-3-flash-preview

Default Models:

  • ChatGPT: gpt-5-nano (fastest, most cost-efficient - $0.05/1M input, $0.40/1M output)
  • Gemini: gemini-3-flash-preview (balanced speed and intelligence)

Upgrade Options for Better Collaboration:

*OpenAI models (ordered by capability and cost):*

  • gpt-5-nano - Fastest, most cost-efficient ($0.05/1M in, $0.40/1M out) - DEFAULT
  • gpt-5-mini - Faster, cost-efficient for well-defined tasks ($0.25/1M in, $2.00/1M out)
  • gpt-5.2 - Best for coding and agentic tasks ($1.75/1M in, $14.00/1M out)
  • gpt-5.2-pro - Smarter, more precise for complex problems ($21.00/1M in, $168.00/1M out)

All models support reasoning tokens, 400K context window, and image input.

*Gemini models (ordered by capability):*

  • gemini-2.5-flash-lite - Ultra-fast, optimized for throughput
  • gemini-2.5-flash - Best price-performance, large-scale processing
  • gemini-3-flash-preview - Balanced speed and frontier intelligence (default)
  • gemini-3-pro-preview - Most intelligent multimodal model, best for complex reasoning

Higher-tier models provide more sophisticated analysis but cost more per API call.

If the .env file doesn't exist or keys are missing, inform the user and provide setup instructions.

Usage Example

User input: "Consult the council: How should I architect a real-time data pipeline for IoT sensors?"

Claude's process:

  • Execute: python3 scripts/query_llms.py "How should I architect a real-time data pipeline for IoT sensors?"
  • Parse JSON responses from ChatGPT and Gemini
  • Analyze their suggestions (e.g., ChatGPT suggests Kafka, Gemini recommends considering edge computing)
  • Synthesize final plan incorporating valuable insights from all models
  • Present the adapted plan to user with attribution

Output Format

Present the final implementation plan naturally, mentioning key insights from other models inline where relevant. For example:

"Based on consultation with ChatGPT and Gemini, here's the recommended architecture:

[Implementation plan with inline references like "ChatGPT highlighted the importance of..." or "Gemini suggested..."]

Key contributions:

  • ChatGPT: [brief summary]
  • Gemini: [brief summary]"

Error Handling

  • If API keys are missing, inform user and provide setup instructions
  • If an API call fails, note which model's perspective is unavailable and proceed with available responses
  • If both APIs fail, inform user and offer to provide Claude's own analysis without external consultation

Other skills for the same job

different authors, same section of the catalogue
AgentDB Learning Plugins
by Microck
×1

Create and train AI learning plugins with AgentDB's 9 reinforcement learning algorithms. Includes Decision Transformer, Q-Learning, SARSA, Actor-Critic, and more. Use when building self-learning agents, implementing RL, or optimizing agent behavior through experience.

3k tokens
Agentdb Learning Plugins
by ComeOnOliver
×1

Create and train AI learning plugins with AgentDB's 9 reinforcement learning algorithms. Includes Decision Transformer, Q-Learning, SARSA, Actor-Critic, and more. Use when building self-learning agents, implementing RL, or optimizing agent behavior through experience.

7k tokens
Langchain Architecture
by wshobson

Design LLM applications using LangChain 1.x and LangGraph for agents, memory, and tool integration. Use when building LangChain applications, implementing AI agents, or creating complex LLM workflows.

5k tokens
Nemoclaw Maintainer Policies
by NVIDIA
vendor

Provide read-only NemoClaw maintainer policy. Use for questions about Issue Type, labels, Project fields, release labels, triage, duplicates, blocked items, and maintainer decisions. Trigger keywords - maintainer policy, workflow policy, project workflow, issue type, labels, label taxonomy, needs labels, project status, blocked issue, duplicate issue, daily release label, release train, triage policy.

25k tokens
HTTP Load Profiler
by zebbern

Run stepped HTTP load tests with ab/wrk, ramping concurrency levels to collect p50/p90/p99 latency, detect performance inflection points, and recommend optimal concurrency. Triggered by requests like 'load test this URL', 'benchmark my API', 'find the max concurrency', or mentions of p99 latency, throughput saturation, or capacity planning.

5k tokens scripts
I4h Workflow Dataset Teleop
by NVIDIA
vendor

Record episodes for an agentic env via teleoperation (keyboard, SO-ARM leader, or VR) into HDF5. Use when the user wants to teleop or record human demos.

5k tokens
Detecting Data Anomalies
by foryourhealth111-pixel

| Investigate outliers, rare events, spikes, and suspicious records in datasets. Use as an explicit anomaly-analysis helper when you want concrete anomaly-detection workflow guidance, not generic data validation or end-to-end ML ownership.

4k tokens scripts
LLM Council
by aiwithremy

Run any question, idea, or decision through a council of 5 AI advisors who independently analyze it, peer-review each other anonymously, and synthesize a final verdict. Based on Karpathy's LLM Council methodology. MANDATORY TRIGGERS: 'council this', 'run the council', 'war room this', 'pressure-test this', 'stress-test this', 'debate this'. STRONG TRIGGERS (use when combined with a real decision or tradeoff): 'should I X or Y', 'which option', 'what would you do', 'is this the right move', 'validate this', 'get multiple perspectives', 'I can't decide', 'I'm torn between'. Do NOT trigger on simple yes/no questions, factual lookups, or casual 'should I' without a meaningful tradeoff (e.g. 'should I use markdown' is not a council question). DO trigger when the user presents a genuine decision with stakes, multiple options, and context that suggests they want it pressure-tested from multiple angles.

5k tokens

How to use it

Copy the folder

Take gcpdev/llm-council from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.