Claude Code Skills·Claude Skills·The open SKILL.md registry for Claude
ClaudSkills / Engineering / devops

medagentbench-setup

Category: Engineering  ·  Sub-category: devops  ·  Last updated:
lang:pythontool:dockerai:gemini
在 Windows 11 / PowerShell 上從零部署 Stanford MedAgentBench(虛擬 EHR / FHIR 環境下評測 LLM agent)。一條龍流程:clone repo、conda Python 3.9 環境、Docker 拉 HAPI FHIR 模擬病歷資料庫、Stanford Box 下載 refsol.py、設定 OpenAI/Gemini/Claude API key、起 20 個 task worker、跑 assigner 派 100 道題、讀 overall.json。也內建所有實測踩坑(importlib_metadata 缺、agent_test 是互動 REPL、PowerShell 剪貼簿黏連、Docker container 命名陷阱、FHIR 啟動訊號判讀)。觸發詞:「裝 MedAgentBench」「部署 MedAgentBench」「Stanford MedAgent」「FHIR LLM agent benchmark」「跑 medagentbench」「測試醫療 LLM agent」「虛擬 EHR 評測」「HAPI FHIR 模擬病歷」「benchmark medical agent」「跑 LLM on FHIR」「s41591 medagent」。不適用:其他醫療 benchmark(MedQA、MedMCQA — 那是純 QA,不是 agent)、真實 EHR 接入(本 skill 只用模擬資料)、自訓醫療 LLM(本 skill 是評測,不是訓練)。先決條件:Docker Desktop 已裝(沒裝先用 [[install-docker-desktop-windows]])、有 OpenAI/Gemini/Claude API key、C 槽 >= 15 GB。
Security AStatic scan found no risk patternsHow grading works ›

From the source SKILL.md

從 git clone 到 overall.json 出來,Stanford 的 [MedAgentBench](https://github.com/stanfordmlgroup/MedAgentBench) 完整部署流程。NEJM AI 2025 publication 對應 codebase。

What this skill does

medagentbench-setup is a community-contributed Claude Code skill in the devops sub-category. It ships as a SKILL.md file that Claude Code auto-discovers under ~/.claude/skills/medagentbench-skill/ and loads when your prompt matches the skill's trigger.

Who uses this skill

The medagentbench-setup Claude Code skill is built for software engineers, backend developers, full-stack teams, and technical leads building and maintaining production systems. It's part of ClaudSkills (also referred to as Claude Skills or Claude Code Skills) — the open community-curated registry of 161,000+ SKILL.md files for Anthropic's Claude Code agent and the wider Claude ecosystem (Claude API, Claude Agent SDK).

How to install

Free

Manual install (2 steps)

mkdir -p ~/.claude/skills/medagentbench-skill
curl -L https://claudskills.com/skills/medagentbench-skill/SKILL.md \
  -o ~/.claude/skills/medagentbench-skill/SKILL.md

Or just download SKILL.md directly and drop it into ~/.claude/skills/medagentbench-skill/. Claude Code auto-discovers it on next session.

Skills live at ~/.claude/skills/medagentbench-skill/SKILL.md on macOS/Linux, or %USERPROFILE%\.claude\skills\medagentbench-skill\SKILL.md on Windows. See the full install guide for step-by-step instructions.

Telegram

📱 Install from your phone or desktop Telegram

Open @claudskills_bot on Telegram, tap Open Desktop App, and the desktop app installs this skill for you. Or share the bot link with a colleague — they get the same one-tap install. Learn more →

Pro

One-click install via the desktop app

The ClaudSkills desktop app installs any skill directly into ~/.claude/skills/ with one click — no terminal required. Pro starts at $9/mo or $149 lifetime.

Pro

For the full experience including quality scoring and one-click install features for each skill — upgrade to Pro.

Frequently asked questions

How do I install the medagentbench-setup Claude Code skill?
Install via the ClaudSkills desktop app (one click) or copy SKILL.md from the source repository to ~/.claude/skills/medagentbench-skill/SKILL.md and restart Claude Code. Both flows are detailed at claudskills.com/install/.
What does the medagentbench-setup skill do?
在 Windows 11 / PowerShell 上從零部署 Stanford MedAgentBench(虛擬 EHR / FHIR 環境下評測 LLM agent)。一條龍流程:clone repo、conda Python 3.9 環境、Docker 拉 HAPI FHIR 模擬病歷資料庫、Stanford Box 下載 refsol.py、設定 OpenAI/Gemini/Claude API key、起 20 個 task worker、跑 assigner 派 100 道題、讀 overall.json。也內建所有實測踩坑(importlib_metadata 缺、agent_test 是互動 REPL、PowerShell 剪貼簿黏連、Docker container 命名陷阱、FHIR 啟動訊號判讀)。觸發詞:「裝 MedAgentBench」「部署 MedAgentBench」「Stanford MedAgent」「FHIR LLM agent benchmark」「跑 medagentbench」「測試醫療 LLM agent」「虛擬 EHR 評測」「HAPI FHIR 模擬病歷」「benchmark medical agent」「跑 LLM on FHIR」「s41591 medagent」。不適用:其他醫療 benchmark(MedQA、MedMCQA — 那是純 QA,不是 agent)、真實 EHR 接入(本 skill 只用模擬資料)、自訓醫療 LLM(本 skill 是評測,不是訓練)。先決條件:Docker Desktop 已裝(沒裝先用 [[install-docker-desktop-windows]])、有 OpenAI/Gemini/Claude API key、C 槽 >= 15 GB。
Is this skill free to install?
Yes. ClaudSkills is an open registry — every skill keeps its source repository's license, and manual install via copy is free. ClaudSkills Pro ($9/mo, $79/yr, or $149 one-time) adds one-click install via the desktop app and a multi-signal Quality Score.
When should I use the medagentbench-setup skill?
Use medagentbench-setup when your Claude Code task falls under the Engineering category — specifically in the devops area. Claude Code auto-discovers installed skills and invokes the right one based on the task description, so you can also ask Claude directly (e.g. "use medagentbench-setup" or describe the task and let Claude pick). Browse related skills at /category/engineering/.
What is a Claude Code skill and how does the medagentbench-setup skill fit in?
A Claude Code skill is a SKILL.md file that lives under ~/.claude/skills/<name>/ and tells the Claude Code CLI agent how to perform a specific task (instructions, prompts, allowed tools). Skills are auto-discovered at session start. medagentbench-setup is one of 67,000+ skills indexed in the open ClaudSkills catalog, classified under the Engineering category. Learn more at /learn/what-is-a-claude-skill/.

Attribution & license

Cite this skill

If you reference this skill in a blog post, paper, or documentation, you can cite it as:

APA
ckt520728. (2026). medagentbench-setup [Claude Code skill]. ClaudSkills. https://claudskills.com/skills/medagentbench-skill/
BibTeX
@misc{medagentbench-skill-2026,
  author    = {ckt520728},
  title     = {medagentbench-setup [Claude Code skill]},
  year      = {2026},
  publisher = {ClaudSkills},
  url       = {https://claudskills.com/skills/medagentbench-skill/}
}

Embed this skill

Promote, attribute, or link this skill from your own README, blog post, or documentation. All three snippets are free to use — no sign-up, no API key. More distribution surfaces →

Badge
[![ClaudSkills](https://claudskills.com/badge/medagentbench-skill.svg)](https://claudskills.com/skills/medagentbench-skill/?utm_source=badge&utm_medium=readme&utm_campaign=skill_badge)
<script>
<script src="https://claudskills.com/embed/medagentbench-skill.js" async></script>
<iframe>
<iframe src="https://claudskills.com/embed/medagentbench-skill.html" width="100%" height="160" frameborder="0" loading="lazy" title="ClaudSkills: medagentbench-setup"></iframe>

Security scan

Grade A · scanned 2026-07-27 — free static scan against the OWASP Agentic Skills Top 10.

No risk patterns were found in any of the ten OWASP-aligned categories. How grading works ›

Show this grade on your repo (click to copy):

[![Security: A](https://img.shields.io/badge/Security-A-2e7d32)](https://claudskills.com/skills/medagentbench-skill/#security)

Free. No spam. Unsubscribe in one click.

More Engineering skills

Browse all Engineering skills in the ClaudSkills registry, or explore these other picks from the same category:

Browse all Engineering skills → Top 100 skills
Part of ClaudSkills — the open registry for Claude Skills & Claude Code Skills.  ·  What's New  ·  Install guide  ·  About  ·  llms.txt

Part of Acreator Store — Adam Lankamer's AI tools: PerfectStudio · Ucaption · UTagger · AutoXPoster · TestYourSkills · AutomationFlows · Au Naturel · Telegram @acreatorstore