medagentbench-setup
在 Windows 11 / PowerShell 上從零部署 Stanford MedAgentBench(虛擬 EHR / FHIR 環境下評測 LLM agent)。一條龍流程:clone repo、conda Python 3.9 環境、Docker 拉 HAPI FHIR 模擬病歷資料庫、Stanford Box 下載 refsol.py、設定 OpenAI/Gemini/Claude API key、起 20 個 task worker、跑 assigner 派 100 道題、讀 overall.json。也內建所有實測踩坑(importlib_metadata 缺、agent_test 是互動 REPL、PowerShell 剪貼簿黏連、Docker container 命名陷阱、FHIR 啟動訊號判讀)。觸發詞:「裝 MedAgentBench」「部署 MedAgentBench」「Stanford MedAgent」「FHIR LLM agent benchmark」「跑 medagentbench」「測試醫療 LLM agent」「虛擬 EHR 評測」「HAPI FHIR 模擬病歷」「benchmark medical agent」「跑 LLM on FHIR」「s41591 medagent」。不適用:其他醫療 benchmark(MedQA、MedMCQA — 那是純 QA,不是 agent)、真實 EHR 接入(本 skill 只用模擬資料)、自訓醫療 LLM(本 skill 是評測,不是訓練)。先決條件:Docker Desktop 已裝(沒裝先用 [[install-docker-desktop-windows]])、有 OpenAI/Gemini/Claude API key、C 槽 >= 15 GB。
Security AStatic scan found no risk patternsHow grading works ›
From the source SKILL.md
從 git clone 到 overall.json 出來,Stanford 的 [MedAgentBench](https://github.com/stanfordmlgroup/MedAgentBench) 完整部署流程。NEJM AI 2025 publication 對應 codebase。
What this skill does
medagentbench-setup is a community-contributed Claude Code skill in the devops sub-category. It ships as a SKILL.md file that Claude Code auto-discovers under ~/.claude/skills/medagentbench-skill/ and loads when your prompt matches the skill's trigger.
Who uses this skill
The medagentbench-setup Claude Code skill is built for software engineers, backend developers, full-stack teams, and technical leads building and maintaining production systems. It's part of ClaudSkills (also referred to as Claude Skills or Claude Code Skills) — the open community-curated registry of 161,000+ SKILL.md files for Anthropic's Claude Code agent and the wider Claude ecosystem (Claude API, Claude Agent SDK).
How to install
Free
Manual install (2 steps)
mkdir -p ~/.claude/skills/medagentbench-skill
curl -L https://claudskills.com/skills/medagentbench-skill/SKILL.md \
-o ~/.claude/skills/medagentbench-skill/SKILL.md
Or just download SKILL.md directly and drop it into ~/.claude/skills/medagentbench-skill/. Claude Code auto-discovers it on next session.
Skills live at ~/.claude/skills/medagentbench-skill/SKILL.md on macOS/Linux, or %USERPROFILE%\.claude\skills\medagentbench-skill\SKILL.md on Windows. See the full install guide for step-by-step instructions.
Telegram
📱 Install from your phone or desktop Telegram
Open @claudskills_bot on Telegram, tap Open Desktop App, and the desktop app installs this skill for you. Or share the bot link with a colleague — they get the same one-tap install. Learn more →
Pro
One-click install via the desktop app
The ClaudSkills desktop app installs any skill directly into ~/.claude/skills/ with one click — no terminal required. Pro starts at $9/mo or $149 lifetime.
Pro
For the full experience including quality scoring and one-click install features for each skill — upgrade to Pro.
Frequently asked questions
How do I install the medagentbench-setup Claude Code skill?
Install via the ClaudSkills desktop app (one click) or copy
SKILL.md from the source repository to
~/.claude/skills/medagentbench-skill/SKILL.md and restart Claude Code. Both flows are detailed at
claudskills.com/install/.
What does the medagentbench-setup skill do?
在 Windows 11 / PowerShell 上從零部署 Stanford MedAgentBench(虛擬 EHR / FHIR 環境下評測 LLM agent)。一條龍流程:clone repo、conda Python 3.9 環境、Docker 拉 HAPI FHIR 模擬病歷資料庫、Stanford Box 下載 refsol.py、設定 OpenAI/Gemini/Claude API key、起 20 個 task worker、跑 assigner 派 100 道題、讀 overall.json。也內建所有實測踩坑(importlib_metadata 缺、agent_test 是互動 REPL、PowerShell 剪貼簿黏連、Docker container 命名陷阱、FHIR 啟動訊號判讀)。觸發詞:「裝 MedAgentBench」「部署 MedAgentBench」「Stanford MedAgent」「FHIR LLM agent benchmark」「跑 medagentbench」「測試醫療 LLM agent」「虛擬 EHR 評測」「HAPI FHIR 模擬病歷」「benchmark medical agent」「跑 LLM on FHIR」「s41591 medagent」。不適用:其他醫療 benchmark(MedQA、MedMCQA — 那是純 QA,不是 agent)、真實 EHR 接入(本 skill 只用模擬資料)、自訓醫療 LLM(本 skill 是評測,不是訓練)。先決條件:Docker Desktop 已裝(沒裝先用 [[install-docker-desktop-windows]])、有 OpenAI/Gemini/Claude API key、C 槽 >= 15 GB。
Is this skill free to install?
Yes. ClaudSkills is an open registry — every skill keeps its source repository's license, and manual install via copy is free. ClaudSkills Pro ($9/mo, $79/yr, or $149 one-time) adds one-click install via the desktop app and a multi-signal Quality Score.
When should I use the medagentbench-setup skill?
Use medagentbench-setup when your Claude Code task falls under the Engineering category — specifically in the devops area. Claude Code auto-discovers installed skills and invokes the right one based on the task description, so you can also ask Claude directly (e.g. "use medagentbench-setup" or describe the task and let Claude pick). Browse related skills at
/category/engineering/.
What is a Claude Code skill and how does the medagentbench-setup skill fit in?
A Claude Code skill is a
SKILL.md file that lives under
~/.claude/skills/<name>/ and tells the Claude Code CLI agent how to perform a specific task (instructions, prompts, allowed tools). Skills are auto-discovered at session start. medagentbench-setup is one of 67,000+ skills indexed in the open ClaudSkills catalog, classified under the Engineering category. Learn more at
/learn/what-is-a-claude-skill/.
Attribution & license
Cite this skill
If you reference this skill in a blog post, paper, or documentation, you can cite it as:
APA
ckt520728. (2026). medagentbench-setup [Claude Code skill]. ClaudSkills. https://claudskills.com/skills/medagentbench-skill/
BibTeX
@misc{medagentbench-skill-2026,
author = {ckt520728},
title = {medagentbench-setup [Claude Code skill]},
year = {2026},
publisher = {ClaudSkills},
url = {https://claudskills.com/skills/medagentbench-skill/}
}
Embed this skill
Promote, attribute, or link this skill from your own README, blog post, or documentation. All three snippets are free to use — no sign-up, no API key. More distribution surfaces →
Badge
[](https://claudskills.com/skills/medagentbench-skill/?utm_source=badge&utm_medium=readme&utm_campaign=skill_badge)
<script>
<script src="https://claudskills.com/embed/medagentbench-skill.js" async></script>
<iframe>
<iframe src="https://claudskills.com/embed/medagentbench-skill.html" width="100%" height="160" frameborder="0" loading="lazy" title="ClaudSkills: medagentbench-setup"></iframe>
Security scan
Grade A · scanned 2026-07-27 — free static scan against the OWASP Agentic Skills Top 10.
No risk patterns were found in any of the ten OWASP-aligned categories. How grading works ›
- ✓ Prompt injection
- ✓ Data exfiltration
- ✓ Supply chain
- ✓ Reverse shell
- ✓ Credentials
- ✓ Execution
- ✓ Filesystem
- ✓ Persistence
- ✓ Obfuscation
- ✓ Network
Show this grade on your repo (click to copy):
[](https://claudskills.com/skills/medagentbench-skill/#security)
More Engineering skills
Browse all Engineering skills in the ClaudSkills registry, or explore these other picks from the same category:
Part of Acreator Store — Adam Lankamer's AI tools:
PerfectStudio ·
Ucaption ·
UTagger ·
AutoXPoster ·
TestYourSkills ·
AutomationFlows ·
Au Naturel ·
Telegram @acreatorstore