FHIR-AgentBench

AI benchmark · Healthcare

Publisher
Paper Arxiv
Domains
Healthcare
Data source
benchmarklist
Snapshot

Open benchmark ↗benchmarklist.com/benchmarks/fhir_agentbench/

About FHIR-AgentBench

FHIR-AgentBench evaluates LLM agents on realistic interoperable EHR question answering over HL7 FHIR resources. It grounds 2,931 real-world clinical questions in FHIR and compares retrieval strategies, interaction patterns, and reasoning approaches such as direct FHIR API calls, specialized tools, single-turn versus multi-turn interaction, and natural-language versus code-generation reasoning.

Where FHIR-AgentBench sits

Its domains, and the nearest entries sharing them. Click any node to open its page.

Select a node to trace its connections. Zoom in for more space; drag to pan.

Next: browse all benchmarks. Entry from the RL Research daily scrape of public sources, 2026-09-21 snapshot.