FHIR-AgentBench
AI benchmark · Healthcare
- Publisher
- Paper Arxiv
- Domains
- Healthcare
- Data source
- benchmarklist
- Snapshot
About FHIR-AgentBench
FHIR-AgentBench evaluates LLM agents on realistic interoperable EHR question answering over HL7 FHIR resources. It grounds 2,931 real-world clinical questions in FHIR and compares retrieval strategies, interaction patterns, and reasoning approaches such as direct FHIR API calls, specialized tools, single-turn versus multi-turn interaction, and natural-language versus code-generation reasoning.
Where FHIR-AgentBench sits
Its domains, and the nearest entries sharing them. Click any node to open its page.
Select a node to trace its connections. Zoom in for more space; drag to pan.
Next: browse all benchmarks. Entry from the RL Research daily scrape of public sources, 2026-09-21 snapshot.