Pith. sign in

Ai hospital: Benchmarking large language models in a multi-agent medical interaction simulator

8 Pith papers cite this work. Polarity classification is still indexing.

8 Pith papers citing it

citation-role summary

background 2

citation-polarity summary

years

2026 2 2025 6

roles

background 2

polarities

background 2

representative citing papers

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit

cs.MA · 2025-07-13 · conditional · novelty 6.0

The paper introduces TinyTroupe, an open-source LLM-powered multiagent persona simulation library supporting detailed persona definitions, population sampling, experimentation, and validation, with preliminary evidence that it can approximate some aspects of real consumer behavior.

R2MED: A Benchmark for Reasoning-Driven Medical Retrieval

cs.IR · 2025-05-20 · accept · novelty 6.0

R2MED is the first benchmark for reasoning-driven medical retrieval, where even top models reach only 41.4 nDCG@10 on queries requiring inference beyond lexical or semantic overlap.

A Survey of Scaling in Large Language Model Reasoning

cs.AI · 2025-04-02 · unverdicted · novelty 3.0

A survey categorizing scaling in LLM reasoning across input size, steps, rounds, training, and future directions, noting that scaling can negatively affect performance.

citing papers explorer

Showing 8 of 8 citing papers.