Model evaluation
Evaluate your model on AnesTRACE-Bench
Our inference code, agent framework, and evaluation tools are publicly available. To request evaluation on AnesTRACE-Bench, submit your model's Hugging Face repository or cloud-storage link through the model submission form. The team evaluates fixed-version model weights locally on the private benchmark. Benchmark data are not released due to license issues.
Clinical teams can review the case, protocol and evidence boundaries to assess the study design. The current public route supports model submissions; there is no public benchmark-data request process. Submissions are public GitHub Issues: do not upload case records, patient information, passwords or access tokens.