ModelCap

Current AI model comparison

Exact identities · public evidence

Allen Institute

Llama 3.1 Tulu 3 70B

Position

#200

Index

22.0

Measured

1 public benchmark observation across 1 board

Comparison side 1

Microsoft

MediPhi PubMed

Position

#197

Index

22.9

Inherited

finetune of microsoft/Phi-3.5-mini-instruct · finetune (100%)

Comparison side 2

Output / 1M

Unavailable

vs Unavailable

Context

vs

Weights

Restricted license

vs Open weights