OpenMedOpen weights model0.6B parameters
OpenMed QwenMed XL benchmarks
- General rankToo few tasks3 non-coding tasks ranked5 needed for a rank
- CostSelf-hostedSmall GPU, ~2.9 GBNo hosted price
- Multilingual#29 / 4372.6%Best task · PII left invs PII detectors
Comparison summary
- Strongest
- Safety and guardrails
- Leads 0 of 3 (0%)
Strongest at safety. It is ranked on 3 non-coding tasks, too few for a general rank (5 needed).
Specifications
- Lab
- OpenMed
- Weights
- Open weights
- Parameters
- 0.6B
- Licence
- Apache-2.0
- Type
- encoder-NER
Safety and guardrails
Share of tasks led · #33 of 46
- Gemma 4 31BLeads 3 of 4
- Perplexity PII TracerLeads 2 of 4
- Amazon ComprehendLeads 1 of 4
- Qwen3.8 27BLeads 1 of 5
- Gemma 4 12BLeads 0 of 4
- Ministral 3 14BLeads 0 of 5
- Gemma 4 26B A4BLeads 0 of 4
- OpenMed QwenMed XLLeads 0 of 3
See the 5 tasks
Traces
Tier 12 of 19#31 / 46PII left in, % · Lower is better
- Amazon Comprehend10.6, interval 9.4 to 11.9, ahead alone
- GLiNER Multi PII v115.3, interval 13.2 to 17.3, tier 2
- NVIDIA GLiNER PII19.7, interval 17.4 to 22.0, tier 3
- Gemma 4 26B A4B19.8, interval 15.8 to 24.2, tier 3
- Qwen3.8 27B20.3, interval 16.5 to 24.3, tier 3
- GLiNER2 Large20.4, interval 18.0 to 22.8, tier 3
- Gemma 4 12B23.3, interval 19.0 to 27.6, tier 3
- OpenMed QwenMed XL70.7, interval 68.4 to 73.0, tier 12
Business
Tier 12 of 18#31 / 46PII left in, % · Lower is better
- Perplexity PII Tracer2.1, interval 1.2 to 3.0, tied for the lead
- Qwen3.8 27B2.4, interval 1.7 to 3.3, tied for the lead
- Gemma 4 31B2.9, interval 1.9 to 4.0, tied for the lead
- GLiNER Multi PII v14.1, interval 3.1 to 5.3, tier 2
- Ministral 3 14B5.4, interval 4.0 to 6.9, tier 2
- Gemma 4 12B5.8, interval 4.4 to 7.3, tier 3
- Amazon Comprehend5.9, interval 4.6 to 7.2, tier 3
- OpenMed QwenMed XL54.7, interval 52.0 to 57.5, tier 12
Multilingual
Tier 19 of 25#29 / 43PII left in, % · Lower is better
- Perplexity PII Tracer2.9, interval 2.5 to 3.4, ahead alone
- Qwen3.8 27B4.2, interval 3.7 to 4.8, tier 2
- Gemma 4 31B4.5, interval 3.9 to 5.1, tier 2
- GLiNER Multi PII v15.7, interval 5.2 to 6.3, tier 3
- bardsai EU PII Multilang6.9, interval 6.3 to 7.5, tier 4
- Qwen3.6 35B A3B7.7, interval 7.0 to 8.4, tier 4
- Gemma 4 12B8.9, interval 8.1 to 9.7, tier 5
- OpenMed QwenMed XL72.6, interval 71.4 to 73.8, tier 19
Look-alikes
Not rankedLook-alikes redacted, % · Lower is better
- OpenMed QwenMed XL0.4, interval 0.1 to 0.8, not ranked
- Gemma 4 31B3.5, interval 2.1 to 5.1, ahead alone
- Qwen3.8 27B10.3, interval 7.9 to 12.9, tier 2
- NuExtract 319.8, interval 15.9 to 23.6, tier 3
- Qwen3.5 9B20.5, interval 16.8 to 24.3, tier 3
- Ministral 3 14B30.5, interval 26.5 to 34.8, tier 4
- NVIDIA GLiNER PII42.0, interval 38.2 to 46.1, tier 5
- Gemma 3 4B IT45.0, interval 41.0 to 49.0, tier 5
LangWatch policy
Not rankedPII left in, % · Lower is better
- Gemma 4 31B6.9, interval 2.8 to 11.3, ahead alone
- Ministral 3 14B10.2, interval 5.8 to 15.4, tier 2
- Gemma 4 12B16.5, interval 10.3 to 23.2, tier 3
- Qwen3.8 27B16.7, interval 9.8 to 25.0, tier 3
- Gemma 4 26B A4B20.5, interval 13.3 to 28.8, tier 4
- Qwen3.5 9B21.5, interval 13.4 to 30.6, tier 4
- Gemma 3 4B IT22.4, interval 14.3 to 32.1, tier 4
- OpenMed QwenMed XL84.1, interval 78.7 to 89.4, not ranked
Cost
Self-hosted · GPU memory needed, among encoder-NERs · lower is better
- Gravitee BERT Small PII~0.6 GB
- Secret Masker v3.3a~0.8 GB
- Anonym-IA CamemBERT PII FR~0.9 GB
- BERT base NER (dslim)~0.9 GB
- Veil PII KO Lite~0.9 GB
- bardsai EU PII Multilang~1.6 GB
- XLM-RoBERTa NER HRL~1.6 GB
- OpenMed QwenMed XL~2.9 GB
How we count
A model leads a task when it is in the task's leading tie tier. Each task counts inside its own benchmark, against that benchmark's models; no score is averaged. An area chart shows the models ranked on at least 2 of the area's tasks in a benchmark this model is in too. The overall standing counts every non-coding tasks and needs 5 for a rank. Latency is compared only on one hardware tier.
OpenMed QwenMed XLOther modelsWhisker: 95% interval
