Skip to content
r/LocalLLaMA

OpenAI's Privacy Filter vs GLiNER on 600 PII Samples

OpenAI's Privacy Filter vs GLiNER on 600 PII samples

A Reddit user compared openai/privacy-filter and GLiNER large-v2.1 on 600 PII samples. On CPU, OpenAI's model ran 2.8 samples/s versus 1.1 for GLiNER; English boundary macro F1 was 0.498 versus 0.416. The key issue is tokenizer offset: strict matching drops openai/privacy-filter to 0.155.

Why it matters: HKR-H/K/R all pass: the Reddit test has a clear matchup, 600 PII samples, speed/F1 numbers, and a tokenizer-offset caveat. Source authority is limited, so it stays in the low featured band.

Read the original ↗Export Markdown