@MaziyarPanahi: We got 755 tokens per second! That's OpenMed privacy-filter v2 (nemotron, MLX 8-bit) reading a 13,000-token clinical fi…

X AI KOLs Timeline Models

Summary

OpenMed privacy-filter v2 using nemotron and MLX 8-bit achieves 755 tokens per second on a Mac, redacting 1,152 PII identifiers across 22 categories from a 13,000-token clinical file without data leaving the machine.

We got 755 tokens per second! 🔥 That's OpenMed privacy-filter v2 (nemotron, MLX 8-bit) reading a 13,000-token clinical file and redacting every identifier as it streams by. 1,152 caught across 22 PII categories, on a Mac. Nothing left the machine. https://t.co/SeoCvSIiGA
Original Article
View Cached Full Text

Cached at: 07/04/26, 08:53 PM

We got 755 tokens per second! 🔥

That’s OpenMed privacy-filter v2 (nemotron, MLX 8-bit) reading a 13,000-token clinical file and redacting every identifier as it streams by. 1,152 caught across 22 PII categories, on a Mac. Nothing left the machine. https://t.co/SeoCvSIiGA

Similar Articles

maziyarpanahi/openmed

GitHub Trending (daily)

OpenMed is an open-source local-first healthcare AI toolkit that provides entity extraction, PII de-identification, and over 1,000 specialized medical models, all running on-device with no cloud dependency.