Measuring synthesis quality when doing text de-identification where utility of de-identified data matters. How would you measure the utility of de-identified (also known as synthesized) text?

작성자

카테고리:

← 피드로
r/programming · /u/duke-45 · 2026-07-14 개발(SW)

I'm an AI scientist at Tonic.ai. We open-sourced a benchmark called PrivacyBench and I wanted to share it here because it’s an interesting way to look at and assess LLMs’ ability to de-identify unstructured text data. This matters because a growing number of organizations are…

원문 보기 ↗

코멘트

답글 남기기

이메일 주소는 공개되지 않습니다. 필수 필드는 *로 표시됩니다