每天早上一封邮件,把昨天的 AI 梳理好订阅邮件

METAL LAB

When Do LLMs Actually Help? Evaluating LLMs as Data Quality Annotators

arXiv:2608.181582026-08-20

LLMs have been increasingly used to catch data quality issues automatically, but we know very little about how consistent these judgments actually are. This study tests an LLM on two e-commerce data quality tasks, entity matching and brand mislabeling, against rule based baselines and human verified ground truth, under both zero-shot and few-shot prompting. On entity matching while using the Abt Buy benchmark (2,194 labeled pairs), a simple rule ba

作者 · Praphulla Lal Shrestha

在 arXiv 阅读

最新论文

全部论文 →

METAL LAB 最新报道