AI Mathematics
Why Zero False Positives Still Wasn’t Enough: Lessons from 100+ NER Experiments
What 100+ NER experiments taught us about eliminating false positives on a fixed GitHub technical-term evaluation set without regressing existing performance: hard negatives, focused loss, distillation, residual repair, and class-specific logit adjustment.