| Both sides previous revisionPrevious revision | |
| design:ownership_resolution [2026/09/11 22:40] – Generic review pass: publish both metrics for the pairwise tests (current+granularity keeps Disconnect's lead, current-only does not), correct 'all four errors are new' to three, state the km0trk/i.ua scoring rule, publish the same-bar and i.ua sensitivit karel.kubicek.claude | design:ownership_resolution [2026/09/11 22:42] (current) – Site-wide sweep: the methodology bullet's second copy of '29 of the 525' hand-changed verdicts was still stale after round four fixed the first. Authored by Claude karel.kubicek.claude |
|---|
| * The coverage denominator is **Tracker Radar's own crawl output**. There is no neutral census of third-party domains, so this favours Tracker Radar; the page says so wherever a coverage figure appears rather than pretending otherwise. | * The coverage denominator is **Tracker Radar's own crawl output**. There is no neutral census of third-party domains, so this favours Tracker Radar; the page says so wherever a coverage figure appears rather than pretending otherwise. |
| * The 28 adjudicated rows are the highest-prevalence **disagreements**, deliberately not a random sample, so they characterise the shape of disagreement and not any list's accuracy. | * The 28 adjudicated rows are the highest-prevalence **disagreements**, deliberately not a random sample, so they characterise the shape of disagreement and not any list's accuracy. |
| * The **random sample** is 60 domains per list from that list's own coverage, stratified into prevalence quartiles with allocation 24/12/12/12, seed 20260905. Its limits, in the order they matter: every rate is conditional on the entry being adjudicable, and after the 2026-09-11 second pass 4–6 of each 60 still are not (11–18 before it); the frame is Tracker Radar's ''domain_summary.json'', so no list is scored on domains that crawl never saw; and the encounter-weighted column is a ratio estimator concentrated in a handful of head domains, which is why its intervals reach 68 percentage points wide. **29 of the 525 verdicts were changed after the adjudicators returned them**, and **the direction is overwhelmingly favourable to the lists** — discount accordingly. Every change is listed with its reason on the provenance pages. | * The **random sample** is 60 domains per list from that list's own coverage, stratified into prevalence quartiles with allocation 24/12/12/12, seed 20260905. Its limits, in the order they matter: every rate is conditional on the entry being adjudicable, and after the 2026-09-11 second pass 4–6 of each 60 still are not (11–18 before it); the frame is Tracker Radar's ''domain_summary.json'', so no list is scored on domains that crawl never saw; and the encounter-weighted column is a ratio estimator concentrated in a handful of head domains, which is why its intervals reach 68 percentage points wide. **30 of the 525 verdicts were changed after the adjudicators returned them**, and **the direction is overwhelmingly favourable to the lists** — discount accordingly. Every change is listed with its reason on the provenance pages. |
| * The corpus is seven venues and omits EuroS&P, ACSAC, RAID, AsiaCCS, CHI and SOUPS, so "136 papers" is a statement about those seven. Ownership resolution is also published outside them: webXray's own author placed most of his results in The BMJ, JAMA, //Communications of the ACM// and //New Media and Society//, none of which this corpus can see. | * The corpus is seven venues and omits EuroS&P, ACSAC, RAID, AsiaCCS, CHI and SOUPS, so "136 papers" is a statement about those seven. Ownership resolution is also published outside them: webXray's own author placed most of his results in The BMJ, JAMA, //Communications of the ACM// and //New Media and Society//, none of which this corpus can see. |
| * Full-text counts are **paper** counts over ''paper.cols.txt'' with whitespace collapsed, so a term broken across a PDF column boundary still matches. They are mentions, not verified uses, except where the table says hand-verified. | * Full-text counts are **paper** counts over ''paper.cols.txt'' with whitespace collapsed, so a term broken across a PDF column boundary still matches. They are mentions, not verified uses, except where the table says hand-verified. |