Calculate every classification review budget
Map original IDs, binary labels and scores, choose ranking rules, and save the complete six-file result.
Prepare and select the complete input
Use one CSV header and complete records. Choose distinct one-based columns; IDs and labels are strings, scores are finite decimal or exponent tokens. CSV quotes, BOM, line endings, original score tokens and source ordinals remain in the evidence. Extra columns remain in the byte-exact original and count toward source cells.
id,label,score
a,1,1
b,0,-0
c,1,0
- Choose exactly two distinct labels; other labels reject the whole source.
- File mode processes the selected original, text mode processes the paste box.
- Nonzero decimal underflow, duplicate IDs and malformed UTF-8 are errors.
Inspect every K and save all evidence
Inspect any integer K from 1 to N. TP is the positive prefix count, FP=K−TP, FN=P−TP. The table preview is small; the complete report, CSV and SVG retain every K, including tied scores and undefined recall.
| File | Evidence |
|---|---|
| ranked-records.csv | All original records in rank order with raw score tokens and spans |
| budgets.csv | Every K, integer counts, numeric values and exact denominators |
| classification-report.json | Full records, all budgets, ties, signed-zero flags and states |
| budget-curves.svg | All K curve points plus complete exact JSON metadata |
| settings.json | Actual source metadata, SHA-256, frozen parameters and localized plot labels |
| source-original.csv | Byte-exact original, including unused columns |
Recover after a complete-operation failure
Changing input or settings cancels the run and clears old downloads. After cancellation or timeout, retry the same original in a fresh worker; fix malformed input or reduce a budget-exceeding scope first.
One 120-second deadline covers capture, read, validation, calculation, encoding, cleanup and first publication; a limit failure returns no partial curve.
- Keep the original and settings with all downloads.
- Capacity is the intersection of all budgets; the formal million-record axis is not a measured simultaneous maximum.
References
- Original every-budget classification question
Complete question, all four comments and six timeline items reviewed; plot only, no author sample values.
Tools in this category
Expand a tool to see its steps, options and supported formats, then open its workspace.
Classification Rank BudgetRank complete labeled scores and inspect every top-K review budget with exact prefix counts.
Choose a positive class, score direction and tie policy, then see how a fixed review budget changes precision, recall and F1 across the entire held-out set.
Steps
- Paste complete CSV or choose one file, then select that active source.
- Set distinct one-based ID, label and score columns, positive/negative labels, direction and tie policy.
- Run, inspect any K, and save all six files with the full copied report.
Available options
- Active input
- Pasted CSV · One original CSV file
- Positive label
- 1
- Negative label
- 0
- Score direction
- Higher scores first · Lower scores first
- Tie policy
- Original source order · ID Unicode codepoint order
- ID column (1-based)
- 1
- Label column (1-based)
- 2
- Score column (1-based)
- 3
Capabilities and limits
- Strict UTF-8 CSV with one header and 1–1,000,000 complete records; unique nonempty string IDs, two explicit labels and finite decimal scores.
- 64 MiB source, 10,000,000 source cells, 12,000,000 curve cells and 100,000,000 combined work units; limits apply together.
- Files plus full text ≤512 MiB, typed result ≤512 MiB, their aggregate ≤768 MiB, wire and owned memory ≤1 GiB. A legal record count alone does not guarantee joint capacity.
- One 120-second deadline covers capture, read, validation, calculation, encoding, cleanup and first publication; a limit failure returns no partial curve.
- Score ties preserve every K. No positives gives undefined recall and F1=0 for every K. Preview is limited to 16 budgets and 4,000 codepoints; downloads and full copy are complete.