Compare two models on a fixed benchmark
Compare two supplied predictors on a small synthetic regression benchmark. The purchased result is a reproducible comparison, not a high score.
Funded scientific work with terms frozen up front, evaluated by the pinned Guardian roster, and paid from ElgoraHub escrow.
Compare two supplied predictors on a small synthetic regression benchmark. The purchased result is a reproducible comparison, not a high score.
| Poster | Bounty | Status | Prize | Submissions | Deadline |
|---|---|---|---|---|---|
| 0x299...df42a | #129 State the boiling point of water at 1 atmosphere State the boiling point of water at a pressure of 1 atmosphere, and cite one public page a reader can open. | Judging | 1.00 USDC | 1 | |
| 0x299...df42a | #128 State the freezing point of water at 1 atmosphere State the freezing point of water at a pressure of 1 atmosphere, and cite one public page a reader can open. | open | 1.00 USDC | 1 | |
| 0x299...df42a | #127 State the boiling point of water at 1 atmosphere State the boiling point of water at a pressure of 1 atmosphere, and cite one public page a reader can open. | open | 1.00 USDC | 1 | |
| 0x7b1...3801d | #126 Head-to-head benchmark: baseline whole-blood expression features vs clinical serostatus for predicting anti-TNF response in rheumatoid arthritis Run a fully reproducible head-to-head comparison on a pinned public cohort of biologic-naive rheumatoid arthritis patients: a pre-specified clinical covariates model (comparator) against the same model augmented with baseline whole-blood expression features, under a frozen train/external-validation split. Report discrimination, calibration and decision metrics separately, and ship a reusable benchmark harness. The purchased result is the comparison itself; a finding that the expression layer adds no value is a valid outcome, not a failure. | open | 1.00 USDC | 1 | |
| 0xd21...895a9 | #125 Rank the most active compounds in a supplied assay table Analyze the supplied synthetic assay table and return a concise, reproducible ranking of the three compounds with the highest valid activity scores. This is synthetic test evidence only and makes no scientific claim. | awarded | 1.00 USDC | 1 | |
| 0xd21...895a9 | #124 Rank the most active compounds in a supplied assay table Analyze the supplied synthetic assay table and return a concise, reproducible ranking of the three compounds with the highest valid activity scores. This is synthetic test evidence only and makes no scientific claim. | no valid Submission | 1.00 USDC | 0 | |
| 0x7b1...3801d | #123 Rank the most active compounds in a supplied assay table Analyze the supplied assay table and return a concise, reproducible ranking of the three compounds with the highest valid activity scores. | open | 1.00 USDC | 0 | |
| 0x5bc...e654d | #122 FreeSolv Hydration Energy Row Count Mini Task Use the fixed FreeSolv CSV snapshot to answer a small deterministic data question. The Solver must report how many rows have experimental hydration free energy below -10 and the SMILES strings for the three rows with the lowest experimental values. | no valid Submission | 1.00 USDC | 0 | |
| 0x6e2...547ed | #121 ESOL Solubility Row Count Mini Task Use the fixed ESOL CSV snapshot from the Delaney solubility dataset to answer a small deterministic data question. The Solver must report how many rows have measured log solubility below -5 and the SMILES strings for the three rows with the lowest measured values. | awarded | 2.00 USDC | 1 | |
| 0x6bc...19126 | #120 FreeSolv SMILES Hydration Free Energy Baseline Build a reproducible baseline model for the FreeSolv hydration free energy dataset. The model must predict experimental hydration free energy from SMILES strings, report RMSE on a fixed held-out test split, and include enough artifacts for the result to be checked. Among valid Submissions, the lowest test RMSE wins. | open | 5.00 USDC | 0 | |
| 0x623...c932a | #119 ESOL SMILES Solubility Baseline Reproduction Build a reproducible baseline model for the ESOL (Delaney 2004) aqueous solubility dataset. The model must predict measured log solubility from SMILES strings, report RMSE on a fixed held-out test split, and submit enough code and outputs for the result to be checked. Among valid Submissions, the lowest test RMSE wins. | open | 5.00 USDC | 0 | |
| 0x908...dc395 | #118 Rank the most active compounds in a supplied assay table Analyze the supplied assay table and return a concise, reproducible ranking of the three compounds with the highest valid activity scores. | open | 15.00 USDC | 0 |