Published July 27, 2026
| Version v1
Dataset
Open
OpenADMET PXR blind challenge data
Authors/Creators
Description
PXR Challenge Train/Test Dataset
A high-quality experimental dataset for predicting human Pregnane-X Receptor (PXR) induction, comprising over 11,000 compounds screened using a high-fidelity in-house assay. This is the largest publicly available PXR activity dataset, released as part of the OpenADMET PXR Induction Blind Challenge.
Blog post: Announcing the Next OpenADMET Blind Challenge: Predicting PXR Induction
Challenge Space: openadmet/pxr-challenge
Challenge period: April 1 – July 1, 2026
Produced by: OpenADMET
CHANGELOG
- Updated 2026-07-03 Added phase 2 unblinded labels for the default test set subset in
phase_2_unblindedconfig. - Updated 2026-05-27 Added phase 1 unblinded labels for the default test set subset in
phase_1_unblindedconfig. - Updated 2026-05-27 Additional crude data added, see here for more detail.
- Updated 2026-04-09 dropping some compounds, fixing minor confidence interval issues and improving naming join. See here for more details.
Dataset contents
| Config | Split | Description |
|---|---|---|
default |
train |
Primary assay training set (pEC50, Emax) |
default |
test |
513-compound blinded test set |
counter_assay |
train |
PXR-null counter-assay training data |
structure |
test |
184 molecules with X-ray crystal structures |
single_concentration |
train |
Single-concentration screening data (log2 fold change) |
crudes_htchem |
train |
Direct to biology assays of crudes (see our blog post here) |
phase_1_unblinded |
test |
Phase 1 unblinded subset of the primary blinded test set with assay labels |
phase_2_unblinded |
test |
Phase 2 unblinded subset of the primary blinded test set with assay labels |
Files
pxr-challenge_96-compound-uscale-semi-pure_TRAIN.csv
Files
(26.8 MB)
| Name | Size | Download all |
|---|---|---|
|
md5:c7e713e143f3760551f4f68ca231c0c2
|
35.9 kB | Preview Download |
|
md5:231b8be2a5802b9f3f25aae32390d17f
|
598.1 kB | Preview Download |
|
md5:cf0861e2fd80105fc0f747faafffe6b4
|
138.2 kB | Preview Download |
|
md5:da1b7a2e43572acc7622c5989fdd97f4
|
6.3 MB | Preview Download |
|
md5:fe4a1325fe74e119a8d1cc6084d2a4d1
|
11.3 kB | Preview Download |
|
md5:426d36c670c191a3d5c3a9983d6e1f67
|
2.8 kB | Download |
|
md5:adaff9b4c64e67a11d182305c6ee79b7
|
31.8 kB | Preview Download |
|
md5:e3b86192629fa4e47dadd597d74ad340
|
50.3 kB | Preview Download |
|
md5:cf414c94b23b6a81e4a4b85f623754a2
|
51.8 kB | Preview Download |
|
md5:74ea935079cc0ee83865693d4575d412
|
858.7 kB | Preview Download |
|
md5:113f5dda971df8c24cb49936b32e5376
|
5.9 kB | Preview Download |
|
md5:a60c20cc63b6304d0a02afbd9af8e6f2
|
18.7 MB | Preview Download |
Additional details
Related works
- Is identical to
- Dataset: 10.57967/hf/9731 (DOI)