Published July 27, 2026 | Version v1

OpenADMET PXR blind challenge data

Description

PXR Challenge Train/Test Dataset

A high-quality experimental dataset for predicting human Pregnane-X Receptor (PXR) induction, comprising over 11,000 compounds screened using a high-fidelity in-house assay. This is the largest publicly available PXR activity dataset, released as part of the OpenADMET PXR Induction Blind Challenge.

Blog post: Announcing the Next OpenADMET Blind Challenge: Predicting PXR Induction

Challenge Space: openadmet/pxr-challenge

Challenge period: April 1 – July 1, 2026

Produced by: OpenADMET

CHANGELOG

  • Updated 2026-07-03 Added phase 2 unblinded labels for the default test set subset in phase_2_unblinded config.
  • Updated 2026-05-27 Added phase 1 unblinded labels for the default test set subset in phase_1_unblinded config.
  • Updated 2026-05-27 Additional crude data added, see here for more detail.
  • Updated 2026-04-09 dropping some compounds, fixing minor confidence interval issues and improving naming join. See here for more details.

Dataset contents

Config Split Description
default train Primary assay training set (pEC50, Emax)
default test 513-compound blinded test set
counter_assay train PXR-null counter-assay training data
structure test 184 molecules with X-ray crystal structures
single_concentration train Single-concentration screening data (log2 fold change)
crudes_htchem train Direct to biology assays of crudes (see our blog post here)
phase_1_unblinded test Phase 1 unblinded subset of the primary blinded test set with assay labels
phase_2_unblinded test Phase 2 unblinded subset of the primary blinded test set with assay labels

Files

pxr-challenge_96-compound-uscale-semi-pure_TRAIN.csv

Files (26.8 MB)

Name Size Download all
md5:c7e713e143f3760551f4f68ca231c0c2
35.9 kB Preview Download
md5:231b8be2a5802b9f3f25aae32390d17f
598.1 kB Preview Download
md5:cf0861e2fd80105fc0f747faafffe6b4
138.2 kB Preview Download
md5:da1b7a2e43572acc7622c5989fdd97f4
6.3 MB Preview Download
md5:fe4a1325fe74e119a8d1cc6084d2a4d1
11.3 kB Preview Download
md5:426d36c670c191a3d5c3a9983d6e1f67
2.8 kB Download
md5:adaff9b4c64e67a11d182305c6ee79b7
31.8 kB Preview Download
md5:e3b86192629fa4e47dadd597d74ad340
50.3 kB Preview Download
md5:cf414c94b23b6a81e4a4b85f623754a2
51.8 kB Preview Download
md5:74ea935079cc0ee83865693d4575d412
858.7 kB Preview Download
md5:113f5dda971df8c24cb49936b32e5376
5.9 kB Preview Download
md5:a60c20cc63b6304d0a02afbd9af8e6f2
18.7 MB Preview Download

Additional details

Related works

Is identical to
Dataset: 10.57967/hf/9731 (DOI)