{"id":752125,"date":"2026-07-09T07:41:13","date_gmt":"2026-07-09T07:41:13","guid":{"rendered":"https:\/\/www.newsbeep.com\/us\/752125\/"},"modified":"2026-07-09T07:41:13","modified_gmt":"2026-07-09T07:41:13","slug":"chromatin-landscape-and-epigenetic-heterogeneity-of-acute-myeloid-leukaemia","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/us\/752125\/","title":{"rendered":"Chromatin landscape and epigenetic heterogeneity of acute myeloid leukaemia"},"content":{"rendered":"<p>Patients and samples<\/p>\n<p>This study was reviewed and approved by the regional ethics review board in Stockholm (2008\/1330-31\/2 and 2017\/2085-31\/2) and the institutional ethics committees of Kyoto University (G608 and G1110) and participating institutions (Hyogo Prefectural Amagasaki General Medical Center, Chugoku Central Hospital, Dokkyo Medical University Saitama Medical Center, Gifu University Hospital, Gifu Municipal Hospital, Hyogo Medical University, Hokkaido University, Japanese Red Cross Kyoto Daini Hospital, Kobe City Medical Center General Hospital, Kurashiki Central Hospital, Kitano Hospital, Kyoto Medical Center, Kyoto City Hospital, Matsushita Memorial Hospital, Japanese Red Cross Nagano Hospital, National Cancer Center Hospital, NTT Medical Center Tokyo, Osaka International Cancer Institute, Japanese Red Cross Osaka Hospital, University of Osaka, Otsu Red Cross Hospital, Shizuoka City Shizuoka Hospital, Shinko Hospital, Shiga General Hospital, Sumitomo Hospital, Takeda General Hospital, Takatsuki Red Cross Hospital, Uji Tokushukai Medical Center and Japanese Red Cross Society Wakayama Medical Center) (no. G608), and was performed in accordance with the Declaration of Helsinki. Informed consent was obtained from all participants at participating institutions.<\/p>\n<p>A total of 1,563 patients diagnosed with AML and related neoplasms were consecutively enrolled from the participating institutes and hospitals between 1997 and 2022 in Sweden and between 2011 and 2022 in Japan. Although no preselection criteria were applied with respect to clinical characteristics, sample inclusion was contingent on the availability of viable cryopreserved tumour cells suitable for ATAC-seq for consecutively diagnosed patients with AML. Diagnoses were made at each participating institution and hospital in accordance with the WHO classification in use at the time. In Sweden, patients were registered through the national AML registry, which prospectively collects clinical and genomic data on all newly diagnosed cases. In Japan, patients were diagnosed and treated at participating hospitals, with biospecimens and clinical information sent to and managed by the Kyoto University biobank, where clinical annotations were updated annually. Treatment was administered according to institutional standards of care and, for Swedish individuals, according to the national treatment guidelines for AML, with a subset of patients participating in clinical trials. All samples analysed in this study were collected at the time of initial diagnosis before treatment. Detailed clinical annotations, including diagnosis, demographic and laboratory data, treatment regimens and outcomes, were extracted from electronic medical records and the Swedish AML registry. No statistical methods were used to predetermine sample size. Patient characteristics and diagnoses are summarized in Supplementary Tables <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">1<\/a> and <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">2<\/a>.<\/p>\n<p>Tumour samples, such as bone marrow or peripheral blood, and matched control buccal samples, were obtained from patients. We also analysed normal bone-marrow samples from 25 individuals without haematological malignancies who underwent hip joint replacement surgery, serving as controls for the ATAC-seq analysis. Bone marrow and peripheral blood cells were isolated, subjected to erythrolysis or mononuclear cell isolation by Ficoll gradient centrifugation, resuspended in CELLBANKER 1 solution (Nippon Zenyaku Kogyo) or in 10% dimethyl sulfoxide (DMSO) (Merck KGaA) with 90% fetal calf serum (FCS; Thermo Fisher Scientific) and cryopreserved in liquid nitrogen. Genomic DNA was extracted using the QIAamp DNA Mini Kit (QIAGEN), the Gentra PureGene Kit (QIAGEN) or the Maxwell RSC Genomic DNA Kit (Promega). RNA was extracted using the RNeasy Mini Kit (QIAGEN) or the Maxwell RSC simplyRNA Tissue Kit (Promega). A summary of this cohort, including diagnosis, ATAC subgroup, AML classifications, driver genes and available multiomics data, is provided in Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">3<\/a>.<\/p>\n<p>Cell lines<\/p>\n<p>K562, KG-1, THP-1, SKM-1, KY821 and NOMO-1 cells were obtained from the RIKEN BioResource Center Cell Bank (Tsukuba, Japan); Kasumi-1, HL-60, MOLM-13 and TF-1 cells from the American Type Culture Collection (ATCC); and OCI-AML3 cells from the German Collection of Microorganisms and Cell Cultures (DSMZ). These cell lines were authenticated by short tandem repeat profiling and tested for mycoplasma by the providing cell banks.<\/p>\n<p>Targeted-capture sequencing<\/p>\n<p>Targeted-capture sequencing was performed using the SureSelect custom kit (Agilent Technologies), with an in-house gene panel including 331 known AML driver genes (Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">4<\/a>) and an additional 1,158\u20131,317 probes for copy-number detection. Captured targets were sequenced using the NovaSeq 6000 (Illumina) or DNBSEQ-G400 (MGI) with a 150-bp paired-end read protocol.<\/p>\n<p>Sequence alignment and mutation calling were performed using the hg19 reference genome and Genomon pipeline<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 25\" title=\"Okuda, R. et al. Genetic analysis of myeloid neoplasms with der(1;7)(q10;p10). Leukemia 39, 760&#x2013;764 (2025).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR25\" id=\"ref-link-section-d87789345e3018\" rel=\"nofollow noopener\" target=\"_blank\">25<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" title=\"Nannya, Y. et al. Postazacitidine clone size predicts long-term outcome of patients with myelodysplastic syndromes and related myeloid neoplasms. Blood Adv. 7, 3624&#x2013;3636 (2023).\" href=\"#ref-CR51\" id=\"ref-link-section-d87789345e3021\">51<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" title=\"Ochi, Y. et al. Clonal evolution and clinical implications of genetic abnormalities in blastic transformation of chronic myeloid leukaemia. Nat. Commun. 12, 2833 (2021).\" href=\"#ref-CR52\" id=\"ref-link-section-d87789345e3021_1\">52<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 53\" title=\"Yoshizato, T. et al. Genetic abnormalities in myelodysplasia and secondary acute myeloid leukemia: impact on outcome of stem cell transplantation. Blood 129, 2347&#x2013;2358 (2017).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR53\" id=\"ref-link-section-d87789345e3024\" rel=\"nofollow noopener\" target=\"_blank\">53<\/a> (<a href=\"https:\/\/github.com\/Genomon-Project\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/Genomon-Project<\/a>). Unless otherwise stated, sequencing data were aligned to the hg19 reference genome. The called variants were further filtered by assessing the oncogenicity of variants on the basis of an in-house curation program<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 25\" title=\"Okuda, R. et al. Genetic analysis of myeloid neoplasms with der(1;7)(q10;p10). Leukemia 39, 760&#x2013;764 (2025).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR25\" id=\"ref-link-section-d87789345e3035\" rel=\"nofollow noopener\" target=\"_blank\">25<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" title=\"Nannya, Y. et al. Postazacitidine clone size predicts long-term outcome of patients with myelodysplastic syndromes and related myeloid neoplasms. Blood Adv. 7, 3624&#x2013;3636 (2023).\" href=\"#ref-CR51\" id=\"ref-link-section-d87789345e3038\">51<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" title=\"Ochi, Y. et al. Clonal evolution and clinical implications of genetic abnormalities in blastic transformation of chronic myeloid leukaemia. Nat. Commun. 12, 2833 (2021).\" href=\"#ref-CR52\" id=\"ref-link-section-d87789345e3038_1\">52<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 53\" title=\"Yoshizato, T. et al. Genetic abnormalities in myelodysplasia and secondary acute myeloid leukemia: impact on outcome of stem cell transplantation. Blood 129, 2347&#x2013;2358 (2017).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR53\" id=\"ref-link-section-d87789345e3041\" rel=\"nofollow noopener\" target=\"_blank\">53<\/a> that uses the COSMIC database (v.96), an in-house blacklist of error calls and public SNP databases, including the 1000 Genomes Project (October 2014 release), NCBI dbSNP build 138, National Heart, Lung, and Blood Institute (NHLBI) Exome Sequencing Project (ESP) 6500, the Human Genetic Variation Database (HGVD) and our in-house dataset.<\/p>\n<p>For copy number analysis,\u00a0SNP probes included in the target bait for targeted-capture sequencing were used to allow the detection of copy-number changes and allelic imbalances<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 25\" title=\"Okuda, R. et al. Genetic analysis of myeloid neoplasms with der(1;7)(q10;p10). Leukemia 39, 760&#x2013;764 (2025).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR25\" id=\"ref-link-section-d87789345e3048\" rel=\"nofollow noopener\" target=\"_blank\">25<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" title=\"Nannya, Y. et al. Postazacitidine clone size predicts long-term outcome of patients with myelodysplastic syndromes and related myeloid neoplasms. Blood Adv. 7, 3624&#x2013;3636 (2023).\" href=\"#ref-CR51\" id=\"ref-link-section-d87789345e3051\">51<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" title=\"Ochi, Y. et al. Clonal evolution and clinical implications of genetic abnormalities in blastic transformation of chronic myeloid leukaemia. Nat. Commun. 12, 2833 (2021).\" href=\"#ref-CR52\" id=\"ref-link-section-d87789345e3051_1\">52<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 53\" title=\"Yoshizato, T. et al. Genetic abnormalities in myelodysplasia and secondary acute myeloid leukemia: impact on outcome of stem cell transplantation. Blood 129, 2347&#x2013;2358 (2017).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR53\" id=\"ref-link-section-d87789345e3054\" rel=\"nofollow noopener\" target=\"_blank\">53<\/a>. This program (CNACS) is available at <a href=\"https:\/\/github.com\/papaemmelab\/toil_cnacs\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/papaemmelab\/toil_cnacs<\/a>. A total copy number (TCN) of 2.22 or higher was defined as gain, and a TCN lower than 1.88 was defined as loss. Copy-number-neutral loss of heterozygosity was called with a B-allele frequency lower than 0.90 and a TCN between 1.88 and 2.22. Arm-level changes were called for regions with a total length greater than 1\u2009Mb. Detected copy-number changes were manually curated.<\/p>\n<p>Structural variations (SVs) were detected using the Genomon SV pipeline<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 54\" title=\"Kataoka, K. et al. Integrated molecular analysis of adult T cell leukemia\/lymphoma. Nat. Genet. 47, 1304&#x2013;1315 (2015).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR54\" id=\"ref-link-section-d87789345e3068\" rel=\"nofollow noopener\" target=\"_blank\">54<\/a>, which uses both breakpoint-containing junction read pairs and improperly aligned read pairs. Detected putative SVs were filtered by removing (i) those with fewer than four supporting tumour reads and fewer than ten supporting tumour\/normal reads and (ii) those present in control normal samples, whose breakpoints were manually inspected using the Integrative Genomics Viewer (IGV).\u00a0For KMT2A and MECOM rearrangements, cases other than t(9;11)(p21.3;q23.3) (KMT2A::MLLT3) and inv(3)(q21.3q26.2) or t(3;3)(q21.3;q26.2) (GATA2::MECOM) were classified as SVs with rare partners and annotated with a distinct colour, reflecting their separate classification from KMT2A::MLLT3 and GATA2::MECOM rearrangements in the ICC classification.<\/p>\n<p>WGS<\/p>\n<p>WGS data were obtained as part of the Japanese national cancer genomics initiative (Genome Research in Cancers and Rare Diseases; G-CARD). A sequencing library was generated using the Illumina DNA PCR-Free Prep Tagmentation Kit according to the manufacturer\u2019s protocol, and sequenced using the NovaSeq 6000 (Illumina) with a 150-bp paired-end read protocol at a target depth of 100\u00d7 for tumours and 30\u00d7 for normal controls.<\/p>\n<p>The reads were aligned to the human hg38 reference genome by Parabricks. Somatic variants were called in tumour\u2013normal paired mode using Mutect2 (GATK v.4.5.0.0), retaining variants with tumour log-odds (TLOD)\u2009\u2265\u200920 and variant allele frequency\u2009\u2264\u20090.3 in the matched normal. Variants listed in germline databases, such as gnomAD (v.3.1.2), the 1000 Genomes Project and Tohoku Medical Megabank Organization (ToMMo), with a minor allele frequency of 0.001 or higher, were removed. SVs were called using GRIDSS v.2.12.0 and Genomon SV v.0.8.0. Filtered outputs from GRIDSS (\u2018FILTER\u2009=\u2009PASS, QUAL\u2009\u2265\u2009500, AS\u2009&gt;\u20090, or RAS\u2009&gt;\u20090\u2019) and Genomon SV (\u2018overhang\u2009\u2265\u2009150\u2019) were combined. CNAs were inferred from tumour\u2013normal paired WGS data using Battenberg (v.3.0.0) with default parameters, based on log R ratios and B-allele frequencies of germline heterozygous SNPs.<\/p>\n<p>Bulk RNA-seq experiments and analysis<\/p>\n<p>Libraries for RNA-seq were prepared using the NEBNext Single Cell\/Low Input RNA Library Prep Kit for Illumina (New England BioLabs) and were subjected to sequencing using the NovaSeq 6000 (Illumina) with a paired-end protocol.<\/p>\n<p>The sequencing reads were preprocessed by fastp<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 55\" title=\"Chen, S. Ultrafast one-pass FASTQ data preprocessing, quality control, and deduplication using fastp. iMeta 2, e107 (2023).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR55\" id=\"ref-link-section-d87789345e3114\" rel=\"nofollow noopener\" target=\"_blank\">55<\/a> with \u2018&#8211;detect_adapter_for_pe -q 15 -n 10 -u 40\u2019 parameters, and aligned to the reference genome using STAR<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 56\" title=\"Dobin, A. et al. STAR: ultrafast universal RNA-seq aligner. Bioinformatics 29, 15&#x2013;21 (2013).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR56\" id=\"ref-link-section-d87789345e3118\" rel=\"nofollow noopener\" target=\"_blank\">56<\/a>. Reads on each gene defined in the University of California, Santa Cruz (UCSC) hg19 gene annotation were\u00a0counted with featureCounts<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 57\" title=\"Liao, Y., Smyth, G. K. &amp; Shi, W. featureCounts: an efficient general purpose program for assigning sequence reads to genomic features. Bioinformatics 30, 923&#x2013;930 (2014).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR57\" id=\"ref-link-section-d87789345e3122\" rel=\"nofollow noopener\" target=\"_blank\">57<\/a>. The quality of sequencing data was assessed using mapping and count statistics. Samples were excluded from the analysis if any of the following criteria were met: \u2018uniquely_mapped_percent\u2019 (STAR)\u2009&lt;\u200930%, \u2018percent_assigned\u2019 (featureCounts)\u2009&lt;\u200930% or \u2018assigned\u2019 (featureCounts)\u2009&lt;\u20093\u2009million. The edgeR package<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 58\" title=\"Chen, Y., Lun, A. T. L. &amp; Smyth, G. K. From reads to genes to pathways: differential expression analysis of RNA-seq experiments using Rsubread and the edgeR quasi-likelihood pipeline. F1000Res. 5, 1438 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR58\" id=\"ref-link-section-d87789345e3126\" rel=\"nofollow noopener\" target=\"_blank\">58<\/a> was used to normalize read counts and calculate counts per million (CPM) values of genes. Genes with CPM\u2009&gt;\u20091 in at least two samples and located on the autosomal chromosomes were kept and used for downstream analysis. To adjust for cohort-specific technical effects while preserving biological variation, batch correction between the Swedish and Japanese cohorts was performed using linear modelling in limma<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 59\" title=\"Ritchie, M. E. et al. limma powers differential expression analyses for RNA-sequencing and microarray studies. Nucleic Acids Res. 43, e47 (2015).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR59\" id=\"ref-link-section-d87789345e3130\" rel=\"nofollow noopener\" target=\"_blank\">59<\/a>. Differentially expressed genes (DEGs) for each subgroup were identified using the eBayes test in limma through a one-versus-rest comparison, with thresholds of FDR\u2009&lt;\u20090.05 and |log2-transformed fold change|\u2009&gt;\u20090.5. The top 3,000 DEGs in each subgroup are provided in Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">8<\/a>. GSEA analysis was performed on genes ranked by fold change using the GSEA function in the clusterProfiler package<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 60\" title=\"Xu, S. et al. Using clusterProfiler to characterize multiomics data. Nat. Protoc. 19, 3292&#x2013;3320 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR60\" id=\"ref-link-section-d87789345e3140\" rel=\"nofollow noopener\" target=\"_blank\">60<\/a> (minGSSize\u2009=\u200920, pAdjustMethod\u2009=\u2009\u201cBH\u201d) and a curated set of genes associated with haematopoietic cells and AML derived from the Human Molecular Signatures Database (MSigDB)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 61\" title=\"Subramanian, A. et al. Gene set enrichment analysis: a knowledge-based approach for interpreting genome-wide expression profiles. Proc. Natl Acad. Sci. USA 102, 15545&#x2013;15550 (2005).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR61\" id=\"ref-link-section-d87789345e3144\" rel=\"nofollow noopener\" target=\"_blank\">61<\/a> (Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">10<\/a>). Gene fusions were detected using the Genomon fusion pipeline (<a href=\"https:\/\/github.com\/Genomon-Project\/fusionfusion\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/Genomon-Project\/fusionfusion<\/a>) and filtered for known drivers of AML.<\/p>\n<p>Bulk ATAC-seq experiments<\/p>\n<p>ATAC-seq experiments were performed using the Fast-ATAC protocol<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 21\" title=\"Corces, M. R. et al. Lineage-specific and single-cell chromatin accessibility charts human hematopoiesis and leukemia evolution. Nat. Genet. 48, 1193&#x2013;1203 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR21\" id=\"ref-link-section-d87789345e3166\" rel=\"nofollow noopener\" target=\"_blank\">21<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 62\" title=\"Ochi, Y. et al. Combined cohesin&#x2013;RUNX1 deficiency synergistically perturbs chromatin looping and causes myelodysplastic syndromes. Cancer Discov. 10, 836&#x2013;853 (2020).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR62\" id=\"ref-link-section-d87789345e3169\" rel=\"nofollow noopener\" target=\"_blank\">62<\/a>. Cryopreserved tumour cells were thawed, and 50,000 cells were pelleted. Fifty microlitres of transposase mixture (comprising 25\u2009\u00b5l 2\u00d7 TD buffer, 2.5\u2009\u00b5l TDE1\u00a0(Illumina, FC-121-1031), 0.5\u2009\u00b5l 1% digitonin (Sigma-Aldrich, D141) and 22\u2009\u00b5l nuclease-free water) was added to the cells. After transposition reactions at 37\u2009\u00b0C for 30\u2009min, transposed DNA was purified using the QIAGEN MinElute Reaction Cleanup Kit or Sera-Mag Select (Cytiva) magnetic beads, and PCR-amplified using the NEBNEXT Q5 Hot Start HiFi PCR Master Mix and custom primers<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 63\" title=\"Buenrostro, J. D. et al. Single-cell chromatin accessibility reveals principles of regulatory variation. Nature 523, 486&#x2013;490 (2015).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR63\" id=\"ref-link-section-d87789345e3173\" rel=\"nofollow noopener\" target=\"_blank\">63<\/a>. The TapeStation 4200 (Agilent) with High Sensitivity D5000 ScreenTape was used to assess fragment size and\u00a0confirm nucleosomal periodicity, a characteristic\u00a0feature of ATAC-seq libraries<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 20\" title=\"Buenrostro, J. D., Giresi, P. G., Zaba, L. C., Chang, H. Y. &amp; Greenleaf, W. J. Transposition of native chromatin for fast and sensitive epigenomic profiling of open chromatin, DNA-binding proteins and nucleosome position. Nat. Methods 10, 1213&#x2013;1218 (2013).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR20\" id=\"ref-link-section-d87789345e3177\" rel=\"nofollow noopener\" target=\"_blank\">20<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 64\" title=\"Grandi, F. C., Modi, H., Kampman, L. &amp; Corces, M. R. Chromatin accessibility profiling by ATAC-seq. Nat. Protoc. 17, 1518&#x2013;1552 (2022).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR64\" id=\"ref-link-section-d87789345e3180\" rel=\"nofollow noopener\" target=\"_blank\">64<\/a>. The resulting library was sequenced using the NovaSeq 6000 (Illumina) with a paired-end protocol.<\/p>\n<p>Bulk ATAC-seq analysis<\/p>\n<p>Reads were aligned to the reference genome using Bowtie2<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 65\" title=\"Langmead, B. &amp; Salzberg, S. L. Fast gapped-read alignment with Bowtie 2. Nat. Methods 9, 357&#x2013;359 (2012).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR65\" id=\"ref-link-section-d87789345e3192\" rel=\"nofollow noopener\" target=\"_blank\">65<\/a> with \u2018-X 2000 \u2013no-mixed \u2013very-sensitive\u2019 parameters, after adapter trimming using Skewer<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 66\" title=\"Jiang, H., Lei, R., Ding, S.-W. &amp; Zhu, S. Skewer: a fast and accurate adapter trimmer for next-generation sequencing paired-end reads. BMC Bioinformatics 15, 182 (2014).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR66\" id=\"ref-link-section-d87789345e3196\" rel=\"nofollow noopener\" target=\"_blank\">66<\/a>. The quality of sequencing data was assessed using ataqv<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 67\" title=\"Orchard, P., Kyono, Y., Hensley, J., Kitzman, J. O. &amp; Parker, S. C. J. Quantification, dynamic visualization, and validation of bias in ATAC-seq data with ataqv. Cell Syst. 10, 298&#x2013;306 (2020).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR67\" id=\"ref-link-section-d87789345e3200\" rel=\"nofollow noopener\" target=\"_blank\">67<\/a> and ATACseqQC<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 68\" title=\"Ou, J. et al. ATACseqQC: a Bioconductor package for post-alignment quality assessment of ATAC-seq data. BMC Genomics 19, 169 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR68\" id=\"ref-link-section-d87789345e3204\" rel=\"nofollow noopener\" target=\"_blank\">68<\/a>. Samples were excluded from the analysis if any of the following criteria were met: fewer than six million uniquely mapped reads; mapping rate lower than 40%; more than 30% of reads on mitochondrial chromosome; or transcription start site (TSS) enrichment scores lower than 2 in both 1,000 and 2,000 bp\u00a0windows. After removing duplicates and reads on the mitochondrial genome or blacklisted regions (ENCODE), peaks were called using HMMRATAC<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 69\" title=\"Tarbell, E. D. &amp; Liu, T. HMMRATAC: a Hidden Markov ModeleR for ATAC-seq. Nucleic Acids Res. 47, e91 (2019).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR69\" id=\"ref-link-section-d87789345e3208\" rel=\"nofollow noopener\" target=\"_blank\">69<\/a> with the \u2018&#8211;window 2500000\u2019 parameter and peaks were fixed to a width of 501\u2009bp centred on the peak summit<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 64\" title=\"Grandi, F. C., Modi, H., Kampman, L. &amp; Corces, M. R. Chromatin accessibility profiling by ATAC-seq. Nat. Protoc. 17, 1518&#x2013;1552 (2022).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR64\" id=\"ref-link-section-d87789345e3213\" rel=\"nofollow noopener\" target=\"_blank\">64<\/a>. When extended or merged peaks overlapped, the region with the highest peak score was retained. For each of the six datasets\u2014including our own datasets (Swedish AML, Japanese AML, AML cell lines and normal bone marrow) as well as public datasets<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 21\" title=\"Corces, M. R. et al. Lineage-specific and single-cell chromatin accessibility charts human hematopoiesis and leukemia evolution. Nat. Genet. 48, 1193&#x2013;1203 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR21\" id=\"ref-link-section-d87789345e3217\" rel=\"nofollow noopener\" target=\"_blank\">21<\/a> (sorted normal blood and AML cells)\u2014we merged peaks from all samples in each dataset. Peaks that overlapped in two or more samples were considered the recurrent peak set for each dataset. For samples from patients with AML, Swedish and Japanese peak sets were further merged and filtered for those recurrently identified in both the Swedish and the Japanese samples. We then merged five peak sets generated as above (samples from patients with AML\u00a0(eCHROMA), AML cell lines, normal bone marrow, public normal blood and public AML) to generate a union peak set fixed to a width of 501\u2009bp, which was used for downstream analysis (Extended Data Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#Fig6\" rel=\"nofollow noopener\" target=\"_blank\">1b<\/a>). Annotation of peaks was done using the annotatePeaks function in HOMER<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 70\" title=\"Heinz, S. et al. Simple combinations of lineage-determining transcription factors prime cis-regulatory elements required for macrophage and B cell identities. Mol. Cell 38, 576&#x2013;589 (2010).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR70\" id=\"ref-link-section-d87789345e3224\" rel=\"nofollow noopener\" target=\"_blank\">70<\/a>. Reads on peaks were counted and normalized to calculate CPM values using edgeR<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 58\" title=\"Chen, Y., Lun, A. T. L. &amp; Smyth, G. K. From reads to genes to pathways: differential expression analysis of RNA-seq experiments using Rsubread and the edgeR quasi-likelihood pipeline. F1000Res. 5, 1438 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR58\" id=\"ref-link-section-d87789345e3228\" rel=\"nofollow noopener\" target=\"_blank\">58<\/a>. Differentially expressed ATAC peaks (DEPs) for each subgroup were identified using the eBayes test in limma through a one-versus-rest comparison, with thresholds of FDR\u2009&lt;\u20090.05 and |log2-transformed fold change|\u2009&gt;\u20090.5. The top 3,000 DEPs in each subgroup are provided in Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">9<\/a>. Inference of cellular contribution to each ATAC-seq data was evaluated by the CIBERSORTx<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 27\" title=\"Newman, A. M. et al. Determining cell type abundance and expression from bulk tissues with digital cytometry. Nat. Biotechnol. 37, 773&#x2013;782 (2019).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR27\" id=\"ref-link-section-d87789345e3238\" rel=\"nofollow noopener\" target=\"_blank\">27<\/a> with CPM values of our AML data and public data for 13 normal blood cell types<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 21\" title=\"Corces, M. R. et al. Lineage-specific and single-cell chromatin accessibility charts human hematopoiesis and leukemia evolution. Nat. Genet. 48, 1193&#x2013;1203 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR21\" id=\"ref-link-section-d87789345e3242\" rel=\"nofollow noopener\" target=\"_blank\">21<\/a> as input files.<\/p>\n<p>Clustering analysis using bulk ATAC-seq<\/p>\n<p>ATAC log2 CPM were quantile normalized using preprocessCore (<a href=\"https:\/\/github.com\/bmbolstad\/preprocessCore\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/bmbolstad\/preprocessCore<\/a>), followed by the exclusion of peaks on chromosomes X and Y. Batch correction was applied as described above. To focus on leukaemia-intrinsic chromatin variation and minimize contamination from non-malignant cells, peak variance was calculated using samples with at least 75% bone-marrow blasts. The 3,000 most variable peaks were selected to capture dominant inter-sample chromatin heterogeneity while reducing noise from low-variance regions. The first 50 principal components were determined by prcomp in R, a nearest-neighbour graph was generated using buildSNNGraph in the scran package<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 71\" title=\"Lun, A. T. L., McCarthy, D. J. &amp; Marioni, J. C. A step-by-step workflow for low-level analysis of single-cell RNA-seq data with Bioconductor. F1000Res. 5, 2122 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR71\" id=\"ref-link-section-d87789345e3263\" rel=\"nofollow noopener\" target=\"_blank\">71<\/a> with a \u2018k\u2009=\u20097\u2019 parameter and Leiden clustering<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 72\" title=\"Traag, V. A., Waltman, L. &amp; van Eck, N. J. From Louvain to Leiden: guaranteeing well-connected communities. Sci. Rep. 9, 5233 (2019).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR72\" id=\"ref-link-section-d87789345e3270\" rel=\"nofollow noopener\" target=\"_blank\">72<\/a> was performed using \u2018cluster_leiden\u2019 from the igraph package with parameters of \u2018resolution\u2009=\u20090.2\u2019 and \u2018n_iterations\u2009=\u2009100\u2019. Clusters with less than 1% of samples (fewer than 16 out of 1,563 cases) were reassigned to the cluster of the nearest sample by calculating the squared Euclidean distance between samples in the two-dimensional UMAP space. ATAC subgroups were annotated with representative genetic alterations and\/or differentiation states (for example, A, PML::RARA) to aid interpretability. Labels were assigned on the basis of enriched and relatively specific features within each subgroup, including mutation patterns and differentiation signatures, but do not indicate\u00a0that subgroup identity is defined solely by these alterations. For high-resolution clustering, we used a resolution of 0.3 with all other parameters kept the same.<\/p>\n<p>Clustering stability<\/p>\n<p>To evaluate the robustness of the Leiden clustering results, we performed a resampling-based stability analysis. First, we repeatedly subsampled 90% of the samples and re-applied the same clustering parameters (resolution\u2009=\u20090.2, k\u2009=\u20097 neighbours) for 100 iterations, calculating the adjusted Rand index between the original and each resampled clustering to quantify overall agreement; adjusted Rand index\u00a0values close to 1 indicate strong stability. Second, we computed a sample-wise consistency score for each sample, defined as the fraction of iterations in which the sample clustered together with its original cluster members, conditional on co-occurrence, thus measuring the stability of cluster membership at the individual sample level. Third, we constructed a cluster\u2013cluster similarity matrix that captures the probability that samples from different clusters co-cluster across all resampling runs, providing a summary of within-cluster consistency and potential cross-cluster mixing.<\/p>\n<p>Decision-tree modelling based on genomic alterations<\/p>\n<p>To test whether ATAC-defined subgroups could be explained by simple combinations of genomic alterations, we constructed one-versus-rest decision-tree classifiers using gene mutations, CNAs and SVs as binary input features. For each of the 16 ATAC subgroups, a binary classifier was trained using a CART framework implemented in the rpart package in R. Rare alterations (present in fewer than four positive samples) were excluded from model training. Samples were randomly split into training (80%) and test (20%) sets. Tree depth was fixed at a maximum depth of 3 to constrain model complexity and emphasize interpretability. The splitting criterion was Gini impurity. Class predictions were derived using a fixed probability threshold of 0.5. Model performance was evaluated on the independent test set using sensitivity, positive predictive value (PPV), F1 score and balanced accuracy. Model complexity was quantified as the number of internal splits in the final tree. To examine how performance changed with increasing model complexity, we also trained models across a range of maximum tree depths (2\u201312), while keeping all other parameters fixed.<\/p>\n<p>Bulk-RNA-seq-based prediction of ATAC subgroups<\/p>\n<p>Normalized gene-count data for four external adult AML cohorts were obtained from a previous study<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 10\" title=\"Severens, J. F. et al. Mapping AML heterogeneity&#x2014;multi-cohort transcriptomic analysis identifies novel clusters and divergent ex-vivo drug responses. Leukemia 38, 751&#x2013;761 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR10\" id=\"ref-link-section-d87789345e3308\" rel=\"nofollow noopener\" target=\"_blank\">10<\/a>. The TARGET cohort was excluded from the analysis because it consisted mainly of paediatric AML. The count matrix was quantile normalized and merged with our RNA-seq count matrix, followed by batch correction using \u2018removeBatchEffect\u2019. Prediction models for ATAC subgroups were generated using the ClaNC algorithm, a nearest-centroid classifier that ranks genes by standard t-statistics and selects subgroup-specific genes<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 30\" title=\"Dabney, A. R. Classification of microarrays to nearest centroids. Bioinformatics 21, 4148&#x2013;4154 (2005).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR30\" id=\"ref-link-section-d87789345e3315\" rel=\"nofollow noopener\" target=\"_blank\">30<\/a>. Fifteen genes per subgroup\u00a0were selected for subgroup-specific markers and incorporated into the prediction model. Gene lists used for each prediction model are summarized in Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">11<\/a>. The accuracy of prediction models was calculated by fivefold cross-validation, separating 80% of samples for training and the remaining 20% for validation.<\/p>\n<p>DNA methylation experiments and analysis<\/p>\n<p>DNA methylation profiles were analysed using the Infinium MethylationEPIC BeadChip Kit according to the manufacturer\u2019s protocol. The raw files were processed using the ChAMP package in R, filtering probes and generating the normalized \u03b2-value matrix. Common probes across samples were used for downstream analysis. To correct batch effects between different cohorts and probe sets, the \u2018removeBatchEffect\u2019 command from the limma package<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 59\" title=\"Ritchie, M. E. et al. limma powers differential expression analyses for RNA-sequencing and microarray studies. Nucleic Acids Res. 43, e47 (2015).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR59\" id=\"ref-link-section-d87789345e3330\" rel=\"nofollow noopener\" target=\"_blank\">59<\/a> was used.<\/p>\n<p>For comparison between chromatin accessibility and DNA methylation, ATAC peak regions overlapping with methylation probes were analysed. DNA methylation levels in a given region were calculated as the average \u03b2-values of all probes in the region. For each ATAC subgroup, average DNA methylation levels were computed, and the top 3,000 regions with the most variable DNA methylation levels across subgroups were visualized in heat maps.<\/p>\n<p>GRN analysis by bulk ATAC-seq and RNA-seq<\/p>\n<p>GRNs were inferred using the ANANSE software<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 35\" title=\"Xu, Q. et al. ANANSE: an enhancer network-based computational approach for predicting key transcription factors in cell fate determination. Nucleic Acids Res. 49, 7966&#x2013;7985 (2021).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR35\" id=\"ref-link-section-d87789345e3345\" rel=\"nofollow noopener\" target=\"_blank\">35<\/a>. In brief, this software integrates chromatin accessibility and gene-expression data to infer enhancer-based GRNs. Input data included the consensus ATAC peak set, a merged ATAC-seq BAM file, mean RNA expression (log2 CPM) and the default motif database (GimmeMotifs<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 73\" title=\"van Heeringen, S. J. &amp; Veenstra, G. J. C. GimmeMotifs: a de novo motif prediction pipeline for ChIP-sequencing experiments. Bioinformatics 27, 270&#x2013;271 (2011).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR73\" id=\"ref-link-section-d87789345e3351\" rel=\"nofollow noopener\" target=\"_blank\">73<\/a>).<\/p>\n<p>For each TF, ANANSE estimates genome-wide binding potential, predicts target genes and incorporates expression levels of the TF and its targets to model regulatory interactions. The regulatory importance of each TF is quantified by two key metrics: out-degree (the number and strength of predicted regulatory connections) and link score (the strength of each individual connection). GRNs were constructed for each ATAC subgroup and for the entire AML cohort. To identify subgroup-specific regulators, we compared each subgroup GRN with the global AML GRN and calculated TF influence scores, which measure how much a TF explains differences in expression between the two groups (Extended Data Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#Fig11\" rel=\"nofollow noopener\" target=\"_blank\">6a<\/a>). The top 20 differential TFs per subgroup were visualized.<\/p>\n<p>ChIP\u2013seq experiments and analysis<\/p>\n<p>ChIP\u2013seq experiments were performed according to the SimpleChIP Plus Sonication Chromatin IP Kit (Cell Signaling Technology)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 62\" title=\"Ochi, Y. et al. Combined cohesin&#x2013;RUNX1 deficiency synergistically perturbs chromatin looping and causes myelodysplastic syndromes. Cancer Discov. 10, 836&#x2013;853 (2020).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR62\" id=\"ref-link-section-d87789345e3369\" rel=\"nofollow noopener\" target=\"_blank\">62<\/a> with minor modifications. Cryopreserved cells were thawed, and more than one million cells were fixed with 1% formaldehyde\u00a0(Thermo Fisher Scientific) in phosphate-buffered saline (PBS) for 10\u2009min at room temperature with gentle mixing. The reaction was stopped by adding glycine solution (10\u00d7) (Cell Signaling Technology) and incubated for 5\u2009min at room temperature, and the cells were washed twice in cold PBS. The cells were then processed with the SimpleChIP Plus Sonication Chromatin IP Kit (Cell Signaling Technology) and Covaris E220 (Covaris) according to the manufacturer\u2019s protocol. The antibodies used for ChIP were as follows: SMC1 (Abcam, ab9262), CTCF (Cell Signaling Technology, D31H2), RPB1 (Cell Signaling Technology, D8L4Y), H3K27ac (Cell Signaling Technology, D5E4) and H3K27me3 (Cell Signaling Technology, C36B11). After purification of the precipitated DNA, libraries were constructed using the ThruPLEX DNA-seq Kit (Takara) as per the manufacturer\u2019s protocol, and subjected to sequencing using the NovaSeq 6000 (Illumina). ChIP\u2013seq experiments were performed with input controls. The sequencing reads were aligned to the reference genome using Bowtie<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 74\" title=\"Langmead, B., Trapnell, C., Pop, M. &amp; Salzberg, S. L. Ultrafast and memory-efficient alignment of short DNA sequences to the human genome. Genome Biol. 10, R25 (2009).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR74\" id=\"ref-link-section-d87789345e3373\" rel=\"nofollow noopener\" target=\"_blank\">74<\/a>, after adapter trimming with Skewer<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 66\" title=\"Jiang, H., Lei, R., Ding, S.-W. &amp; Zhu, S. Skewer: a fast and accurate adapter trimmer for next-generation sequencing paired-end reads. BMC Bioinformatics 15, 182 (2014).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR66\" id=\"ref-link-section-d87789345e3377\" rel=\"nofollow noopener\" target=\"_blank\">66<\/a> and read tail trimming to a total length of 50\u2009bp using Cutadapt<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 75\" title=\"Martin, M. Cutadapt removes adapter sequences from high-throughput sequencing reads. EMBnet J. 17, 10&#x2013;12 (2011).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR75\" id=\"ref-link-section-d87789345e3381\" rel=\"nofollow noopener\" target=\"_blank\">75<\/a>. The quality of sequencing data was assessed using \u2018plotFingerprint\u2019 in deepTools<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 76\" title=\"Ram&#xED;rez, F. et al. deepTools2: a next generation web server for deep-sequencing data analysis. Nucleic Acids Res. 44, W160&#x2013;W165 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR76\" id=\"ref-link-section-d87789345e3385\" rel=\"nofollow noopener\" target=\"_blank\">76<\/a>. Samples were excluded from the analysis if any of the following criteria were met: \u2018X-intercept\u2019 (deepTools)\u2009&gt;\u20090.85 or \u2018Synthetic JS Distance\u2019 (deepTools)\u2009&lt;\u20090.225. After removing duplicates and reads on blacklisted regions (ENCODE)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 77\" title=\"Amemiya, H. M., Kundaje, A. &amp; Boyle, A. P. The ENCODE blacklist: identification of problematic regions of the genome. Sci. Rep. 9, 9354 (2019).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR77\" id=\"ref-link-section-d87789345e3390\" rel=\"nofollow noopener\" target=\"_blank\">77<\/a>, peaks were called using MACS2<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 78\" title=\"Zhang, Y. et al. Model-based analysis of ChIP&#x2013;seq (MACS). Genome Biol. 9, R137 (2008).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR78\" id=\"ref-link-section-d87789345e3394\" rel=\"nofollow noopener\" target=\"_blank\">78<\/a> and a P\u2009value threshold of 1\u2009\u00d7\u200910\u20133 with an input control for each sample. For each ChIP\u2013seq (SMC1, CTCF, RPB1, H3K27ac and H3K27me3), peaks were merged for all AML samples, and recurrently identified peaks were regarded as a consensus peak set.<\/p>\n<p>SE analysis<\/p>\n<p>To identify SEs, recurrent enhancers were first identified in all AML samples using H3K27ac ChIP\u2013seq data. Identified enhancers were stitched and ranked with H3K27ac ChIP\u2013seq and input data, using ROSE<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 34\" title=\"Whyte, W. A. et al. Master transcription factors and mediator establish super-enhancers at key cell identity genes. Cell 153, 307&#x2013;319 (2013).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR34\" id=\"ref-link-section-d87789345e3411\" rel=\"nofollow noopener\" target=\"_blank\">34<\/a> with a \u2018-t 2500\u2019 parameter. Mean signals were used to calculate enhancer ranks across all AML samples. To identify SEs for each ATAC subgroup, mean signals were calculated to identify the top 750 enhancers in each ATAC subgroup, which were regarded as SEs. SEs were separated into three categories on the basis of their distribution across subgroups: common SEs (present in 12 or more subgroups), partially shared SEs (present in 2 to 11 subgroups) and unique SEs (specific to a single subgroup). Subgroup-specific SEs were defined as those found in each subgroup and categorized as partially shared or unique SEs. Known SEs for various cell types and cancers were obtained from previous reports<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 34\" title=\"Whyte, W. A. et al. Master transcription factors and mediator establish super-enhancers at key cell identity genes. Cell 153, 307&#x2013;319 (2013).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR34\" id=\"ref-link-section-d87789345e3415\" rel=\"nofollow noopener\" target=\"_blank\">34<\/a>. Annotation of SEs was done using annotatePeaks in HOMER<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 70\" title=\"Heinz, S. et al. Simple combinations of lineage-determining transcription factors prime cis-regulatory elements required for macrophage and B cell identities. Mol. Cell 38, 576&#x2013;589 (2010).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR70\" id=\"ref-link-section-d87789345e3419\" rel=\"nofollow noopener\" target=\"_blank\">70<\/a>, filtering for genes expressed in our AML cohort (logCPM\u2009&gt;\u20091 in more than 1% of patients). AML driver genes and TF genes were determined using databases, including the Catalogue of Somatic Mutations in Cancer (COSMIC) Cancer Gene Census (CGC) (as of 6 November 2024)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 79\" title=\"Sondka, Z. et al. The COSMIC Cancer Gene Census: describing genetic dysfunction across all human cancers. Nat. Rev. Cancer 18, 696&#x2013;705 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR79\" id=\"ref-link-section-d87789345e3423\" rel=\"nofollow noopener\" target=\"_blank\">79<\/a>, the database of leukaemia gene literature (dbLGL)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 80\" title=\"Liu, Y., Luo, M., Jin, Z., Zhao, M. &amp; Qu, H. dbLGL: an online leukemia gene and literature database for the retrospective comparison of adult and childhood leukemia genetics with literature evidence. Database 2018, bay062 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR80\" id=\"ref-link-section-d87789345e3427\" rel=\"nofollow noopener\" target=\"_blank\">80<\/a>, a list of TFs from a previous study<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 81\" title=\"Lambert, S. A. et al. The human transcription factors. Cell 172, 650&#x2013;665 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR81\" id=\"ref-link-section-d87789345e3432\" rel=\"nofollow noopener\" target=\"_blank\">81<\/a> and the JASPAR<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 21\" title=\"Corces, M. R. et al. Lineage-specific and single-cell chromatin accessibility charts human hematopoiesis and leukemia evolution. Nat. Genet. 48, 1193&#x2013;1203 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR21\" id=\"ref-link-section-d87789345e3436\" rel=\"nofollow noopener\" target=\"_blank\">21<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 82\" title=\"Rauluseviciute, I. et al. JASPAR 2024: 20th anniversary of the open-access database of transcription factor binding profiles. Nucleic Acids Res. 52, D174&#x2013;D182 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR82\" id=\"ref-link-section-d87789345e3439\" rel=\"nofollow noopener\" target=\"_blank\">82<\/a> database, and were manually selected. A list of SEs identified in this study is provided in Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">6<\/a>.<\/p>\n<p>SE-regulated gene signatures<\/p>\n<p>Enriched gene ontologies in SE-associated genes were identified using enricher in the clusterProfiler package<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 60\" title=\"Xu, S. et al. Using clusterProfiler to characterize multiomics data. Nat. Protoc. 19, 3292&#x2013;3320 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR60\" id=\"ref-link-section-d87789345e3454\" rel=\"nofollow noopener\" target=\"_blank\">60<\/a> (pAdjustMethod\u2009=\u2009\u201cBH\u201d, qvalueCutoff\u2009=\u20090.25) and ontology geneset (\u2018hallmark\u2019, \u2018c2.cp.reactome\u2019 and \u2018c5.go.bp\u2019) from MSigDB (v.2024.1)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 61\" title=\"Subramanian, A. et al. Gene set enrichment analysis: a knowledge-based approach for interpreting genome-wide expression profiles. Proc. Natl Acad. Sci. USA 102, 15545&#x2013;15550 (2005).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR61\" id=\"ref-link-section-d87789345e3458\" rel=\"nofollow noopener\" target=\"_blank\">61<\/a>. Expression levels of SE-associated genes in each haematopoietic cell type were calculated using public gene-expression data from DMAP<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 83\" title=\"Novershtern, N. et al. Densely interconnected transcriptional circuits control cell states in human hematopoiesis. Cell 144, 296&#x2013;309 (2011).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR83\" id=\"ref-link-section-d87789345e3462\" rel=\"nofollow noopener\" target=\"_blank\">83<\/a>, and the average expression levels were determined for each cell type.<\/p>\n<p>SE-based TF network analysis<\/p>\n<p>To analyse SE-regulated TF networks, we applied the Coltron software<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 36\" title=\"Ott, C. J. et al. Enhancer architecture and essential core regulatory circuitry of chronic lymphocytic leukemia. Cancer Cell 34, 982&#x2013;995 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR36\" id=\"ref-link-section-d87789345e3475\" rel=\"nofollow noopener\" target=\"_blank\">36<\/a> with ROSE<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 34\" title=\"Whyte, W. A. et al. Master transcription factors and mediator establish super-enhancers at key cell identity genes. Cell 153, 307&#x2013;319 (2013).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR34\" id=\"ref-link-section-d87789345e3479\" rel=\"nofollow noopener\" target=\"_blank\">34<\/a> outputs generated from the merged\u00a0H3K27ac BAM files for each ATAC subgroup. This software computes the inward binding (in-degree) of other SE-associated TFs to a given SE-associated TF, as well as the outward binding (out-degree) of the TF to other SEs. The Coltron score for each TF was determined as the sum of its in-degree and out-degree (Extended Data Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#Fig11\" rel=\"nofollow noopener\" target=\"_blank\">6c<\/a>).<\/p>\n<p>scRNA\/ATAC-seq experiments<\/p>\n<p>Single-cell matched RNA-seq and ATAC-seq experiments were performed using the Next GEM Single Cell Multiome ATAC + Gene Expression Reagent Kit (10x Genomics), according to the manufacturer\u2019s protocols (CG000365 for nuclei isolation and CG000338 for library generation). Cryopreserved cells were thawed, dead cells were stained with DAPI and live mononuclear cells were sorted using FACS Aria III (BD Biosciences). The samples were resuspended in lysis buffer (10\u2009mM Tris-HCl (pH\u20097.4), 10\u2009mM NaCl, 3\u2009mM MgCl2, 1% bovine serum albumin (BSA; Miltenyi Biotec, 130-091-376), 0.1% Tween-20 (Bio-Rad, 1610781), 0.1% IGEPAL (Sigma-Aldrich, i8896), 0.01% digitonin (Thermo Fisher Scientific, BN2006) and 1\u2009mM DTT (Sigma-Aldrich, 646563), plus 1\u2009U\u2009\u03bcl\u22121 RNase inhibitor (Thermo Fisher Scientific, 10777019)), incubated on ice for 3\u2009min, washed three times with wash buffer (10\u2009mM Tris-HCl (pH\u20097.4), 10\u2009mM NaCl, 3\u2009mM MgCl2, 1% BSA, 0.1% Tween-20 and 1\u2009mM DTT, plus 1\u2009U\u2009\u2009\u03bcl\u22121 RNase inhibitor) and passed through a 40-\u03bcm Flowmi Cell Strainer (Bel-Art). After microscopy inspection and counting of nuclei using a Countess II FL Automated Cell Counter (Thermo Fisher Scientific), nucleus suspensions were prepared in a concentration targeting a maximum of 10,000 nuclei recovery, and incubated with transposase to add adapter sequences to the DNA fragments. The suspensions containing transposed nuclei were subjected to gel bead in emulsion (GEM) generation, incubation, and clean-up, using the Chromium Next GEM Chip J Single Cell Kit (10x Genomics) and Chromium Controller. The resulting suspensions contained ATAC fragments and cDNA with the same cell barcodes. Pre-amplification of cDNA was performed, and the amplified product was split and used as the input for ATAC and gene-expression library construction. Libraries were generated using the 10x Genomics Single Index N Set for ATAC and the 10x Genomics Dual Index TT Set A for RNA. The scATAC libraries were sequenced using the DNBSEQ-G400 (MGI) with a custom protocol (read 1: 50 cycles, read 2: 49 cycles, i5 Index: 24 cycles, i7 Index: 8 cycles). The scRNA libraries were sequenced using the DNBSEQ-G400 (MGI) with a custom protocol (read 1: 28 cycles, read 2: 90 cycles).<\/p>\n<p>scRNA\/ATAC-seq analysis<\/p>\n<p>Cell Ranger ARC (10x Genomics)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 84\" title=\"Zheng, G. X. Y. et al. Massively parallel digital transcriptional profiling of single cells. Nat. Commun. 8, 14049 (2017).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR84\" id=\"ref-link-section-d87789345e3510\" rel=\"nofollow noopener\" target=\"_blank\">84<\/a> was used for data processing to generate BAM files and count matrices for scRNA and scATAC with the reference genome and GENCODE human v.19 annotation<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 85\" title=\"Frankish, A. et al. GENCODE: reference annotation for the human and mouse genomes in 2023. Nucleic Acids Res. 51, D942&#x2013;D949 (2023).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR85\" id=\"ref-link-section-d87789345e3514\" rel=\"nofollow noopener\" target=\"_blank\">85<\/a>. The report generated by Cell Ranger was manually evaluated and samples were filtered to remain with no errors or with only warnings. Basic quality-control reports including analysed cell numbers from Cell Ranger for each sample are summarized in Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM3\" rel=\"nofollow noopener\" target=\"_blank\">7<\/a>. ArchR<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 86\" title=\"Granja, J. M. et al. ArchR is a scalable software package for integrative single-cell chromatin accessibility analysis. Nat. Genet. 53, 403&#x2013;411 (2021).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR86\" id=\"ref-link-section-d87789345e3521\" rel=\"nofollow noopener\" target=\"_blank\">86<\/a> was used to filter for cells with the following thresholds: number of unique molecular identifiers\u2009&gt;\u2009100 and proportion of mitochondria genes\u2009&lt;\u20090.05 for scRNA; TSS enrichment score\u2009&gt;\u20094 and number of fragments\u2009&gt;\u20091,000 for scATAC. Doublet cells were removed by the \u2018filterDoublets\u2019 function in ArchR with a filterRatio of 1. The gene-expression count matrix was stored in the Seurat<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 87\" title=\"Hao, Y. et al. Dictionary learning for integrative, multimodal and scalable single-cell analysis. Nat. Biotechnol. 42, 293&#x2013;304 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR87\" id=\"ref-link-section-d87789345e3525\" rel=\"nofollow noopener\" target=\"_blank\">87<\/a> object, and mitochondrial genes were excluded from the following analysis. The consensus peak set generated from bulk ATAC-seq was filtered for peaks on autosomal chromosomes and used to count reads on peaks in the scATAC analysis, using the FeatureMatrix function from the Signac package<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 88\" title=\"Stuart, T., Srivastava, A., Madad, S., Lareau, C. A. &amp; Satija, R. Single-cell chromatin state analysis with Signac. Nat. Methods 18, 1333&#x2013;1341 (2021).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR88\" id=\"ref-link-section-d87789345e3530\" rel=\"nofollow noopener\" target=\"_blank\">88<\/a>.<\/p>\n<p>Merging individual data, normalization and clustering for scRNA\/ATAC-seq<\/p>\n<p>Raw count matrix data for single-cell gene expression and chromatin accessibility was written to disk using the BPCells package (<a href=\"https:\/\/github.com\/bnprks\/BPCells\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/bnprks\/BPCells<\/a>) to enable high-throughput data processing and merged across samples. The merged gene-count matrix was normalized for sequencing depth using the SCTransform function in Seurat<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 87\" title=\"Hao, Y. et al. Dictionary learning for integrative, multimodal and scalable single-cell analysis. Nat. Biotechnol. 42, 293&#x2013;304 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR87\" id=\"ref-link-section-d87789345e3549\" rel=\"nofollow noopener\" target=\"_blank\">87<\/a>. scRNA-based clustering was done using the first 50 principal component analysis (PCA) dimensions with the FindNeighbors and FindClusters functions (resolution\u2009=\u20090.15, algorithm\u2009=\u20091). scRNA UMAP was computed using the first 50 PCA dimensions with RunUMAP (n.neighbours\u2009=\u200930, min.dist\u2009=\u20090.1). The merged count matrix for scATAC was normalized using the term frequency inverse document frequency (TF-IDF) normalization function in BPCells. Cells were clustered on the basis of scATAC using the PCA dimensions 2 to 20 with knn_hnsw (ef\u2009=\u20093000), knn_to_snn_graph and cluster_graph_louvain (resolution\u2009=\u20090.15) functions in BPCells. Each scATAC cluster was classified as AML-predominant if it met both of the following criteria: (i) the average proportion of cells derived from remission samples was less than 10%; and (ii) at least one AML subgroup contributed 25% or more of its cells to the cluster. Clusters not satisfying these criteria were classified as normal-mixed. scATAC UMAP was computed using the PCA dimensions 2 to 20 with the umap function in the uwot package (n.neighbours\u2009=\u200930, min.dist\u2009=\u20090.1).<\/p>\n<p>Estimation of differentiation, pseudotime and LSC scores using scRNA-seq<\/p>\n<p>The BoneMarrowMap package in R was used to project AML cells onto reference scRNA-seq data of human bone-marrow haematopoiesis<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 41\" title=\"Zeng, A. G. X. et al. Single-cell transcriptional atlas of human hematopoiesis reveals genetic and hierarchy-based determinants of aberrant AML differentiation. Blood Cancer Discov. 6, 307&#x2013;324 (2025).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR41\" id=\"ref-link-section-d87789345e3561\" rel=\"nofollow noopener\" target=\"_blank\">41<\/a>. Cells with low mapping quality were excluded from the analysis according to the following criteria: (1) mean absolute deviation of mapping error scores\u2009\u2267\u20092; (2) assigned to \u2018orthochromatic erythroblast\u2019 and AUC score of haemoglobin genes\u2009&lt;\u20090.2, calculated using the AUCell package. Pseudotime analysis was performed using the \u2018predict_Pseudotime\u2019 function in BoneMarrowMap. The LSC score was computed as the AUC for gene set enrichment of LSC signatures, defined by DEGs specific to sorted LSC+ fractions<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 41\" title=\"Zeng, A. G. X. et al. Single-cell transcriptional atlas of human hematopoiesis reveals genetic and hierarchy-based determinants of aberrant AML differentiation. Blood Cancer Discov. 6, 307&#x2013;324 (2025).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR41\" id=\"ref-link-section-d87789345e3567\" rel=\"nofollow noopener\" target=\"_blank\">41<\/a>,<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 42\" title=\"Ng, S. W. K. et al. A 17-gene stemness score for rapid determination of risk in acute leukaemia. Nature 540, 433&#x2013;437 (2016).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR42\" id=\"ref-link-section-d87789345e3570\" rel=\"nofollow noopener\" target=\"_blank\">42<\/a>.<\/p>\n<p>Single-cell GRN analysis<\/p>\n<p>TF activity was inferred using SCENIC+<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 43\" title=\"Bravo Gonz&#xE1;lez-Blas, C. et al. SCENIC+: single-cell multiomic inference of enhancers and gene regulatory networks. Nat. Methods 20, 1355&#x2013;1367 (2023).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR43\" id=\"ref-link-section-d87789345e3582\" rel=\"nofollow noopener\" target=\"_blank\">43<\/a>, which integrates matched scRNA-seq and scATAC-seq data to construct GRNs at single-cell resolution. The analysis followed the standard SCENIC+ pipeline with minor modifications (Extended Data Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#Fig13\" rel=\"nofollow noopener\" target=\"_blank\">8a<\/a>). We used a multiome mode with 5 cells per metacell, and empirically set the number of topics to 100. The search space for regulatory elements was defined as 0\u2013500\u2009kb from each gene. Established eRegulons were filtered with the following parameters: \u2018rho_threshold\u2019\u2009=\u20090.03, \u2018min_regions_per_gene\u2019\u2009=\u20090 and \u2018min_target_genes\u2019\u2009=\u200910. For group-level comparisons, cells were downsampled to 2,000 per ATAC subgroup when calculating region-based and gene-based specificity scores. For pseudotime analysis, cells were downsampled to 300 per pseudotime bin in each ATAC subgroup if more than 300 cells were present in that bin. Direct positive eRegulons (those with positive TF-to-gene and region-to-gene links) were used for downstream analyses. Extended eRegulons were included only when direct ones were unavailable.<\/p>\n<p>Drug sensitivity screening in samples from patients with AML<\/p>\n<p>The procedure for drug sensitivity and resistance testing in the samples used for the study has been described in detail previously<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 46\" title=\"Struyf, N. et al. Delineating functional and molecular impact of ex vivo sample handling in precision medicine. NPJ Precis. Oncol. 8, 38 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR46\" id=\"ref-link-section-d87789345e3598\" rel=\"nofollow noopener\" target=\"_blank\">46<\/a>. Biobanked mononuclear cells were thawed and added to pre-spotted drug plates (FIMM HTB)<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 89\" title=\"Malani, D. et al. Implementing a functional precision medicine tumor board for acute myeloid leukemia. Cancer Discov. 12, 388&#x2013;401 (2022).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR89\" id=\"ref-link-section-d87789345e3602\" rel=\"nofollow noopener\" target=\"_blank\">89<\/a>. The mononuclear cells were incubated in HS-5 conditioned (12.5%; ATCC) complete RPMI (10% FBS (Thermo Fisher Scientific), 2 mM l-glutamine (Sigma-Aldrich), 100 IU\u2009ml\u22121 penicillin and 0.1\u2009mg\u2009ml\u22121 streptomycin (Pen-Strep; Sigma-Aldrich)), for 72\u2009h (37\u2009\u00b0C, 5% CO2). Cell viability was measured by CellTiter-Glo (CTG; Promega) on an EnSight plate reader (PerkinElmer). Drug sensitivity scores (DSSs) were calculated with Breeze<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 90\" title=\"Potdar, S. et al. Breeze 2.0: an interactive web-tool for visual analysis and comparison of drug response data. Nucleic Acids Res. 51, W57&#x2013;W61 (2023).\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#ref-CR90\" id=\"ref-link-section-d87789345e3616\" rel=\"nofollow noopener\" target=\"_blank\">90<\/a>. Selective DSS (sDSS) values were calculated by subtracting the DSSs of healthy bone-marrow control samples from the DSSs of samples from patients with AML.<\/p>\n<p>Survival analysis<\/p>\n<p>Survival analysis was performed for patients who were treated with standard intensive chemotherapy, and observations were censored at the last follow-up. For clinical parameters, the median values were used as the threshold unless otherwise specified. Overall survival was estimated using the Kaplan\u2013Meier method, and differences between groups were evaluated by the log-rank test.<\/p>\n<p>To assess the prognostic relevance of deconvolution patterns, we performed feature selection using LASSO-penalized Cox proportional hazards regression (glmnet R package, alpha\u2009=\u20091), incorporating the estimated proportions of 13 haematopoietic cell populations derived from ATAC-seq\u2013based deconvolution as explanatory variables, with overall survival as the outcome. The optimal penalty parameter (lambda.min) was selected by cross-validation. The resulting coefficients were used to compute a LASSO-based risk score for each patient, defined as the weighted sum of the corresponding cell fractions.<\/p>\n<p>To determine whether ATAC subgroups provided extra prognostic information beyond established genomic risk categories, multivariable Cox proportional hazards models were fitted separately within each ELN 2022 risk group. These models included age, sex, WBC, haemoglobin level, platelet count, blast percentage and ATAC subgroup membership as covariates. Subgroups that were significantly associated with worse overall survival within each ELN stratum were designated as high-risk ATAC subgroups. We stratified patients in each ELN category into two groups according to the presence or absence of these risk ATAC subgroups and compared overall survival in Kaplan\u2013Meier analyses. To evaluate the additive prognostic value of ATAC subgroups, we built a Cox model for ELN 2022 risk classification and compared it with a combined model of ELN and ATAC risk groups (Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#Fig5\" rel=\"nofollow noopener\" target=\"_blank\">5b<\/a>). Model performance was assessed by calculating the C-index using 1,000 bootstrap replicates in both the training (Swedish) and the validation (Japanese) cohort. The incremental prognostic value of ATAC subgroups was quantified as the difference in C-index between models, assessed across the same bootstrap replicates.<\/p>\n<p>Statistical analysis<\/p>\n<p>Experimental replication was not performed. The robustness of the findings is supported by the large cohort size (n\u2009=\u20091,563) and the consistency of results across several independent analytical approaches, data types and independent cohorts. Statistical analyses were performed in R. Comparisons between groups were based on the two-sided Wilcoxon rank-sum test for continuous data and the Fisher\u2019s exact test for categorical data, unless otherwise specified. Explained variances in several features by ICC, WHO and ATAC classifications were evaluated as R2 values calculated by the adonis function in R, with Euclidean distances. For gene-expression profiles, the first 50 principal components were used as input.<\/p>\n<p>Reporting summary<\/p>\n<p>Further information on research design is available in the\u00a0<a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41586-026-10703-4#MOESM2\" rel=\"nofollow noopener\" target=\"_blank\">Nature Portfolio Reporting Summary<\/a> linked to this article.<\/p>\n","protected":false},"excerpt":{"rendered":"Patients and samples This study was reviewed and approved by the regional ethics review board in Stockholm (2008\/1330-31\/2&hellip;\n","protected":false},"author":2,"featured_media":752126,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[32],"tags":[104222,322530,157070,57337,1159,1160,79],"class_list":["post-752125","post","type-post","status-publish","format-standard","has-post-thumbnail","category-science","tag-acute-myeloid-leukaemia","tag-cancer-epigenetics","tag-classification-and-taxonomy","tag-epigenomics","tag-humanities-and-social-sciences","tag-multidisciplinary","tag-science"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/752125","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/comments?post=752125"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/752125\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media\/752126"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media?parent=752125"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/categories?post=752125"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/tags?post=752125"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}