{"id":718164,"date":"2026-07-30T00:25:23","date_gmt":"2026-07-30T00:25:23","guid":{"rendered":"https:\/\/www.newsbeep.com\/uk\/718164\/"},"modified":"2026-07-30T00:25:23","modified_gmt":"2026-07-30T00:25:23","slug":"paired-mutation-calling-and-spatial-transcriptomics-identify-cellular-neighborhoods-associated-with-the-neoplastic-outcome-of-mouse-colitis","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/uk\/718164\/","title":{"rendered":"Paired mutation calling and spatial transcriptomics identify cellular neighborhoods associated with the neoplastic outcome of mouse colitis"},"content":{"rendered":"<p>Mice<\/p>\n<p>Mice were of C57BL\/6 background. The Muc2KO line used was described by Velcich et al.<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 16\" title=\"Velcich, A. et al. Colorectal cancer in mice genetically deficient in the mucin Muc2. Science 295, 1726&#x2013;1729 (2002).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR16\" id=\"ref-link-section-d54364004e2186\" rel=\"nofollow noopener\" target=\"_blank\">16<\/a>. Mice containing inter-crossed R26Confetti and VillinCreERt2 alleles<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 18\" title=\"Snippert, H. J. et al. Intestinal crypt homeostasis results from neutral competition between symmetrically dividing Lgr5 stem cells. Cell 143, 134&#x2013;144 (2010).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR18\" id=\"ref-link-section-d54364004e2203\" rel=\"nofollow noopener\" target=\"_blank\">18<\/a> were bred to Muc2KO mice to generate VillinCreERt2; Muc2KO; R26Confetti mice. To model sporadic colon carcinogenesis, the intestinal epithelium-specific inducible Cre VillinCreERt2 line was crossed with Apcfl\/+(ref. <a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 58\" title=\"Colnot, S. et al. Colorectal cancers in a new mouse model of familial adenomatous polyposis: influence of genetic and environmental modifiers. Lab. Invest. 84, 1619&#x2013;1630 (2004).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR58\" id=\"ref-link-section-d54364004e2234\" rel=\"nofollow noopener\" target=\"_blank\">58<\/a>). All genotyping was outsourced to Transnetyx.<\/p>\n<p>Animal husbandry<\/p>\n<p>Animal care and procedures were performed at the Cancer Research UK Cambridge Institute Biological Resource Unit according to the UK Home Office under the authority of a Home Office project licence (PD5F099BE) approved by the Animal Welfare and Ethical Review Body at the CRUK Cambridge Institute, University of Cambridge. Mice were housed under controlled conditions (temperature (21\u2009\u00b1\u20092\u2009\u00b0C), humidity (55%\u2009\u00b1\u200910%), 12\u2009h light\u2013dark cycle) in a specific-pathogen-free facility (tested according to the recommendations for health monitoring by the Federation of European Laboratory Animal Science Associations). Animals had unrestricted access to food and water. None of the mice had been involved in any procedure before the study. To trigger acute inflammation, 2% DSS was provided in drinking water for 5\u2009days. To trigger colon carcinogenesis, 200\u2009mg\u2009kg\u22121 ENU was injected intraperitoneally. Mice which showed clinical signs of tumor burden (anemia, hunching and loss of body condition) before experimental endpoints were culled and removed from the study.<\/p>\n<p>Human tissue<\/p>\n<p>Colon tissue samples were obtained from patients with IBD and\/or CAC at St James University Hospital Leeds under full local research ethical committee approval (12\/LO\/1217 approved by London\u2013Bloomsbury Research Ethics Committee) according to the Health Research Authority institution. Patient consent was obtained at the time of surgery, and they did not receive compensation. Colectomy specimens were fixed in 10% neutral buffered formalin and embedded en face in paraffin (FFPE) blocks.<\/p>\n<p>Three-dimensional imaging and clone counting<\/p>\n<p>The whole colon of VillinCreERt2; Muc2KO; R26Confetti mice was dissected, flushed with cold PBS, cut longitudinally and whole-mounted. Following fixation in 4% paraformaldehyde overnight at 4\u2009\u00b0C, the tissue was washed in PBS and selected colon segments were excised. Optical clearing was performed using the CUBIC protocol<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 59\" title=\"Susaki, E. A. et al. Whole-brain imaging with single-cell resolution using chemical cocktails and computational analysis. Cell 157, 726&#x2013;739 (2014).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR59\" id=\"ref-link-section-d54364004e2273\" rel=\"nofollow noopener\" target=\"_blank\">59<\/a>. In brief, excised segments were incubated with CUBIC-1a solution (10% urea, 5% N,N,N\u2032,N\u2032-tetrakis(2-hydroxypropyl) ethyl-enediamine, 10% Triton X-100 and 25\u2009mM NaCl in distilled water) at 37\u2009\u00b0C for 7\u201310\u2009days with alternate day solution changes. 4\u2032,6-diamidino-2-phenylindole\u2009(DAPI) was used for nuclear counterstaining at a dilution of 1:1,000. The cleared tissue was then washed in PBS for 24\u2009h. Additional clearing and refractive index matching were performed with Rapiclear 1.52 (152002, SunJin Labs) for 24\u2009h. Samples were mounted in a 0.25-mm i-Spacer (Sunjin Labs) for confocal imaging on a Leica SP5 TCS confocal microscope (LAS software v2.8.0, Leica Biosystems) with a 10\u00d7 objective, 1.4\u20131.7 optical zoom and 8\u201312\u2009\u03bcm z-steps throughout the whole thickness of the tissue. Clone counting was performed using the cell counter tool in the ImageJ software.<\/p>\n<p>Immunohistochemistry<\/p>\n<p>Mouse colons were opened longitudinally and cut in approximately 15-mm2 sections placed in cassettes and fixed overnight at 4\u2009\u00b0C in 4% paraformaldehyde. H&amp;E staining was performed using an automated ST5020 Multistainer (Leica Biosystems). Staining for \u03b2-catenin was performed on Leica\u2019s automated Bond-III platform in conjunction with the Polymer Refine Detection System (DS9800, Leica Biosystems). In brief, epitope retrieval was performed using Leica Epitope Retrieval Solution 1 (AR9961, Leica Biosystems) at 100\u2009\u00b0C. Blocking was performed with Protein Block Buffer (X090930-2, Dako). Following incubation with primary antibody against \u03b2-catenin (0.25\u2009\u03bcg\u2009ml\u22121, mouse, 610154, BD Biosciences), sections were incubated with secondary antibody (1:1,500, rabbit anti-mouse IgG1, ab125913, Abcam) before development and mounting. For \u03b2-catenin staining, an additional mouse-on-mouse blocking step was performed.<\/p>\n<p>Immunofluorescence<\/p>\n<p>Heat-mediated epitope retrieval was performed on rehydrated 3-\u03bcm-thick paraffin sections in 10\u2009mM sodium citrate (pH 6.0). The sections were then incubated in blocking solution (10% donkey serum and 0.05% Tween-20 in PBS) at room temperature for 30\u2009min. Primary antibodies against RFP (1:100, rabbit, <a href=\"https:\/\/www.ncbi.nlm.nih.gov\/nuccore\/R10367\" rel=\"nofollow noopener\" target=\"_blank\">R10367<\/a>, Thermo Fisher), GFP (1:100, chicken, ab13970, Abcam), E-cadherin (1:200, mouse, 610182, BD Biosciences), Trop2 (1:100, goat, AF1122, R&amp;D systems) and Arid1a (1:100, rabbit, 12354S, Cell Signaling) were diluted in blocking solution, in which sections were then incubated in the dark overnight at 4\u2009\u00b0C. Sections were washed and incubated with fluorophore-conjugated secondary antibodies (donkey anti-rabbit <a href=\"https:\/\/www.ncbi.nlm.nih.gov\/nuccore\/A31572\" rel=\"nofollow noopener\" target=\"_blank\">A31572<\/a>, donkey anti-goat <a href=\"https:\/\/www.ncbi.nlm.nih.gov\/nuccore\/A21447\" rel=\"nofollow noopener\" target=\"_blank\">A21447<\/a> and\/or donkey anti-chicken 703-6-5-155, Thermo Fisher and\/or donkey anti-mouse ab150109, Abcam) diluted 1:200 in 0.05% Tween-20 in PBS for 45\u2009min at room temperature. DAPI (1:1,000) was used for nuclear counterstaining. After washing, the stained sections were mounted using ProLong Gold Antifade Mountant (<a href=\"https:\/\/www.uniprot.org\/uniprot\/P36930\" rel=\"nofollow noopener\" target=\"_blank\">P36930<\/a>, Thermo Fisher).<\/p>\n<p>RNAscope<\/p>\n<p>Simultaneous detection of Notum and Reg4 was performed on paraffin embedded sections using Advanced Cell Diagnostics (ACD) RNAscope 2.5 LS Duplex Reagent Kit (322440), RNAscope 2.5 LS Probe- Mm- Notum C1 (428988) and RNAscope 2.5 LS Probe-Mm-Reg4-C2 (409608). The 3-\u00b5m-thick sections were baked for 1\u2009h at 60\u2009\u00b0C before loading onto a Bond RX instrument (Leica Biosystems). Slides were deparaffinized and rehydrated on board before pre-treatments using Epitope Retrieval Solution 2 (AR9640, Leica Biosystems) at 95\u2009\u00b0C for 15\u2009min and ACD Enzyme from the Duplex Reagent kit at 40\u2009\u00b0C for 15\u2009min. Probe hybridization and signal amplification were performed according to manufacturer\u2019s instructions. Fast red detection of C2 was performed on the Bond Rx using the Bond Polymer Refine Red Detection Kit (DS9390, Leica Biosystems) according to ACD protocol. Slides were then removed from the Bond Rx and detection of the C1 signal was performed using the RNAscope 2.5 LS Green Accessory Pack (322550, ACD) according to kit instructions. Slides were heated at 60\u2009\u00b0C for 1\u2009h, dipped in xylene and mounted using VectaMount Permanent Mounting Medium (H-5000, Vector Laboratories). The slides were imaged on the Aperio AT2 (Leica Biosystems) to create whole slide images. Images were captured at 40\u00d7 magnification, with a resolution of 0.25\u2009\u03bcm per pixel.<\/p>\n<p>Inference of effective fission rate from lineage tracing data<\/p>\n<p>The statistical model for crypt fission described in Nicholson et al.<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 22\" title=\"Nicholson, A. M. et al. Fixation and spread of somatic mutations in adult human colonic epithelium. Cell Stem Cell 22, 909&#x2013;918 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR22\" id=\"ref-link-section-d54364004e2364\" rel=\"nofollow noopener\" target=\"_blank\">22<\/a> and implemented in the R package RHClones (<a href=\"https:\/\/github.com\/ElEd2\/RHClones\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/ElEd2\/RHClones<\/a>) was adapted to infer effective fission rates using counts of clone sizes from n\u2009=\u20093 equal size tissue areas from n\u2009=\u20093 Muc2het and n\u2009=\u20093 Muc2hom mice with different average Trop2 expression levels, selected at random. Clone sizes were counted manually using the ImageJ cell counter tool, then the number of pixels stained with Trop2 was obtained using the Area tool in the measure function in ImageJ. Each tissue region was assigned a local fission rate \\({\\rho }_{i}\\) (units, per day) that is sampled from a hierarchical Student\u2019s T prior<\/p>\n<p>$${\\rho }_{i}\\, \\sim \\,{StudentT}\\left(\\nu ,\\mu ,\\sigma \\right)$$<\/p>\n<p>with population-level parameters<\/p>\n<p>$$\\sigma \\,\\approx \\,\\mathrm{normal}\\left(0,{10}^{-2}\\right)$$<\/p>\n<p>$$\\sigma \\,\\approx \\,\\mathrm{normal}\\,\\left(0,{10}^{-2}\\right)$$<\/p>\n<p>$$\\nu \\,\\approx \\,\\mathrm{gamma}\\left(2,{10}^{-1}\\right).$$<\/p>\n<p>The vector of observed clone sizes, \\({{\\boldsymbol{g}}}_{i}\\), for region i is then distributed according to<\/p>\n<p>$${{\\boldsymbol{g}}}_{i}\\, \\sim \\,{Multinomial}\\left({\\boldsymbol{f}}({\\rho }_{i},{t}_{i})\\right)$$<\/p>\n<p>where<\/p>\n<p>$${{\\bf{f}}}_{n}\\left(\\rho ,t\\right)={{\\rm{e}}}^{-\\rho t}{(1-{{\\rm{e}}}^{-\\rho t})}^{n-1}$$<\/p>\n<p>is the solution of the Yule\u2013Furry pure birth process, and \\({t}_{i}\\) is the time (in days) since clone induction in region i.<\/p>\n<p>The model for fission inference in this study differs from the model for fission after continuous clone induction described in Nicholson et al.<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 22\" title=\"Nicholson, A. M. et al. Fixation and spread of somatic mutations in adult human colonic epithelium. Cell Stem Cell 22, 909&#x2013;918 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR22\" id=\"ref-link-section-d54364004e2861\" rel=\"nofollow noopener\" target=\"_blank\">22<\/a>, where this solution is integrated over the age of the individual. However, as in the aforementioned study, a correction was applied to \\({f}_{1}\\) and \\({f}_{2}\\) to correct for the chance occurrence of neighboring, nonclonally related crypts. This model does not incorporate crypt fusion, nor does it account for tissue remodeling that may partly explain clone sizes in tissue areas associated with repair following inflammation. In this respect, inferred fission rates should be interpreted as \u2018effective\u2019 in the sense that they only capture the value of \\({\\rho }_{i}\\) that is required to explain clone sizes using the Yule\u2013Furry pure birth process introduced in Nicholson et al.<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 22\" title=\"Nicholson, A. M. et al. Fixation and spread of somatic mutations in adult human colonic epithelium. Cell Stem Cell 22, 909&#x2013;918 (2018).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR22\" id=\"ref-link-section-d54364004e2934\" rel=\"nofollow noopener\" target=\"_blank\">22<\/a>.<\/p>\n<p>Simulations of crypt dynamics<\/p>\n<p>Crypt dynamics were simulated using an on-lattice, voter-type model. The square lattice is of shape \\(M\\times M\\). For a site \\((i,{j})\\) on the lattice, the neighboring sites are defined to be the set<\/p>\n<p>$${N}_{i,j}\\,=\\left\\{\\left(i,\\,j\\,+\\,1\\right),\\,\\left(i,j-\\,1\\right),\\,\\left(i\\,+\\,1,\\,j\\right),\\,\\left(i-1,j\\right),\\,\\left(i\\,+\\,1,\\,j\\,+\\,1\\right),\\,\\left(i\\,-\\,1,\\,j-1\\right)\\right\\}$$<\/p>\n<p>resembling a configuration equivalent to a hexagonal lattice in two dimensions with periodic boundary conditions. The state of a lattice site \\((i,{j})\\), denoted by \\({s}_{i,j}\\), is defined as \\({s}_{i,j}=-1\\) for unlabeled (UL) sites, \\({s}_{i,j}=0\\) for empty sites and \\({s}_{i,j}=+1\\) for labeled (L) sites. Assuming a labeling efficiency of 20%, the initial population on the lattice consists of 80% of UL and 20% of L sites randomly assigned to each site. The parameters governing the simulations are the fission (\\({\\rho }_{s}\\)), fusion (\\({f}_{s}\\)) and fixation (\\({P}_{s}\\)) probabilities of the state \\(s=\\pm 1\\). All are assumed to be neutral and equal for L and UL crypts as the baseline scenario (\\({\\rho }_{s}\\), \\({f}_{s},\\,{{P}}_{s}=\\,0.5\\) for \\(s=\\,\\pm 1\\), undefined otherwise). To determine the impact of fission bias in tissue repair, the fission probability of L sites was increased to \\({\\rho }_{1}=0.95\\), with all other parameters remaining the same.<\/p>\n<p>The simulation rules are that, at each time step, we choose a random site \\(\\left(i,{j}\\right)\\) and update the lattice state according to<\/p>\n<p>                  1.<\/p>\n<p>If \\(\\left(i,\\,j\\right)\\) is an empty site \\(\\left({s}_{i,j}=\\,0\\right):\\)<\/p>\n<p>                  i.<\/p>\n<p>All neighbors are empty sites \\(({s}_{k,l}\\,=\\,0\\,\\forall \\,\\left(k,{l}\\right)\\in \\,{N}_{i,j}\\,),\\) do nothing;<\/p>\n<p>                  ii.<\/p>\n<p>Otherwise, choose a nonempty neighbor site \\(\\left(k,{l}\\right)\\in {N}_{i,{j}}:\\,{s}_{k,{l}}\\ne 0\\) at random to undergo fission by setting \\({s}_{i,{j}}={s}_{k,l}\\) with probability \\({\\rho }_{{s}_{k,l}}\\) (Fig.\u2009<a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#Fig1\" rel=\"nofollow noopener\" target=\"_blank\">1a<\/a> and Supplementary Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#MOESM1\" rel=\"nofollow noopener\" target=\"_blank\">1<\/a>).<\/p>\n<p>                  2.<\/p>\n<p>If \\((i,{j})\\) is nonempty\\(\\,{(s}_{i,j}\\ne 0)\\), then choose a neighbor \\(\\left(k,{l}\\right)\\in {N}_{i,{j}}\\) at random:<\/p>\n<p>                  i.<\/p>\n<p>If \\(\\left(k,{l}\\right)\\) is empty, then \\((i,{j})\\) undergoes fission by setting \\({s}_{k,l}={s}_{i,j}\\) with probability \\({\\rho }_{{s}_{i,j}}\\);<\/p>\n<p>                  ii.<\/p>\n<p>Otherwise, fusion occurs with probability \\(\\frac{{f}_{{s}_{i,j}}+{f}_{{s}_{k,l}}}{2}\\) by choosing \\(\\left(h,{w}\\right)\\in \\{\\left(i,j\\right),(k,l)\\}\\) with equal probability and setting \\({s}_{h,w}=0\\). The outcome of monoclonal fixation following fusion is then determined to be the state \\(s{\\prime}\\) according to the fixation probabilities \\({P}_{{s}_{i,j}}\\) and \\({P}_{{s}_{k,l}}\\) as outlined in Fig.\u2009<a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#Fig1\" rel=\"nofollow noopener\" target=\"_blank\">1b,c<\/a> and Supplementary Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#MOESM1\" rel=\"nofollow noopener\" target=\"_blank\">1<\/a> and then setting \\({s}_{p,q}=s{\\prime}\\) where \\(\\left(p,{q}\\right)\\in \\{\\left(i,j\\right),\\left(k,l\\right)\\}\\backslash \\{\\left(h,{w}\\right)\\}\\).<\/p>\n<p>We note that our definition of homeostasis in in silico models is not equivalent to an equilibrium since, while the average numbers of labeled and unlabeled crypts remain constant, their spatial distribution across the lattice undergoes domain coarsening as captured by a \u2018site aggregation\u2019 metric that slowly increases across the course of simulations (Fig. <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"figure anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#Fig1\" rel=\"nofollow noopener\" target=\"_blank\">1f,l<\/a>). Furthermore, fission-biased conditions lead to a nonhomeostatic system even in the absence of damage, as higher expansion rates generally outpace the creation of empty space by crypt fusion. This is reminiscent of a regime where additional mechanisms, such as diffusion (not considered here), may be required to spatially accommodate mutant crypts with higher rates of fission (Olpe et al.<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 21\" title=\"Olpe, C. et al. A diffusion-like process enables expansion of advantaged gene mutations in human colonic epithelium. Gastroenterology 161, 548&#x2013;559 (2021).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR21\" id=\"ref-link-section-d54364004e4665\" rel=\"nofollow noopener\" target=\"_blank\">21<\/a>).<\/p>\n<p>Calculation of site aggregation<\/p>\n<p>The number of distinct neighbor pairs (DNPs) was quantified to represent aggregation of labeled and unlabeled sites in the lattice. The number of the DNPs for a \\((i,{j})\\) on the lattice is defined to be the cardinality of the set<\/p>\n<p>$${G}_{i,j}=\\left\\{\\left({s}_{i,j}\\,,\\,{s}_{k,l}\\right):\\left(k,\\,l\\right)\\in \\,{N}_{i,j}\\,,\\,{s}_{i,j}\\,\\ne \\,{s}_{k,l},\\,{s}_{i,j}\\,\\ne 0,\\,{s}_{k,l}\\ne 0\\right\\}.$$<\/p>\n<p>The neighbor pairs of a site (i, j) is the set:<\/p>\n<p>$${K}_{i,j}=\\left\\{\\left({s}_{i,j}\\,,\\,{s}_{k,l}\\right):\\left(k,\\,l\\right)\\in \\,{N}_{i,j}\\,,\\,{s}_{i,j}\\,\\ne 0,\\,{s}_{k,l}\\ne 0\\right\\}.$$<\/p>\n<p>To obtain a measure of DNPs independent of grid size, we consider the concentration of DNPs, defined as<\/p>\n<p>$$C=\\,\\frac{\\alpha }{\\beta }$$<\/p>\n<p>where \\(\\alpha\\) and \\(\\beta\\) are the total number of DNPs and neighbor pairs over the whole lattice, respectively. Importantly, we considered aggregation of labeled and unlabeled sites, excluding empty sites and plotted site aggregation as \\(1-C\\), so that increasing values denote a decreasing concentration of DNPs.<\/p>\n<p>Bulk RNA sequencing<\/p>\n<p>RNA was collected from n\u2009=\u20093 Muc2hom and n\u2009=\u20093 Muc2het mouse colons from 10-month-old littermates. mRNA extraction was performed following instructions from the extraction kit (180244, Qiagen). Library prep was performed by the CRUK Cambridge Institute Genomics Core using the Illumina Stranded mRNA Prep kit (20040532, Illumina) according to manufacturer\u2019s instructions. Samples were sequenced using the Illumina Novaseq platform with 50-bp paired end reads. Differential expression analysis was performed using DESeq2. Expressed genes were ranked by descending log2 fold change for use in gene set enrichment analysis<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 60\" title=\"Subramanian, A. et al. Gene set enrichment analysis: a knowledge-based approach for interpreting genome-wide expression profiles. Proc. Natl Acad. Sci. USA 102, 15545&#x2013;15550 (2005).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR60\" id=\"ref-link-section-d54364004e5188\" rel=\"nofollow noopener\" target=\"_blank\">60<\/a>.<\/p>\n<p>Setup for paired spatial transcriptomic and DNA sequencing<\/p>\n<p>A cohort of five Muc2hom mice were treated with ENU 8\u2009weeks post birth and aged for 16 more weeks before collection. Whole-mounted colons were fixed, and six serial tissue sections of 3\u20135-\u03bcm thickness were cut at the crypt base, for at least three areas per colon representing varied levels of pathology. Levels were used for: (1) spatial transcriptomics profiling, (2) pathology assessment with H&amp;E, (3) tumor counting with \u03b2-catenin immunohistochemistry, (4) clone counting with GFP, RFP\u2009+\u2009DAPI and E-cadherin counterstain, and (5) Arid1a-mutated clone counting with Arid1a\u2009+\u2009DAPI counterstain.<\/p>\n<p>Spatial transcriptomics<\/p>\n<p>Slide processing and library preparation were performed according to the 10x Visium Cytassist FFPE protocol. In brief, 5-\u03bcm sections of tissue were transferred to fit each of the two 11-m2 oligo-barcoded capture areas on Visium 10x Genomics slides. Libraries were processed by the CRUK Cambridge Institute Genomics Core according to the manufacturer\u2019s instructions and sequenced on Illumina\u2019s NovaSeqX Plus to an average depth of 8\u2009million mapped reads per sample. Fastq files were processed using the Spaceranger command line tool (10x Genomics v3.0.1) and mapped to the pre-built mm10 reference genome. Processed gene expression matrices for each slide were converted to a Seurat object. Number of counts and features were capped at a minimum of 500 and 200, respectively, to remove low quality spots. After normalization by variance stabilizing transformation using SCTransform, objects corresponding to each slide were integrated into a merged Seurat object using Harmony<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 61\" title=\"Korsunsky, I. et al. Fast, sensitive and accurate integration of single-cell data with Harmony. Nat. Methods 16, 1289&#x2013;1296 (2019).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR61\" id=\"ref-link-section-d54364004e5214\" rel=\"nofollow noopener\" target=\"_blank\">61<\/a>. Clusters were identified using FindNeighbors using integration anchors at a UMAP dimension of 0.5. Genes differentially expressed in each cluster compared with all other clusters were derived using FindMarkers. To select top markers for each cluster, differentially expressed (DE) genes were filtered for expression in at least 50% of barcoded spots in one cluster (pct.1\u2009&gt;0.5) and expression in less than 50% of other clusters (pct.2\u2009&lt;0.3). This filtered list was ranked on descending log2 fold change and ascending P value and the top two markers per cluster were selected. For pseudo-bulk analysis, LoupeBrowser was used to extract barcoded spots assigned to each of the 2-mm-diameter biopsies from which DNA was sequenced.<\/p>\n<p>Transfer of cluster labels<\/p>\n<p>Cluster identities from a reference sample (SITSA1) were transferred to query samples (Muc2het and Muc2hom) using Seurat\u2019s label transfer workflow. Reference and query objects were merged, SCT-normalized and PCA and UMAP dimensionality reduction were performed. The merged object was split by sample, and for each query, transfer anchors were computed on shared features using FindTransferAnchors. Cluster labels from the reference were then predicted with TransferData and added to each query as metadata. Query samples were saved by cohorts for downstream analysis.<\/p>\n<p>DWLS cell-type deconvolution<\/p>\n<p>Cell-type proportions in each barcoded spot were estimated using DWLS<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 38\" title=\"Tsoucas, D. et al. Accurate estimation of cell-type composition from gene expression data. Nat. Commun. 10, 2975 (2019).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR38\" id=\"ref-link-section-d54364004e5247\" rel=\"nofollow noopener\" target=\"_blank\">38<\/a>. Cell-type signatures were derived from a single-cell RNA-sequencing dataset of ten mouse colons in different health conditions: n\u2009=\u20093 healthy, n\u2009=\u20093 acute DSS colitis and n\u2009=\u20094 chronic DSS colitis<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 39\" title=\"Hong, D. et al. Integrative analysis of single-cell RNA-seq and gut microbiome metabarcoding data elucidates macrophage dysfunction in mice with DSS-induced ulcerative colitis. Commun. Biol. 7, 731 (2024).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR39\" id=\"ref-link-section-d54364004e5260\" rel=\"nofollow noopener\" target=\"_blank\">39<\/a> (accession <a href=\"https:\/\/www.ncbi.nlm.nih.gov\/geo\/query\/acc.cgi?acc=GSE264408\" rel=\"nofollow noopener\" target=\"_blank\">GSE264408<\/a>), using the author curated \u2018major\u2019 cell-type annotations with the top ten markers for each cell type. The average proportion of each cell type in each cluster was calculated by taking the mean of the DWLS-estimated fractions across all spots assigned to that cluster. Of note, the cellular origin of the tumor cluster 10 could not be appropriately defined using this healthy cell-type reference.<\/p>\n<p>GSVA<\/p>\n<p>Variation in pathway activity between clusters was quantified using GSVA. Mouse Hallmark gene sets were obtained from MSigDB, excluding sets with fewer than five genes. SCT-normalized expression data was used, after removing genes with near-zero variance (expressed in \u22641 spot). Pathway scores were scaled before downstream analyses. Differential pathway activity for each cluster was identified by comparing that cluster against all other clusters, using FindAllMarkers without thresholds for log fold change or minimum expression.<\/p>\n<p>Correlation in cluster coverage<\/p>\n<p>Correlations between proportions of barcoded spots mapping to the tumor cluster and other clusters were calculated based on symmetric balances for compositional parts data, as described by Kynclova, Hron and Filzmoser<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 41\" title=\"Kyn&#x10D;lov&#xE1;, P., Hron, K. &amp; Filzmoser, P. Correlation between compositional parts based on symmetric balances. Math. Geosci. 49, 777&#x2013;796 (2017).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR41\" id=\"ref-link-section-d54364004e5289\" rel=\"nofollow noopener\" target=\"_blank\">41<\/a>. For each tumor and nontumor cluster pair, two sets of orthonormal coordinate systems were constructed according to symmetric balances of log ratios of either cluster relative to all other clusters in each sample. Pearson correlation coefficients for the first coordinate in each of the two coordinate systems then captures the association of nontumor cluster with the tumor cluster across samples: a positive correlation coefficient implies that enrichment of the two clusters over the respective \u2018average representatives\u2019 of other clusters increase simultaneously and vice versa for negative correlation. A coefficient of zero implies that enrichment of these two clusters is controlled by uncorrelated processes.<\/p>\n<p>Pseudo-spatiotemporal mapping<\/p>\n<p>SpaceFlow<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 42\" title=\"Ren, H., Walker, B. L., Cang, Z. &amp; Nie, Q. Identifying multicellular spatiotemporal organization of cells with SpaceFlow. Nat. Commun. 13, 4976 (2022).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR42\" id=\"ref-link-section-d54364004e5301\" rel=\"nofollow noopener\" target=\"_blank\">42<\/a> was used to combine expression matrices and spatial coordinates from spatial transcriptomics data in a graph-convolutional deep model, to produce low-dimensional embeddings reflecting both gene expression similarity and spatial proximity between barcoded spots. Pseudo-spatiotemporal maps (pSM) are a pseudotime-like ordering of the clusters according to the embeddings, as described in the original SpaceFlow paper.<\/p>\n<p>Cluster correlation and matching for additional tumorigenesis models<\/p>\n<p>Gene expression was correlated between clusters identified in the Muc2hom\u2009+\u2009ENU dataset, and the dataset integrating the Muc2hom\u2009+\u2009ENU cohort, as well as DSS\u2009+\u2009ENU, Apc\u2009+\u2009ENU, Muc2hom and Muc2het cohorts. For each dataset, average expression profiles were computed using Seurat (v5) and restricted to the set of DE genes detected in both datasets. DE genes (P value\u2009&lt;0.05) were identified using Seurat\u2019s FindAllMarkers function. To ensure comparability, the number of DE genes used per cluster was limited to 124, the smallest number of DE genes observed in any cluster in the reference dataset.<\/p>\n<p>Pairwise Spearman correlation coefficients were computed between all \u2018old\u2019 (Muc2hom\u2009+\u2009ENU) and \u2018new\u2019 (full dataset) cluster mean expression vectors. For each old cluster, new clusters exhibiting a correlation coefficient &lt;0.8 were identified as matched. Cluster correspondences were visualized using an alluvial plot generated with the ggalluvial R package (v0.12.5).<\/p>\n<p>NGS library preparation with targeted DNA sequencing library assays<\/p>\n<p>Genes of interest were imported in the Fluidigm D3 Assay design platform and dual coverage primers were designed (Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#MOESM1\" rel=\"nofollow noopener\" target=\"_blank\">7<\/a>). Eight amplicon pools were prepared and used in the LP 8.8.6 integrated fluidic circuit (101-7663) according to manufacturer\u2019s instructions. Samples were sequenced in two separate runs by the CRUK Cambridge Institute Genomics Core on Illumina NovaSeqX Plus for 150-bp paired end reads.<\/p>\n<p>Mutation calling<\/p>\n<p>RePlow<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 62\" title=\"Kim, J. et al. The use of technical replication for detection of low-level somatic mutations in next-generation sequencing. Nat. Commun. 10, 1047 (2019).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR62\" id=\"ref-link-section-d54364004e5362\" rel=\"nofollow noopener\" target=\"_blank\">62<\/a> was used for mutation calling based on dual replicate amplicon coverage using combinatorial pooling of amplicons and applied to the first 20\u2009million reads of each sample (Supplementary Table <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#MOESM4\" rel=\"nofollow noopener\" target=\"_blank\">7<\/a>). In brief, RePlow exploits replicate library preparations to separate the contribution of background errors occurring in library preparation from sequencing errors occurring in sequencing. This is done with a statistical model that calculates the total log ratio of probabilities for variant candidates compared with background errors across all replicates simultaneously. In doing so, adjusted VAFs are constructed by subtracting the contribution from sequencing errors, whereas background error profiles for the error model are calculated independently for the six base pair substitution types (A\u2009&gt;\u2009C, A\u2009&gt;\u2009G, A\u2009&gt;\u2009T, C\u2009&gt;\u2009A, C\u2009&gt;\u2009G, C\u2009&gt;\u2009T) across the targeted regions. An additional stringent filtering step was applied whereby only variants with a positive log ratio of probabilities in both replicates individually were retained, to exclude false positive calls at low VAF values. Orthogonal validation of mutation calls was done using the Ampliconseq pipeline (<a href=\"https:\/\/github.com\/crukci-bioinformatics\/ampliconseq\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/crukci-bioinformatics\/ampliconseq<\/a>). In brief, variants are called using Vardict from sequence reads aligned to the reference genome (GRCm39). The pipeline models the background substitution noise at each amplicon position to identify and filter SNV calls with an allele fraction indistinguishable from noise. A minimum of 0.01% VAF was set as the lower detection threshold. Outliers in the VAF\/noise threshold plane, corresponding to remaining artefacts and germline SNVs were removed post hoc.<\/p>\n<p>Tessellation algorithm for clone calling<\/p>\n<p>To parsimoniously assign multiple calls of the same mutation made from the same piece of colon tissue to individual clones, as well as resolve the spatial context of the colon that was opened longitudinally before biopsy sampling, a Voronoi tessellation with periodic boundary conditions along the radial axis of each tissue section was calculated using the spatial coordinates of the corresponding biopsies. A depth-first-search algorithm was then used to call single clones by enumerating all connected components of the graph with edges defined by adjacent Voronoi tiles (or those found within 2\u2009mm from one another) containing the given mutation.<\/p>\n<p>Quantifying selection using dN\/dS<\/p>\n<p>The latest version of the maximum likelihood implementation of dN\/dS, as originally described in Martincorena et al.<a data-track=\"click\" data-track-action=\"reference anchor\" data-track-label=\"link\" data-test=\"citation-ref\" aria-label=\"Reference 46\" title=\"Martincorena, I. et al. Universal patterns of selection in cancer and somatic tissues. Cell 171, 1029&#x2013;1041 (2017).\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#ref-CR46\" id=\"ref-link-section-d54364004e5393\" rel=\"nofollow noopener\" target=\"_blank\">46<\/a> and available in the R package dNdScv (<a href=\"https:\/\/github.com\/im3sanger\/dndscv\" rel=\"nofollow noopener\" target=\"_blank\">https:\/\/github.com\/im3sanger\/dndscv<\/a>), was applied to identify genes under positive selection. A custom reference object was built from the assembly GRCm39 by subsetting the trinucleotide context-dependent substitution consequence matrix to the set of possible mutations based on the region of the genome selected for targeted sequencing in this study. Counts of mutations fed into the algorithm were those based on individual clones, as defined above.<\/p>\n<p>Statistical analyses and reproducibility<\/p>\n<p>Visualization and statistical analysis of data were performed in the R statistical computing environment (version 2024.04.0\u2009+\u2009735). All experiments were performed on at least three independent biological replicates (three different mice). Micrographs depict representative data derived from at least three independent biological replicates. Normality was assessed using Shapiro\u2013Wilk\u2019s tests and relevant statistical tests applied. Tests and corresponding P values are indicated in the figure legends and figures, respectively. Box plots display the distribution of data using the following components: lower whisker show the smallest observation greater than or equal to lower hinge\u2009minus 1.5\u00d7\u2009IQR; lower hinge shows the 25% quantile; the center line shows the median, 50% quantile; the upper hinge shows the 75% quantile; the upper whisker shows the largest observation less than or equal to upper hinge plus 1.5\u00d7\u2009IQR.<\/p>\n<p>Reporting summary<\/p>\n<p>Further information on research design is available in the <a data-track=\"click\" data-track-label=\"link\" data-track-action=\"supplementary material anchor\" href=\"http:\/\/www.nature.com\/articles\/s41588-026-02673-0#MOESM2\" rel=\"nofollow noopener\" target=\"_blank\">Nature Portfolio Reporting Summary<\/a> linked to this article.<\/p>\n","protected":false},"excerpt":{"rendered":"Mice Mice were of C57BL\/6 background. The Muc2KO line used was described by Velcich et al.16. Mice containing&hellip;\n","protected":false},"author":2,"featured_media":718165,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[10],"tags":[5083,5085,3251,5082,25648,8815,59,5084,3250,102,5081,60830,11803,56,54,55],"class_list":["post-718164","post","type-post","status-publish","format-standard","has-post-thumbnail","category-health","tag-agriculture","tag-animal-genetics-and-genomics","tag-biomedicine","tag-cancer-research","tag-colon-cancer","tag-dna-sequencing","tag-gb","tag-gene-function","tag-general","tag-health","tag-human-genetics","tag-inflammatory-bowel-disease","tag-rna-sequencing","tag-uk","tag-united-kingdom","tag-unitedkingdom"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts\/718164","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/comments?post=718164"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts\/718164\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/media\/718165"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/media?parent=718164"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/categories?post=718164"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/tags?post=718164"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}