Search arXiv⌕ Search

arXiv · 1903.01652

ColourQuant: a high-throughput technique to extract and quantify colour phenotypes from plant images

Abstract

Colour patterning contributes to important plant traits that influence ecological interactions, horticultural breeding, and agricultural performance. High-throughput phenotyping of colour is valuable for understanding plant biology and selecting for traits related to colour during plant breeding. Here we present ColourQuant, an automated high-throughput pipeline that allows users to extract colour phenotypes from images. This pipeline includes methods for colour phenotyping using mean pixel values, Gaussian density estimator of Lab colour, and the analysis of shape-independent colour patterning by circular deformation.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Mao Li, Margaret H. Frank, Zoë Migicovsky. 2019-03-05. ColourQuant: a high-throughput technique to extract and quantify colour phenotypes from plant images. https://arxiv.org/abs/1903.01652

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Predicting Mutational Signature Exposures from H&E Whole Slide Images: A Pan-Cancer Feasibility Study

Mutational signatures reveal cancer-driving processes with clinical relevance: MMR-deficient tumors respond better to immunotherapy, HRD tumors are sensitive to PARP inhibitors, and POLE-mutant tumors often have high mutation burdens influencing treatment response. Yet signature profiling remains limited by the cost, complexity, and turnaround time of whole-genome sequencing (WGS) or whole-exome sequencing. Histopathology is widely available and cost-effective, and H&E morphology has been linked to MSI, HRD, POLE-related processes, and driver mutations. Whether histology can recover the broader landscape of mutational signature exposures in a pan-cancer setting remains unknown. Here we introduce Hist2Sig, a deep learning framework predicting exposures to 30 COSMIC SBS signatures from H&E-stained slides. Trained on matched WGS and histology data from 7,063 TCGA patients across 29 cancer types, Hist2Sig was compared with a tumor type-only baseline to separate morphology-derived signal from tissue-of-origin priors. In internal cross-validation, it recovered relative signature compositions across most tumor types and achieved higher mean top-three overlap than the baseline in 19 of 29, with the largest gains in COAD, PRAD, GBM, and UCEC. It also recurrently recovered signatures of poorly characterized etiology, including SBS8, SBS12, SBS39, and SBS40a. In an external CPTAC cohort of 193 patients across five tumor types, Hist2Sig retained moderate concordance with observed compositions and outperformed the baseline in glioblastoma and pancreatic adenocarcinoma. These results support the feasibility of inferring mutational signature exposures from routine histology, while highlighting variation across tumor types. Hist2Sig provides a foundation for tumor-specific models that could complement sequencing in pre-sequencing triage or support decisions where WGS is unavailable.

q-bio.QM↗

Multimodal AI predicts clinical outcomes of drug combinations from preclinical data

Predicting clinical outcomes from preclinical data is essential for selecting safe and effective drug combinations and for reducing late-stage failures. AI models use molecular structure and target annotations, and do not leverage the perturbation readouts that report how a compound acts in a cellular context. Here we introduce Madrigal, a multimodal AI model that learns from structural, pathway, cell-viability, and transcriptomic data. Madrigal aligns these modalities across 21,842 compounds into a shared latent space and predicts combination outcomes even for drugs observed in only a subset of the data modalities. Trained on 158 expert-curated and 795 patient-reported combination outcomes, Madrigal outperforms single-modality and state-of-the-art multimodal methods. Ablations show that modality alignment and multimodal input each improve predictive performance. Madrigal predicts elevated risk for combinations that share membrane transporters. In head-to-head trials that compare two combination arms,the arm with the higher observed incidence of neutropenia, anemia, alopecia, or hypoglycemia receives the higher predicted risk in 25 of 28 comparisons. In MASH, Madrigal ranks resmetirom among the candidates with favorable predicted safety when paired with type 2 diabetes drugs. Madrigal also improves adverse-event prediction in a longitudinal patient cohort and an independent oncology cohort and predicts efficacy in primary acute myeloid leukemia samples and patient-derived xenografts.

q-bio.QM↗

ProteoEM: probabilistic protein abundance estimation from iterative affinity traces

Single-molecule affinity mapping enables molecular-level measurement of proteins and proteoforms, but imperfect and nonspecific probe binding makes individual affinity traces compatible with multiple molecular identities. Accurate abundance estimation therefore requires apportionment of ambiguous traces by weight rather than assignment to a single candidate. We developed ProteoEM, an expectation-maximization framework for weighted proteoform quantification, inspired by transcript abundance estimation methods for RNA sequencing and released as an open-source Python package. ProteoEM evaluates each molecule against every candidate using fixed, pre-calibrated probe-response rates held separate from the abundance estimate, while retaining the full likelihood of the observed affinity features. The framework estimates proteoform abundances, reports indistinguishable proteoforms as groups when measurements cannot separate them, and accounts for differential observation yields to distinguish the composition of observed molecules from that of the source sample. In simulations, ProteoEM accurately recovered the underlying molecular composition where approaches that reduce each trace to a hard yes/no call introduced substantial errors. ProteoEM's performance was insensitive to a moderate, uniform calibration error but was biased by informative missing data and by proteoforms absent from the reference. When observation yields were known, it also recovered source-sample composition from observed molecular counts. ProteoEM provides an open-source, reproducible framework for quantitative analysis of single-molecule affinity measurements, and these results motivate validation on experimental molecule-level data.

q-bio.QM↗