Table of Contents
Advances in database resources and computational methods for predicting antibody polyreactivity
Antibody polyreactivity refers to the ability of a monoclonal antibody to non-specifically bind to a diverse range of antigens. While this property may be an intrinsic mechanism of immune response, it poses significant challenges in therapeutic antibody ...
More.Antibody polyreactivity refers to the ability of a monoclonal antibody to non-specifically bind to a diverse range of antigens. While this property may be an intrinsic mechanism of immune response, it poses significant challenges in therapeutic antibody development, often leading to off-target effects, poor pharmacokinetics, and potential toxicity. This review compiles the data resources related to polyreactive antibodies and places a particular emphasis on computational models for predicting antibody polyreactivity. The latter includes empirical models based on physicochemical properties, traditional machine learning models, deep learning networks, and protein language models. Through delineating the complexity of antibody polyreactivity, this review emphasizes the critical role and growing potential of computational prediction tools in selecting and engineering antibody drug candidates at early stages, thereby reducing development risks and accelerating the development of safer and better therapeutic antibodies.
Less.Haoxiang Tang, ... Jian Huang
DOI:https://doi.org/10.70401/cbm.2026.0021 - July 14, 2026
Identification of potential associations between circRNAs and diseases based on meta relation aware
Aims: Circular RNAs (circRNAs) have been shown to be closely associated with the occurrence and progression of various diseases. However, most existing circRNA-disease association prediction methods are limited to homogeneous networks and are ...
More.Aims: Circular RNAs (circRNAs) have been shown to be closely associated with the occurrence and progression of various diseases. However, most existing circRNA-disease association prediction methods are limited to homogeneous networks and are unable to effectively capture deep semantic associations through high-order meta-paths. This study aims to develop an efficient computational method for accurately predicting potential circRNA-disease associations.
Methods: We propose a meta-relation-aware heterogeneous graph learning framework for circRNA-disease association prediction. Specifically, known circRNA-disease associations are first used to compute Gaussian interaction profile kernel similarity and extract node attribute features, based on which a heterogeneous graph network is constructed. A graph neural network is then employed to perform multi-layer message passing on the heterogeneous graph, aggregating neighborhood information to achieve deep fusion of multi-source features and generate node embeddings that encode both local and global structural information. Finally, the learned embeddings are fed into a gradient boosting decision tree classifier, and an ensemble strategy is adopted to improve prediction accuracy. Five-fold cross-validation is used for performance evaluation.
Results: Experimental results on three benchmark datasets, CircR2Disease V2.0, circAtlas 3.0, and circRNADisease V2.0, show that the proposed model achieves area under the receiver operating characteristic curve (AUC) values of 92.17%, 91.83%, and 91.73%, respectively. The model outperforms traditional methods in terms of accuracy, precision, and recall. Furthermore, ablation studies validate the effectiveness of the meta-relation-aware strategy.
Conclusions: Overall, this work provides an efficient and reliable computational framework for molecular association prediction and biomarker discovery in the biomedical domain.
Less.Xingyu Tan, ... Zhuhong You
DOI:https://doi.org/10.70401/cbm.2026.0020 - June 22, 2026
Isoform function prediction via knowledge distillation from alternative splicing
Aims: Alternative splicing serves as a primary mechanism for diversifying the proteome, making the prediction of distinct isoform functions critical for understanding complex disease mechanisms. However, determining the specific functional ...
More.Aims: Alternative splicing serves as a primary mechanism for diversifying the proteome, making the prediction of distinct isoform functions critical for understanding complex disease mechanisms. However, determining the specific functional roles of isoforms remains hindered by high sequence homology among variants and the sparsity of isoform-level annotations.
Methods: In this study, we propose SpliceEM, a deep learning framework for isoform function prediction at single-cell resolution. SpliceEM utilizes a splicing event-aware encoder with cross-modal attention to separate functional signals from global protein sequences. A Heterogeneous Graph Transformer captures the dependencies among isoforms, genes, and Gene Ontology terms. To bridge the annotation gap, we incorporate a self-distillation framework guided by an Exponential Moving Average teacher model and Multi-Instance Learning, optimized by an Asymmetric Loss and hierarchical constraints.
Results: Benchmarking on human datasets demonstrates that SpliceEM outperforms existing methods in isoform function prediction, particularly in identifying rare functional terms under data-sparse conditions. Furthermore, splicing-function analysis reveals that specific splicing events, such as skipped exons and alternative first exons, act as prominent drivers in oncogenic signaling cascades and context-specific functional switching.
Conclusion: SpliceEM provides a computational foundation for exploring transcriptomic functional diversity. By shifting the focus from global sequences to localized splicing events and utilizing hierarchical biological priors, it offers high-resolution insights into cell-type-specific molecular mechanisms and potential therapeutic targets.
Less.Tong Gu, Jun Wang
DOI:https://doi.org/10.70401/cbm.2026.0019 - June 15, 2026
scAdaptAnno: Target graph domain adaptation for cross-patient single-cell annotation transfer in tumor microenvironments
Aims: Single-cell RNA sequencing (scRNA-seq) has emerged as a cornerstone technology in tumor microenvironment research. Accurate cell-type annotation is fundamental to downstream scRNA-seq analysis. However, automated tools are often highly ...
More.Aims: Single-cell RNA sequencing (scRNA-seq) has emerged as a cornerstone technology in tumor microenvironment research. Accurate cell-type annotation is fundamental to downstream scRNA-seq analysis. However, automated tools are often highly sensitive to dataset noise and show limited adaptability in cross-patient scenarios. To address these challenges, we propose scAdaptAnno, a graph-based target domain adaptation framework for cross-patient single-cell annotation.
Methods: In the graph construction phase, scAdaptAnno integrates both gene expression similarity and biological prior knowledge to build a more biologically meaningful cell graph. By leveraging cell representations enriched with biological priors to mitigate noise in gene expression data and by implementing a bidirectional adaptation mechanism, the model achieves source-free target domain alignment.
Results: We performed comprehensive benchmarking against nine leading methods across multiple datasets spanning various cancer types. The results demonstrate that scAdaptAnno achieves state-of-the-art performance.
Conclusion: scAdaptAnno is a robust and accurate single-cell annotation tool that excels in cross-patient cell-type annotation transfer. By integrating biologically informed graph construction and bidirectional source-free domain adaptation, it delivers reliable, noise-resistant performance across diverse tumor microenvironments, providing an effective solution for automated cell-type annotation in multi-patient scRNA-seq studies.
Less.Xi-Yue Cao, ... Yu-An Huang
DOI:https://doi.org/10.70401/cbm.2026.0018 - June 12, 2026