Biomolecules

2026-09-09 | | Total: 8

#1 Multi-ligand simultaneous docking of Carica papaya leaf phytochemicals, Carpaine and Rutin, reveals multi-mechanism inhibition of cancer proteins BCL-2 and WWP1 [PDF] [Copy] [Kimi] [REL]

Authors: Merla Sudha, Asmita Saha, Belaguppa Manjunath Ashwin Desai, Anil Ranu Mhashal, Pronama Biswas

Cancer remains a major global health concern due to chemotherapy resistance and toxicity from high-dose treatments. To overcome these challenges, new therapeutic strategies targeting key proteins in cancer progression are essential. This study evaluates two phytochemicals, Carpaine (Car) and Rutin (Rut), from Carica papaya leaves, for their potential in enhancing cancer therapy by targeting B-cell lymphoma 2 (BCL-2) and WW domain-containing protein 1 (WWP1) proteins. We assessed their additive, allosteric, and synergistic effects using molecular docking, multi-ligand simultaneous docking (MLSD), molecular dynamics (MD) simulations, and MMPBSA analysis. Car and Rut showed an additive effect on BCL-2 by binding at distinct regions within the same pocket. MLSD revealed an improved binding affinity of -13.13 +/- 0.08 kcal/mol, compared with individual ligands or the commercial inhibitor Venetoclax. For WWP1, Car bound near the H-site and Rut near the Le-site, exhibiting an allosteric effect that increased Car's binding affinity in MLSD to -15.59 +/- 0.39 kcal/mol. Furthermore, Rut combined with bortezomib (Bort) demonstrated a synergistic interaction with WWP1. Binding energies were -7.64 +/- 0.156 kcal/mol for Bort, -10.26 +/- 0.07 kcal/mol for Rut, and -15.59 +/- 0.39 kcal/mol for MLSD, suggesting a more stable complex through synergy. These results suggest Car and Rut, particularly in combination with Bort, as promising candidates against cancer-related proteins BCL-2 and WWP1. Further experimental validation is warranted to explore their therapeutic potential.

Subject: Biomolecules

Publish: 2026-09-08 10:29:10 UTC


#2 Predicting directional flexibility in proteins [PDF] [Copy] [Kimi] [REL]

Authors: Vsevolod Viliuga, Leif Seute, Matteo Tadiello, Nicolas Wolf, Frauke Gräter, Arne Elofsson

Predicting protein dynamics is a long-standing problem in computational structural biology. Often, protein function critically depends on local directed motions, such as hinge movements, catalytic loop rearrangements and domain reorientations, which can be characterized by directional flexibility and correlated structural motions of the protein backbone. While Molecular Dynamics (MD) simulations provide an established but often prohibitively expensive approach, recent deep generative models aim to reduce this cost by directly predicting conformational ensembles, emulating MD. However, due to their large size and the need to generate several states until the derived dynamical properties converge, these models remain expensive. In this work, we propose BackFlip-2: a fast SE(3)-equivariant graph neural network trained to directly predict dynamical descriptors, such as directional backbone flexibility and pairwise dynamic correlations, from an equilibrium structure. In a series of experiments, we show that our model matches the accuracy of substantially larger ensemble generation models while being orders of magnitude faster, and demonstrate that the proposed equivariant architecture is especially well-suited for capturing anisotropic motions in proteins. BackFlip-2 model weights, training and inference code are available at https://github.com/graeter-group/backflip.

Subject: Biomolecules

Publish: 2026-09-08 09:18:46 UTC


#3 PocketVE: Stable and Property-Guided Structure-Based Drug Design with Variance-Exploding Diffusion [PDF] [Copy] [Kimi] [REL]

Authors: Peining Zhang, Jinbo Bi

Protein-conditioned 3D molecule generation is a central challenge in structure-based drug design, requiring a balance between pocket compatibility, molecular properties, and physical geometry. We propose \textbf{PocketVE}, a protein-pocket-conditioned variance-exploding (VE) diffusion framework that couples stable coordinate denoising with inference-time property guidance. Specifically, PocketVE combines an EDM-style training and sampling setup for 3D denoising, classifier-free guidance for multi-property steering without external property classifiers, and adaptive protein perturbation as a training-time pocket regularizer. Evaluated on CrossDocked2020 under the GenBench3D protocol, PocketVE improves Valid$_{3\text{D}}$ from 58.6 to 80.6 and reduces strain energy from 457.4 to 127.9 relative to its TAGMol architectural baseline, while retaining competitive docking and molecular-property scores under moderate guidance. A guidance-scale study shows that moderate guidance gives a favorable balance between target-related objectives and geometric quality, whereas stronger guidance can degrade geometry and distributional fidelity. Pocket-permutation and PoseCheck diagnostics further support pocket-specific spatial compatibility with reduced steric conflicts. Overall, the results suggest that geometric stability and inference-time property guidance should be considered as coupled design objectives.

Subjects: Biomolecules , Machine Learning

Publish: 2026-09-08 01:19:17 UTC


#4 PSLL: Persistent Sheaf Laplacian Learning for Protein-Ligand Binding Affinity Prediction [PDF] [Copy] [Kimi] [REL]

Authors: Mushal Zia, Benjamin Jones, Guo-Wei Wei

Accurate prediction of protein-ligand binding affinity remains a central challenge in computational drug discovery due to the complex interplay among molecular geometry, physicochemical interactions, and atom-specific charge information. In this work, we introduce a Persistent Sheaf Laplacian learning (PSLL) framework for protein-ligand binding affinity prediction. The proposed approach constructs multiscale topological representations from three-dimensional protein-ligand complexes by incorporating atomic partial charges into sheaf restriction maps over Vietoris-Rips and alpha complex filtrations. To capture chemically diverse protein-ligand interactions, we introduce element-specific and category-specific atom-pair representations within the PSLL framework. Harmonic and non-harmonic spectra extracted from the resulting persistent sheaf Laplacians are used as molecular descriptors. To complement the PSLL-derived molecular representation, we incorporate transformer-based protein embeddings and SMILES-derived ligand descriptors for binding affinity prediction. The scoring power of the proposed multiscale PSLL model is validated against existing state-of-the-art methods on three widely used PDBbind benchmark datasets, including PDBbind-v2007, PDBbind-v2013, and PDBbind-v2016. The computational results indicate that the proposed PSLL model achieves strong predictive performance across benchmark datasets, highlighting its potential as an interpretable and mathematically grounded framework with promising generalizability for molecular machine learning and drug discovery.

Subject: Biomolecules

Publish: 2026-08-22 04:07:02 UTC


#5 Condition aware learning enables robust prediction of oligonucleotide melting behavior across diverse chemistries and assay conditions [PDF] [Copy] [Kimi] [REL]

Authors: Danielle L. Ferreira, Lifeng Lin, Adam Aslam, Nicholas Chang, Rebekah G. Baig, Edgar Baculi, Zoey Cao, Melanie Senn

Oligonucleotide melting temperature is a fundamental determinant of nucleic acid hybridization and underpins the design of molecular diagnostics, polymerase chain reaction assays, and many other biotechnology applications. However, accurately predicting melting behavior remains difficult because it depends not only on sequence composition, but also on experimental conditions and chemical modifications commonly used in modern assay design. Existing thermodynamic models rely on fixed parameterizations that are often difficult to extend across diverse reaction environments and nucleotide chemistries. Here we show that a condition-aware nucleotide language model can accurately predict oligonucleotide melting behavior across diverse experimental conditions and both unmodified and chemically modified oligonucleotides. By combining contextual sequence representations with explicit information describing the reaction environment, the framework achieves sub-degree prediction accuracy and reduces prediction error for locked nucleic acid-modified oligonucleotides by up to 25% relative to nearest-neighbor thermodynamic approaches. The model also more accurately captures the thermal effects introduced by nucleotide modification and maintains strong performance on independent benchmark datasets spanning experimental conditions substantially different from those represented during training. Our results demonstrate that learned sequence representations can complement classical thermodynamic models by capturing context-dependent effects that are difficult to encode using fixed parameter tables alone. More broadly, this work provides a scalable framework for predicting oligonucleotide melting behavior across diverse chemistries and assay conditions, supporting more reliable molecular assay design.

Subjects: Biomolecules , Machine Learning

Publish: 2026-08-05 22:58:56 UTC


#6 ZetaDial: dialing net charge of protein binders at inference time for therapeutic developability [PDF] [Copy] [Kimi] [REL]

Authors: Mohammed Sameer Syed, Tamara Dinneen

Net charge is a developability-relevant property of therapeutic binders, linked to viscosity, clearance, nonspecific interaction and aggregation, and antibody screens already use charge-related criteria. Yet inverse-folding pipelines expose no way to set it to a target value. ProteinMPNN and BindCraft offer amino-acid biases, weight choices and custom losses, but neither supplies a per-protein feedback loop that measures realised charge after sampling and corrects it to a setpoint. ZetaDial contributes a post-sampling, per-protein secant controller around fixed-backbone ProteinMPNN. On matched stochastic benchmarks the secant loop reduced mean absolute error relative to a fixed-slope loop on RCSB complexes (5.17 vs 6.46 charge units) and Cas13 monomers (5.57 vs 8.23). Relative to the optimised matched global bias, it cut RCSB error from 11.71 to 5.17 (cluster bootstrap p < 0.001) and was statistically indistinguishable on Cas13 (5.47 vs 5.57). Across 800 eight-protein subsets, sensitivity heterogeneity was associated with calibration gain (Pearson r = 0.79); this is descriptive resampling, not a prospective decision rule. Foldability deteriorated as bias magnitude increased. In the full 52-complex seed-0 analysis, reference-based DockQ declined clearly at +/-3 but not at +/-1.5; a selected five-seed replication on eight complexes showed paired declines at every nonzero setting, but does not estimate the effect for all 52. In exploratory BindCraft sweeps, PD-L1 designs moved toward near-neutral charge at similar maximum interface pTM but with overlapping success-rate intervals; IL-7R-alpha responses were non-monotonic and RBD produced no strong designs. A fixed-backbone C-alpha-neighbour analysis found smaller same-sign charge-patch proxies near neutral charge, but this proxy is not a measured electrostatic surface or experimental developability endpoint.

Subjects: Biomolecules , Machine Learning

Publish: 2026-08-04 18:36:14 UTC


#7 Precise Positional Readout of Molecular Barcode Structures using Solid-State Nanopores [PDF] [Copy] [Kimi] [REL]

Authors: Wouter Botermans, Eric Beamish, Matteo Cartiglia, Liam Vandekerckhove, Wouter Renckens, Natan Biesmans, Koen Ongena, Wannes Peeters, Pol Van Dorpe, Sanjin Marion

Fast and nonuniform translocation through solid-state nanopores limits both the detection of small molecular labels and their precise localization along molecular carriers. In this work we report the detection and localization performance of nucleotide-based molecular labels along double-stranded DNA scaffolds using solid-state nanopores in thin planar membranes. For small labels that are challenging to resolve individually, we introduce an anchoring strategy where readily detectable bulky labels serve as reference points to align multiple translocation events, enabling population-based detection and localization of smaller molecular features. The measured anchor positions constrain a probabilistic model of translocation velocity, identifying the most probable velocity profile for each event and enabling nonlinear trace "unwarping" for improved multi-event alignment. A complementary window-based evidence aggregation procedure accumulates weak but consistent label signatures across events, enabling detection of features that are individually masked by noise. These approaches enable robust recovery of single-dumbbell labels (DB1) on the order of 28 nucleotides and reduce mean localization errors to as low as 10 base pairs for DB3 labels and 40 base pairs for DB1 labels when averaging over multiple events. Stronger fractional DNA-associated current blockades, used as proxy for smaller pore geometries, are additionally associated with improved detection and lower localization error across membrane-based nanopore fabrication techniques. Overall, anchor-guided alignment provides a route to higher-density molecular information readout without compromising throughput via controlled translocation approaches.

Subjects: Biomolecules , Biological Physics , Chemical Physics , Data Analysis, Statistics and Probability , Instrumentation and Detectors

Publish: 2026-07-31 08:53:41 UTC


#8 Novel hybrid protein scaffold gap filling using weighted machine learning ensemble, beam search, and mass-constrained reranking [PDF] [Copy] [Kimi] [REL]

Authors: Tahmid Enam Shrestha, Md. Manzurul Hasan, Md. Rafiqul Islam

Protein scaffold gap filling is an important computational task in protein sequence reconstruction, where missing amino acid regions must be inferred from incomplete scaffold information. This study proposes a hybrid machine learning and mass constrained reranking framework for protein scaffold gap filling under known-gap-size and known-gapmass settings. Homologous protein sequences from MabCampath, P5A proteoform, and carbonic anhydrase 2 were used to generate masked 11-mer residue-level samples and fullgap evaluation cases. The residue prediction task was formulated as a 20-class amino acid classification problem using first-, middle-, and last-position masking. Multiple classical machine learning models were trained using raw encoded, row-average, and SVD-reduced features, and the strongest models were combined through a validation-accuracy-weighted ensemble. For known-size gap reconstruction, beam search was used to generate complete missing peptide sequences from residue-level probability estimates. For known-mass reconstruction, mass-constrained homologous candidate retrieval was combined with hybrid reranking based on mass validity, homologous frequency, context support, ensemble likelihood, mass error, and length penalty. The proposed framework achieved 95.41% residue-level validation accuracy, 87.50% known-size exact-match accuracy, and 100% top-5 recovery on seven CAH2 known-mass benchmark cases. These results indicate that the proposed framework can effectively reconstruct missing protein regions by integrating local sequence learning, homologous evidence, peptide mass constraints, and biochemical validation.

Subjects: Biomolecules , Machine Learning

Publish: 2026-07-16 11:22:04 UTC