STELLAR: A Fragment-Based Docking Workflow for Highly Flexible Long-Chain Biopolymers

Abstract Docking methods have improved significantly through optimization using generative approaches and systematic exploration of parameter configurations within docking software. However, docking large molecules such as peptides remains challenging, while the docking of large non-peptidic biopolymers is still insufficiently explored. Existing tools often exhibit limitations in both accuracy and computational efficiency when applied to longer peptides or other polymers. Moreover, although generative deep learning and machine learning approaches have shown promise, they still frequently lack robust ranking metrics, resulting in suboptimal performance. To address these limitations, we developed STELLAR (Score-Tuning for Efficient Ranking of Large Ligands using an Accurate and Refined Docking Configuration), a workflow designed to enable efficient and accurate docking of highly flexible long-chain biopolymers exceeding 10 subunits. STELLAR implements a fragment-based strategy, decomposing the biopolymer into smaller units for individual docking, followed by recomposition into full-length structures. It also includes optimized algorithms for handling long-chain biopolymers. The pipeline integrates structural optimization steps using tools such as GNINA, RDKit, and GROMACS, ensuring physically realistic poses. STELLAR achieved RMSD values below 5 Å in validation experiments using benchmark complexes from Propedia and the Protein Data Bank (PDB). It offers a faster and more efficient alternative to several state-of-the-art tools across a range of peptide lengths. Additionally, it scales linearly in computational time, runs primarily on CPUs with minimal GPU usage, avoids reliance on generative approximations, provides an accurate metric for ligand ranking, and reduces computational cost compared to similar tools. Its design also supports high-throughput screening of linear polymer–protein interactions by efficiently reducing the complexity associated with high conformational flexibility.

Authors

Institutions

Publication Details

Journal
Journal of Chemical Information and Modeling
Published
2026-09-16
DOI
https://doi.org/10.1021/acs.jcim.6c00987
Primary Topic
Chemical Synthesis and Analysis
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

STELLAR: A Fragment-Based Docking Workflow for Highly Flexible Long-Chain Biopolymers

Jochem Nelen, Carlos Martínez-Cortés, Alejandro Rodríguez‐Martínez, Horacio Pérez‐Sánchez et al.
Journal of Chemical Information and Modeling
Chemical Synthesis and Analysis
article

STELLAR: A Fragment-Based Docking Workflow for Highly Flexible Long-Chain Biopolymers

Jochem Nelen, Carlos Martínez-Cortés, Alejandro Rodríguez‐Martínez, Horacio Pérez‐Sánchez, Miguel Carmena‐Bargueño
article en

Abstract

Abstract Docking methods have improved significantly through optimization using generative approaches and systematic exploration of parameter configurations within docking software. However, docking large molecules such as peptides remains challenging, while the docking of large non-peptidic biopolymers is still insufficiently explored. Existing tools often exhibit limitations in both accuracy and computational efficiency when applied to longer peptides or other polymers. Moreover, although generative deep learning and machine learning approaches have shown promise, they still frequently lack robust ranking metrics, resulting in suboptimal performance. To address these limitations, we developed STELLAR (Score-Tuning for Efficient Ranking of Large Ligands using an Accurate and Refined Docking Configuration), a workflow designed to enable efficient and accurate docking of highly flexible long-chain biopolymers exceeding 10 subunits. STELLAR implements a fragment-based strategy, decomposing the biopolymer into smaller units for individual docking, followed by recomposition into full-length structures. It also includes optimized algorithms for handling long-chain biopolymers. The pipeline integrates structural optimization steps using tools such as GNINA, RDKit, and GROMACS, ensuring physically realistic poses. STELLAR achieved RMSD values below 5 Å in validation experiments using benchmark complexes from Propedia and the Protein Data Bank (PDB). It offers a faster and more efficient alternative to several state-of-the-art tools across a range of peptide lengths. Additionally, it scales linearly in computational time, runs primarily on CPUs with minimal GPU usage, avoids reliance on generative approximations, provides an accurate metric for ligand ranking, and reduces computational cost compared to similar tools. Its design also supports high-throughput screening of linear polymer–protein interactions by efficiently reducing the complexity associated with high conformational flexibility.

Journal of Chemical Information and Modeling
Universidad Católica de Murcia (ES)
Openalex Percentile: Top 18%
Chemical Synthesis and Analysis
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.