Back

CCK* (Convex Closure K*): A Suite of Algorithms for De Novo L- and D-peptide Design

Childs, H.; McBride, A. C.; Donald, B. R.

2026-06-01 bioinformatics
10.1101/2025.11.21.689740 bioRxiv
Show abstract

The computational design of L-peptides and their mirror-image counterparts, D-peptides, is an active area in drug design. Peptide therapeutics offer exceptional structural diversity and high binding specificity, while D-peptides additionally confer critical advantages such as proteolytic resistance. Progress in de novo D-peptide design has been hindered by the absence of evolutionary context and limited structural data, both of which underpin the deep learning methods widely used in L-peptide design. Consequently, a robust framework capable of designing both L- and D-peptides should integrate data-driven inference with first-principles, physics-based modeling. Here, we introduce a unified computational framework that supports de novo design of both L- and D-peptides, thereby expanding the accessible design space across both chiral spaces. Convex Closure K* (CCK*) is a suite of chirality-agnostic algorithms: SCOPE, MONTAGE, and ARISE. SCOPE uses geometry as a proxy for chemical energetics, computing convex hull representations of rotameric states to rapidly generate multi-sequence protein contact maps. MONTAGE employs geometric hashing in conjunction with the K* algorithm to generate and rank backbone scaffolds according to their suitability for sequence design. ARISE is a K*-based sequence design algorithm that performs iterative residue assignment in an undirected graph to design high-affinity peptide sequences. We apply the full CCK* suite to six de novo design tasks, benchmarking chirality-preserving and chirality-inverting designs in both homochiral and heterochiral complexes.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Journal of Chemical Information and Modeling
238 papers in training set
Top 0.5%
11.7%
2
Proteins: Structure, Function, and Bioinformatics
88 papers in training set
Top 0.1%
11.7%
3
PLOS Computational Biology
1863 papers in training set
Top 4%
9.5%
4
Journal of Chemical Theory and Computation
140 papers in training set
Top 0.2%
8.8%
5
Bioinformatics
1204 papers in training set
Top 4%
6.6%
6
Protein Science
246 papers in training set
Top 0.5%
6.6%
50% of probability mass above
7
Journal of Cheminformatics
29 papers in training set
Top 0.1%
4.8%
8
Bioinformatics Advances
203 papers in training set
Top 1%
4.2%
9
Briefings in Bioinformatics
354 papers in training set
Top 2%
3.5%
10
Protein Engineering, Design and Selection
15 papers in training set
Top 0.1%
3.2%
11
Journal of Molecular Biology
232 papers in training set
Top 1%
2.7%
12
PLOS ONE
5266 papers in training set
Top 43%
2.4%
13
ImmunoInformatics
12 papers in training set
Top 0.1%
1.7%
14
Journal of Computational Chemistry
13 papers in training set
Top 0.1%
1.4%
15
Nature Communications
5641 papers in training set
Top 49%
1.3%
16
International Journal of Molecular Sciences
494 papers in training set
Top 12%
1.1%
17
Nucleic Acids Research
1281 papers in training set
Top 11%
1.1%
18
ACS Omega
105 papers in training set
Top 2%
1.1%
19
mAbs
32 papers in training set
Top 0.5%
0.8%
20
BMC Bioinformatics
457 papers in training set
Top 6%
0.8%
21
Artificial Intelligence in the Life Sciences
13 papers in training set
Top 0.3%
0.8%
22
Computational and Structural Biotechnology Journal
242 papers in training set
Top 7%
0.8%
23
The Journal of Physical Chemistry B
167 papers in training set
Top 2%
0.8%
24
Scientific Reports
3612 papers in training set
Top 75%
0.8%
25
Frontiers in Immunology
638 papers in training set
Top 11%
0.6%
26
eLife
5828 papers in training set
Top 70%
0.6%
27
Chemical Science
73 papers in training set
Top 2%
0.6%