Design and characterization of SAKe, a new building block for protein self-assembly

preprint OA: closed
📄 Open PDF Full text JSON View at publisher

Abstract

ABSTRACT The nanofabrication of functional protein-based surfaces is challenging due to the chemical complexity of proteins and their unpredictable behavior at the solid-liquid interface. Many proteins of interest –such as antibodies or large enzymatic complexes – lack strong and dynamic protein-protein and protein-surface interactions necessary to drive self-assembly of stable arrays with high surface coverage. Additionally, adsorption-induced conformational changes at the solid-liquid interface could lead to a loss of activity and increase the risk of undesirable interfacial processes. Here we introduce SAKe, a kelch-like designer protein, as a versatile platform to address these challenges. Ancestral sequence reconstruction led to high thermal stability, and the high symmetry allowed modularity of the protein’s core. Rational engineering of the bottom side allowed SAKe to form large (up to 5 micrometers in length), well-defined and pH-dependent two-dimensional assemblies while maintaining structural integrity, which is key for further development of functional materials. SAKe self-assembly was investigated through in-liquid atomic force microscopy on muscovite mica. High resolution imaging confirmed the integrity of the SAKe protein upon adsorption on the solid-liquid interface. These results showcase the SAKe protein as a platform for the further engineering of functional protein-based two-dimensional materials.
Full text 76,304 characters · extracted from oa-pdf · 9 sections · click to expand

Keywords

protein design, assembly, nanomaterials (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint

Abstract

The nanofabrication of functional protein -based surfaces is challenging due to the chemical complexity of proteins and their unpredictable behavior at the solid-liquid interface. Many proteins of interest –such as antibodies or large enzymatic complexes – lack strong and dynamic protein -protein and protein -surface interactions necessary to drive self-assembly of stable arrays with high surface coverage. Additionally, adsorption- induced conformational changes at the solid-liquid interface could lead to a loss of activity and increase the risk of undesirable interfacial processes. Here we introduce SAKe, a kelch-like designer protein, as a versatile platform to address these challenges. Ancestral sequence reconstruction led to high t hermal stability, and the high symmetry allowed modularity of the protein’s core. Rational engineering of the bottom side allowed SAKe to form large (up to 5 micrometers in length ), well-defined and pH -dependent two- dimensional assemblies while maintaining structural integrity , wh ich is key for further development of functional materials. SAKe self-assembly was investigated through in - liquid atomic force microscopy on muscovite mica. High resolution imaging confirmed the integrity of the SAKe protein upon adsorption on the solid -liquid interface. These results showcase the SAKe protein as a platform for the further engineering of functional protein- based two-dimensional materials. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint

Introduction

Nanotechnology aims to achieve precise control over functional nanomaterials at the atomic scale.1 Biological systems, particularly proteins, offer a promising starting point due to their chemical variability. Common strategies to functionalize relevant surfaces such as muscovite mica and graphene are direct adsorption or covalent linkage to reactive moieties such as thiols, amines or carboxylic groups.2–6 For example, the group of Duncan covalently immobilized plasma fibronectin onto polystyrene via thiol or amine groups. While this resulted in improved chemical stability, it did not produce a well - organized protein layer , illustrated by high heterogeneity in the z -plane.5 To improve spatial control over the localization of reactive groups on the surface, a scaffolding protein layer can be used. A well-studied model system of this kind is the S-layer protein (SLP).6– 8 S-layer proteins re adily form two -dimensional arrays with tunable periodicity . These proteins have been successfully used for the functionalization of surfaces with both biomolecules and inorganic materials such as metal nanoparticles.9 While covalent linkage results in great stability of the proteins at the interface, it relies on highly reactive groups that are broadly distributed on the protein surface. Such little control over the interaction between the protein and the interface leads to unpredictable and heterogeneous protein orientation at the solid -liquid interface, which relates to low surface coverage.5,10,11 Moreover, covalent attachment of proteins at interfaces directly impacts a key determinant of crystallinity at the solid -liquid interface. Covalent bonds prevent molecular diffusion and molecular rearrangement on the surface, impeding proteins from adopting the correct orientation with respect to nucleation sites, thereby hindering the growth of defect -free crystalline arrays .12 The need of such dynamic behavior is clearly seen in the formation of SLP arrays. The group of Magalí described such dynamic behavior imaging three different phases for the formation of a crystalline layer of SLP on muscovite mica. They reported that initially proteins are randomly adsorbed on the surface, then the nucleation sites start growing reaching a final stage where proteins suff er a conformational change to stabilize the assemblies. These processes required up to one hour to reach the final stable assembly.13 Physisorption of the proteins onto the surface , on the other hand, often leads to protein denaturation as well as heterogeneity of possible protein orientations at the solid -liquid interface; especially when using hydrophobic surfaces .14,15 The group of Craighead directly adsorbed ConA (a lectin protein) on graphene. This immobilization resulted in the loss of the binding capacity to oligosaccharides.10 Thus, showcasing how arbitrary orientation of the proteins at the interface limits accessibility of their functional sites . Additionally, inefficient surface coverage and protein unfolding leave open areas for non- specific interactions which may decrease the sensitivity of biosensing, or lead to undesired by-products in catalysis.6,14–16 (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint Rational p rotein engineering can increase surface coverage and homogenize protein orientation. For example, the group s of Tezcan and Zhan achieved highly symmetrical protein assemblies by introducing disulfide bonds or metal coordination sites to L - Rhamnulose-1-phosphate aldolase and tobacco virus coat protein respectively.17,18 Ju Song and coworkers used non-canonical amino acids to drive protein self-assembly. 2D crystals were assembled in solution by placing bipyridyl -alanine groups at geometrically interesting positions in a D3 homohexamer .19 Lastly, efficient protein surface assembly can be achieved by computational redesign of either protein-protein interfaces or protein- surface interfaces. This was demonstrated by the design of an α -helical repeat protein capable of highly efficient self-assembling on mica by the Baker group.20,21 A common trait in rationally designed self -assembling systems is the use of protein symmetry. Near-perfect symmetry is commonly found in proteins and especially in oligomers.22–24 It is hypothesized that symmetry facilitates protein folding and has a strong correlation with protein function. The geometry of protein complexes can be controlled by tuning the symmetry of the building blocks.22,25–27 Symmetric placement of protein contact points largely facilitates the interaction between molecules as well as reduces the possible conformations that proteins can adopt within the complex.25 While protein engineering has made significant advances in the nanofabrication of two - dimensional materials, current approaches still depend on complex protein modifications or on the use of intermediate molecules and/or scaffolding proteins to achieve high surface coverage. Despite intensive effort, protein scaffolds capable of forming densely packed two-dimensional arrays combined with the ability to implement a desired function remain underexplored. Our group previously designed and engineered a fully symmetric beta propeller, Pizza.28 This scaffold was further engineered for polyoxometalate coordination ,29 metal-free catalysis30 as well as to scaffold the crystallization of salt nanocrystals.31 Despite these varied applications, the application potential of Pizza was limited given the lack of a large flexible protein surface. Following the same design principles used in the engineering of Pizza,28 a new pseudo-symmetric protein scaffold was used as a starting point, this time targeting proteins mediating protein-protein interactions as their natural function. Here, we report on the computational design and biophysical characterization of a modular protein building block (SAKe) inspired by the kelch protein family. SAKe proteins were designed as candidate proteins for achieving this dual requirement of self-assembly and function . These proteins not only show great self -assembling properties both in solution and on surface but also can tolerate extensive mutations to further develop them into functional building blocks. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint

Results

AND DISCUSSION Design and characterization of the SAKe Protein Scaffold The Kelch folding domain was used as a starting point for the design of a stable symmetric scaffold. Kelch repeat proteins are β -propeller proteins composed of 6 nearly identical tandem sequence repeats that each fold into four -stranded anti-parallel sheets (referred to as blades) around a central cavity.32,33 While the core of these blades is well conserved, the top-side loops vary both in sequence and length (Figure 1), influencing the protein's function. The majority of natural Kelch domains modulate protein -protein interactions (PPIs), especially in humans as an adaptor protein of the CUL3 complex.34 However, the loops which mediate the protein interactions as observed in the CUL complexes on some Kelch domains are also found to harbor an enzymatic site, demonstrating the versatility of this protein fold.32,33,35 The Kelch domain of the Keap1 protein was selected as the starting point for proteins design as RADAR analysis 36 revealed it to consist of the most conserved repeats of all crystalized Kelch domains exhibiting a highly symmetrical tertiary structure. Proteins with a high degree of symmetry have not only been shown to have superior stability to their non-symmetric counterparts, but it has shown also to play a key role in molecular self - assembly.25,37,38 Designing symmetrical building blocks is important as the number of available contact points to drive and stabilize the complex increases equally around the molecule. This is key as the structure of the final oligomer is often determined by the initial seeding of the monomers.37,39,40 Next, an ensemble of SAKe proteins was generated from the human Keap1 β -propeller using the RE3Volutionary design procedure (Figure 1).28 Because the Keap1 template is composed of blades carrying loops of three different lengths, Rosetta symmetry docking was used to generate three different C 6 symmetry starting backbones .41 All models retained a velcro closure type identical to the parent Keap1 protein42 (Figure S1). Initially, the backbones from the 2 nd and 6 th blades of the Keap1 kelch domain were chosen to generate the type A and type B SAKe backbones, respectively. These two blades were chosen as initial design backbones since they accommodate the most common loop length (six amino acids) observed in human propellers. This way, the two most common loop lengths of the Keap1 kelch repeats could be incorporated into the new SAKe designs: ten amino acids for type A SAKe (S6A) and six for type B (S6B)s. Three different S6A (S6AE, S6AR and S6AC) and S6B (S6BE, S6BR and S6BC) scaffolds were selected for experimental evaluation. The appended letters refer to the (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint selection criterium used: Rosetta Talaris2013 energy score (E) ,43 RMSD deviation from the idealized symmetric backbone architecture (R), or a combined rank score from criterium E and R (C). Analysis of the energy landscapes after RE 3Volutionary design showed that the original Keap1 repeat sequences had worse scores compared to most inferred sequences. All six proteins were successfully expressed and purified from E. coli BL21 (DE3) and circular dichroism (CD) spectroscopy showed spectra matching that of the parent template Keap1 β -propeller (Figure S2) . We identified a positive 233 nm CD signal, representing tertiary structure features, that was used to derive SAKe's melting temperatures. The results show exceptional thermal stability, with S6BE exhibiting a melting temperature (T m) of over 95 °C. This is significantly higher than the melting temperature of the template Keap1 protein (Tm of 44.1 ° C) (Figure 1). The methodology used for the design of the SAKe backbone relies on ancestral sequence reconstruction. This method is known for yielding very stable proteins as it biases the sequence reconstruction on structurally relevant amino acids. Repetition of such sequences into a symmetric globular protein explains the increase in thermal stability of the different SAKe scaffolds respectively to their natural counterpart (Keap1).28,44–47 In order to confirm correct folding of the SAKe proteins, all purified proteins were subjected to crystallography. All proteins crystallized within days, diffracted with a resolution ranging from 1.3 to 1.95 Å and were successfully phased with their design er templates, confirming in this way the accuracy of the initial predictions (Figure S3). (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint Figure 1: SAKe design strategy. The six blades of Keap 1 kelch domain (1ZGK) were isolated (A) and used to construct an MSA (B) in order to create a fully symmetrical kelch protein using the RE3Volutionary method. The B-factors of the different loop lengths indicate the high flexibility of the loops as the length increases (C). The high stability of the SAKe protein is reflected by very high melting temperatures. As the loop length increases, the flexibility of the top loops increases, what notably reduces the thermal stability of the proteins (D). SAKe loop variability. Loop modifiability is necessary for designing a protein binding interface, for example antibody-like protein binders, and for later functionalization of protein -based surface assemblies.48 To assess loop modifiability, three loop variants of the two most stable C 6 symmetry SAKes: S6AC and S6BE were created. The L1 loop was created by conserving the intersection of S6A-type and S6B-type loops, while adding four amino acids. For L2, a larger part of the S6A-type loop was conserved with the insertion of four amino acids. In both cases, these inserted four AA were randomly generated following the amino acid occurrence derived from a database of known nanobody CDR motifs .49 Only sequences containing at least one tyrosine and histidine were accepted, as these residues are often found to mediate interactions between antibody and antigen ,50 while symmetric histidine arrangements were used before to scaffold metals and metaloxo clusters .29,31,51 The resultant sequence motifs are not (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint commonly found in the natural human Kelch proteins. L3 describes a 3 -fold symmetric variant with two long loops of various lengths: The longest loop in the L3 configuration was extracted from the crystal structure of the Kelch domain of human KBTBD5 .52 S6AC/S6BE-L1, S6AC/S6BE -L2 and S6AC/S6BE -L3 have loop lengths of 9, 12 and alternating 10 -15 amino acids, respectively. Finally, to assess the influence of loop composition on protein stability, we also transferred the original S6BE loop to S6AC, yielding S6AC-LB (Table S1 ). All new variants were successfully purified, and the structures of S6BE loop variants and S6AC -LB were confirmed via X -ray diffraction following an identical approach described above (Figure 1). With a Tm of 51.7 °C, the SAKe with the longest loops (10-15 AA) is still more stable than the natural Keap1 β-propeller (4-9 AA loops, Tm of 44.1 °C) (Figure 1). The shortest loop (6AA) shows a Tm exceeding 95 °C and longer loops progressively lower thermostability. Interestingly, S6AC seemed the most moldable scaffold, retaining high stability even when carrying the loops which significantly destabilized S6BE (Figure S4). Hence, S6AC core proved to accommodate a larger range of loops lengths. S6AC was designed from blade 1 of the Keap 1 protein with a loop length of ten amino acids unlike S6BE which initial blades contained only six amino acids. This small difference biased the S6AC core to accept longer loop sequences. For the continuation of this work we however continued with S6BE as S6AC derivatives precipitated in a buffer screening experiment at even mildly acidic pH, whereas the S6BE appeared to crystallize spontaneously. Engineering of the self-assembling SAKe While protein crystallization cannot directly point towards protein self -assembly, there is a strong interplay between self -assembling forces and the crystallization process .53 Particles that self -assemble deliver faster crystal growth as the energy barrier for nucleation is substantially reduced .54 This idea was exploited to engineer new protein - protein and protein -surface interfaces for S6BE to improve self-assembly on mica surfaces. As a first step to engineer a SAKe protein capable of self-assembling on the mica surface, the PPIs needed to be understood to find key residues stabilizing such a process. For this, S6BE was submitted to a pH titration experiment which highlighted t he importance of neutralizing the protein charge to trigger spontaneous self-assembly. The high order of symmetry of S6BE together with its high pH stability led to spontaneous assembly of mm sized crystals in mildly acidic conditions (pH < pI), disassembling only at 1 pH point above the pI of the protein (Figure 2). The self-assembled crystals showed a hydrogen bonding network stabilizing lateral contacts (growth on the longitudinal direction) and vertical contacts (stacking proteins one on top of another). The charge residues at the bottom (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint loops (asparagine 30 and aspartic acid 31) stabilized the PPIs laterally interacting with the backbone of the neighboring protein. At the same time, proteins stack on top of each other stabilized by the tyrosine 57 at the top loops and the arginine at the bottom loops of the first blade (and equivalent for the other five blades) (Figure 2). This led to the hypothesis that introduction of residues favoring surface -based assembly while preventing this vertical stacking of proteins would favor the formation of monolayers, by reducing the chances of having a multi -layered assembly or an on -surface adsorbed crystal. Histidines are very versatile amino acid s and are often used to drive self -assembly by metal coordination through the protonated π nitrogen.55–58 Another great trait of histidine residues is the imidazole ring, which behaves as an aromatic ring and is known to interact via π - π stacking with surrounding imidazole rings.59–61 These two interaction strategies are of great advantage when designing a self -assembling system. In the case of the π - π stacking not being strong enough to drive the self-assembly, metals could be introduced in the system to trigger the coordination and hence ease the formation of nucleation points. Therefore, two histidine residues were place d in exchange of the glutamic acid and the arginine residue creating a bis-his clamp to coordinate Zinc cations in order to gain control over the self-assembling process of the SAKe protein (Figure 2). Histidine residues were introduced by stepwise substitution of the residues involved in the hydrogen bonding with the backbone of the neighboring proteins stabilizing the lateral interaction (R30 and E31 and subsequent residues for each blade). S6BE -3HH (7OPA) with substitutions at the bottom protrusions of the 2nd (E76H and R77H), 4th (E170H and R171H) and 6th (E264H and R265H) blades (C3 rotational symmetry) , and S6BE-6HH where the substitutions were done at the bottom protrusions of all the blades (C6 rotational symmetry) were designed to optimize the interaction between the proteins and the mica surface (Figure S6). As a control to investigate whether both histidine residues are required for the self -assembly on the mica surface of the SAKe protein, two intermediate mutants were created, where the blades 2nd, 4th and 6th were mutated back either the first histidine residues (S6BE -3EH) or the second histidine residues (S6BE - 3HR). The contact points of the S6BE crystals were of electrostatic nature as the protein-protein interface was stabilized by hydrogen bonds, indi cating a possible charge dependency (Figure 2). In order to test such hypothesis, the in-solution self-assembling properties of the new variants (S6BE -3HH, S6BE-6HH, S6BE-3EH and S6BE -3HR) and S6BE were tested through dialysis experiments conducted at various pH values. The pH at which S6BE and S6BE -3HH assembled was found to be around 4 and 4.5 for S6BE -6HH. Notably, S6BE-3HH crystals required a week to grow, while the others assembled within a day (Figure S5). This delay in crystal growth could be attributed to the loss of the (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint rotational symmetry and hence a higher entropy change upon assembly .62 The S6BE crystals remained stable up to pH 5 while the crystals of S6BE -6HH remained up to pH 7. This difference on the disassembling pH could be understood from two fronts. First, the isoelectric point of S6BE-6HH is one point higher than that of S6BE. In a system that is highly controlled by the electrostatics of the monomers, maintaining charge neutrality increases the likelihood of monomers interacting in a stable manner. Second, replacing two charged residues (the asparagine and glutamic acid) at the sides of the proteins by histidines which remain neutral at pH 7, reduces the electrostatic repulsion and increases the pH of disassembly.63 Figure 2: Design of the self -assembling variants: Dialysis experiments showcasing the pH sensitivity of S6BE crystallization and the pH sensitivity of the assemblies (A). Similarly to (7ONC) a P1 symmetry was obtained, where proteins form hydrogen bonding laterally but also vertically. The electrostatic surface at pH 4 (left) and pH 5.5 (right) are displayed (B). Based on the contact points of the S6BE crystals, 4 different SA variants were designed. Histidine residues were introduced in the 2 nd, 4th and 6 th blade or in all blades (C). Schematic representation of the p3 intended monolayer 2D lattice. When the crystals formed at pH 4 were studied, a C121 symmetry was obtained, where proteins interacted laterally similar as in 7ONE, but also vertically, this time through the top loops. Still, both proteins readily reassembled upon lowering the pH again, highlighting the reversibility. To study the interactions driving the self -assembly, the crystals were analyzed using X -ray diffraction (Figure S11). The packing arrangement was nearly (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint identical to the one obtained with vapor diffusion at pH 7.0 and at pH 8.0 for S6BE (PDB 7ONE). The symmetry and dipole -like character of S6BE along with the shape complementarity likely facilitated self-assembly into a tightly packed hexagonal structure where a network of hydrogen bonds and hydrophobic interactions further stabilizes the assembly. A s pH increases above the pI of the proteins, residues Asp11, Asp23, Glu20, Glu28 and His15 (and the equivalent for the other blades) will lose the protonation state and the negative net charge of the protein is restored, triggering disassembly. One of the engineered contact points was a hydrogen bond between the arginine residue and the backbone of the neighboring protein. This hydrogen bond was expected to not be altered by the change of pH as the p Ka of the side chain of the arginine residue is around 10. Hence, in order to test whether both residues, the glutamic acid and the arginine were required to be mutated off, two intermediate SAKe mutants were designed, S6BE-3EH and S6BE -3HR (Figure S6). These new mutants were submitted to the dialysis experiment as well and followed a similar trend as their parent protein. This time, the crystals could only be visualized after 48h incubation at pH 4. These crystals, for either protein, dissolved at pH 5 .5 agreeing with the hypothesis of a pH driven crystallization (Figure S5). These results indicate that the high order symmetry of the SAKe scaffold is sufficient for the formation of the crystals. The addition of double histidine moieties in each blade su bstantially increases the stability of these crystals and renders them less sensitive to pH variations. Moreover, the ir introduction was expected to impact the interaction between the protein and the mica surface, having a key effect on stabilizing the assemblies on the surface. Metal-induced self-assembly Metal-mediated protein self -assembly is not a new technique. It is commonly found in nature55,56 and has been used by many to engineer new PPIs points in protein building blocks. Working with a symmetric protein it is theorized to bring an advantage when it comes to engineering these new metal -mediating points, as it substantially reduces the search space for designing metal coordination points. Metals offer a list of benefits when compared to noncovalent interactions. Metal coordination bonds are deemed to be of stronger nature and highly controllable when working with metals like Zn 2+. Moreover, they offer a certain degree of directionality implied by the coordination geometry preferred by the metal, which is advantageous when designing protein assemblies. Bis-his clamps were introduced in order to coordinate Zn 2+ with the purpose of forming open honeycombs on the mica surface (S6BE-3HH) (Figure 2). The effect of adding zinc (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint cations was first studied by dynamic light scattering (DLS) to assess how long and at which ratios bigger molecules would form (Figure S7). As metal coordination is pH - sensitive, three different values were used in these experiments, pH 5, 6 and 7 (Figure S7). These values were chosen as beneath pH 5 already led to the formation of crystals in solution, so the appearance of peaks representing larger complexes would not indicate a direct effect of metal coordination. Another reason is that the histidine residues need to be neutral in order to coordinate metal cations .64,65 When zinc nitrate was added in the solution, there was a clear formation of large complexes already after 10 minutes incubation time. It is worth mentioning that at 1:1 ratio, a large fraction of protein remained in the monomer state, most likely due to insufficient number of metal cations to interact with the proteins. S6BE-3HH has, potentially, three coordination points per protein. At a ratio of 1:1, not all these coordination points will be occupied. At excess of metal cations, it is more clear how proteins start to easily coordinate these metals and form what could be protein aggregates or p rotein assemblies. To invest igate whether these macrostructures were ordered, the 1:20 excess metal condition was visualized in the AFM (Figure S8). A protein concentration of 16 µM led to direct aggregation and so, the protein concentration was reduced to 1 µM to perform the analysis. After 20 minutes of incubation (the time required to prepare the solution and to calibrate the AFM parameters), the surface exhibited no discernible order, indicating the formation of protein aggregates rather than protein assemblies. Interestingly, the control experiment led to the formation of ordered assemblies on the surface , what prompted the hypothesis that metal coordination was not necessary for the self-assembling of the SAKe proteins. On surface self-assembly On-surface assembly on mica was investigated through amplitude -modulated atomic force microscopy (AFM). Since pH was anticipated to act as the self -assembly trigger, protein deposition was carried out using solutions spanning a pH range from 4 to 7. S6BE showed a large number of proteins adsorbed onto the surface at pH 4, although no ordered assemblies were observed. At pH 7, no adsorbed protein was detected, most likely due to electrostatic repulsion between the negatively charged protein and the negatively charged mica substrate. In contrast , S6BE -3HH and S6BE -6HH exhibited enhanced on-surface self-assembly. Unlike S6BE, both variants formed assemblies on the substrate, though with distinct features. For S6BE-3HH, self-assembled arrays were observed across all tested pH values except pH 6 , which is slightly above the isoelectric points of S6BE -3HH (5.32) . Notably, the (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint arrays observed at pH 7 exhibited more uniformity, with most proteins oriented along their vertical axes, an interaction that is showcased by the fiber -like structures imaged at the surface. At pH 4 and 5, the arrays imaged did not exhibit long-range order nor uniformity. This diversity on morphologies highlights the critical role of electrostatic interactions. In the case of pH 7 favoring radial over tangential growth, while at pH 4 and 5 the re is no clear preference of growth ( Figure 4). As discussed for S6BE, the protonation state of the protein governs the strength of the interaction between the proteins and the mica surface. Protonation of Nπ and N τ nitrogens is expected to stabilize the protein on the mica surface by interacting with the negatively charged hydroxyl groups. With increasing pH, deprotonation of these nitrogens weakens the protein –surface interaction, thereby limiting the formation or stabilization of ordered assemblies on the surface. The necessity of this double histidine moiety for the formation of assemblies is showcased by the lack of assemblies found on the mica surface for S6BE -3EH and S6BE -3HR (Figure S9). These two proteins failed to form assemblies on the surface while crystals successfully formed on solution at pH 4. S6BE-3EH does not have the first histidine mutation which is expected to point towards the mica interface, stabilizing the proteins and the assemblies on the surface. S6BE-3HR on the other hand has a relatively large and flexible residue at the interface between neighboring proteins, destabilizing the PPIs and hence reducing the stability of the assemblies on the surface as the degrees -of-freedom of the proteins are reduced. In contrast to S6BE -3HH, S6BE-6HH formed extended hexagonal assemblies at pH 4 and 5 that were readily visualized and resulted in high surface coverage. Upon increasing the pH to 6, these assemblies began to deteriorate and completely disappeared at pH 7 (Figure 4). The observed structures followed the packing geometry anticipated from the design of the self-assembling variants (Figure 5). The proteins most likely anchor to the mica surface via their bottom loops, while lateral interactions between the protein sides establish the contact points that stabilize the lattice. This arrangement would promote on- surface assembly with the top loops exposed to the solvent, rendering them accessible for subsequent functionalization (Figure 5). To achieve sub-monolayer coverage of the mica surface, the protein concentration was reduced to 0.01 uM, which is 100-fold lower than the working concentrations used for S6BE and S6BE -3HH, further highlighting the robustness and stability of the S6BE-6HH assemblies on the surface (Figure S10). Interestingly, the on-surface assemblies do not fully mirror the trends observed in solution. Although increasing the pH beyond the protein’s pI markedly affects the overall charge , producing pronounced effects for S6BE and S6BE -6HH, the loss of protein –surface interactions, and thus on -surface self-assembly, occurs at lower pH values than those required to disrupt assembly in solution. With respect to assembly kinetics, no discernible differences were observed between S6BE -3HH and S6BE -6HH. For both variants, assemblies formed within the first 10 minutes (corresponding to the time needed for AFM (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint cantilever calibration) and remained stable throughout imaging. However, the two - dimensional assemblies formed by S6BE -3HH do not replicate the crystalline lattice observed in its self -assembled crystals, whereas those formed by S6BE -6HH closely resemble the lattice architecture obtained in the crystalline state (Figure S11, S12). Figure 4: SAKe self -assembly on muscovite mica: Overview of AFM topography images acquired on muscovite mica for S6BE, S6BE-3HH, and S6BE -6HH at progressively increasing pH values (pH 4 – pH 7), highlighting the critical role of the presence and protonation state of the incorporated histidine residues in governing on-surface assembly. The scale bar is 50 nm length. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint As only single-layer assemblies are observed, with no evidence of multilayered or overlaid crystal precipitation, it’s unlikely that these on-surface assemblies are the result of the precipitation of preformed crystals in solution (Figure S1 2). Moreover, if crystal precipitation was the dominant mechanism , S6BE would be expected to form large assemblies, which is inconsistent with the experimental observations (Figure 4). Figure 5: High resolution imaging and assembly mechanism on Mica: High resolution images of assemblies of S6BE -6HH on muscovite mica showcasing the vast surface coverage as well as the lattice periodicity . The scale bars in the figure represent from left to right 200, 50 and 20 nm respectively (A). Independent domains were consistently found on the surface arising from different directions of growth . Scale bar represents 50 nm (B). Analysis of the lattice periodicity shows unit vectors characteristic of C6 symmetry, with a repeat distance corresponding to the dimension of a single protein. These experimentally derived vectors are in agreement with those predicted by the structural model (C-F). A schematic representation illustrates the proposed interaction between the mica surface and the histidine residues located in the bottom loops of S6BE-6HH, highlighting their role in anchoring and stabilizing the on-surface assembly (G). The proposed mechanism underlying SAKe protein self -assembly extends beyond protein–protein interactions (PPIs) and incorporates a crucial contribution from protein – surface interactions. Muscovite mica becomes negatively charged upon immersion in aqueous solution due to the desorption of interfacial potassium ions. During overnight incubation, small protein nuclei are expected to form in solution; these nuclei subsequently precipitate and interact with the mica surface. This interaction is likely mediated by the histidine residues engineered on the bottom face of the protein (Figure S6). Only under conditions where the histidines remain protonated favor (i) stable (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint anchoring these nuclei to the surface and (ii) continued adsorption of monomers from the bulk solution, followed by lateral assembly growth (Figure 4). Because nucleation occurs randomly across the surface, growing arrays expand outward until they encounter neighboring domains. When two misaligned arrays meet, a domain boundary is formed (Figure S13). The templating influence of the surface is further evidenced by the angular relationship between adjacent domains. A consistent angle of 160° ± 5° is observe d, reflecting the pseudo -C6 symmetry imposed by the mica lattice (Figure 5). Additional support for a surface-mediated assembly mechanism is provided by the clear differences in packing between the S6BE-3HH crystal structure (Figure S11) and the smaller surface- confined arrays (Figure S12), indicating that the mica substrate modulates the packing orientation of the protein.

Conclusion

Our design strategy highlights the value of using pseudo -symmetric natural scaffolds for the design of lattice forming building blocks. The variability of the Kelch fold implies a wide variety of loop motifs are supported while the bottom remains modifiable for surface interaction. High-resolution AFM confirmed that addition of surface -facing histidines successfully formed ordered arrays on muscovite mica (up to 5 µm 2). The intrinsic symmetry of the SAKe protein together with its charge distribution strongly contributed to the stabilization a nd the in -solution and on -surface morphology of the assemblies . The additional modifiability of this scaffold makes it a valuable building block for technologies such as biosensing or on-surface catalysis. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint

Acknowledgements

We thank Prof. Tatjana P. Vogt for granting us use of the CD spectrometer. We thank the beamline scientists at Diamond Light Source, Swiss Light Source, European Synchrotron Radiation Facility and Elettra Synchrotron and their scientists for their assistan ce. This work was supported by Research Foundation Flanders (FWO) (1S89918N, G0F9316N and G051917N, ZKE-1919-04-W01) and by KU Leuven-Internal Funds (C14/23/090). (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint

Materials

& Methods Computational design SAKe was designed using the RE3Volutionary protein design method28 using the KELCH domain of human Keap1 (PDB code: 1ZGK) as a template.33 Rosetta symmetry docking was used to construct C 6 symmetric backbone models, using blade structures extracted from the human Keap1 β-propeller.41,66 We used the 2nd blade for backbone modeling of A-type SAKe, and the 6th blade to design the B type SAKe backbone. For each run, 20000 backbone models were generated. The models were filtered on Rosetta symmetry docking energy scores and RMSD from a manually constructed symmetric backbone. Clustal Omega was used to generate multiple sequence alignments (MSAs) from the six repeats of the Keap1 β-propeller67. When designing B type SAKe, the conserved VAPM motive was erroneously replaced by VAPL during MSA. The MSAs and their accompanying unrooted phylogenetic trees were used to construct lists of putative ancestral sequences using the FastML server,68 with 250 sequences per node for a total of 1000 sequences for each SAKe construct. The ancestral sequences were repeated 6- fold to fit the full length of the symmetric backbones. The putative ancestral sequences were mapped on their corresponding backbone models using a custom PyRosetta script, which was written by prof. D. Simoncini and is available on our laboratory GitHub repository (https://github.com/kullbmd/kullbmd).28 Talaris2013 energy scores and RMSD calculated against the input backbone model were used to filter for viable designs .43 For all SAKe constructs except S6AR, cysteine residues were then mutated to serine or alanine, depending on their locations. Randomized loop generation First, a Python2.7 script was used to calculate the percentage of occurrence for each amino acid in the iCAN database file, which contains various nanobody CDR motifs.49 As input, we provided a TXT file containing all iCAN database sequence strings. Then, another Python2.7 script was used to randomly generate `b' number of sequences with length `l'. The script uses the previously determined percentages of occurrence and selects only sequences which contain both a Tyr and a His. These scripts are available in the supplementary information (Supplementary Scripts). Cloning Amino acid sequences were first reverse translated into DNA sequences, using a codon optimization tool by IDT (Integrated DNA Technologies, Haasrode, Belgium). DNA was ordered as gBlocks from IDT and subsequently subcloned in pET -28a(+) via NdeI and XhoI restriction sites. Alternatively, DNA was pre -cloned in pET-28a(+) plasmids by BGI (BGI genomics, China). Primers were bought from IDT. All PCR reactions followed the PhusionTM High-Fidelity DNA Polymerase protocol (Thermo Fisher Scientific, (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint Massachusetts, United States). All restriction enzymes were bought from ThermoFisher (FastDigest product line). Ligations were performed with the T4 DNA Ligase from Promega (Promega, Wisconsin, United States). DNA was purified using the GeneJET Plasmid Miniprep Kit, GeneJET PCR Purification Kit and GeneJET Gel Extraction kit from Thermo Fisher Scientific. Before Sanger sequencing by LGC (LGC genomics, Teddington, United Kingdom), recombinant vectors were transformed to chemically competent E. coli DH5α using standard heat shock transformation. For protein expression, DNA isolated from single colonies was transformed to chemically competent E. coli BL21 (DE3) via heat shock transformation. Protein expression E. coli BL21 (DE3) transformed with protein-encoding pET28 plasmids, were grown in 1 L LB (with 0.05 mg/mL kanamycin) at 37 °C, while shaking. At an OD 600 of 0.6, the cells were left to cool on ice for 20 min. After adding 0.5 mM IPTG incubation continued at 20 °C for 16-18 h while shaking. Cells were harvested via centrifugation at 6721 g. The supernatant was discarded, and pellets were immediately stored at -24 °C. Protein purification SAKe proteins: Pellets were thawed on ice and suspended in lysis buffer (40 mL, 50 mM NaH2PO4, 200 mM NaCl, 10 mM imidazole, 1 mM phenylmethylsulfonyl fluoride (PMSF), 30 mg hen eggwhite lysozyme (HEWL), pH 8.0). They were then incubated for 30 min at 15 °C while rotating head-over-head. After, they were lysed via sonication. Lysates were centrifuged at 20216 g for 30 min. The Supernatant was filtered (0.45 µm) and loaded on a 5 mL Ni-NTA column equilibrated with buffer A (5 CV, 50 mM NaH2PO4, 200 mM NaCl, 10 mM imidazole, pH 8.0). The column was washed with buffer A (10 CV, 50 mM NaH2PO4, 200 mM NaCl, 10 mM imidazole, pH 8.0) and buffer B (10 CV, 50 mM NaH2PO4, 200 mM NaCl, 20 mM imidazole, pH 8.0) and the proteins eluted with buffer C (10 CV, 50 mM NaH2PO4, 200 mM NaCl, 300 mM imidazole, pH 8.0). SDS PAGE was used to verify presence of the proteins in the eluate fractions. The fractions containing the proteins of interest were collected and dialyzed overnight in phosphate buffer (50 mM NaH2PO4, 200 mM NaCl, pH 8.0), using 7000 kDa cut-off SnakeSkinTM Dialysis tubing (Thermo Fischer Scientific). At the same time, histidine tags were removed via thrombin digestion (100 U). The dialyzed sampl es were subjected to an additional Ni -NTA chromatography step, where the proteins now appear in the flow through. Next, the proteins were concentrated via ultrafiltration and injected on a Superdex 200pg 16/600 or Superdex 75pg 16/600 column (Cytiva, Hoegaarden, Belgium) equilibrated with elution buffer. UV280 absorbance peaks were collected and concentrated via ultrafiltration to stocks of 20 mg/mL or more. These stock proteins were stored at 4 °C. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint S6BE, S6BE-3HH and S6BE-6HH for AFM: were purified as described above with the exception that 5 mM EDTA was added before SEC. Keap1: Because of the solvent - exposed cysteine, 1 mM freshly prepared DTT is supplemented in every step. This includes the sample preparation steps for SDS PAGE and crystallization. Circular dichroism spectroscopy CD spectroscopy was performed with a JASCO J -1500 spectrometer (JASCO Inc., Maryland, United States). To measure the CD spectra, protein samples were diluted (400 µL, 0.1 mg/mL) in phosphate buffer (20 mM mM NaH 2PO4, pH 7.6). Ellipticity was measured at 20 °C from 260 nm to 200 nm, using 1 mm cuvettes. Five accumulations were averaged. To estimate melting temperatures, protein samples were first diluted (400 µL, 0.05 mg/mL) in phosphate buffer (20 mM mM NaH2PO4, pH 7.6). Then, ellipticity was measured from 260 nm to 200 nm, from 5 to 95 °C with intervals of 5 °C, using sealable 2 mm cuvettes. A Python script was used to plot the data as a heatmap. Three accumulations were averaged. For accurate determination of melting temperatures, samples were diluted (400 µL, 0.25 mg/mL) in phosphate buffer (20 mM mM NaH 2PO4, pH 7.6). The signal at 233 nm was followed from 5 to 95 °C with intervals of 0.2 °C, using sealable 2 mm cuvettes. The data was analyzed with a Python script, which fits a Boltzmann-sigmoid equation and extracts the midpoint Tm. The parameters are described in a previous publication by Mylemans et al.38 Dynamic light scattering Concentrated S6BE, S6BE-3HH, S6BE-3EH or S6BE-3HR (20 mg/mL, 20 mM HEPES, pH 8.0) were diluted with either MES (20 mM, pH 5.6) or MilliQ to a concentration of 1 mg/mL. Metal suspensions (Cu(NO 3)2 or Zn(NO3)2, MilliQ) were titrated to achieve the desired ratio of protein:metal. Size measurements were obtained at 25 °C using a Zetasizer Nano ZS instrument (Malvern Panalytical, Malvern, United Kingdom) and quartz cuvette (ZEN2112). Data analysis was performed us ing the Zetasizer software 7.11 (Malvern Panalytical). Atomic force microscopy Proteins for AFM experiments were obtained from stock solutions (in 20mM HEPES, 200 mM NaCl, pH8) with 50% glycerol and stored at -80 °C. The samples were dialyzed overnight in order to prepare them with the imaging buffer, and further diluted to the working concentration. The buffers for the experiments were prepared according to the working pH. Buffers at pH 4 and 5 were prepared with 20 mM Na -Acetate, equilibrating the pH with 100% acetate. Buffers at pH 6 were prepared with 20 mM MES (2 -(N- morpholino)ethanesulfonic acid). The pH was equilibrated with 6N NaOH. And the buffer (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint at pH 7 was prepared with 20 mM of HEPES (4 -(2-hydroxyethyl)-1- piperazineethanesulfonic acid). The pH was equilibrated with 6N NaOH. To prevent cross contamination, all the utensils required to perform the analysis were cleaned by five minutes of sonication in isopropanol, water MQ quality and 100% ethanol spectroscopy quality and dried with N 2 or Argon gas. All the images were taken in a Cypher S AFM Microscope from Oxford Instruments Asylum Research with a liquid perfusion cell (Cypher ES) and Fast -scanning high -frequency silicon probes (FS - 1500AUD) from Oxford Instruments. Excitation of the t ip was done by blue Drive photothermal excitation. AFM images were processed using Scanning Probe Image Processor (SPIP, Image Metrology ApS), and Gwyddion software. Protein crystallization Crystals were grown via sitting-drop vapour diffusion using NeXtal Crystal Screening kits (Qiagen), SWISSCI MRC 96 -well plates (Molecular Dimensions Inc) and a Crystal Gryphon robot (Art Robbins Instruments). For native crystallography droplets consisted of 0.3 µL buffer and 0.3 µL protein (10 mg/mL in 20 mM HEPES pH 8.0, 200 mM NaCl). Plates were incubated at 20 °C. Protein crystals were vitrified after single -step soaking. PEG 400 or glycerol were used as cryoprotectant. X -ray diffraction experiments were performed at Diamond Light Source (United Kingdom), European Synchrotron Radiation Facility (France), Elletra Synchrotron (Italy) and Swiss Light Source (Switzerland). pH induced crystallization To determine the pH tipping point of assembly of S6BE, S6BE -3HH and S6BE -6HH (5mg/mL), the three proteins were dialyzed at 20 °C in 500 mL of 50 mM citrate buffer at pH 4.0, pH 4.5, pH 5.0, pH 5.5, pH 6.0, pH 6.5; and in 500 mL of 20 mM HEPES buffer at pH 7.0, pH 7.5 an pH 8.0. To test the reversibility of the self -assembly of the three proteins, two experiments were carried out c oncurrently. One where the three proteins were dialyzed from pH 4.0 to pH 8.0, and one from pH 6.5 to pH 4.0. The latter would indicate the pH at which the crystals were readily visible. Such pH value (pH 4.0 for S6BE and S6BE -3HH and pH 4.5 for S6BE -6HH) was then used to dialyze the respective proteins upon dissolution of the crystals. For each protein, a self -assembled crystal was soaked in cryo -protectant, vitrified and shipped to Diamond Light Source (United Kingdom) and Elettra Synchrotron (Italy) for X -ray diffraction. Pictures of the self - assembled crystals were taken with a Nikon SMZ800N microscope, outfitted with a TV Lens C 0.45x (Nikon, Tokyo, Japan). Xray diffraction analysis (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint Diffraction patterns were indexed using XDS or DIALS.69,70 Data reduction was done with Aimless in CCP4 .71,72 Resolution cut -offs were chosen to achieve an outer shell completeness of at least 92.5 %, while also satisfying the requirements of outer shell Rpim below 45 %, CC1/2 above 30 %, and I/sigma(I) above 1.0. Molecular Replacement phasing was done with PHASER, using our computationally designed models as search ensemble.18 Refinement was done manually with phenix.refine and Coot .73,74 The final structures were validated using Molprobity and the PDB validation tool ,75 before being deposited to RCSB PDB. Determination of pI and surface electrostatics Surface electrostatics were calculated at pH 4.0 and 8.0 via PDB2PQR, using a full-length SAKe as input. The calculated surfaces were visualized via PyMOL. pI values were calculated via PROPKA.76–78 (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint

Bibliography

(1) Ariga, K. Nanoarchitectonics: The Method for Everything in Materials Science. Bull. Chem. Soc. Jpn. 2024, 97 (1), uoad001. https://doi.org/10.1093/bulcsj/uoad001. (2) Lim, C. Y.; Owens, N. A.; Wampler, R. D.; Ying, Y.; Granger, J. H.; Porter, M. D.; Takahashi, M.; Shimazu, K. Succinimidyl Ester Surface Chemistry: Implications of the Competition between Aminolysis and Hydrolysis on Covalent Protein Immobilization. Langmuir 2014, 30 (43), 12868–12878. https://doi.org/10.1021/la503439g. (3) Cifuentes, A.; Borrós, S. Comparison of Two Different Plasma Surface-Modification Techniques for the Covalent Immobilization of Protein Monolayers. Langmuir 2013, 29 (22), 6645–6651. https://doi.org/10.1021/la400597e. (4) Vescovi, V.; Kopp, W.; Guisán, J. M.; Giordano, R. L. C.; Mendes, A. A.; Tardioli, P. W. Improved Catalytic Properties of Candida Antarctica Lipase B Multi-Attached on Tailor- Made Hydrophobic Silica Containing Octyl and Multifunctional Amino- Glutaraldehyde Spacer Arms. Process Biochem. 2016, 51 (12), 2055–2066. https://doi.org/10.1016/j.procbio.2016.09.016. (5) Ba, O. M.; Hindie, M.; Marmey, P.; Gallet, O.; Anselme, K.; Ponche, A.; Duncan, A. C. Protein Covalent Immobilization via Its Scarce Thiol versus Abundant Amine Groups: Effect on Orientation, Cell Binding Domain Exposure and Conformational Lability. Colloids Surf. B Biointerfaces 2015, 134, 73–80. https://doi.org/10.1016/j.colsurfb.2015.06.009. (6) Picher, M. M.; Küpcü, S.; Huang, C.-J.; Dostalek, J.; Pum, D.; Sleytr, U. B.; Ertl, P. Nanobiotechnology Advanced Antifouling Surfaces for the Continuous Electrochemical Monitoring of Glucose in Whole Blood Using a Lab-on-a-Chip. Lab. Chip 2013, 13 (9), 1780. https://doi.org/10.1039/c3lc41308j. (7) Pum, D.; Breitwieser, A.; Sleytr, U. B. Patterns in Nature—S-Layer Lattices of Bacterial and Archaeal Cells. Crystals 2021, 11 (8), 869. https://doi.org/10.3390/cryst11080869. (8) Pum, D.; Sleytr, U. B. Reassembly of S-Layer Proteins. Nanotechnology 2014, 25 (31), 312001. https://doi.org/10.1088/0957-4484/25/31/312001. (9) Ilk, N.; Egelseer, E. M.; Sleytr, U. B. S-Layer Fusion Proteins—Construction Principles and Applications. Curr. Opin. Biotechnol. 2011, 22 (6), 824–831. https://doi.org/10.1016/j.copbio.2011.05.510. (10) Alava, T.; Mann, J. A.; Théodore, C.; Benitez, J. J.; Dichtel, W. R.; Parpia, J. M.; Craighead, H. G. Control of the Graphene–Protein Interface Is Required To Preserve Adsorbed Protein Function. Anal. Chem. 2013, 85 (5), 2754–2759. https://doi.org/10.1021/ac303268z. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint (11) Roach, P.; Farrar, D.; Perry, C. C. Interpretation of Protein Adsorption:  Surface- Induced Conformational Changes. J. Am. Chem. Soc. 2005, 127 (22), 8168–8173. https://doi.org/10.1021/ja042898o. (12) Tekin, C.; Caroprese, V.; Bastings, M. M. C. Dynamic Surface Interactions Enable the Self-Assembly of Perfect Supramolecular Crystals. ACS Appl. Mater. Interfaces 2024, 16 (43), 59040–59048. https://doi.org/10.1021/acsami.4c11813. (13) Stel, B.; Cometto, F.; Rad, B.; De Yoreo, J. J.; Lingenfelder, M. Dynamically Resolved Self-Assembly of S-Layer Proteins on Solid Surfaces. Chem. Commun. 2018, 54 (73), 10264–10267. https://doi.org/10.1039/C8CC04597F. (14) Saha, B.; Evers, T. H.; Prins, M. W. J. How Antibody Surface Coverage on Nanoparticles Determines the Activity and Kinetics of Antigen Capturing for Biosensing. Anal. Chem. 2014, 86 (16), 8158–8166. https://doi.org/10.1021/ac501536z. (15) Barbosa, A. I.; Edwards, A. D.; Reis, N. M. Antibody Surface Coverage Drives Matrix Interference in Microfluidic Capillary Immunoassays. ACS Sens. 2021, 6 (7), 2682– 2690. https://doi.org/10.1021/acssensors.1c00704. (16) Schoenbaum, C. A.; Schwartz, D. K.; Medlin, J. W. Controlling the Surface Environment of Heterogeneous Catalysts Using Self-Assembled Monolayers. Acc. Chem. Res. 2014, 47 (4), 1438–1445. https://doi.org/10.1021/ar500029y. (17) Suzuki, Y.; Cardone, G.; Restrepo, D.; Zavattieri, P. D.; Baker, T. S.; Tezcan, F. A. Self-Assembly of Coherently Dynamic, Auxetic, Two-Dimensional Protein Crystals. Nature 2016, 533 (7603), 369–373. https://doi.org/10.1038/nature17633. (18) Zhang, J.; Wang, X.; Zhou, K.; Chen, G.; Wang, Q. Self-Assembly of Protein Crystals with Different Crystal Structures Using Tobacco Mosaic Virus Coat Protein as a Building Block. ACS Nano 2018, 12 (2), 1673–1679. https://doi.org/10.1021/acsnano.7b08316. (19) Yang, M.; Song, W. J. Diverse Protein Assembly Driven by Metal and Chelating Amino Acids with Selectivity and Tunability. Nat. Commun. 2019, 10 (1), 5545. https://doi.org/10.1038/s41467-019-13491-w. (20) Pyles, H.; Zhang, S.; De Yoreo, J. J.; Baker, D. Controlling Protein Assembly on Inorganic Crystals through Designed Protein Interfaces. Nature 2019, 571 (7764), 251– 256. https://doi.org/10.1038/s41586-019-1361-6. (21) Chen, Z.; Johnson, M. C.; Chen, J.; Bick, M. J.; Boyken, S. E.; Lin, B.; De Yoreo, J. J.; Kollman, J. M.; Baker, D.; DiMaio, F. Self-Assembling 2D Arrays with de Novo Protein Building Blocks. J. Am. Chem. Soc. 2019, 141 (22), 8891–8895. https://doi.org/10.1021/jacs.9b01978. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint (22) André, I.; Strauss, C. E. M; Kaplan, D. B; Bradley, P; Baker, D. Emergence of symmetry in homooligomeric biological assemblies. Proceedings of the National Academy of Science 2008, 105 (42), 16148-16152. https://doi.org/10.1073/pnas.0807576105. (23) Kojić-Prodić, B.; Štefanić, Z. Symmetry versus Asymmetry in the Molecules of Life: Homomeric Protein Assemblies. Symmetry 2010, 2 (2), 884–906. https://doi.org/10.3390/sym2020884. (24) Blundell, T. L.; Srinivasan, N. Symmetry, Stability, and Dynamics of Multidomain and Multicomponent Protein Systems. Proc. Natl. Acad. Sci. 1996, 93 (25), 14243– 14248. https://doi.org/10.1073/pnas.93.25.14243. (25) Workum, K. V.; Douglas, J. F. Symmetry, Equivalence, and Molecular Self-Assembly. Physical Review E. 2006, 73, (3), 031502-031502. https://doi.org/10.1103/physreve.73.031502. (26) Laniado, J; Yeates, T. O. A complete rule set for designing symmetry combination

Materials

from protein molecules. Proceedings of the National Academy of Science 2020, 117 (50), 31817-31823. https://doi.org/10.1073/pnas.2015183117. (27) Goodsell, D. S.; Olson, A. J. Structural Symmetry and Protein Function. Annu. Rev. Biophys. Biomol. Struct. 2000, 29 (1), 105–153. https://doi.org/10.1146/annurev.biophys.29.1.105. (28) Voet, A. R. D.; Noguchi, H.; Addy, C.; Simoncini, D.; Terada, D.; Unzai, S.; Park, S.-Y.; Zhang, K. Y. J.; Tame, J. R. H. Computational Design of a Self-Assembling Symmetrical β-Propeller Protein. Proc. Natl. Acad. Sci. 2014, 111 (42), 15102–15107. https://doi.org/10.1073/pnas.1412768111. (29) Vandebroek, L.; Noguchi, H.; Kamata, K.; Tame, J. R. H.; Meervelt, L. V.; Parac-Vogt, T. N.; Voet, A. R. D. Hybrid Assemblies of a Symmetric Designer Protein and Polyoxometalates with Matching Symmetry. Chem. Commun. 2020, 56 (78), 11601– 11604. https://doi.org/10.1039/D0CC05071G. (30) Clarke, D. E.; Noguchi, H.; Gryspeerdt, J.-L. A. G.; Feyter, S. D.; Voet, A. R. D. Artificial β-Propeller Protein-Based Hydrolases. Chem. Commun. 2019, 55 (60), 8880– 8883. https://doi.org/10.1039/C9CC04388H. (31) Voet, A. R. D.; Noguchi, H.; Addy, C.; Zhang, K. Y. J.; Tame, J. R. H. Biomineralization of a Cadmium Chloride Nanocrystal by a Designed Symmetrical Protein. Angew. Chem. Int. Ed. 2015, 54 (34), 9857–9860. https://doi.org/10.1002/anie.201503575. (32) Adams, J.; Kelso, R.; Cooley, L. The Kelch Repeat Superfamily of Proteins: Propellers of Cell Function. Trends Cell Biol. 2000, 10 (1), 17–24. https://doi.org/10.1016/S0962-8924(99)01673-6. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint (33) Beamer, L. J.; Li, X.; Bottoms, C. A.; Hannink, M. Conserved Solvent and Side-Chain Interactions in the 1.35 Å Structure of the Kelch Domain of Keap1. Acta Crystallogr. D Biol. Crystallogr. 2005, 61 (10), 1335–1342. https://doi.org/10.1107/S0907444905022626. (34) Jiang, Z.-Y.; Lu, M.-C.; You, Q.-D. Discovery and Development of Kelch-like ECH- Associated Protein 1. Nuclear Factor Erythroid 2-Related Factor 2 (KEAP1:NRF2) Protein–Protein Interaction Inhibitors: Achievements, Challenges, and Future Directions. J. Med. Chem. 2016, 59 (24), 10837–10858. https://doi.org/10.1021/acs.jmedchem.6b00586. (35) Ito, N.; Phillips, S. E. V.; Yadav, K. D. S.; Knowles, P. F. Crystal Structure of a Free Radical Enzyme, Galactose Oxidase. J. Mol. Biol. 1994, 238 (5), 794–814. https://doi.org/10.1006/jmbi.1994.1335. (36) Heger, A.; Holm, L. Rapid Automatic Detection and Alignment of Repeats in Protein Sequences. Proteins Struct. Funct. Bioinforma. 2000, 41 (2), 224–237. https://doi.org/10.1002/1097-0134. (37) Nieckarz, D.; Rżysko, W.; Szabelski, P. On-Surface Self-Assembly of Tetratopic Molecular Building Blocks. Phys. Chem. Chem. Phys. 2018, 20 (36), 23363–23377. https://doi.org/10.1039/C8CP03820A. (38) Mylemans, B.; Noguchi, H.; Deridder, E.; Lescrinier, E.; Tame, J. R. H.; Voet, A. R. D. Influence of Circular Permutations on the Structure and Stability of a Six-Fold Circular Symmetric Designer Protein. Protein Sci. 2020, 29 (12), 2375–2386. https://doi.org/10.1002/pro.3961. (39) Laniado, J.; Meador, K.; Yeates, T. O. A Fragment-Based Protein Interface Design Algorithm for Symmetric Assemblies. Protein Engineering, Design and Selection 2021, 34. https://doi.org/10.1093/protien/gzab008. (40) Damasceno, P. F.; Engel, M.; Glotzer, S. C. Predictive Self-Assembly of Polyhedra into Complex Structures. Science 2012, 337 (6093), 453–457. https://doi.org/10.1126/science.1220869. (41) André, I.; Bradley, P.; Wang, C.; Baker, D. Prediction of the Structure of Symmetrical Protein Assemblies. Proc. Natl. Acad. Sci. 2007, 104 (45), 17656–17661. https://doi.org/10.1073/pnas.0702626104. (42) Yadid, I.; Kirshenbaum, N.; Sharon, M.; Dym, O.; Tawfik, D. S. Metamorphic Proteins Mediate Evolutionary Transitions of Structure. Proc. Natl. Acad. Sci. 2010, 107 (16), 7287–7292. https://doi.org/10.1073/pnas.0912616107. (43) Leaver-Fay, A.; O’Meara, M. J.; Tyka, M.; Jacak, R.; Song, Y.; Kellogg, E. H.; Thompson, J.; Davis, I. W.; Pache, R. A.; Lyskov, S.; Gray, J. J.; Kortemme, T.; (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint Richardson, J. S.; Havranek, J. J.; Snoeyink, J.; Baker, D.; Kuhlman, B. Scientific Benchmarks for Guiding Macromolecular Energy Function Improvement. Methods Enzymol. 2013, 523, 109–143. https://doi.org/10.1016/B978-0-12-394292-0.00006-0. (44) Trudeau, D. L.; Kaltenbach, M.; Tawfik, D. S. On the Potential Origins of the High Stability of Reconstructed Ancestral Proteins. Mol. Biol. Evol. 2016, 33 (10), 2633–2641. https://doi.org/10.1093/molbev/msw138. (45) Spence, M. A.; Kaczmarski, J. A.; Saunders, J. W.; Jackson, C. J. Ancestral Sequence Reconstruction for Protein Engineers. Curr. Opin. Struct. Biol. 2021, 69, 131–141. https://doi.org/10.1016/j.sbi.2021.04.001. (46) Broom, A.; Doxey, A. C.; Lobsanov, Y. D.; Berthin, L. G.; Rose, D. R.; Howell, P. L.; McConkey, B. J.; Meiering, E. M. Modular Evolution and the Origins of Symmetry: Reconstruction of a Three-Fold Symmetric Globular Protein. Structure 2012, 20 (1), 161–171. https://doi.org/10.1016/j.str.2011.10.021. (47) Smock, R. G.; Yadid, I.; Dym, O.; Clarke, J.; Tawfik, D. S. De Novo Evolutionary Emergence of a Symmetrical Protein Is Shaped by Folding Constraints. Cell 2016, 164 (3), 476–486. https://doi.org/10.1016/j.cell.2015.12.024. (48) Skerra, A. Alternative Non-Antibody Scaffolds for Molecular Recognition. Curr. Opin. Biotechnol. 2007, 18 (4), 295–304. https://doi.org/10.1016/j.copbio.2007.04.010. (49) Zuo, J.; Li, J.; Zhang, R.; Xu, L.; Chen, H.; Jia, X.; Su, Z.; Zhao, L.; Huang, X.; Xie, W. Institute Collection and Analysis of Nanobodies (iCAN): A Comprehensive Database and Analysis Platform for Nanobodies. BMC Genomics 2017, 18 (1), 797. https://doi.org/10.1186/s12864-017-4204-6. (50) Fellouse, F. A.; Wiesmann, C.; Sidhu, S. S. Synthetic Antibodies from a Four-Amino- Acid Code: A Dominant Role for Tyrosine in Antigen Recognition. Proc. Natl. Acad. Sci. 2004, 101 (34), 12467–12472. https://doi.org/10.1073/pnas.0401786101. (51) Vrancken, J. P. M.; Noguchi, H.; Zhang, K. Y. J.; Tame, J. R. H.; Voet, A. R. D. The Symmetric Designer Protein Pizza as a Scaffold for Metal Coordination. Proteins Struct. Funct. Bioinforma. 2021, 89 (8), 945–951. https://doi.org/10.1002/prot.26072. (52) Canning, P.; Cooper, C. D. O.; Krojer, T.; Murray, J. W.; Pike, A. C. W.; Chaikuad, A.; Keates, T.; Thangaratnarajah, C.; Hojzan, V.; Marsden, B. D.; Gileadi, O.; Knapp, S.; von Delft, F.; Bullock, A. N. Structural Basis for Cul3 Protein Assembly with the BTB-Kelch Family of E3 Ubiquitin Ligases. J. Biol. Chem. 2013, 288 (11), 7803–7814. https://doi.org/10.1074/jbc.M112.437996. (53) Jia, Y.; Liu, X.-Y. Self-Assembly of Protein at Aqueous Solution Surface in Correlation to Protein Crystallization. Appl. Phys. Lett. 2005, 86 (2). https://doi.org/10.1063/1.1846153. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint (54) Zhang, T.-D.; Chen, L.-L.; Lin, W.-J.; Shi, W.-P.; Wang, J.-Q.; Zhang, C.-Y.; Guo, W.- H.; Deng, X.; Yin, D.-C. Searching for Conditions of Protein Self-Assembly by Protein Crystallization Screening Method. Appl. Microbiol. Biotechnol. 2021, 105 (7), 2759– 2773. https://doi.org/10.1007/s00253-021-11188-z. (55) Brodin, J. D.; Ambroggio, X. I.; Tang, C.; Parent, K. N.; Baker, T. S.; Tezcan, F. A. Metal-Directed, Chemically Tunable Assembly of One-, Two- and Three-Dimensional Crystalline Protein Arrays. Nat. Chem. 2012, 4 (5), 375–382. https://doi.org/10.1038/nchem.1290. (56) Salgado, E. N.; Lewis, R. A.; Faraone-Mennella, J.; Tezcan, F. A. Metal-Mediated Self-Assembly of Protein Superstructures: Influence of Secondary Interactions on Protein Oligomerization and Aggregation. J. Am. Chem. Soc. 2008, 130 (19), 6082– 6084. https://doi.org/10.1021/ja8012177. (57) Aldeek, F.; Safi, M.; Zhan, N.; Palui, G.; Mattoussi, H. Understanding the Self- Assembly of Proteins onto Gold Nanoparticles and Quantum Dots Driven by Metal- Histidine Coordination. ACS Nano 2013, 7 (11), 10197–10210. https://doi.org/10.1021/nn404479h. (58) Ludden, M. J. W.; Mulder, A.; Schulze, K.; Subramaniam, V.; Tampé, R.; Huskens, J. Anchoring of Histidine-Tagged Proteins to Molecular Printboards: Self-Assembly, Thermodynamic Modeling, and Patterning. Chem. – Eur. J. 2008, 14 (7), 2044–2051. https://doi.org/10.1002/chem.200701478. (59) Calinsky, R.; Levy, Y. Histidine in Proteins: pH-Dependent Interplay between π–π, Cation–π, and CH–π Interactions. J. Chem. Theory Comput. 2024, 20 (15), 6930–6945. https://doi.org/10.1021/acs.jctc.4c00606. (60) Al-Madhagi, L. H.; Callear, S. K.; Schroeder, S. L. M. Hydrophilic and Hydrophobic Interactions in Concentrated Aqueous Imidazole Solutions: A Neutron Diffraction and Total X-Ray Scattering Study. Phys. Chem. Chem. Phys. 2020, 22 (9), 5105–5113. https://doi.org/10.1039/C9CP05993H. (61) Liao, S.-M.; Du, Q.-S.; Meng, J.-Z.; Pang, Z.-W.; Huang, R.-B. The Multiple Roles of Histidine in Protein Interactions. Chem. Cent. J. 2013, 7 (1), 44. https://doi.org/10.1186/1752-153X-7-44. (62) Newton, A. C.; Groenewold, J.; Kegel, W. K.; Bolhuis, P. G. Rotational Diffusion Affects the Dynamical Self-Assembly Pathways of Patchy Particles. Proc. Natl. Acad. Sci. 2015, 112 (50), 15308–15313. https://doi.org/10.1073/pnas.1513210112. (63) Yang, B.; Adams, D. J.; Marlow, M.; Zelzer, M. Surface-Mediated Supramolecular Self-Assembly of Protein, Peptide, and Nucleoside Derivatives: From Surface Design to (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint the Underlying Mechanism and Tailored Functions. Langmuir 2018, 34 (50), 15109– 15125. https://doi.org/10.1021/acs.langmuir.8b01165. (64) Barraud, P.; Schubert, M.; Allain, F. H.-T. A Strong 13C Chemical Shift Signature Provides the Coordination Mode of Histidines in Zinc-Binding Proteins. J. Biomol. NMR 2012, 53 (2), 93–101. https://doi.org/10.1007/s10858-012-9625-6. (65) Zhou, L.; Li, S.; Su, Y.; Yi, X.; Zheng, A.; Deng, F. Interaction between Histidine and Zn(II) Metal Ions over a Wide pH as Revealed by Solid-State NMR Spectroscopy and DFT Calculations. J. Phys. Chem. B 2013, 117 (30), 8954–8965. https://doi.org/10.1021/jp4041937. (66) Chaudhury, S.; Lyskov, S.; Gray, J. J. PyRosetta: A Script-Based Interface for Implementing Molecular Modeling Algorithms Using Rosetta. Bioinformatics 2010, 26 (5), 689–691. https://doi.org/10.1093/bioinformatics/btq007. (67) Madeira, F.; Park, Y. mi; Lee, J.; Buso, N.; Gur, T.; Madhusoodanan, N.; Basutkar, P.; Tivey, A. R. N.; Potter, S. C.; Finn, R. D.; Lopez, R. The EMBL-EBI Search and Sequence Analysis Tools APIs in 2019. Nucleic Acids Res. 2019, 47 (W1), W636–W641. https://doi.org/10.1093/nar/gkz268. (68) Ashkenazy, H.; Penn, O.; Doron-Faigenboim, A.; Cohen, O.; Cannarozzi, G.; Zomer, O.; Pupko, T. FastML: A Web Server for Probabilistic Reconstruction of Ancestral Sequences. Nucleic Acids Res. 2012, 40 (W1), W580–W584. https://doi.org/10.1093/nar/gks498. (69) Kabsch, W. XDS. Acta Crystallogr. D Biol. Crystallogr. 2010, 66 (2), 125–132. https://doi.org/10.1107/S0907444909047337. (70) Winter, G.; Waterman, D. G.; Parkhurst, J. M.; Brewster, A. S.; Gildea, R. J.; Gerstel, M.; Fuentes-Montero, L.; Vollmar, M.; Michels-Clark, T.; Young, I. D.; Sauter, N. K.; Evans, G. DIALS: Implementation and Evaluation of a New Integration Package. Acta Crystallogr. Sect. Struct. Biol. 2018, 74 (2), 85–97. https://doi.org/10.1107/S2059798317017235. (71) Winn, M. D.; Ballard, C. C.; Cowtan, K. D.; Dodson, E. J.; Emsley, P.; Evans, P. R.; Keegan, R. M.; Krissinel, E. B.; Leslie, A. G. W.; McCoy, A.; McNicholas, S. J.; Murshudov, G. N.; Pannu, N. S.; Potterton, E. A.; Powell, H. R.; Read, R. J.; Vagin, A.; Wilson, K. S. Overview of the CCP 4 Suite and Current Developments. Acta Crystallogr. D Biol. Crystallogr. 2011, 67 (4), 235–242. https://doi.org/10.1107/S0907444910045749. (72) Evans, P. R.; Murshudov, G. N. How Good Are My Data and What Is the Resolution? Acta Crystallogr. D Biol. Crystallogr. 2013, 69 (7), 1204–1214. https://doi.org/10.1107/S0907444913000061. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint (73) Adams, P. D.; Afonine, P. V.; Bunkóczi, G.; Chen, V. B.; Davis, I. W.; Echols, N.; Headd, J. J.; Hung, L.-W.; Kapral, G. J.; Grosse-Kunstleve, R. W.; McCoy, A. J.; Moriarty, N. W.; Oeffner, R.; Read, R. J.; Richardson, D. C.; Richardson, J. S.; Terwilliger, T. C.; Zwart, P. H. PHENIX : A Comprehensive Python-Based System for Macromolecular Structure Solution. Acta Crystallogr. D Biol. Crystallogr. 2010, 66 (2), 213–221. https://doi.org/10.1107/S0907444909052925. (74) Emsley, P.; Lohkamp, B.; Scott, W. G.; Cowtan, K. Features and Development of Coot. Acta Crystallogr. D Biol. Crystallogr. 2010, 66 (4), 486–501. https://doi.org/10.1107/S0907444910007493. (75) Chen, V. B.; Arendall, W. B.; Headd, J. J.; Keedy, D. A.; Immormino, R. M.; Kapral, G. J.; Murray, L. W.; Richardson, J. S.; Richardson, D. C. MolProbity : All-Atom Structure Validation for Macromolecular Crystallography. Acta Crystallogr. D Biol. Crystallogr. 2010, 66 (1), 12–21. https://doi.org/10.1107/S0907444909042073. (76) Olsson, M. H. M.; Søndergaard, C. R.; Rostkowski, M.; Jensen, J. H. PROPKA3: Consistent Treatment of Internal and Surface Residues in Empirical pKa Predictions. J. Chem. Theory Comput. 2011, 7 (2), 525–537. https://doi.org/10.1021/ct100578z. (77) Dolinsky, T. J.; Czodrowski, P.; Li, H.; Nielsen, J. E.; Jensen, J. H.; Klebe, G.; Baker, N. A. PDB2PQR: Expanding and Upgrading Automated Preparation of Biomolecular Structures for Molecular Simulations. Nucleic Acids Res. 2007, 35 (suppl_2), W522– W525. https://doi.org/10.1093/nar/gkm276. (78) Li, H.; Robertson, A. D.; Jensen, J. H. Very Fast Empirical Prediction and Rationalization of Protein pKa Values. Proteins Struct. Funct. Bioinforma. 2005, 61 (4), 704–721. https://doi.org/10.1002/prot.20660. (which was not certified by peer review) is the author/funder. All rights reserved. No reuse allowed without permission. The copyright holder for this preprintthis version posted March 18, 2026. ; https://doi.org/10.64898/2026.03.16.710736doi: bioRxiv preprint

Text is read by the "Ask this paper" AI Q&A widget below. Extraction quality varies by source — PMC NXML preserves structure cleanly, OA-HTML may include some navigation residue, and OA-PDF can have broken hyphenation. The publisher copy (via DOI) is the canonical version.

My notes (saved in your browser only)

Ask this paper AI returns verbatim quotes from the full text · source: oa-pdf

Answers must be backed by verbatim quotes from this paper's full text. Hallucinated quotes are dropped automatically; if no verbatim passage answers the question, we say so. How this works

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2026) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.

Source provenance

europepmc
last seen: 2026-05-20T01:45:00.602351+00:00
unpaywall
last seen: 2026-06-05T02:00:03.366016+00:00