Next Article in Journal
Quercus acuta Acorn Bran Extract Enhances Wound Healing by Promoting Human Dermal Fibroblast Migration and Antioxidant Activity
Previous Article in Journal
Something Old Brings New Insights: Activity of Cerivastatin Against Thermally Dimorphic Fungi and Its Potential as an Antifungal Scaffold
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Review

AI-Driven Design of Miniproteins as Potential Allosteric Modulators

1
Zhejiang Key Laboratory of Soft Matter Biomedical Materials, Wenzhou Institute, University of Chinese Academy of Sciences, Wenzhou 325000, China
2
School of Physical Science and Technology, Ningbo University, Ningbo 315211, China
*
Authors to whom correspondence should be addressed.
Pharmaceuticals 2026, 19(3), 480; https://doi.org/10.3390/ph19030480
Submission received: 9 January 2026 / Revised: 10 March 2026 / Accepted: 12 March 2026 / Published: 14 March 2026
(This article belongs to the Section Biopharmaceuticals)

Abstract

Allosteric modulation has emerged as a powerful strategy for achieving superior selectivity and safety in drug discovery and protein function regulation. Unlike highly conserved orthosteric sites, allosteric pockets are structurally diverse and less evolutionarily constrained, making them particularly suitable for modulation by designed miniproteins. Miniproteins can provide extended binding interfaces and high affinity for shallow, dynamic, or cryptic regulatory surfaces that are often inaccessible to small molecules. Recent advances in artificial intelligence (AI) are transforming this field through deep learning-based structure prediction and generative modeling. These AI-driven approaches enable the identification of allosteric hotspots, characterization of conformational ensembles, and de novo design of structured miniprotein binders. They are rapidly expanding the landscape for designing selective modulators across diverse allosteric targets, including GPCRs, receptor tyrosine kinases, nuclear receptors, ion channels, and other protein–protein interaction systems. This review summarizes state-of-the-art AI-driven computational methodologies for designing miniproteins as potential allosteric modulators and discusses their current challenges and future opportunities in allosteric drug discovery.

Graphical Abstract

1. Introduction

Allosteric modulation is a fundamental mechanism of protein function in which ligand binding at a site distinct from the orthosteric pocket induces conformational or dynamic changes that modulate activity [1,2]. While orthosteric sites are usually highly conserved due to functional constraints [3], allosteric sites are more topologically diverse and less conserved, enabling greater selectivity. Several modes of allosteric modulation have been described (Figure 1A), including negative allosteric modulation (NAM), positive allosteric modulation (PAM), and silent allosteric modulation (SAM), which decrease, enhance, or lack intrinsic efficacy toward orthosteric signaling, respectively [4]. These properties make allosteric sites attractive targets for achieving improved selectivity and reduced off-target effects [4,5,6].
The molecular basis of allostery is intimately tied to protein dynamics and is best understood through the modern ensemble model [2]. It views proteins not as fixed structures but as dynamic ensembles of interconverting conformations governed by an energy landscape [7,8]. Allosteric ligands reshape this energy landscape by shifting the population distribution among conformational states. In the classical Monod–Wyman–Changeux framework [9], the conformational populations are reduced to two limiting states, the low-affinity tense (T) state and the high-affinity relaxed (R) state, whose equilibrium shift upon ligand binding gives rise to sigmoidal enzyme kinetics (Figure 1B). This population shift can occur through multiple mechanisms. These include entropic effects, kinetic coupling, and perturbations of residue interaction networks, even in the absence of observable structural changes [7]. Allosteric modulation often exploits transient or cryptic pockets that are accessible only in specific ensemble states [10]. Consequently, understanding allostery demands integrated structural, thermodynamic, and kinetic analyses. It also requires methods capable of capturing protein motions at atomic resolution across the conformational ensemble [11,12].
Allosteric modulators span diverse chemical classes. Metal ions [13] (e.g., Ca 2 + , Mg 2 + , or Na + ) act as endogenous regulators in numerous proteins. Small molecules remain the predominant modality and have produced clinical agents targeting proteins, yet they often struggle to engage shallow, dynamic, or extended allosteric interfaces [14]. Peptides can access larger surfaces but typically suffer from limited conformational stability and poor pharmacokinetic properties [4,5]. Antibody-based binders provide high affinity and specificity at the cost of large size and restricted access to many allosteric sites [6]. Miniproteins (typically 3–8 kDa), by contrast, occupy an intermediate design space, combining larger interaction surfaces than small molecules with greater compactness and adaptability than antibodies, enabling effective recognition and modulation of complex allosteric interfaces and protein–protein interactions [6,15]. A naturally occurring example of miniprotein-mediated regulation is observed in G protein-coupled receptors (GPCRs) [16] (Figure 1C). These attributes make engineered miniproteins particularly well suited for targeting previously intractable allosteric sites, positioning them as promising scaffolds for next-generation allosteric therapeutics [17,18,19].
To date, AI-driven computational approaches are reshaping the investigation of allosteric modulation by enabling systematic analysis of complex conformational ensembles, identification of cryptic allosteric sites, and de novo design of miniproteins [20,21,22,23,24]. Miniproteins as potential allosteric modulators represent a rapidly advancing frontier, offering high specificity and enhanced interface adaptability for targets long considered “undruggable”. In the following sections, we focus specifically on AI-driven design strategies, including structure analysis and generative algorithms, and discuss how these methods are being leveraged to optimize miniprotein binders for allosteric sites with two recent and representative case studies.

2. AI-Driven Pipeline for Designing Allosteric Miniprotein Modulators

Miniproteins occupy a unique intermediate design space between small molecules and large protein biologics [25,26], thereby providing extended and chemically diverse interaction surfaces that are particularly well suited for engaging shallow, dynamic, and weakly conserved allosteric regulatory sites. Unlike orthosteric pockets, allosteric sites are often transient, conformationally heterogeneous, and weakly conserved, posing fundamental challenges for classical structure-based drug discovery, including the reliance on static structures that fail to capture cryptic or transiently open pockets, leading to difficulties in identifying druggable sites and achieving high-affinity binding [23,27,28].
Prior to the advent of modern AI methodologies, the computational design of miniprotein binders for allosteric modulation primarily relied on structure-guided rational design and physics-based de novo modeling [29]. These approaches leveraged experimentally determined or homology models to identify putative allosteric pockets, infer functional hotspots, and engineer binders through motif grafting, interface redesign, or scaffold repurposing. While successful in selected cases, such strategies were inherently constrained by limited conformational sampling, strong dependence on prior structural knowledge, and the difficulty of explicitly modeling long-range allosteric coupling within complex protein energy landscapes.
Recent advances in AI and machine learning are reshaping this design paradigm [24]. AI-driven pipelines enable a more integrated treatment of allostery by combining conformational ensemble modeling, generative backbone construction, sequence optimization, and binding assessment within a unified computational framework [30]. Collectively, AI-enabled design pipelines are transforming the development of allosteric miniprotein modulators from a largely heuristic, case-by-case endeavor into a scalable and systematic process. By coupling data-driven generative models with biophysically grounded evaluation and experimental feedback, these approaches are significantly accelerating the discovery of functional allosteric binders. In parallel, they are expanding the accessible design space beyond what was achievable with traditional computational methods alone. Figure 2 illustrates the AI-driven computational pipeline for the design of allosteric miniprotein modulators.

2.1. Structure Analysis

2.1.1. Allosteric Pocket Identification

Identifying biologically relevant allosteric pockets is a critical first step toward understanding and engineering regulatory control in proteins [31] (Figure 2). In contrast to orthosteric binding sites, allosteric pockets are frequently shallow, weakly conserved, and highly dependent on the underlying conformational ensemble, many of which are cryptic or only transiently populated [28,32]. This intrinsic dynamical nature substantially limits the effectiveness of static, geometry-based analyses. Although experimental techniques—including NMR spectroscopy, cryo-electron microscopy, hydrogen–deuterium exchange mass spectrometry (HDX-MS), site-directed mutagenesis, and functional perturbation assays—can provide high-resolution insights into dynamic regulatory regions and allosteric communication [33,34], they are typically labor-intensive, costly, and low-throughput. These constraints have motivated the widespread adoption of computational approaches as scalable alternatives for systematic allosteric site discovery.
A broad spectrum of computational methods has therefore been developed to interrogate allostery from a dynamic and network-centric perspective [35,36,37,38,39,40,41]. Molecular dynamics simulations, Markov state models, elastic network models, coevolutionary coupling analysis, and residue interaction network theory have been routinely employed to reveal allosteric pathways, dynamic hotspots, and long-range coupling mechanisms. More recently, machine learning-based models have emerged as particularly powerful tools, leveraging curated structural and dynamical datasets to detect cryptic pockets, predict ensemble shifts, and quantify functional coupling with improved accuracy and reduced computational cost (Table 1) [42,43]. Early AI-driven approaches, such as Allosite [35], framed allosteric pocket identification as a supervised classification task using physicochemical and geometric descriptors, whereas subsequent methods, including AlloPred [36], integrated perturbation information from normal mode analysis to implicitly encode allosteric coupling. Together, these developments have established machine learning as a central paradigm for scalable and mechanistically informed allosteric pocket identification.
Further improvements were achieved by integrating richer structural feature representations and more robust learning strategies. AllositePro  [37] expanded the feature space and optimized model training to enhance prediction stability across diverse protein families. Building on curated allosteric datasets, the PASSer [38] framework adopted ensemble learning to improve generalization and scalability, while PASSer 2.0 [39] further advanced this direction through an AutoML architecture that automates feature selection and model optimization. Rather than treating allosteric pocket identification as a binary classification task, PASSerRank [40] reframed the problem as a learning-to-rank task, enabling prioritization of candidate pockets according to their predicted regulatory relevance. This approach achieved superior performance, ranking known allosteric pockets in the top 3 positions for 83.6% (ASD dataset) and 80.5% (CASBench) of test proteins, with higher F1 scores (0.662 on ASD, 0.608 on CASBench) and Matthews correlation coefficients (0.645 on ASD, 0.589 on CASBench) compared to conventional classifiers.
Most recently, ensemble optimization strategies such as MEF-AlloSite [41] have combined multiple machine learning models with optimized feature selection from thousands of pocket descriptors to achieve improved robustness and accuracy in identifying allosteric regions at both pocket and site levels. Compared to state-of-the-art methods like PASSer 2.0 and PASSerRank, MEF-AlloSite demonstrated statistically significant gains (1–6% higher mean average precision across multiple test sets, with p < 0.05 and Cohen’s d > 0.5), along with enhanced classification accuracy (ranging from 0.452 to 0.620 on varied test cases). These AI-driven approaches transform allosteric pocket identification from heuristic, structure-centric analyses into a data-driven inference problem, providing a systematic and scalable foundation for downstream allosteric modulator and binder design.

2.1.2. Structure Prediction and Ensemble Modeling

Accurate structural models of target proteins are critical for designing allosteric miniprotein binders, as allosteric modulation and ligand recognition are governed by protein dynamics and conformational ensembles rather than a single static structure [44]. Allosteric modulation is fundamentally an ensemble phenomenon, operating through shifts in conformational equilibria on a complex energy landscape, which makes ensemble-level structural information essential for rational binder design [1,2,8]. Traditional experimental structures obtained by X-ray crystallography or cryo-EM typically represent low-energy or highly populated conformations and may fail to capture excited or low-population states that harbor functional allosteric or cryptic binding sites [14,32].
Recent advances in AI-based and MSA-based structure prediction have substantially expanded access to high-quality atomic models (Table 2). Deep learning frameworks such as AlphaFold and RoseTTAFold achieve near-experimental accuracy for monomeric proteins and many protein complexes by leveraging evolutionary information encoded in multiple sequence alignments [45,46,47]. Extensions including AlphaFold-Multimer and AlphaFold 3 further enable modeling of oligomeric assemblies and biomolecular interactions that are directly relevant to allosteric signaling and regulation, with some predictions showing agreement with experimental structures of allosteric proteins and ligand-binding sites [48]. These AI-predicted structures often serve as starting points for exploring conformational variability and for generating structural hypotheses in systems lacking experimental data, although they remain speculative and require experimental validation, particularly in dynamic allosteric systems [24,29].
To address the inherently dynamic nature of allostery, structure prediction is increasingly integrated with ensemble modeling techniques [12,49,50]. Molecular dynamics simulations, together with enhanced sampling approaches such as metadynamics or replica-exchange molecular dynamics, enable systematic exploration of conformational landscapes beyond single predicted structures and provide access to transient or functionally relevant states [2,43]. AI-assisted strategies further guide ensemble generation by biasing sampling toward alternative conformational states inferred from evolutionary couplings, energetic frustration, or learned structure–dynamics relationships [10,22].
Such ensemble models are particularly valuable for miniprotein binder design, as they help identify conformations that expose regulatory surfaces compatible with extended protein–protein interaction interfaces [25,26]. In practical design workflows, ensemble-aware structural models guide the selection of target conformations for binder generation, reducing the risk of designing binders that recognize only rare or non-functional states and improving the likelihood of functional allosteric modulation. Recent AI-driven binder design frameworks explicitly benefit from this ensemble perspective, linking structure prediction and conformational sampling with generative design strategies [17,30].
Table 2. Representative MSA-based AI tools for protein structure prediction and validation.
Table 2. Representative MSA-based AI tools for protein structure prediction and validation.
Tool aMethod and Key FeaturesYearRef.
trRosettaDeep learning model predicting inter-residue distances and orientations from MSA-derived features; early high-throughput deep predictor for fold inference.2020[47]
RoseTTAFoldThree-track neural network integrating sequence, pairwise distances, and 3D coordinates; uses MSAs for accurate monomer and multimer predictions.2021[46]
AlphaFold2Deep learning model using MSA and Evoformer architecture; delivers high-accuracy monomer and complex structure predictions with confidence metrics.2021[45]
AlphaFold-MultimerExtension of AlphaFold2 for protein complex modeling; incorporates paired MSAs to capture inter-chain co-evolutionary signals.2022[51]
AlphaFold3Diffusion-based generative architecture for modeling protein–ligand, protein–DNA, and protein–protein complexes, while still using MSA information.2024[52]
a These tools rely primarily on multiple sequence alignment (MSA) to capture evolutionary and co-evolutionary constraints, and have been widely used for structural validation in binder design workflows.

2.2. Generative Design of Binders

Following structural analysis, the generative design stage focuses on constructing miniprotein binders that are compatible with target allosteric sites in terms of geometry, energetics, and dynamics. Unlike small-molecule design, which primarily emphasizes pocket occupancy, miniprotein design seeks to engineer extended protein–protein interfaces. These interfaces are intended to engage regulatory surfaces and stabilize specific conformational states along the target protein’s energy landscape. Accordingly, generative models are required not only to produce binders with high affinity but also to shape interactions that bias conformational equilibria underlying allosteric modulation [53].
AI-driven generative frameworks treat binder design as a conditional generation problem [54]. In this setting, backbone topology and amino acid sequences are sampled under explicit constraints defined by the target structure, binding geometry, and functional objectives (Figure 2). These pipelines typically decompose the design task into backbone generation, sequence design, and integrated co-optimization steps. Such modularization enables efficient exploration of the vast combinatorial space associated with miniprotein binders (Table 3).

2.2.1. Backbone Generation

Backbone generation represents the first and most structurally constrained stage of de novo miniprotein binder design, in which three-dimensional protein scaffolds are generated to geometrically complement a target binding surface. At the miniprotein scale, backbone topology critically determines fold stability, surface curvature, and the spatial organization of interface residues. Consequently, backbone generation models must balance physical realism with sufficient flexibility to accommodate diverse protein–protein interaction geometries.
Recent advances in this area have been driven primarily by generative models operating directly in three-dimensional coordinate space, with diffusion- and flow-based approaches emerging as dominant paradigms [55,56,57,58]. RFdiffusion exemplifies diffusion-based backbone generation, formulating scaffold design as an iterative denoising process that transforms random coordinate noise into structured protein backbones consistent with learned geometric priors [55]. By conditioning generation on target surface geometry, interface residue positions, or rigid-body constraints, RFdiffusion enables the design of miniprotein backbones that are explicitly shaped to engage challenging protein surfaces, including shallow and discontinuous allosteric regions.
In contrast, complementary generative models such as Chroma [59] and FoldFlow [58] primarily focus on general de novo protein backbone generation rather than target-conditioned binder design, and thus currently offer limited direct utility for high-precision binder scaffolding, despite their potential as future extensible frameworks. These backbone-generation methods aim to produce physically plausible and foldable miniprotein scaffolds that define a viable structural substrate for downstream optimization. However, backbone-only generation does not ensure functional binding, as interface chemistry and energetic complementarity are not explicitly resolved at this stage. Backbone generation therefore serves primarily to establish geometric feasibility, while sequence design and integrated structure–sequence optimization are required to achieve functional miniprotein binders.

2.2.2. Sequence Design

Following backbone generation, sequence design assigns amino acid identities that stabilize the intended fold and mediate favorable interactions with the target surface. This inverse folding problem is particularly stringent for miniprotein binders, where limited sequence length balances the coupling between folding stability, interface specificity, and conformational robustness. AI-driven sequence design approaches address this challenge by learning conditional probability distributions over sequence space given a fixed three-dimensional backbone.
Graph-based neural architectures trained on large structural datasets have become central to AI-driven sequence design. ProteinMPNN [60] represents a widely adopted message-passing neural network that encodes residue–residue spatial relationships and predicts amino acid probabilities compatible with a given backbone, enabling efficient and accurate sequence optimization for miniprotein scaffolds. In addition to computational benchmarks, ProteinMPNN-designed sequences have been experimentally validated in functional allosteric modulation contexts, such as redesign of ubiquitin variants that bind and allosterically activate Rsp5 E3 ligase in vitro [61]. These studies support the utility of ProteinMPNN for generating sequences that not only fold correctly but also exert measurable regulatory effects on target proteins. PiFold [62] adopts a related graph neural network paradigm optimized for scalability and computational efficiency, facilitating rapid sequence design cycles for compact proteins.
An alternative class of inverse folding models leverages pretrained protein language models. ESM-IF1 [63] projects structural information into semantically enriched embedding spaces learned from large-scale sequence data, implicitly encoding evolutionary and biophysical constraints. This language-model-based strategy enables the generation of sequences that are not only structurally compatible but also evolutionarily plausible.
Across these approaches, the unifying objective is to generate sequences that reliably fold into the designed backbone while forming specific, energetically favorable contacts at the binder–target interface. Because these models typically assume a fixed backbone, their effectiveness motivates the development of integrated frameworks that jointly optimize structure and sequence, as introduced in the following.

2.2.3. Integrated Binder Generation

While backbone generation and sequence design can be executed as separate stages, integrated binder generation frameworks aim to co-optimize structure and sequence within a unified generative process. This integration is particularly advantageous for miniprotein binders, where the interface geometry, residue composition, and folding stability are tightly coupled.
Several integrated AI-based approaches relying on structure prediction feedback have guided generative optimization. AlphaProteo (2024) [64] and AlphaDesign (2025) [65] employ AlphaFold-assisted evaluation to iteratively refine binder sequences and conformations toward structurally stable and functionally competent states. In these frameworks, structure prediction acts as an implicit physical filter that biases generative sampling toward favorable folding and interaction profiles. O-design [66] emphasizes objective-driven interface refinement by combining energy-based scoring with deep learning–guided sequence optimization.
Some other recent methods adopt fully end-to-end generative paradigms. BindCraft [67] implements an automated one-shot design pipeline that integrates backbone generation, sequence assignment, and confidence-based filtering, achieving experimental hit rates of 10–100% for de novo miniprotein binders across diverse targets (e.g., 13/53 for PD-1, 7/9 for PD-L1, and 4/16 for CD45). BoltzGen [68] extends this concept through an all-atom generative framework that unifies structure and sequence sampling, enabling universal binder design across diverse protein targets, with wet-lab validation yielding nanomolar-affinity binders for 66% of nine novel targets (for both nanobody and general protein designs). Similarly, PXDesign [69] provides an end-to-end pipeline combining generative modeling with rigorous post hoc confidence assessment to prioritize experimentally viable designs.
PPDiff [70] represents a joint sequence–structure diffusion framework capable of directly generating protein–protein complexes, including miniprotein binders, within a single generative process. By modeling interface formation as an emergent property of coupled structure and sequence generation, such approaches further blur the boundary between backbone synthesis and sequence design.
Together, these integrated generative systems represent a significant advance in computational binder design, enabling coherent exploration of structure–sequence space and providing scalable, AI-driven routes to engineer high-affinity and high-specificity miniprotein binders for challenging protein targets.
Table 3. Key AI Methods for Miniprotein Binder Design.
Table 3. Key AI Methods for Miniprotein Binder Design.
CategoryToolCore CapabilityYearRef.
Backbone generationRFdiffusionDiffusion-based backbone generation conditioned on target interfaces for stable miniprotein scaffolds2023[55]
Sequence generationProteinMPNNInverse folding-based sequence design for fixed backbone miniproteins2022[60]
ESM-IF1Protein language model-based inverse folding for sequence design on fixed miniprotein backbones2022[63]
PiFoldGraph neural network-based inverse folding enabling efficient miniprotein sequence design2022[62]
Integrated design
of backbone and sequence
AlphaProteoAlphaFold-assisted binder design emphasizing functional interaction motifs2024[64]
BindCraftAutomated one-shot de novo miniprotein binder design with high experimental hit rates2024[67]
O-designObjective-driven interface refinement via energy-based and deep learning-assisted sequence optimization2025[66]
AlphaDesignAlphaFold-guided hallucination with diffusion-based sequence optimization for multistate binder design2025[65]
BoltzGenAll-atom generative model unifying structure and sequence for universal binder design, including miniproteins2025[68]
PXDesignEnd-to-end de novo binder design pipeline (generation plus confidence filtering) with high experimental success rates2025[69]
PPDiffJoint sequence–structure diffusion framework for direct generation of protein–protein complexes and miniprotein binders2025[70]

2.3. Selection and Optimization

2.3.1. Screening and Structure Validation

After generative modeling, large numbers of de novo miniprotein binders must be computationally screened to identify candidates that are structurally reliable and likely to engage the target (Figure 2) [71]. In this stage, MSA-based structure prediction models, particularly AlphaFold2, serve as high-precision filters. Although these binders typically lack natural evolutionary homologs, AlphaFold’s implicit structural and physical priors provide stringent checks on foldability and topology. Key screening metrics include the predicted Local Distance Difference Test (pLDDT), which estimates residue-level structural confidence, and predicted aligned error (pAE), which evaluates the relative positioning of residue pairs across the binder–target interface [72]. Designs with low pLDDT, high interface pAE, or significant deviations from the intended backbone (e.g., measured via RMSD) are efficiently filtered, ensuring that only geometrically plausible candidates proceed to downstream validation.
Complementary to AI-driven screening, physics-based scoring functions, most commonly implemented in Rosetta, assess atomic-level interactions and interface quality. The Rosetta interface binding free energy change ( Δ Δ G ) estimates the energetic contribution of the binder–target interface, while surface complementarity and hydrophobic packing are evaluated using metrics such as the solvent-accessible surface area penalty (SAP_score) [73]. Integrating AlphaFold2 confidence metrics with Rosetta energy-based evaluations provides a balanced assessment of both structural plausibility and interface competency [18,19]. Designs that satisfy both criteria are prioritized as high-confidence candidates for experimental characterization, ensuring that selected binders are geometrically consistent, energetically favorable, and likely to function as intended.

2.3.2. Partial Diffusion

Partial diffusion is a refinement technique in computational protein design, implemented within the RFdiffusion framework [55]. It works by selectively adding noise to certain regions of a protein structure and then denoising them to generate optimized variations, while keeping other regions fixed. In the context of miniprotein binder design, this approach enhances affinity, diversity, and specificity. RFdiffusion can selectively re-diffuse only a subset of high-ranked backbones, preserving key structural elements such as the overall fold or anchoring interactions at the allosteric site. This targeted resampling enables local structural optimization without disrupting previously identified favorable binding geometries.
Following partial diffusion, redesigned backbones are subjected to sequence redesign, and the resulting models are re-evaluated using AlphaFold2 confidence metrics, including pLDDT and interface pAE. Designs that exhibit improved structural confidence or interface definition are retained and recycled as inputs for subsequent rounds of partial diffusion. By iterating this cycle of constrained backbone resampling, sequence optimization, and AlphaFold2-based evaluation, the overall quality and reliability of allosteric miniprotein designs can be progressively enhanced [30,74].

2.3.3. Refining Interfaces with Molecular Dynamics

Molecular dynamics (MD) simulations complement static structure predictions by explicitly sampling the conformational ensemble of protein–protein complexes, allowing assessment of interface stability and flexibility under near-physiological conditions. Unlike single snapshot models, trajectory-based MD reveals transient fluctuations at the binding interface and helps identify regions that maintain key contacts versus those that undergo rearrangements, thus providing a physics-based check on conformational plausibility. Such dynamic insights are particularly valuable for the design of miniprotein binders or small-molecule protein–protein interaction (PPI) inhibitors, where maintaining precise interfacial contacts is critical for functional inhibition [75].
Beyond qualitative assessment, MD trajectories can be integrated with post-processing methods, such as clustering representative conformations, estimating relative free energies, or re-scoring with ensemble-based metrics, to refine predicted complexes toward more accurate interfacial geometries and energetics. By combining initial AI-driven design models with MD-derived ensembles, researchers can prioritize candidates with both robust dynamic stability and favorable interaction profiles, facilitating the development of effective miniprotein binders or PPI inhibitors for experimental validation [76].

3. Latest Case Study in AI-Driven Design of Miniprotein Modulators

While direct precedents for AI-driven miniprotein design explicitly targeting allosteric sites remain limited, two recent studies employing de novo design against non-orthosteric pockets provide close analogs. These two approaches modulate protein function through conformational perturbations or interface disruptions, akin to allosteric mechanisms. Below, we introduce these two representative examples (Figure 3 and Table 4).

3.1. Case 1: High-Affinity Binders to the Flpp3 Virulence Factor

A compelling example is the de novo design of high-affinity miniprotein binders targeting Flpp3, a virulence factor from Francisella tularensis (Figure 3A) [77]. Notably, Flpp3 lacks deep pockets or known binding partners, rendering it challenging for conventional small-molecule inhibition. The designed binders aim to target two distinct surfaces on Flpp3. Site I corresponds to an electronegative α -helical face that is hypothesized to mediate membrane interaction, whereas Site II is an electropositive β -sheet face. Binding at these surfaces may disrupt immune evasion and bacterial dissemination through induced conformational changes or modulation of protein–protein interactions. Although not explicitly described as allosteric, this strategy closely parallels classical allosteric modulation.
The design pipeline integrated physics-based docking with deep learning tools. Rotamer interaction fields were generated using RIFGen, scaffold placement was performed with PatchDock, and interface refinement was carried out using RIFDock. ProteinMPNN was employed for sequence design, followed by iterative backbone optimization using Rosetta FastRelax. AlphaFold2 was then used for model filtering. Selection criteria emphasized predicted structural confidence and interface quality, including pLDDT values greater than 80–90, interface pAE below 6, Rosetta Δ Δ G values less than 35 to 40  kcal/mol, and SAP scores below 30–35.
The pipeline began with a library of 43,724 miniprotein scaffolds ranging from 25 to 65 amino acids in length and generated approximately 500,000 docked conformations. This process yielded 15,000 α -site and 8817 β -site candidates for experimental screening. Experimental screening and validation confirmed that several designed miniproteins bind Flpp3 with nanomolar to sub-nanomolar affinity. Yeast surface display–based selection identified a small set of enriched α - and β -site binders adopting three-helix bundle topologies, consistent with the intended scaffold designs.
Biophysical characterization demonstrated high binding affinity and exceptional stability of the top candidates, while structural analyses revealed near-atomic agreement between the designed models and experimentally determined complex structures. These results validate the accuracy of the AI-assisted design pipeline in targeting shallow, non-orthosteric protein surfaces and demonstrate its potential for functional modulation through allosteric-like mechanisms.

3.2. Case 2: Miniprotein Inhibitors of Bacterial Adhesins

Another illustrative example is the design of miniprotein inhibitors targeting chaperone–usher pathway (CUP) adhesins from uropathogenic Escherichia coli and Acinetobacter baumannii, which mediate urinary tract infections (UTIs) (Figure 3B) [78]. Here, we focus on designed F7, a miniprotein inhibitor of the FimH adhesin from E. coli. Rather than directly occupying the orthosteric host receptor-binding site, F7 binds to a pocket adjacent to this site that is preferentially accessible in the low-affinity state (LAS) of FimH. Binding at this site induces a conformational population shift that disfavors the high-affinity state (HAS), thereby exerting allosteric-like inhibitory effects. By stabilizing inactive conformations, F7 disrupts host receptor binding and biofilm formation.
F7 was wholly designed using the AI-driven pipeline to explore the large conformational and sequence space, integrating hotspot-conditioned RFdiffusion, ProteinMPNN sequence optimization, Rosetta interface evaluation, and AlphaFold2 (AF2)-based structural filtering. Approximately 10,000 backbone designs were generated per target using crystal structures of FimH in both the high-affinity (HAS; PDB: 1UWF) and low-affinity (LAS; PDB: 3JWN) states, with hotspot residues proximal to the mannose-binding pocket specified as diffusion constraints. For each backbone, multiple sequences were assigned using ProteinMPNN and subjected to initial AF2 filtering, retaining designs with pLDDT values greater than 80 and predicted aligned error (pAE) below 10.
High-confidence designs were further refined through iterative partial diffusion, followed by sequence reassignment and AF2 evaluation. Final candidates were selected using stringent complex- and monomer-level criteria, including AF2 metrics for the complex (binder pLDDT ≥ 90 and interface pAE ≤ 6–6.5), favorable Rosetta interface metrics ( Δ Δ G 30 kcal/mol and SAP score ≤ 40), and AF2 monomer confidence for the isolated minibinder (pLDDT ≥ 90). Top-ranked designs were synthesized and screened by cDNA display, leading to the identification of F7.
Experimental screening and validation identified F7 as a high-affinity binder that selectively stabilizes the low-affinity state of FimH, inhibiting red blood cell aggregation and biofilm formation. Structural and functional analyses, including X-ray crystallography, NMR spectroscopy, and in vivo UTI models, confirmed the designed binding mode and conformational modulation mechanism.

4. Discussion

Despite the rapid progress of artificial intelligence in protein structure prediction and generative design, several important limitations remain in designing miniproteins as potential allosteric modulators.
First, current AI-based structure prediction frameworks, exemplified by AlphaFold-class models, primarily generate a single high-confidence static structure rather than the full conformational ensemble that proteins populate under physiological conditions [20,52]. This representation is fundamentally mismatched with the intrinsically dynamic nature of proteins and therefore limits direct access to folding pathways, energy landscapes, and low-population but functionally relevant conformational states. Training datasets are also heavily biased toward crystallographic and cryo-EM structures, which often capture stabilized or experimentally trapped conformations. As a result, flexible segments and intrinsically disordered regions are frequently assigned low confidence or poorly defined structures despite their central roles in allosteric communication. In addition, current predictors cannot explicitly distinguish ligand-free and ligand-bound conformational preferences, making it difficult to capture induced-fit or conformational-selection mechanisms or to describe how allosteric perturbations reshape the underlying energy landscape [79].
Second, accurate prediction of allosteric pockets and communication pathways remains challenging. Although AI integrated with molecular simulations has improved the identification of potential regulatory sites, predictions still suffer from substantial false positives and misannotations, and experimental validation remains necessary [21]. Moreover, most approaches focus on static pocket detection and lack mechanistic insight into long-range energy propagation or population shifts within conformational ensembles. Predicting allosteric communication networks often requires extensive molecular dynamics simulations, which remain computationally demanding and limit large-scale application [20].
Third, limitations also exist in AI-driven generative frameworks for the design of de novo miniprotein binders. Current scoring models typically learn from datasets of binding energies or structural metrics and may overfit to known complexes, resulting in limited generalization to unseen folds, scaffolds, or functional interfaces [53,60]. Consequently, predicted scores do not always correlate well with experimental binding affinity or kinetics, and different computational models may produce inconsistent rankings.
Furthermore, although generative models can efficiently propose candidate sequences and backbones, many computationally designed structures fail to fold correctly, express efficiently, or form stable complexes with their targets during experimental validation [53]. The design of novel topologies or previously unobserved interface geometries further challenges model generalization. Evaluation metrics used in current pipelines—including ML-based scores, energy functions, and structure prediction confidence—remain imperfect proxies for experimental success. As a result, the wet-lab success rate of de novo binders is often low (typically <20%), requiring extensive experimental screening [67,68].
Overall, these challenges highlight the need for improved ensemble-aware modeling of protein dynamics, more accurate prediction of allosteric communication mechanisms, and enhanced scoring and validation frameworks that integrate computational design with iterative experimental feedback [71].

5. Conclusions and Future Prospects

Allosteric modulation represents a fundamental principle by which biological systems achieve long-range control over protein activity, signal transduction, and gene regulation. Proteins do not function solely through isolated active sites; rather, they operate as dynamic networks, where local perturbations propagate across spatial distances to modulate enzymatic activity, receptor signaling, oligomerization, and transcriptional output. This capacity for remote control underlies critical biological processes, including ion channel gating, receptor activation, chromatin remodeling, and transcriptional regulation, and is thus considered a “second layer” of regulation beyond primary ligand recognition, providing robustness and tunability to complex signaling systems [80].
From a therapeutic perspective, allosteric modulation offers significant advantages over conventional orthosteric drugs. Orthosteric sites are often highly conserved, particularly within kinases, GPCRs, and nuclear receptors, increasing the risk of off-target binding, cross-reactivity, and dose-limiting toxicity. Even highly specific orthosteric ligands may induce adverse effects by perturbing normal physiological signaling. In contrast, allosteric modulators act at less conserved regulatory sites, enabling higher selectivity, reduced systemic toxicity, and precise modulation of protein function rather than simple inhibition or activation [27,81,82,83]. By modulating conformational equilibria, allosteric binders can bias signaling pathways, control oligomeric states, or selectively affect disease-associated functional states, thus expanding the druggable target space.
Despite these advantages, allosteric modulators have historically been challenging to discover [4,6]. Allosteric sites are often shallow, transient, or only populated in specific conformational states, making them poorly suited to traditional high-throughput screening strategies optimized for small molecules. A lack of structural and mechanistic understanding of long-range coupling further limits rational intervention. Recent advances in cryo-electron microscopy, solution-based structural techniques, and the accumulation of large structural databases have begun to illuminate these regulatory landscapes. Nonetheless, experimental approaches alone are insufficient to systematically explore the vast combinatorial space of potential allosteric binders.
In this context, AI-driven protein design has emerged as a transformative enabling technology. By learning from large-scale structural, sequence, and biophysical data, modern AI models capture the principles governing protein folding, interaction geometry, and conformational plasticity. Unlike traditional optimization strategies that primarily refine existing scaffolds, AI enables genuine de novo exploration of protein sequence and structure space, allowing the design of entirely new architectures tailored to specific regulatory sites. Importantly, AI does not replace experimental screening but refines it: computational design narrows the astronomically large sequence space to a high-quality subset, efficiently guiding subsequent experimental selection and optimization.
Within this paradigm, miniproteins occupy a uniquely advantageous position [15,17,18,19,77,78]. Compared to small molecules, they offer larger and more versatile interaction surfaces, enabling high-affinity and highly selective recognition of shallow or extended allosteric interfaces. Their inherent structural diversity and capacity for multivalent interactions allow effective engagement of regulatory surfaces involved in protein–protein interactions and conformational control. At the same time, miniproteins are smaller than antibodies, allowing access to sterically restricted environments and, in some cases, intracellular targets inaccessible to large biologics. Advances in computational stabilization and sequence optimization further enhance their structural robustness and functional reliability, making them ideal candidates for allosteric modulation. Models such as RFdiffusion [55] enable de novo generation of backbone architectures compatible with target regulatory surfaces, while ProteinMPNN [60] provides efficient sequence optimization for stability and interface complementarity. These tools allow rational construction of novel miniprotein scaffolds tailored to specific allosteric pockets, rather than relying on incremental modification of existing proteins.
Taken together, AI-driven de novo design of miniprotein allosteric modulators represents not merely an incremental improvement but a qualitative shift in drug discovery strategy. Rather than competing with endogenous ligands at highly conserved active sites, this approach leverages regulatory architecture, conformational dynamics, and system-level control to achieve precise modulation of protein function. While challenges remain—including accurate modeling of conformational entropy, reliable prediction of in vivo behavior, and integration of developability constraints—the convergence of AI, structural biology, and protein engineering provides a powerful framework to address them.
Ultimately, as AI-driven design methodologies continue to mature, this paradigm promises to deepen mechanistic understanding of allosteric modulation while enabling novel therapeutic strategies. By coupling de novo protein design with iterative experimental validation, it offers not only new avenues for drug development but also a framework for probing how biological systems encode and transmit regulatory information across molecular scales [84].

Funding

This study is financially supported by the National Natural Science Foundation of China (No. 12474192), Zhejiang Provincial Natural Science Foundation of China (No. LZ24A040002), Beijing National Laboratory for Condensed Matter Physics (No. 2023BNLCMPKF009), Zhejiang Key Laboratory of Soft Matter Biomedical Materials (No. 2025ZY01036, 2025E10072) and Wenzhou Institute, University of Chinese Academy of Sciences (No. WIUCASQD2022019).

Institutional Review Board Statement

Not applicable.

Informed Consent Statement

Not applicable.

Data Availability Statement

No new data were created or analyzed in this study. Data sharing is not applicable.

Conflicts of Interest

The authors declare no conflicts of interest.

References

  1. Fenton, A.W. Allostery: An illustrated definition for the ‘second secret of life’. Trends Biochem. Sci. 2008, 33, 420–425. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  2. Astore, M.A.; Pradhan, A.S.; Thiede, E.H.; Hanson, S.M. Protein dynamics underlying allosteric regulation. Curr. Opin. Struct. Biol. 2024, 84, 102768. [Google Scholar] [CrossRef] [Scilit]
  3. Kenakin, T.; Christopoulos, A. Signalling bias in new drug discovery: Detection, quantification and therapeutic impact. Nat. Rev. Drug Discov. 2013, 12, 205–216. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  4. Mannes, M.; Martin, C.; Menet, C.; Ballet, S. Wandering beyond small molecules: Peptides as allosteric protein modulators. Trends Pharmacol. Sci. 2022, 43, 406–423. [Google Scholar] [CrossRef] [Scilit]
  5. Olson, K.M.; Traynor, J.R.; Alt, A. Allosteric modulator leads hiding in plain site: Developing peptide and peptidomimetics as GPCR allosteric modulators. Front. Chem. 2021, 9, 671483. [Google Scholar] [CrossRef] [Scilit]
  6. Fournier, L.; Guarnera, E.; Kolmar, H.; Becker, S. Allosteric antibodies: A novel paradigm in drug discovery. Trends Pharmacol. Sci. 2025, 46, 311–323. [Google Scholar] [CrossRef] [Scilit]
  7. Kar, G.; Keskin, O.; Gursoy, A.; Nussinov, R. Allostery and population shift in drug discovery. Curr. Opin. Pharmacol. 2010, 10, 715–722. [Google Scholar] [CrossRef] [Scilit]
  8. Yan, Z.; Wang, J. Funneled energy landscape unifies principles of protein binding and evolution. Proc. Natl. Acad. Sci. USA 2020, 117, 27218–27223. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  9. Changeux, J.P. Allostery and the Monod-Wyman-Changeux model after 50 years. Annu. Rev. Biophys. 2012, 41, 103–133. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  10. Yan, Z.; Li, Y.; Cao, Y.; Tao, X.; Wang, J.; Jiang, Y. Binding Specificity and Local Frustration in Structure-based Drug Discovery. Curr. Med. Chem. 2025, 32, 1–11. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  11. Nussinov, R. Introduction to protein ensembles and allostery. Chem. Rev. 2016, 116, 6263–6266. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  12. Lewis, S.; Hempel, T.; Jiménez-Luna, J.; Gastegger, M.; Xie, Y.; Foong, A.Y.; Satorras, V.G.; Abdin, O.; Veeling, B.S.; Zaporozhets, I.; et al. Scalable emulation of protein equilibrium ensembles with generative deep learning. Science 2025, 389, eadv9817. [Google Scholar] [CrossRef] [Scilit]
  13. Shen, S.; Zhao, C.; Wu, C.; Sun, S.; Li, Z.; Yan, W.; Shao, Z. Allosteric modulation of G protein-coupled receptor signaling. Front. Endocrinol. 2023, 14, 1137604. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  14. Chatzigoulas, A.; Cournia, Z. Rational design of allosteric modulators: Challenges and successes. Comput. Mol. Sci. 2021, 11, e1529. [Google Scholar] [CrossRef] [Scilit]
  15. Asada, N.; Krebs, C.F.; Panzer, U. Miniproteins may have a big impact: New therapeutics for autoimmune diseases and beyond. Signal Transduct. Target. Ther. 2024, 9, 298. [Google Scholar] [CrossRef] [Scilit]
  16. Maeda, S.; Xu, J.; Kadji, F.M.N.; Clark, M.J.; Zhao, J.; Tsutsumi, N.; Aoki, J.; Sunahara, R.K.; Inoue, A.; Garcia, K.C.; et al. Structure and selectivity engineering of the M1 muscarinic receptor toxin complex. Science 2020, 369, 161–167. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  17. Cao, L.; Goreshnik, I.; Coventry, B.; Case, J.B.; Miller, L.; Kozodoy, L.; Chen, R.E.; Carter, L.; Walls, A.C.; Park, Y.J.; et al. De novo design of picomolar SARS-CoV-2 miniprotein inhibitors. Science 2020, 370, 426–431. [Google Scholar] [CrossRef] [Scilit]
  18. Berger, S.; Seeger, F.; Yu, T.Y.; Aydin, M.; Yang, H.; Rosenblum, D.; Guenin-Macé, L.; Glassman, C.; Arguinchona, L.; Sniezek, C.; et al. Preclinical proof of principle for orally delivered Th17 antagonist miniproteins. Cell 2024, 187, 4305–4317. [Google Scholar] [CrossRef] [Scilit]
  19. Roy, A.; Shi, L.; Chang, A.; Dong, X.; Fernandez, A.; Kraft, J.C.; Li, J.; Le, V.Q.; Winegar, R.V.; Cherf, G.M.; et al. De novo design of highly selective miniprotein inhibitors of integrins αvβ6 and αvβ8. Nat. Commun. 2023, 14, 5660. [Google Scholar] [CrossRef] [Scilit]
  20. Nussinov, R.; Zhang, M.; Liu, Y.; Jang, H. AlphaFold, artificial intelligence (AI), and allostery. J. Phys. Chem. B 2022, 126, 6372–6383. [Google Scholar] [CrossRef] [Scilit]
  21. Agajanian, S.; Alshahrani, M.; Bai, F.; Tao, P.; Verkhivker, G.M. Exploring and learning the universe of protein allostery using artificial intelligence augmented biophysical and computational approaches. J. Chem. Inf. Model. 2023, 63, 1413–1428. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  22. Xiao, S.; Verkhivker, G.M.; Tao, P. Machine learning and protein allostery. Trends Biochem. Sci. 2023, 48, 375–390. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  23. Huang, J.; Tang, G.; Liu, N.; Li, X.; Lu, S. Recent advances in computational strategies for allosteric site prediction: Machine learning, molecular dynamics, and network-based approaches. Drug Discov. Today 2025, 30, 104466. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  24. Koh, H.Y.; Zheng, Y.; Yang, M.; Arora, R.; Webb, G.I.; Pan, S.; Li, L.; Church, G.M. AI-driven protein design. Nat. Rev. Bioeng. 2025, 3, 1034–1056. [Google Scholar] [CrossRef] [Scilit]
  25. Bouchiba, Y.; Ruffini, M.; Schiex, T.; Barbe, S. Computational design of miniprotein binders. In Computational Peptide Science: Methods and Protocols; Springer: Berlin/Heidelberg, Germany, 2022; pp. 361–382. [Google Scholar]
  26. Ożga, K.; Berlicki, Ł. Design and engineering of miniproteins. ACS Bio. Med. Chem. 2022, 2, 316–327. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  27. Govindaraj, R.G.; Thangapandian, S.; Schauperl, M.; Denny, R.A.; Diller, D.J. Recent applications of computational methods to allosteric drug discovery. Front. Mol. Biosci. 2023, 9, 1070328. [Google Scholar] [CrossRef] [Scilit]
  28. Bemelmans, M.P.; Cournia, Z.; Damm-Ganamet, K.L.; Gervasio, F.L.; Pande, V. Computational advances in discovering cryptic pockets for drug discovery. Curr. Opin. Struct. Biol. 2025, 90, 102975. [Google Scholar] [CrossRef] [Scilit]
  29. Kortemme, T. De novo protein design—From new structures to programmable functions. Cell 2024, 187, 526–544. [Google Scholar] [CrossRef] [Scilit]
  30. Fox, D.R.; Taveneau, C.; Clement, J.; Grinter, R.; Knott, G.J. Code to complex: AI-driven de novo binder design. Structure 2025, 33, 1631–1642. [Google Scholar] [CrossRef] [Scilit]
  31. Zhu, R.; Wu, C.; Zha, J.; Lu, S.; Zhang, J. Decoding allosteric landscapes: Computational methodologies for enzyme modulation and drug discovery. RSC Chem. Biol. 2025, 6, 539–554. [Google Scholar] [CrossRef] [Scilit]
  32. Hollingsworth, S.A.; Kelly, B.; Valant, C.; Michaelis, J.A.; Mastromihalis, O.; Thompson, G.; Venkatakrishnan, A.; Hertig, S.; Scammells, P.J.; Sexton, P.M.; et al. Cryptic pocket formation underlies allosteric modulator selectivity at muscarinic GPCRs. Nat. Commun. 2019, 10, 3289. [Google Scholar] [CrossRef] [Scilit]
  33. Wallerstein, J.; Han, X.; Levkovets, M.; Lesovoy, D.; Malmodin, D.; Mirabello, C.; Wallner, B.; Sun, R.; Sandalova, T.; Agback, P.; et al. Insights into mechanisms of MALT1 allostery from NMR and AlphaFold dynamic analyses. Commun. Biol. 2024, 7, 868. [Google Scholar] [CrossRef] [Scilit]
  34. Duan, J.; He, X.H.; Li, S.J.; Xu, H.E. Cryo-electron microscopy for GPCR research and drug discovery in endocrinology and metabolism. Nat. Rev. Endocrinol. 2024, 20, 349–365. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  35. Huang, W.; Lu, S.; Huang, Z.; Liu, X.; Mou, L.; Luo, Y.; Zhao, Y.; Liu, Y.; Chen, Z.; Hou, T.; et al. Allosite: A method for predicting allosteric sites. Bioinformatics 2013, 29, 2357–2359. [Google Scholar] [CrossRef] [Scilit]
  36. Greener, J.G.; Sternberg, M.J. AlloPred: Prediction of allosteric pockets on proteins using normal mode perturbation analysis. BMC Bioinform. 2015, 16, 335. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  37. Song, K.; Liu, X.; Huang, W.; Lu, S.; Shen, Q.; Zhang, L.; Zhang, J. Improved method for the identification and validation of allosteric sites. J. Chem. Inf. Model. 2017, 57, 2358–2363. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  38. Tian, H.; Jiang, X.; Tao, P. PASSer: Prediction of allosteric sites server. Mach. Learn. Sci. Technol. 2021, 2, 035015. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  39. Xiao, S.; Tian, H.; Tao, P. PASSer2. 0: Accurate prediction of protein allosteric sites through automated machine learning. Front. Mol. Biosci. 2022, 9, 879251. [Google Scholar] [CrossRef] [Scilit]
  40. Tian, H.; Xiao, S.; Jiang, X.; Tao, P. PASSerRank: Prediction of allosteric sites with learning to rank. J. Comput. Chem. 2023, 44, 2223–2229. [Google Scholar] [CrossRef] [Scilit]
  41. Ugurlu, S.Y.; McDonald, D.; He, S. Mef-allosite: An accurate and robust multimodel ensemble feature selection for the allosteric site identification model. J. Cheminformatics 2024, 16, 116. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  42. Nerin-Fonz, F.; Cournia, Z. Machine learning approaches in predicting allosteric sites. Curr. Opin. Struct. Biol. 2024, 85, 102774. [Google Scholar] [CrossRef] [Scilit]
  43. Li, R.; He, X.; Wu, C.; Li, M.; Zhang, J. Advances in structure-based allosteric drug design. Curr. Opin. Struct. Biol. 2025, 90, 102974. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  44. Hu, G.; Doruker, P.; Li, H.; Demet Akten, E. Understanding protein dynamics, binding and allostery for drug design. Front. Mol. Biosci. 2021, 8, 681364. [Google Scholar] [CrossRef] [Scilit]
  45. Jumper, J.; Evans, R.; Pritzel, A.; Green, T.; Figurnov, M.; Ronneberger, O.; Tunyasuvunakool, K.; Bates, R.; Žídek, A.; Potapenko, A.; et al. Highly accurate protein structure prediction with AlphaFold. Nature 2021, 596, 583–589. [Google Scholar] [CrossRef] [Scilit]
  46. Baek, M.; DiMaio, F.; Anishchenko, I.; Dauparas, J.; Ovchinnikov, S.; Lee, G.R.; Wang, J.; Cong, Q.; Kinch, L.N.; Schaeffer, R.D.; et al. Accurate prediction of protein structures and interactions using a three-track neural network. Science 2021, 373, 871–876. [Google Scholar] [CrossRef] [Scilit]
  47. Du, Z.; Su, H.; Wang, W.; Ye, L.; Wei, H.; Peng, Z.; Anishchenko, I.; Baker, D.; Yang, J. The trRosetta server for fast and accurate protein structure prediction. Nat. Protoc. 2021, 16, 5634–5651. [Google Scholar] [CrossRef] [Scilit]
  48. He, X.H.; Li, J.R.; Shen, S.Y.; Xu, H.E. AlphaFold3 versus experimental structures: Assessment of the accuracy in ligand-bound G protein-coupled receptors. Acta Pharmacol. Sin. 2025, 46, 1111–1122. [Google Scholar] [CrossRef] [Scilit]
  49. Niazi, S.K.; Yang, J. A comprehensive application of FiveFold for conformation ensemble-based protein structure prediction. Sci. Rep. 2025, 15, 33498. [Google Scholar] [CrossRef] [Scilit]
  50. Cui, X.; Ge, L.; Chen, X.; Lv, Z.; Wang, S.; Zhou, X.; Zhang, G. Beyond static structures: Protein dynamic conformations modeling in the post-AlphaFold era. Briefings Bioinform. 2025, 26, bbaf340. [Google Scholar] [CrossRef] [Scilit]
  51. Evans, R.; O’Neill, M.; Pritzel, A.; Antropova, N.; Senior, A.; Green, T.; Žídek, A.; Bates, R.; Blackwell, S.; Yim, J.; et al. Protein complex prediction with AlphaFold-Multimer. bioRxiv 2021. [Google Scholar] [CrossRef] [Scilit]
  52. Abramson, J.; Adler, J.; Dunger, J.; Evans, R.; Green, T.; Pritzel, A.; Ronneberger, O.; Willmore, L.; Ballard, A.J.; Bambrick, J.; et al. Accurate structure prediction of biomolecular interactions with AlphaFold 3. Nature 2024, 630, 493–500. [Google Scholar] [CrossRef] [Scilit]
  53. Winnifrith, A.; Outeiral, C.; Hie, B.L. Generative artificial intelligence for de novo protein design. Curr. Opin. Struct. Biol. 2024, 86, 102794. [Google Scholar] [CrossRef] [Scilit]
  54. Das, U. Generative AI for drug discovery and protein design: The next frontier in AI-driven molecular science. Med. Drug Discov. 2025, 27, 100213. [Google Scholar] [CrossRef] [Scilit]
  55. Watson, J.L.; Juergens, D.; Bennett, N.R.; Trippe, B.L.; Yim, J.; Eisenach, H.E.; Ahern, W.; Borst, A.J.; Ragotte, R.J.; Milles, L.F.; et al. De novo design of protein structure and function with RFdiffusion. Nature 2023, 620, 1089–1100. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  56. Ahern, W.; Yim, J.; Tischer, D.; Salike, S.; Woodbury, S.M.; Kim, D.; Kalvet, I.; Kipnis, Y.; Coventry, B.; Altae-Tran, H.R.; et al. Atom-level enzyme active site scaffolding using RFdiffusion2. Nat. Methods 2025, 23, 96–105. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  57. Butcher, J.; Krishna, R.; Mitra, R.; Brent, R.I.; Li, Y.; Corley, N.; Kim, P.T.; Funk, J.; Mathis, S.; Salike, S.; et al. De novo design of all-atom biomolecular interactions with rfdiffusion3. bioRxiv 2025. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  58. Bose, A.J.; Akhound-Sadegh, T.; Huguet, G.; Fatras, K.; Rector-Brooks, J.; Liu, C.H.; Nica, A.C.; Korablyov, M.; Bronstein, M.; Tong, A. Se (3)-stochastic flow matching for protein backbone generation. arXiv 2023, arXiv:2310.02391. [Google Scholar]
  59. Ingraham, J.B.; Baranov, M.; Costello, Z.; Barber, K.W.; Wang, W.; Ismail, A.; Frappier, V.; Lord, D.M.; Ng-Thow-Hing, C.; Van Vlack, E.R.; et al. Illuminating protein space with a programmable generative model. Nature 2023, 623, 1070–1078. [Google Scholar] [CrossRef] [Scilit]
  60. Dauparas, J.; Anishchenko, I.; Bennett, N.; Bai, H.; Ragotte, R.J.; Milles, L.F.; Wicky, B.I.; Courbet, A.; de Haas, R.J.; Bethel, N.; et al. Robust deep learning–based protein sequence design using ProteinMPNN. Science 2022, 378, 49–56. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  61. Kao, H.W.; Lu, W.L.; Ho, M.R.; Lin, Y.F.; Hsieh, Y.J.; Ko, T.P.; Danny Hsu, S.T.; Wu, K.P. Robust Design of Effective Allosteric Activators for Rsp5 E3 ligase using the machine learning tool ProteinMPNN. ACS Synth. Biol. 2023, 12, 2310–2319. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  62. Gao, Z.; Tan, C.; Chacón, P.; Li, S.Z. Pifold: Toward effective and efficient protein inverse folding. arXiv 2022, arXiv:2209.12643. [Google Scholar]
  63. Hsu, C.; Verkuil, R.; Liu, J.; Lin, Z.; Hie, B.; Sercu, T.; Lerer, A.; Rives, A. Learning inverse folding from millions of predicted structures. In Proceedings of the International Conference on Machine Learning, PMLR, Baltimore, MA, USA, 17–23 July 2022; pp. 8946–8970. [Google Scholar]
  64. Zambaldi, V.; La, D.; Chu, A.E.; Patani, H.; Danson, A.E.; Kwan, T.O.; Frerix, T.; Schneider, R.G.; Saxton, D.; Thillaisundaram, A.; et al. De novo design of high-affinity protein binders with AlphaProteo. arXiv 2024, arXiv:2409.08022. [Google Scholar] [CrossRef] [Scilit]
  65. Jendrusch, M.A.; Yang, A.L.; Cacace, E.; Bobonis, J.; Voogdt, C.G.; Kaspar, S.; Schweimer, K.; Perez-Borrajero, C.; Lapouge, K.; Scheurich, J.; et al. AlphaDesign: A de novo protein design framework based on AlphaFold. Mol. Syst. Biol. 2025, 21, 1166. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  66. Zhang, O.; Zhang, X.; Lin, H.; Tan, C.; Wang, Q.; Mo, Y.; Feng, Q.; Du, G.; Yu, Y.; Jin, Z.; et al. ODesign: A World Model for Biomolecular Interaction Design. arXiv 2025, arXiv:2510.22304. [Google Scholar] [CrossRef] [Scilit]
  67. Pacesa, M.; Nickel, L.; Schellhaas, C.; Schmidt, J.; Pyatova, E.; Kissling, L.; Barendse, P.; Choudhury, J.; Kapoor, S.; Alcaraz-Serna, A.; et al. BindCraft: One-shot design of functional protein binders. bioRxiv 2024. [Google Scholar] [CrossRef] [Scilit]
  68. Stark, H.; Faltings, F.; Choi, M.; Xie, Y.; Hur, E.; O’Donnell, T.J.; Bushuiev, A.; Uçar, T.; Passaro, S.; Mao, W.; et al. BoltzGen: Toward Universal Binder Design. bioRxiv 2025. [Google Scholar] [CrossRef] [Scilit]
  69. Team, P.; Ren, M.; Sun, J.; Guan, J.; Liu, C.; Gong, C.; Wang, Y.; Wang, L.; Cai, Q.; Ma, W.; et al. PXDesign: Fast, modular, and accurate de novo design of protein binders. bioRxiv 2025. [Google Scholar] [CrossRef] [Scilit]
  70. Song, Z.; Li, T.; Li, L.; Min, M.R. PPDiff: Diffusing in Hybrid Sequence-Structure Space for Protein-Protein Complex Design. arXiv 2025, arXiv:2506.11420. [Google Scholar]
  71. Bennett, N.R.; Coventry, B.; Goreshnik, I.; Huang, B.; Allen, A.; Vafeados, D.; Peng, Y.P.; Dauparas, J.; Baek, M.; Stewart, L.; et al. Improving de novo protein binder design with deep learning. Nat. Commun. 2023, 14, 2625. [Google Scholar] [CrossRef] [Scilit]
  72. Jussupow, A.; Kaila, V.R. Effective molecular dynamics from neural network-based structure prediction models. J. Chem. Theory Comput. 2023, 19, 1965–1975. [Google Scholar] [CrossRef] [Scilit]
  73. Das, R.; Baker, D. Macromolecular modeling with rosetta. Annu. Rev. Biochem. 2008, 77, 363–382. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  74. Sgrignani, J.; Buscarini, S.; Locatelli, P.; Guerra, C.; Furlan, A.; Chen, Y.; Zoppi, G.; Cavalli, A. AI assisted design of ligands for Lipocalin-2. bioRxiv 2025. [Google Scholar] [CrossRef] [Scilit]
  75. Gómez Borrego, J.; Torrent Burgas, M. Evaluating ligand docking methods for drugging protein–protein interfaces: Insights from AlphaFold2 and molecular dynamics refinement. J. Cheminformatics 2025, 17, 144. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  76. Lazim, R.; Suh, D.; Choi, S. Advances in molecular dynamics simulations and enhanced sampling methods for the study of protein systems. Int. J. Mol. Sci. 2020, 21, 6339. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  77. Gokce-Alpkilic, G.; Huang, B.; Liu, A.; Kreuk, L.S.; Wang, Y.; Adebomi, V.; Bueso, Y.F.; Bera, A.K.; Kang, A.; Gerben, S.R.; et al. De Novo Design of High-Affinity Miniprotein Binders Targeting Francisella Tularensis Virulence Factor. Angew. Chem. Int. Ed. 2025, 137, e202516058. [Google Scholar] [CrossRef] [Scilit]
  78. Chazin-Gray, A.M.; Thompson, T.R.; Lopatto, E.D.; Magala, P.; Erickson, P.W.; Hunt, A.C.; Manchenko, A.; Aprikian, P.; Tchesnokova, V.; Basova, I.; et al. De Novo Design of Miniprotein Inhibitors of Bacterial Adhesins. bioRxiv 2025. [Google Scholar] [CrossRef] [Scilit]
  79. Niazi, S.K. Critical assessment of AI-based protein structure prediction: Fundamental challenges and future directions in drug discovery. Comput. Struct. Biotechnol. Rep. 2025, 15, 100064. [Google Scholar] [CrossRef] [Scilit]
  80. Heling, L.W.; van der Veen, J.; Rofe, A.; West, E.; Jiménez-Panizo, A.; Alegre-Martí, A.; Sheikhhassani, V.; Ng, J.; Schmidt, T.; Estébanez-Perpiñá, E.; et al. Deciphering the allosteric control of androgen receptor DNA binding by its disordered N-terminal domain. Mol. Cell. Endocrinol. 2025, 123, 112634. [Google Scholar] [CrossRef] [Scilit]
  81. Yang, D.; Zhou, Q.; Labroska, V.; Qin, S.; Darbalaei, S.; Wu, Y.; Yuliantie, E.; Xie, L.; Tao, H.; Cheng, J.; et al. G protein-coupled receptors: Structure-and function-based drug discovery. Signal Transduct. Target. Ther. 2021, 6, 7. [Google Scholar] [CrossRef] [Scilit]
  82. Moore, M.N.; Person, K.L.; Robleto, V.L.; Alwin, A.R.; Krusemark, C.L.; Foster, N.; Ray, C.; Inoue, A.; Jackson, M.R.; Sheedlo, M.J.; et al. Designing allosteric modulators to change GPCR G protein subtype selectivity. Nature 2025, 648, 229–238. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  83. Meijer, F.A.; Leijten-van de Gevel, I.A.; de Vries, R.M.; Brunsveld, L. Allosteric small molecule modulators of nuclear receptors. Mol. Cell. Endocrinol. 2019, 485, 20–34. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  84. Chang, Y.; Hawkins, B.A.; Du, J.J.; Groundwater, P.W.; Hibbs, D.E.; Lai, F. A guide to in silico drug design. Pharmaceutics 2023, 15, 49. [Google Scholar] [CrossRef] [Scilit] [PubMed]
Figure 1. Illustrations of allosteric modulation. (A) Schematics of negative (NAM), positive (PAM), and silent (SAM) allosteric modulation. The allosteric binder, target protein, and orthosteric binder are shown in salmon, teal, and purple, respectively. (B) Allosteric enzyme kinetics showing reaction velocity versus substrate concentration for the T and R states. The R (relaxed) state exhibits higher substrate affinity and a left-shifted hyperbolic curve, whereas the T (tense) state shows lower affinity and a right-shifted curve. (C) A representative protein exhibiting classical allosteric modulation, illustrated by a GPCR (PDB ID: 6WJC; structure determined by X-ray crystallography at 2.55 Å resolution). The miniprotein (MT7) is highlighted in salmon, the target protein (M1AChR) in teal, and the small molecule (atropine) in purple.
Figure 1. Illustrations of allosteric modulation. (A) Schematics of negative (NAM), positive (PAM), and silent (SAM) allosteric modulation. The allosteric binder, target protein, and orthosteric binder are shown in salmon, teal, and purple, respectively. (B) Allosteric enzyme kinetics showing reaction velocity versus substrate concentration for the T and R states. The R (relaxed) state exhibits higher substrate affinity and a left-shifted hyperbolic curve, whereas the T (tense) state shows lower affinity and a right-shifted curve. (C) A representative protein exhibiting classical allosteric modulation, illustrated by a GPCR (PDB ID: 6WJC; structure determined by X-ray crystallography at 2.55 Å resolution). The miniprotein (MT7) is highlighted in salmon, the target protein (M1AChR) in teal, and the small molecule (atropine) in purple.
Pharmaceuticals 19 00480 g001
Figure 2. The AI-driven computational pipeline for the design of allosteric miniprotein modulators, integrating structural analysis, generative binder design and selection and optimization.
Figure 2. The AI-driven computational pipeline for the design of allosteric miniprotein modulators, integrating structural analysis, generative binder design and selection and optimization.
Pharmaceuticals 19 00480 g002
Figure 3. Representative case studies in AI-driven design of miniprotein modulators, illustrating the targeting of non-orthosteric surfaces and the mechanisms underlying conformational modulation with two examples. (A) A miniprotein binder (ASD1) designed against the α site of the Flpp3 virulence factor, with the α site and β site shown in green and gray, and the designed miniprotein shown in purple, respectively. (B) A miniprotein inhibitor (F7) targeting the mannose-binding lectin domain of FimH; F7 can bind both the high-affinity (HAS) and low-affinity (LAS) conformations and, upon binding to the HAS, induces a conformational transition toward the LAS, with the lectin domain and pilin domain shown in green and gray, and the designed miniprotein shown in purple, respectively.
Figure 3. Representative case studies in AI-driven design of miniprotein modulators, illustrating the targeting of non-orthosteric surfaces and the mechanisms underlying conformational modulation with two examples. (A) A miniprotein binder (ASD1) designed against the α site of the Flpp3 virulence factor, with the α site and β site shown in green and gray, and the designed miniprotein shown in purple, respectively. (B) A miniprotein inhibitor (F7) targeting the mannose-binding lectin domain of FimH; F7 can bind both the high-affinity (HAS) and low-affinity (LAS) conformations and, upon binding to the HAS, induces a conformational transition toward the LAS, with the lectin domain and pilin domain shown in green and gray, and the designed miniprotein shown in purple, respectively.
Pharmaceuticals 19 00480 g003
Table 1. Machine learning-based tools for allosteric pocket identification. The tools were evaluated on different datasets as reported in their original studies.
Table 1. Machine learning-based tools for allosteric pocket identification. The tools were evaluated on different datasets as reported in their original studies.
Tool aMachine-Learning Strategy and Key FeaturesYearRef.
AllositeSupport vector machine classifier trained on static structural descriptors to discriminate allosteric from non-allosteric pockets2013[35]
AlloPredPerturbation-guided machine learning scoring of candidate pockets combined with normal mode analysis2016[36]
AllositeProStructure-based machine learning framework integrating multiple physicochemical and geometric features for improved robustness2017[37]
PASSerEnsemble machine learning approach trained on curated allosteric datasets for large-scale pocket identification2021[38]
PASSer 2.0AutoML-driven framework enabling automated feature selection, model optimization, and improved generalization2022[39]
PASSerRankLearning-to-rank strategy for prioritizing predicted allosteric pockets rather than binary classification2023[40]
MEF-AlloSiteMulti-model ensemble learning with optimized feature selection for accurate identification of allosteric sites and pockets2024[41]
a We include only machine learning-based methods explicitly developed for allosteric pocket identification or prioritization, excluding approaches aimed at inferring allosteric signaling pathways, identifying key allosteric residues, or relying on non-machine learning-based prediction strategies.
Table 4. Experimental results from AI-driven miniprotein design cases. This table summarizes key outcomes from two representative case studies, highlighting the functional and therapeutic impact of the designed miniproteins.
Table 4. Experimental results from AI-driven miniprotein design cases. This table summarizes key outcomes from two representative case studies, highlighting the functional and therapeutic impact of the designed miniproteins.
CategoryCase 1: Flpp3 BindersCase 2: FimH Inhibitor (F7)
Target proteinFlpp3 virulence factor from Francisella tularensisFimH adhesin from uropathogenic Escherichia coli
Binding site Site I ( α -helical face, membrane interaction) and Site II ( β -sheet face); allosteric-like disruption of protein–protein interactionsPocket adjacent to the orthosteric site; stabilizes the low-affinity state (LAS) and disfavors the high-affinity state (HAS), leading to allosteric inhibition
Binding affinity Site I: 24–110 nM; Site II: initially 81 nM, optimized to sub-nanomolar rangeNanomolar affinity confirmed by screening; selectively stabilizes LAS
Structural validation Circular dichroism confirms three-helix bundle topology; X-ray crystal structure RMSD of 0.9 Å relative to the design modelX-ray crystallography confirms the binding mode; NMR spectroscopy validates the induced conformational shift
Functional effects High stability; disrupts immune evasion, bacterial dissemination, and plasminogen binding; validated by yeast surface display enrichment and biolayer interferometryInhibits red blood cell aggregation, biofilm formation, and host receptor binding; blocks pathogenesis without direct orthosteric competition
Practical impact/in vivo results Provides research tools to probe Flpp3 function in tularemia and supports development of therapeutic candidates against antibiotic-resistant strainsEffective in treating and preventing uncomplicated and catheter-associated UTIs in mouse models, representing an antibiotic-sparing strategy for MDR infections
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content.

Share and Cite

MDPI and ACS Style

Liu, X.; Sun, Y.; Xia, Y.; Li, H.; Yan, Z. AI-Driven Design of Miniproteins as Potential Allosteric Modulators. Pharmaceuticals 2026, 19, 480. https://doi.org/10.3390/ph19030480

AMA Style

Liu X, Sun Y, Xia Y, Li H, Yan Z. AI-Driven Design of Miniproteins as Potential Allosteric Modulators. Pharmaceuticals. 2026; 19(3):480. https://doi.org/10.3390/ph19030480

Chicago/Turabian Style

Liu, Xin, Yunxiang Sun, Yulong Xia, Huaqiong Li, and Zhiqiang Yan. 2026. "AI-Driven Design of Miniproteins as Potential Allosteric Modulators" Pharmaceuticals 19, no. 3: 480. https://doi.org/10.3390/ph19030480

APA Style

Liu, X., Sun, Y., Xia, Y., Li, H., & Yan, Z. (2026). AI-Driven Design of Miniproteins as Potential Allosteric Modulators. Pharmaceuticals, 19(3), 480. https://doi.org/10.3390/ph19030480

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop