Many biological processes rely on mechanical forces, with protein molecules acting as key mediators. Understanding how proteins respond to mechanical stress is essential for conditions including cardiomyopathy and muscular dystrophy. Natural proteins such as dystrophin and utrophin are composed of heterogeneous folding domains with distinct mechanical properties; deciphering domain-level behavior provides insights into disease mechanisms and informs therapeutic strategies. Single-molecule force spectroscopy (SMFS) enables probing the mechanical properties of entire proteins, yet current approaches struggle to identify heterogeneous folding domains, particularly without prior knowledge. Here, we present the first automated framework to identify heterogeneous folding domains in SMFS data, applying both existing clustering methods and a novel physics-aware deep clustering architecture, LatentUnfold. LatentUnfold learns complementary latent representations from force magnitude and the force-extension physical relationship through dual autoencoders, jointly optimized for clustering assignments. We apply our framework to experimental SMFS data collected from a synthetic two-domain protein (ddFLN4-Titin I27) as well as natural protein constructs of dystrophin and utrophin, with Monte Carlo simulated datasets serving as controlled validation. For the synthetic protein, we recover mechanical properties consistent with previously reported values for each domain. For the natural proteins, we uncover two mechanically distinct domain populations - corresponding to the N-terminal domain and spectrin-like repeats - with differences in both unfolding force and contour length increase, and reveal different unfolding order between them for the first time. This work enables domain-level biological inference, overcoming prior limitations that relied on averaging and overlooked heterogeneity, thus advancing the understanding of mechanical behavior in protein unfolding.
The comparison of adopter and non-adopter sample reveals three potential adoption inhibitor, security, data privacy, and portability, which underlines the importance of the technical and security perspectives for research investigating the adoption of technology.
Nattakarn Phaphoom, Xiaofeng Wang, S. Samuel et al.· Journal of Systems and Softw...· 111 citations· ⚡8
This study investigates how Lean internal startup facilitates software product innovation in large companies and identifies its enablers and inhibitors, and shows the potential of the method-in-action framework to investigate the Lean startup approach in non-startup context.
Henry Edison, Nina M. Smørsgård, Xiaofeng Wang et al.· Journal of Systems and Softw...· 78 citations· ⚡6
This paper highlights the challenges to conduct proper affect-related studies with psychology, provides a comprehensive literature review in affect theory, and proposes guidelines for conducting psychoempirical software engineering.
D. Graziotin, Xiaofeng Wang, P. Abrahamsson· SSE@SIGSOFT FSE· 56 citations· ⚡4
This study conducts a multiple case study on twenty European software startups and proposes a prototype-centric learning model in early stage software startups, and identifies factors that occur as barriers but also facilitators for prototyping in earlystage software startups.
Anh Nguyen-Duc, Xiaofeng Wang, P. Abrahamsson· International Conference on...· 44 citations· ⚡5
It is demonstrated that linker-free PROTACs can outperform traditional designs, marking a paradigm shift in PROTAC development for targeted protein degradation.
Pinal, a 16-billion-parameter foundation model that produces protein candidates from natural-language functional descriptions, supports natural language as a high-level interface for candidate generation in protein design, enabling programmable exploration with reduced reliance on manually specified structural or sequence constraints.
A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.