Oct 2026· International Journal of Data Science and Analysis· Vol 22· 0 citations· 7 references
Topic Modeling
Abstract
Conversational language model systems persist user constraints, preferences, identity, commitments, and assigned roles across turns. The dominant memory architecture stores this alongside ordinary episodic content in a vector index and retrieves the top-k items by similarity to the active query. We argue that this design is structurally inappropriate for the subset of items we call governing facts: facts whose binding force does not depend on their similarity to the current query. We define governance fidelity, the fraction of governing facts that reach the assembled context, and prove a small impossibility result: any retriever that ranks by a type-blind similarity score and returns a bounded top-k can be defeated by an adversarial topical-noise sequence. The remedy is type-conditional retrieval. We instantiate it as Stratified Context Reconstruction (SCR) in Satchel, a typed graph memory. Across 2880 sparse-and-hybrid and 2160 dense retrieval measurements, governance fidelity collapses from 1.0 to below 0.10 by N = 100 adversarial noise pairs for every similarity retriever tested; dense bi-encoders collapse as well, and a follow-up sweep over larger encoders (bge-large, e5-large) and a cross-encoder reranker confirm that the collapse is intrinsic to similarity ranking and is not avoided by stronger or larger encoders, which postpone but do not prevent it. SCR holds fidelity at 1.0 by construction (Cohen’s d = 1.06, p ≈ 2.6 + 10⁻25). We further show that type-conditional retrieval alone is insufficient once governing facts are revised over time: we prove a second impossibility—governance consistency cannot be guaranteed by any conflict-blind retriever, and the error grows linearly in revision depth—and give SCR-T, a temporally stratified, supersession-resolved retriever that attains both completeness and consistency by construction. A revision-depth benchmark confirms the prediction: naive SCR returns up to fifteen stale superseded facts per query and its governance minimality—the precision of the governing block—decays to 0.25, while SCR-T returns none and holds fidelity, consistency, and minimality jointly at 1.0. We frame these three as governance invariants against which conversational memory should be evaluated, alongside traceability as a structural property of typed memory. On four small open-weight LLMs, SCR improves constraint adherence over a TF-IDF RAG baseline by 17.59 percentage points. A further live-LLM experiment under the paper’s own adversarial noise, with retrieval and prompt formatting separated in a controlled four-arm design, shows that SCR’s downstream advantage is a high-noise phenomenon that grows with noise volume and is not an artefact of formatting.
The results are packaged in the Greenfield Startup Model (GSM), which explains the priority of startups to release the product as quickly as possible, and the need to shorten time-to-market, by speeding up the development through low-precision engineering activities.
Carmine Giardino, Nicolò Paternoster, M. Unterkalmsteiner et al.· IEEE Transactions on Softwar...· 178 citations· ⚡14
Software startup companies develop innovative, software-intensive products within limited timeframes and with few resources, searching for sustainable and scalable business models.
M. Unterkalmsteiner, P. Abrahamsson, Xiaofeng Wang et al.· e-Informatica Software Engin...· 157 citations· ⚡17
This study conducts a case survey study based on the secondary data of the major pivots happened in 49 software startups, and demonstrates that customer need pivot is the most common among all pivot types.
Sohaib Shahid Bajwa, Xiaofeng Wang, Anh Nguyen-Duc et al.· Empirical Software Engineeri...· 127 citations· ⚡15
The comparison of adopter and non-adopter sample reveals three potential adoption inhibitor, security, data privacy, and portability, which underlines the importance of the technical and security perspectives for research investigating the adoption of technology.
Nattakarn Phaphoom, Xiaofeng Wang, S. Samuel et al.· Journal of Systems and Softw...· 111 citations· ⚡8
The ongoing work building a Raspberry Pi cluster consisting of 300 nodes is presented, with potential use cases being an inexpensive and green test bed for cloud computing research and a robust and mobile data center for operating in adverse environments.
P. Abrahamsson, S. Helmer, Nattakarn Phaphoom et al.· IEEE International Conferenc...· 110 citations· ⚡7
The results indicate that software developers are a slightly happy population, but the need for limiting the unhappiness of developers remains, and 219 factors representing causes of unhappiness while developing software are identified.
D. Graziotin, Fabian Fagerholm, Xiaofeng Wang et al.· International Conference on...· 84 citations· ⚡6