Volume 11 • Issue 2 • PP: 87–93 • 2026
MultiUserWiki: Managing Editorial Divergence in Collaborative LLM-Maintained Knowledge Bases
Open Access & Copyright
© 2026 The Author(s). Published by ASPG. This article is licensed under the Creative Commons Attribution 4.0 International License (CC BY 4.0).
Abstract
When multiple LLM-assisted editors maintain a shared knowledge base, the conflicts that arise are seldom outright factual contradictions. More often, two users reading the same source extract compatible assertions that differ in specificity, coverage, or editorial emphasis. This paper asks a prior question: when should a conflict be resolved by selecting a winner, and when should both perspectives be preserved? We study this question in a controlled single-source collaborative setting. Five role-conditioned instances of GLM-4-Flash—simulating a senior researcher, a teacher, a student, a clinician, and a conservative editor—independently extract assertions from AI-Wiki-21, a 21-page AI knowledge base. Their outputs yield 4,556 candidate conflict pairs; annotation of 138 selected pairs produces a 100-conflict gold standard. Three findings emerge, each with scope limited to the single-source, single- LLM-family, simulated-role setting studied. First, confirmed conflicts are exclusively editorial—54% Granularity and 46% Coverage—with zero Factual conflicts, suggesting that in this controlled setting the dominant challenge is editorial coordination, not truth arbitration. Second, direct LLM judgment achieves 97.5% agreement with human annotators on this editorial-conflict gold standard, while a DPO-inspired quality heuristic (DIQS) reaches 81.0% without API overhead; collective voting plateaus at 41.8%, echoing social-choice instability. Third, sensitivity analysis of the fiber split threshold across τ ∈ [0.1,0.9] identifies τ∗ = 0.5 as the stable optimum, partitioning 3.1% of conflicts while preserving diverse user perspectives. The results also raise ethical questions about automated knowledge governance.
Keywords
References
[1] S. Pan, L. Luo, Y. Wang, C. Chen, J. Wang, and X. Wu, “Unifying large language models and knowledge graphs: A roadmap,” IEEE Transactions on Knowledge and Data Engineering, vol. 36, no. 7, pp. 3580–3599, 2024.
[2] B. J. Gutiérrez, Y. Shu, W. Qi, S. Zhou, and Y. Su, “From RAG to memory: Non-parametric continual learning for large language models,” in Proceedings of the 42nd International Conference on Machine Learning, vol. 267 of Proceedings of Machine Learning Research, pp. 21497–21515, 2025.
[3] K. J. Arrow, Social Choice and Individual Values. Yale University Press, 2 ed., 1963.
[4] A. Halfaker, R. S. Geiger, J. T. Morgan, and J. Riedl, “The rise and decline of an open collaboration system,”American Behavioral Scientist, vol. 57, no. 5, pp. 664– 688, 2013.
[5] L. de Alfaro, B. T. Adler, A. Kulshreshtha, and I. Pye, “Reputation systems for open collaboration,” Communications of the ACM, vol. 54, no. 8, pp. 81–87, 2011.
[6] D. Vrandeˇci´c and M. Krötzsch, “Wikidata: A free collaborative knowledgebase,” Communications of the ACM, vol. 57, no. 10, pp. 78–85, 2014.
[7] F. Petroni, T. Rocktäschel, P. Lewis, A. Bakhtin, Y. Wu, A. H. Miller, and S. Riedel, “Language models as knowledge bases?,” in Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing, pp. 2463–2473, 2019.
[8] L. Zheng, W.-L. Chiang, Y. Sheng, S. Zhuang, Z. Wu, Y. Zhuang, Z. Lin, Z. Li, D. Li, E. Xing, H. Zhang, J. E. Gonzalez, and I. Stoica, “Judging LLM-as-a-judge with MT-bench and chatbot arena,” in Advances in Neural Information Processing Systems, vol. 36, 2023.
[9] Y. Yao, P. Wang, B. Tian, S. Cheng, Z. Li, S. Deng, H. Chen, and N. Zhang, “Editing large language models: Problems, methods, and opportunities,” in Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, pp. 10222–10240, 2023.
[10] Y. Hu, T.-P. Nguyen, S. Ghosh, and S. Razniewski, “Enabling LLM knowledge analysis via extensive materialization,” in Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp. 16189–16202, Association for Computational Linguistics, 2025.
[11] R. Rafailov, A. Sharma, E. Mitchell, S. Ermon, C. D. Manning, and C. Finn, “Direct preference optimization: Your language model is secretly a reward model,” in Advances in Neural Information Processing Systems, vol. 36, 2023.
Cite This Article
Choose your preferred format
Publisher's Note
The statements, opinions, and data presented in this article are solely those of the author(s) and do not necessarily represent those of ASPG, the journal, or its editors. ASPG and the editors disclaim responsibility for any harm arising from the use of any ideas, methods, instructions, or products described in this article, to the fullest extent permitted by applicable law.