The increasing complexity and scale of technical document corpora present challenges for consistency verification, particularly in politically sensitive or high-stakes contexts. This paper proposes an iterative approach that integrates long-context large language models (LLMs), human expertise, and hybrid clustering mechanisms to address these challenges. The approach focuses on two types of inconsistencies: real inconsistencies, such as contradictory statements or omissions, and fabricated inconsistencies, which are plausible yet artificially introduced. This paper uses the Swiss National Cooperative for the Disposal of Radioactive Waste (Nagra) and its corpus of up to 300 technical documents as a case study. Experimental results suggest that targeted structuring of document contexts improves recall in inconsistency detection. The findings highlight the potential of combining structured human input with LLM-based reasoning for improving document integrity and trustworthiness. Future work will focus on refining the approach, including automated clustering strategies and optimization of prompt engineering.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Discovering Inconsistencies in Documents with Long-Context LLMs

  • Andreas Martin,
  • Hans Friedrich Witschel,
  • Mona Stockhecke,
  • Patrick Nydegger,
  • Kyrylo Buga

摘要

The increasing complexity and scale of technical document corpora present challenges for consistency verification, particularly in politically sensitive or high-stakes contexts. This paper proposes an iterative approach that integrates long-context large language models (LLMs), human expertise, and hybrid clustering mechanisms to address these challenges. The approach focuses on two types of inconsistencies: real inconsistencies, such as contradictory statements or omissions, and fabricated inconsistencies, which are plausible yet artificially introduced. This paper uses the Swiss National Cooperative for the Disposal of Radioactive Waste (Nagra) and its corpus of up to 300 technical documents as a case study. Experimental results suggest that targeted structuring of document contexts improves recall in inconsistency detection. The findings highlight the potential of combining structured human input with LLM-based reasoning for improving document integrity and trustworthiness. Future work will focus on refining the approach, including automated clustering strategies and optimization of prompt engineering.