| International Journal of Computer Applications |
| Foundation of Computer Science (FCS), NY, USA |
| Volume 187 - Number 131 |
| Year of Publication: 2026 |
| Authors: Rohan Kummaraguntla, Andrew J. Ouderkirk |
10.5120/ijca1159a412d939
|
Rohan Kummaraguntla, Andrew J. Ouderkirk . Enhancing Patent Readability: Leveraging Large Language Model-Generated Taxonomies for Prior Art Analysis. International Journal of Computer Applications. 187, 131 ( Aug 2026), 1-9. DOI=10.5120/ijca1159a412d939
Patent documents are notoriously difficult to read because of their technical jargon, strict formatting, and lack of semantic structure. This paper studies the application of large language models (LLMs) to produce multi-level hierarchical taxonomies as a strategy to make patents more readable and applicable. By converting unstructured language into structured hierarchies, automated taxonomies offer an intuitive and scalable solution to navigating dense legal text for inventors, researchers, and intellectual property professionals. Patents from various fields including software, medical devices, and materials science were analyzed to evaluate the consistency, depth, and readability of the generated taxonomies. The results indicate that LLMs can effectively restructure complex legal documents into understandable, layered forms— providing an accurate and consistent tool for enhancing information retrieval, prior-art analysis, and a step towards human-artificial intelligence collaboration in intellectual property applications.