Back to top

Structured Extraction of Terms and Conditions from German and English Online Shops

Last modified May 19, 2022
   No tags assigned

The automated analysis of Terms and Conditions has gained attention in recent years, mainly due to its relevance to consumer protection. Well-structured data sets are the base for every analysis. While content extraction, in general, is a well-researched field and many open source libraries are available, our evaluation shows, that existing solutions cannot extract Terms and Conditions in sufficient quality, mainly because of their special structure. In this paper, we present an approach to extract the content and hierarchy of Terms and Conditions from German and English online shops. Our evaluation shows, that the approach outperforms the current state of the art. A python implementation of the approach is made available under an open license.

Files and Subpages

Name Type Size Last Modification Last Editor
2022.ecnlp-1.21.pdf 182 KB 19.05.2022