A method for extracting data from semis-tructured documents
Linguistic method to solve the problem of data extraction from weakly structured documents is developed, approved, and described in detail in the paper. Sample data were taken from thesis catalogue of Vernadsky National Library of Ukraine. The sequence of all stages is described: document collection...
Gespeichert in:
Datum: | 2020 |
---|---|
Hauptverfasser: | Kudim, K.A., Proskudina, G.Yu. |
Format: | Artikel |
Sprache: | rus |
Veröffentlicht: |
Інститут програмних систем НАН України
2020
|
Schlagworte: | |
Online Zugang: | https://pp.isofts.kiev.ua/index.php/ojs1/article/view/388 |
Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Назва журналу: | Problems in programming |
Institution
Problems in programmingÄhnliche Einträge
-
Methods and tools for extracting personal data from theses abstracts
von: Kudim, K.A., et al.
Veröffentlicht: (2019) -
Extracting structure from text documents based on machine learning
von: Kudim, K.A., et al.
Veröffentlicht: (2023) -
About technologies of use of external data on creating and editing of encyclopedic texts
von: Proskudina, G.Yu., et al.
Veröffentlicht: (2018) -
Mixed topic-entity ontology for enhanced topic vector-spaced model
von: Shabinskiy, A.S.
Veröffentlicht: (2025) -
Overview of global open access resource aggregation services and their requirements for data providers
von: Proskudina, G.Yu., et al.
Veröffentlicht: (2025)