Combining Programming-by-Example with Transformation Discovery from large Databases
dc.contributor.author | özmen, Aslihan | |
dc.contributor.author | Esmailoghli, Mahdi | |
dc.contributor.author | Abedjan, Ziawasch | |
dc.contributor.editor | Kai-Uwe Sattler | |
dc.contributor.editor | Melanie Herschel | |
dc.contributor.editor | Wolfgang Lehner | |
dc.date.accessioned | 2021-03-16T07:57:10Z | |
dc.date.available | 2021-03-16T07:57:10Z | |
dc.date.issued | 2021 | |
dc.description.abstract | Data transformation discovery is one of the most tedious tasks in data preparation. In particular, the generation of transformation programs for semantic transformations is tricky because additional sources for look-up operations are necessary. Current systems for semantic transformation discovery face two major problems: either they follow a program synthesis approach that only scales to a small set of input tables, or they rely on extraction of transformation functions from large corpora, which requires the identification of exact transformations in those resources and is prone to noisy data. In this paper, we try to combine approaches to benefit from large corpora and the sophistication of program synthesis. To do so, we devise a retrieval and pruning strategy ensemble that extracts the most relevant tables for a given transformation task. The extracted resources can then be processed by a program synthesis engine to generate more accurate transformation results than state-of-the-art. | en |
dc.identifier.doi | 10.18420/btw2021-16 | |
dc.identifier.isbn | 978-3-88579-705-0 | |
dc.identifier.pissn | 1617-5468 | |
dc.identifier.uri | https://dl.gi.de/handle/20.500.12116/35799 | |
dc.language.iso | en | |
dc.publisher | Gesellschaft für Informatik, Bonn | |
dc.relation.ispartof | BTW 2021 | |
dc.relation.ispartofseries | Lecture Notes in Informatics (LNI) - Proceedings, Volume P-311 | |
dc.title | Combining Programming-by-Example with Transformation Discovery from large Databases | en |
gi.citation.endPage | 324 | |
gi.citation.startPage | 313 | |
gi.conference.date | 13.-17. September 2021 | |
gi.conference.location | Dresden | |
gi.conference.sessiontitle | Data Integration, Semantic Data Management, Streaming |
Dateien
Originalbündel
1 - 1 von 1