Efficient Frequent Subtree Mining Beyond Forests
Saved in:
| Title: | Efficient Frequent Subtree Mining Beyond Forests |
|---|---|
| Description: | A common paradigm in distance-based learning is to embed the instance space into a feature space equipped with a metric and define the dissimilarity between instances by the distance of their images in the feature space. Frequent connected subgraphs are sometimes used to define such feature spaces if the instances are graphs, but identifying the set of frequent connected subgraphs and subsequently computing embeddings for graph instances is computationally intractable. As a result, existing frequent subgraph mining algorithms either restrict the structural complexity of the instance graphs or require exponential delay between the output of subsequent patterns, meaning that distance-based learners lack an efficient way to operate on arbitrary graph data. This book presents a mining system that gives up the demand on the completeness of the pattern set, and instead guarantees a polynomial delay between subsequent patterns. To complement this, efficient methods devised to compute the embedding of arbitrary graphs into the Hamming space spanned by the pattern set are described. As a result, a system is proposed that allows the efficient application of distance-based learning methods to arbitrary graph databases. In addition to an introduction and conclusion, the book is divided into chapters covering: preliminaries; related work; probabilistic frequent subtrees; boosted probabilistic frequent subtrees; and fast computation, with a further two chapters on Hamiltonian path for cactus graphs and Poisson binomial distribution. |
| Authors: | Pascal Welke |
| Resource Type: | eBook. |
| Subjects: | Data mining |
| Categories: | COMPUTERS / Artificial Intelligence / General |
| Database: | eBook Collection (EBSCOhost) |
| FullText | Links: – Type: ebook-pdf Text: Availability: 0 |
|---|---|
| Header | DbId: nlebk DbLabel: eBook Collection (EBSCOhost) An: 2512584 RelevancyScore: 1097 AccessLevel: 6 PubType: eBook PubTypeId: ebook PreciseRelevancyScore: 1096.64697265625 |
| IllustrationInfo | |
| ImageInfo | – Size: thumb Target: https://rps2images.ebscohost.com/rpsweb/othumb?id=NL$2512584$PDF&s=r – Size: medium Target: https://rps2images.ebscohost.com/rpsweb/othumb?id=NL$2512584$PDF&s=d |
| Items | – Name: Title Label: Title Group: Ti Data: Efficient Frequent Subtree Mining Beyond Forests – Name: Abstract Label: Description Group: Ab Data: A common paradigm in distance-based learning is to embed the instance space into a feature space equipped with a metric and define the dissimilarity between instances by the distance of their images in the feature space. Frequent connected subgraphs are sometimes used to define such feature spaces if the instances are graphs, but identifying the set of frequent connected subgraphs and subsequently computing embeddings for graph instances is computationally intractable. As a result, existing frequent subgraph mining algorithms either restrict the structural complexity of the instance graphs or require exponential delay between the output of subsequent patterns, meaning that distance-based learners lack an efficient way to operate on arbitrary graph data. This book presents a mining system that gives up the demand on the completeness of the pattern set, and instead guarantees a polynomial delay between subsequent patterns. To complement this, efficient methods devised to compute the embedding of arbitrary graphs into the Hamming space spanned by the pattern set are described. As a result, a system is proposed that allows the efficient application of distance-based learning methods to arbitrary graph databases. In addition to an introduction and conclusion, the book is divided into chapters covering: preliminaries; related work; probabilistic frequent subtrees; boosted probabilistic frequent subtrees; and fast computation, with a further two chapters on Hamiltonian path for cactus graphs and Poisson binomial distribution. – Name: Author Label: Authors Group: Au Data: <searchLink fieldCode="AR" term="%22Pascal+Welke%22">Pascal Welke</searchLink> – Name: TypePub Label: Resource Type Group: TypPub Data: eBook. – Name: Subject Label: Subjects Group: Su Data: <searchLink fieldCode="DE" term="%22Data+mining%22">Data mining</searchLink> – Name: SubjectBISAC Label: Categories Group: Su Data: <searchLink fieldCode="ZK" term="%22COMPUTERS+%2F+Artificial+Intelligence+%2F+General%22">COMPUTERS / Artificial Intelligence / General</searchLink> |
| PLink | https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=nlebk&AN=2512584 |
| RecordInfo | BibRecord: BibEntity: Classifications: – Code: 006.312 Scheme: ddc Type: prePub Languages: – Code: eng Text: English Subjects: – SubjectFull: Data mining Type: general Titles: – TitleFull: Efficient Frequent Subtree Mining Beyond Forests Type: main BibRelationships: HasContributorRelationships: – PersonEntity: Name: NameFull: Pascal Welke – PersonEntity: Name: NameFull: Pascal Welke IsPartOfRelationships: – BibEntity: Dates: – D: 01 M: 01 Type: published Y: 2020 – D: 08 M: 07 Type: profile Y: 2020 Identifiers: – Type: isbn-print Value: 9781643680781 – Type: isbn-electronic Value: 9781643680798 Titles: – TitleFull: Efficient Frequent Subtree Mining Beyond Forests Type: main |
| ResultId | 1 |