On the Consistency of k-means++ algorithm.
Saved in:
| Title: | On the Consistency of k-means++ algorithm. |
|---|---|
| Authors: | Kłopotek, Mieczysław A.1 (AUTHOR) klopotek@ipipan.waw.pl |
| Source: | Fundamenta Informaticae. 2020, Vol. 172 Issue 4, p361-377. 17p. |
| Subjects: | Big data, Expected returns, Databases, Algorithms |
| Abstract: | We prove in this paper that the expected value of the objective function of the k-means++ algorithm for samples converges to population expected value. As k-means++, for samples, provides with constant factor approximation for k-means objectives, such an approximation can be achieved for the population with increase of the sample size. This result is of potential practical relevance when one is considering using subsampling when clustering large data sets (large data bases). [ABSTRACT FROM AUTHOR] |
| Copyright of Fundamenta Informaticae is the property of Polskie Towarzystwo Matematyczne and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.) | |
| Database: | Engineering Source |
| FullText | Links: – Type: pdflink Text: Availability: 0 |
|---|---|
| Header | DbId: egs DbLabel: Engineering Source An: 141686041 AccessLevel: 6 PubType: Academic Journal PubTypeId: academicJournal PreciseRelevancyScore: 0 |
| IllustrationInfo | |
| Items | – Name: Title Label: Title Group: Ti Data: On the Consistency of k-means++ algorithm. – Name: Author Label: Authors Group: Au Data: <searchLink fieldCode="AR" term="%22Kłopotek%2C+Mieczysław+A%2E%22">Kłopotek, Mieczysław A.</searchLink><relatesTo>1</relatesTo> (AUTHOR)<i> klopotek@ipipan.waw.pl</i> – Name: TitleSource Label: Source Group: Src Data: <searchLink fieldCode="JN" term="%22Fundamenta+Informaticae%22">Fundamenta Informaticae</searchLink>. 2020, Vol. 172 Issue 4, p361-377. 17p. – Name: Subject Label: Subjects Group: Su Data: <searchLink fieldCode="DE" term="%22Big+data%22">Big data</searchLink><br /><searchLink fieldCode="DE" term="%22Expected+returns%22">Expected returns</searchLink><br /><searchLink fieldCode="DE" term="%22Databases%22">Databases</searchLink><br /><searchLink fieldCode="DE" term="%22Algorithms%22">Algorithms</searchLink> – Name: Abstract Label: Abstract Group: Ab Data: We prove in this paper that the expected value of the objective function of the k-means++ algorithm for samples converges to population expected value. As k-means++, for samples, provides with constant factor approximation for k-means objectives, such an approximation can be achieved for the population with increase of the sample size. This result is of potential practical relevance when one is considering using subsampling when clustering large data sets (large data bases). [ABSTRACT FROM AUTHOR] – Name: AbstractSuppliedCopyright Label: Group: Ab Data: <i>Copyright of Fundamenta Informaticae is the property of Polskie Towarzystwo Matematyczne and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.) |
| PLink | https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=egs&AN=141686041 |
| RecordInfo | BibRecord: BibEntity: Identifiers: – Type: doi Value: 10.3233/FI-2020-1909 Languages: – Code: eng Text: English PhysicalDescription: Pagination: PageCount: 17 StartPage: 361 Subjects: – SubjectFull: Big data Type: general – SubjectFull: Expected returns Type: general – SubjectFull: Databases Type: general – SubjectFull: Algorithms Type: general Titles: – TitleFull: On the Consistency of k-means++ algorithm. Type: main BibRelationships: HasContributorRelationships: – PersonEntity: Name: NameFull: Kłopotek, Mieczysław A. IsPartOfRelationships: – BibEntity: Dates: – D: 15 M: 03 Text: 2020 Type: published Y: 2020 Identifiers: – Type: issn-print Value: 01692968 Numbering: – Type: volume Value: 172 – Type: issue Value: 4 Titles: – TitleFull: Fundamenta Informaticae Type: main |
| ResultId | 1 |