A Monte Carlo Study of Parallel Analysis, Minimum Average Partial, Indicator Function, and Modified Average Roots for Determining the Number of Dimensions with Binary Variables in Test Data: Impact of Sample Size and Factor Structure

Saved in:
Bibliographic Details
Title: A Monte Carlo Study of Parallel Analysis, Minimum Average Partial, Indicator Function, and Modified Average Roots for Determining the Number of Dimensions with Binary Variables in Test Data: Impact of Sample Size and Factor Structure
Language: English
Authors: Pornchanok Ruengvirayudh
Source: ProQuest LLC. 2018Ph.D. Dissertation, Ohio University.
Availability: ProQuest LLC. 789 East Eisenhower Parkway, P.O. Box 1346, Ann Arbor, MI 48106. Tel: 800-521-0600; Web site: http://www.proquest.com/en-US/products/dissertations/individuals.shtml
Peer Reviewed: N
Page Count: 368
Publication Date: 2018
Document Type: Dissertations/Theses - Doctoral Dissertations
Descriptors: Monte Carlo Methods, Tests, Data, Sample Size, Factor Analysis, Construct Validity, Shift Studies, Criteria, Research Design
ISBN: 979-88-02-75503-7
Abstract: Determining the number of dimensions underlying many variables in the data or many items in the test is a crucial process prior to performing exploratory factor analysis. Failure to do so leads to serious consequences concerning construct validity. Parallel analysis (PA) has been found to be useful to determine the number of dimensions (i.e., components or factors) in many conditions. As computational power of computers is much advanced, novel procedures have been developed to improve the accuracy of PA. Authors of a number of previous studies have investigated the use of parallel analysis with scale data (e.g., questionnaires). However, little research has been conducted on the performance of PA when applied to existing real test data. This present study, therefore, compared the consistency of PA and other criteria (i.e., minimum average partial, broken stick, average root and modified average roots, imbedded error, and indicator function) in extracting the number of dimensions from large existing real test data at the population level (approximately 400,000 cases) based on these studied variables: sample size, factor structure, number of randomly generated data sets, threshold, and type of input correlation matrices. R scripts in the R program were written to repeatedly sample from a population's data in a Monte Carlo simulation procedure and to run the analyses. Consistent methods yielding precise results under most studied conditions were: MAP, IND, PA[subscript COR95] (i.e., PA using the original correlation matrices with 1s on the diagonal) with the 95th percentile as a threshold and 100 randomly generated data sets, and MAR[subscript 1.4] (i.e., 1.4*Average Root), respectively. When practitioners have small sample sizes of at least 100, MAP is recommended for use. PA performed consistently with sample sizes of at least 200. However, MAP and PA are not incorporated in commercial statistical software (e.g., SPSS, SAS). Therefore, alternative methods are recommended for use in place of or in conjunction with recommended methods to compare the results. IND and MAR[subscript 1.4] are recommended for use with sample sizes of at least 200 and 300, respectively. However, BS and IE are not recommended due to large errors and unvaried results of other than one dimension when n [greater than or equal to] 100, respectively. In general, a sample size of 100 is not recommended for use because it is not sufficient to yield precise results. Sample sizes of 200, 300, and 400 are recommended to be minimum, acceptable, and desirable sample sizes to yield consistent results. In this study, an unbalanced factor structure (i.e., unequal numbers of items in each dimension) showed a negative impact on the precision of the factor extraction results. Tutorials on how to perform PA and the other five criteria with examples were presented. [The dissertation citations contained here are published with the permission of ProQuest LLC. Further reproduction is prohibited without permission. Copies of dissertations may be obtained by Telephone (800) 1-800-521-0600. Web page: http://www.proquest.com/en-US/products/dissertations/individuals.shtml.]
Abstractor: As Provided
Entry Date: 2024
Access URL: https://gateway.proquest.com/openurl?url_ver=Z39.88-2004&rft_val_fmt=info:ofi/fmt:kev:mtx:dissertation&res_dat=xri:pqm&rft_dat=xri:pqdiss:29282482
Accession Number: ED646416
Database: ERIC
FullText Text:
  Availability: 0
Header DbId: eric
DbLabel: ERIC
An: ED646416
AccessLevel: 3
PubType: Dissertation/ Thesis
PubTypeId: dissertation
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: A Monte Carlo Study of Parallel Analysis, Minimum Average Partial, Indicator Function, and Modified Average Roots for Determining the Number of Dimensions with Binary Variables in Test Data: Impact of Sample Size and Factor Structure
– Name: Language
  Label: Language
  Group: Lang
  Data: English
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Pornchanok+Ruengvirayudh%22">Pornchanok Ruengvirayudh</searchLink>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="SO" term="%22ProQuest+LLC%22"><i>ProQuest LLC</i></searchLink>. 2018Ph.D. Dissertation, Ohio University.
– Name: Avail
  Label: Availability
  Group: Avail
  Data: ProQuest LLC. 789 East Eisenhower Parkway, P.O. Box 1346, Ann Arbor, MI 48106. Tel: 800-521-0600; Web site: http://www.proquest.com/en-US/products/dissertations/individuals.shtml
– Name: PeerReviewed
  Label: Peer Reviewed
  Group: SrcInfo
  Data: N
– Name: Pages
  Label: Page Count
  Group: Src
  Data: 368
– Name: DatePubCY
  Label: Publication Date
  Group: Date
  Data: 2018
– Name: TypeDocument
  Label: Document Type
  Group: TypDoc
  Data: Dissertations/Theses - Doctoral Dissertations
– Name: Subject
  Label: Descriptors
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Monte+Carlo+Methods%22">Monte Carlo Methods</searchLink><br /><searchLink fieldCode="DE" term="%22Tests%22">Tests</searchLink><br /><searchLink fieldCode="DE" term="%22Data%22">Data</searchLink><br /><searchLink fieldCode="DE" term="%22Sample+Size%22">Sample Size</searchLink><br /><searchLink fieldCode="DE" term="%22Factor+Analysis%22">Factor Analysis</searchLink><br /><searchLink fieldCode="DE" term="%22Construct+Validity%22">Construct Validity</searchLink><br /><searchLink fieldCode="DE" term="%22Shift+Studies%22">Shift Studies</searchLink><br /><searchLink fieldCode="DE" term="%22Criteria%22">Criteria</searchLink><br /><searchLink fieldCode="DE" term="%22Research+Design%22">Research Design</searchLink>
– Name: ISBN
  Label: ISBN
  Group: ISBN
  Data: 979-88-02-75503-7
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: Determining the number of dimensions underlying many variables in the data or many items in the test is a crucial process prior to performing exploratory factor analysis. Failure to do so leads to serious consequences concerning construct validity. Parallel analysis (PA) has been found to be useful to determine the number of dimensions (i.e., components or factors) in many conditions. As computational power of computers is much advanced, novel procedures have been developed to improve the accuracy of PA. Authors of a number of previous studies have investigated the use of parallel analysis with scale data (e.g., questionnaires). However, little research has been conducted on the performance of PA when applied to existing real test data. This present study, therefore, compared the consistency of PA and other criteria (i.e., minimum average partial, broken stick, average root and modified average roots, imbedded error, and indicator function) in extracting the number of dimensions from large existing real test data at the population level (approximately 400,000 cases) based on these studied variables: sample size, factor structure, number of randomly generated data sets, threshold, and type of input correlation matrices. R scripts in the R program were written to repeatedly sample from a population's data in a Monte Carlo simulation procedure and to run the analyses. Consistent methods yielding precise results under most studied conditions were: MAP, IND, PA[subscript COR95] (i.e., PA using the original correlation matrices with 1s on the diagonal) with the 95th percentile as a threshold and 100 randomly generated data sets, and MAR[subscript 1.4] (i.e., 1.4*Average Root), respectively. When practitioners have small sample sizes of at least 100, MAP is recommended for use. PA performed consistently with sample sizes of at least 200. However, MAP and PA are not incorporated in commercial statistical software (e.g., SPSS, SAS). Therefore, alternative methods are recommended for use in place of or in conjunction with recommended methods to compare the results. IND and MAR[subscript 1.4] are recommended for use with sample sizes of at least 200 and 300, respectively. However, BS and IE are not recommended due to large errors and unvaried results of other than one dimension when n [greater than or equal to] 100, respectively. In general, a sample size of 100 is not recommended for use because it is not sufficient to yield precise results. Sample sizes of 200, 300, and 400 are recommended to be minimum, acceptable, and desirable sample sizes to yield consistent results. In this study, an unbalanced factor structure (i.e., unequal numbers of items in each dimension) showed a negative impact on the precision of the factor extraction results. Tutorials on how to perform PA and the other five criteria with examples were presented. [The dissertation citations contained here are published with the permission of ProQuest LLC. Further reproduction is prohibited without permission. Copies of dissertations may be obtained by Telephone (800) 1-800-521-0600. Web page: http://www.proquest.com/en-US/products/dissertations/individuals.shtml.]
– Name: AbstractInfo
  Label: Abstractor
  Group: Ab
  Data: As Provided
– Name: DateEntry
  Label: Entry Date
  Group: Date
  Data: 2024
– Name: URL
  Label: Access URL
  Group: URL
  Data: <link linkTarget="URL" linkTerm="https://gateway.proquest.com/openurl?url_ver=Z39.88-2004&rft_val_fmt=info:ofi/fmt:kev:mtx:dissertation&res_dat=xri:pqm&rft_dat=xri:pqdiss:29282482" linkWindow="_blank">http://gateway.proquest.com/openurl?url_ver=Z39.88-2004&rft_val_fmt=info:ofi/fmt:kev:mtx:dissertation&res_dat=xri:pqm&rft_dat=xri:pqdiss:29282482</link>
– Name: AN
  Label: Accession Number
  Group: ID
  Data: ED646416
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=eric&AN=ED646416
RecordInfo BibRecord:
  BibEntity:
    Languages:
      – Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 368
    Subjects:
      – SubjectFull: Monte Carlo Methods
        Type: general
      – SubjectFull: Tests
        Type: general
      – SubjectFull: Data
        Type: general
      – SubjectFull: Sample Size
        Type: general
      – SubjectFull: Factor Analysis
        Type: general
      – SubjectFull: Construct Validity
        Type: general
      – SubjectFull: Shift Studies
        Type: general
      – SubjectFull: Criteria
        Type: general
      – SubjectFull: Research Design
        Type: general
    Titles:
      – TitleFull: A Monte Carlo Study of Parallel Analysis, Minimum Average Partial, Indicator Function, and Modified Average Roots for Determining the Number of Dimensions with Binary Variables in Test Data: Impact of Sample Size and Factor Structure
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Pornchanok Ruengvirayudh
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 01
              Type: published
              Y: 2018
          Identifiers:
            – Type: isbn-print
              Value: 979-88-02-75503-7
          Titles:
            – TitleFull: ProQuest LLC
              Type: main
ResultId 1