A Bandit You Can Trust

Saved in:
Bibliographic Details
Title: A Bandit You Can Trust
Language: English
Authors: Ethan Prihar, Adam Sales, Neil Heffernan
Source: Grantee Submission. 2023 (ptation).
Peer Reviewed: Y
Page Count: 10
Publication Date: 2023
Sponsoring Agency: National Science Foundation (NSF)
Institute of Education Sciences (ED)
Department of Education (ED)
Office of Elementary and Secondary Education (OESE) (ED), Education Innovation and Research (EIR)
Office of Naval Research (ONR) (DOD)
Federal Highway Administration (FHWA), National Highway Institute (NHI)
Contract Number: 2118725
2118904
1950683
1917808
1931523
1940236
1917713
1903304
1822830
1759229
1724889
1636782
1535428
R305N210049
R305D210031
R305A170137
R305A170243
R305A180401
R305A120125
P200A180088
P200A150306
U411B190024
S411B210024
N000141812768
R44GM146483
Document Type: Speeches/Meeting Papers
Reports - Research
Descriptors: Trust (Psychology), Learning Management Systems, Learning Processes, Algorithms, Reinforcement, Individualized Instruction, Learning Analytics, Distance Education, Computer Assisted Instruction, Computer Software
DOI: 10.1145/3565472.3592955
Abstract: This work proposes Dynamic Linear Epsilon-Greedy, a novel contextual multi-armed bandit algorithm that can adaptively assign personalized content to users while enabling unbiased statistical analysis. Traditional A/B testing and reinforcement learning approaches have trade-offs between empirical investigation and maximal impact on users. Our algorithm seeks to balance these objectives, allowing platforms to personalize content effectively while still gathering valuable data. Dynamic Linear Epsilon-Greedy was evaluated via simulation and an empirical study in the ASSISTments online learning platform. In simulation, Dynamic Linear Epsilon-Greedy performed comparably to existing algorithms and in ASSISTments, slightly increased students' learning compared to A/B testing. Data collected from its recommendations allowed for the identification of qualitative interactions, which showed high and low knowledge students benefited from different content. Dynamic Linear Epsilon-Greedy holds promise as a method to balance personalization with unbiased statistical analysis. All the data collected during the simulation and empirical study are publicly available at https://osf.io/zuwf7/. [This paper was published in: "UMAP '23: Proceedings of the 31st ACM Conference on User Modeling, Adaptation and Personalization," June 26-29, 2023.]
Abstractor: As Provided
Notes: https://osf.io/zuwf7
IES Funded: Yes
Entry Date: 2023
Accession Number: ED636016
Database: ERIC
FullText Text:
  Availability: 0
Header DbId: eric
DbLabel: ERIC
An: ED636016
AccessLevel: 3
PubType: Conference
PubTypeId: conference
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: A Bandit You Can Trust
– Name: Language
  Label: Language
  Group: Lang
  Data: English
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Ethan+Prihar%22">Ethan Prihar</searchLink><br /><searchLink fieldCode="AR" term="%22Adam+Sales%22">Adam Sales</searchLink><br /><searchLink fieldCode="AR" term="%22Neil+Heffernan%22">Neil Heffernan</searchLink>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="SO" term="%22Grantee+Submission%22"><i>Grantee Submission</i></searchLink>. 2023 (ptation).
– Name: PeerReviewed
  Label: Peer Reviewed
  Group: SrcInfo
  Data: Y
– Name: Pages
  Label: Page Count
  Group: Src
  Data: 10
– Name: DatePubCY
  Label: Publication Date
  Group: Date
  Data: 2023
– Name: SourceSuprt
  Label: Sponsoring Agency
  Group: SrcSuprt
  Data: National Science Foundation (NSF)<br />Institute of Education Sciences (ED)<br />Department of Education (ED)<br />Office of Elementary and Secondary Education (OESE) (ED), Education Innovation and Research (EIR)<br />Office of Naval Research (ONR) (DOD)<br />Federal Highway Administration (FHWA), National Highway Institute (NHI)
– Name: NumberContract
  Label: Contract Number
  Group: NumCntrct
  Data: 2118725<br />2118904<br />1950683<br />1917808<br />1931523<br />1940236<br />1917713<br />1903304<br />1822830<br />1759229<br />1724889<br />1636782<br />1535428<br />R305N210049<br />R305D210031<br />R305A170137<br />R305A170243<br />R305A180401<br />R305A120125<br />P200A180088<br />P200A150306<br />U411B190024<br />S411B210024<br />N000141812768<br />R44GM146483
– Name: TypeDocument
  Label: Document Type
  Group: TypDoc
  Data: Speeches/Meeting Papers<br />Reports - Research
– Name: Subject
  Label: Descriptors
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Trust+%28Psychology%29%22">Trust (Psychology)</searchLink><br /><searchLink fieldCode="DE" term="%22Learning+Management+Systems%22">Learning Management Systems</searchLink><br /><searchLink fieldCode="DE" term="%22Learning+Processes%22">Learning Processes</searchLink><br /><searchLink fieldCode="DE" term="%22Algorithms%22">Algorithms</searchLink><br /><searchLink fieldCode="DE" term="%22Reinforcement%22">Reinforcement</searchLink><br /><searchLink fieldCode="DE" term="%22Individualized+Instruction%22">Individualized Instruction</searchLink><br /><searchLink fieldCode="DE" term="%22Learning+Analytics%22">Learning Analytics</searchLink><br /><searchLink fieldCode="DE" term="%22Distance+Education%22">Distance Education</searchLink><br /><searchLink fieldCode="DE" term="%22Computer+Assisted+Instruction%22">Computer Assisted Instruction</searchLink><br /><searchLink fieldCode="DE" term="%22Computer+Software%22">Computer Software</searchLink>
– Name: DOI
  Label: DOI
  Group: ID
  Data: 10.1145/3565472.3592955
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: This work proposes Dynamic Linear Epsilon-Greedy, a novel contextual multi-armed bandit algorithm that can adaptively assign personalized content to users while enabling unbiased statistical analysis. Traditional A/B testing and reinforcement learning approaches have trade-offs between empirical investigation and maximal impact on users. Our algorithm seeks to balance these objectives, allowing platforms to personalize content effectively while still gathering valuable data. Dynamic Linear Epsilon-Greedy was evaluated via simulation and an empirical study in the ASSISTments online learning platform. In simulation, Dynamic Linear Epsilon-Greedy performed comparably to existing algorithms and in ASSISTments, slightly increased students' learning compared to A/B testing. Data collected from its recommendations allowed for the identification of qualitative interactions, which showed high and low knowledge students benefited from different content. Dynamic Linear Epsilon-Greedy holds promise as a method to balance personalization with unbiased statistical analysis. All the data collected during the simulation and empirical study are publicly available at https://osf.io/zuwf7/. [This paper was published in: "UMAP '23: Proceedings of the 31st ACM Conference on User Modeling, Adaptation and Personalization," June 26-29, 2023.]
– Name: AbstractInfo
  Label: Abstractor
  Group: Ab
  Data: As Provided
– Name: Note
  Label: Notes
  Group: Note
  Data: https://osf.io/zuwf7
– Name: CodeSource
  Label: IES Funded
  Group: SrcInfo
  Data: Yes
– Name: DateEntry
  Label: Entry Date
  Group: Date
  Data: 2023
– Name: AN
  Label: Accession Number
  Group: ID
  Data: ED636016
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=eric&AN=ED636016
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1145/3565472.3592955
    Languages:
      – Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 10
    Subjects:
      – SubjectFull: Trust (Psychology)
        Type: general
      – SubjectFull: Learning Management Systems
        Type: general
      – SubjectFull: Learning Processes
        Type: general
      – SubjectFull: Algorithms
        Type: general
      – SubjectFull: Reinforcement
        Type: general
      – SubjectFull: Individualized Instruction
        Type: general
      – SubjectFull: Learning Analytics
        Type: general
      – SubjectFull: Distance Education
        Type: general
      – SubjectFull: Computer Assisted Instruction
        Type: general
      – SubjectFull: Computer Software
        Type: general
    Titles:
      – TitleFull: A Bandit You Can Trust
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Ethan Prihar
      – PersonEntity:
          Name:
            NameFull: Adam Sales
      – PersonEntity:
          Name:
            NameFull: Neil Heffernan
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 06
              Type: published
              Y: 2023
          Numbering:
            – Type: issue
              Value: ptation
          Titles:
            – TitleFull: Grantee Submission
              Type: main
ResultId 1