REVIEW 3 cited by
RLTutor: Reinforcement Learning Based Adaptive Tutoring System by Modeling Virtual Student with Fewer Interactions
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
A major challenge in the field of education is providing review schedules that present learned items at appropriate intervals to each student so that memory is retained over time. In recent years, attempts have been made to formulate item reviews as sequential decision-making problems to realize adaptive instruction based on the knowledge state of students. It has been reported previously that reinforcement learning can help realize mathematical models of students learning strategies to maintain a high memory rate. However, optimization using reinforcement learning requires a large number of interactions, and thus it cannot be applied directly to actual students. In this study, we propose a framework for optimizing teaching strategies by constructing a virtual model of the student while minimizing the interaction with the actual teaching target. In addition, we conducted an experiment considering actual instructions using the mathematical model and confirmed that the model performance is comparable to that of conventional teaching methods. Our framework can directly substitute mathematical models used in experiments with human students, and our results can serve as a buffer between theoretical instructional optimization and practical applications in e-learning systems.
Forward citations
Cited by 3 Pith papers
-
Constructing a Question-Answering Simulator through the Distillation of LLMs
LDSim distills an LLM's concept-prerequisite knowledge and mastery reasoning into a lightweight simulator that beats LLM-based and LLM-free baselines on four knowledge-tracing datasets.
-
Personalized Education with Ranking Alignment Recommendation
Ranking Alignment Recommendation adds a collaborative ranking loss to RL-based question recommenders, improving simulated learning effects across five environments.
-
GraphRAG-Induced Dual Knowledge Structure Graphs for Personalized Learning Path Recommendation
KnowLP combines LLM-generated prerequisite and similarity knowledge graphs with reinforcement learning to recommend personalized learning paths, reporting state-of-the-art results on Junyi, MOOCCubeX, and ASSISTments2009.
Discussion (0). Continue with ORCID to comment.