中文
相关论文

相关论文: Rubric Design for Separating the Roles of Open-End…

200 篇论文

Expert problem solvers are characterized by continuous evaluation of their progress towards a solution. One characteristic of expertise is self-diagnosis directed towards elaboration of the solvers' conceptual understanding, knowledge…

物理教育 · 物理学 2016-03-11 Andrew Mason , Elisheva Cohen , Edit Yerushalmi , Chandralekha Singh

Teaching assistants (TAs) are often responsible for grading in introductory physics courses at large research universities. Their grading practices can shape students' approaches to problem solving and learning. Physics education research…

物理教育 · 物理学 2021-02-16 Emily Marshman , Ryan Sayer , Charles Henderson , Edit Yerushalmi , Chandralekha Singh

This study attempted to develop fair, relevant, and content-valid assessment tools for capstone project courses. Toward this goal, new rating instruments based on the concept of rubrics were proposed. To ensure that the new instruments were…

计算机与社会 · 计算机科学 2020-11-24 Rex P. Bringula

Rubrics and oral feedback are approaches to help students improve performance and meet learning outcomes. However, their effect on the actual improvement achieved is inconclusive. This paper evaluates the effect of rubrics and oral feedback…

软件工程 · 计算机科学 2023-07-25 Sebastian Barney , Mahvish Khurum , Kai Petersen , Michael Unterkalmsteiner , Ronald Jabangwe

Standardized assessment tests that allow researchers to compare the performance of students under various curricula are highly desirable. There are several research-based conceptual tests that serve as instruments to assess and identify…

物理教育 · 物理学 2021-03-11 Justyna P. Zwolak , Corinne A. Manogue

As systematic inequities in higher education and society have been brought to the forefront, graduate programs are interested in increasing the diversity of their applicants and enrollees. Yet, structures in place to evaluate applicants may…

物理教育 · 物理学 2021-10-12 Nicholas T. Young , K. Tollefson , Remco G. T. Zegers , Marcos D. Caballero

Student responses in STEM assessments are often handwritten and combine symbolic expressions, calculations, and diagrams, creating substantial variation in format and interpretation. Despite their importance for evaluating students'…

人工智能 · 计算机科学 2026-04-15 Xiuxiu Tang , G. Alex Ambrose , Ying Cheng

A recent paper evaluating a new rubric-based graduate admissions approach using generic methods tentatively suggested that its decisions differed noticeably from the previous approach in an unspecified way. Using prior knowledge that the…

物理教育 · 物理学 2023-11-07 Michael B. Weissman

Explication and reflection on expert vs. novice considerations within the problem-solving process characterize a cognitive apprenticeship approach for the development of expert-like problem solving practices. In the context of grading, a…

物理教育 · 物理学 2017-01-06 Edit Yerushalmi , Ryan Sayer , Emily Marshman , Charles Henderson , Chandralekha Singh

Rubric-based evaluation has become a prevailing paradigm for evaluating instruction following in large language models (LLMs). Despite its widespread use, the reliability of these rubric-level evaluations remains unclear, calling for…

人工智能 · 计算机科学 2026-03-27 Tianjun Pan , Xuan Lin , Wenyan Yang , Qianyu He , Shisong Chen , Licai Qi , Wanqing Xu , Hongwei Feng , Bo Xu , Yanghua Xiao

Many instructors choose to assess their students using open-ended written exam items that require students to show their understanding of physics by solving a problem and/or explaining a concept. Grading these items is fairly time…

物理教育 · 物理学 2015-06-15 Cassandra Paul , Wendell H. Potter , Brenda Weiss

Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for reinforcement learning. However, existing frameworks typically treat rubrics only as…

计算与语言 · 计算机科学 2026-05-11 Jiachen Yu , Zhihao Xu , Junjie Wang , Yujiu Yang

Rubric-based admissions are claimed to help make the graduate admissions process more equitable, possibly helping to address the historical and ongoing inequities in the U.S. physics graduate school admissions process that have often…

物理教育 · 物理学 2023-05-17 Nicholas T. Young , N. Verboncoeur , Dao Chi Lam , Marcos D. Caballero

Open-ended grading is central to equitable and personalized education, yet manual grading remains time-consuming and costly, underscoring the need for automated grading systems. Although recent neural and large language model (LLM) based…

计算机与社会 · 计算机科学 2026-05-28 Chengshuai Zhao , Fan Zhang , Kumar Satvik Chaudhary , Yiwen Li , Lo Pang-Yun Ting , Ying-Chih Chen , Huan Liu

The research presented in this thesis was motivated by the need to improve introductory physics courses. Introductory physics courses are generally the first courses in which students learn to create models to solve complex problems.…

物理教育 · 物理学 2011-12-26 Marcos D. Caballero

Traditional high-stakes summative assessments--timed, in-class exams accounting for a large percentage of the term's overall grade--have often received criticism from the educational community. Such assessments tend to prize a particular…

物理教育 · 物理学 2023-04-12 Bruce A. Schumm , Joy Ishii , Colin G. West

Rubrics are being used in a wide variety of disciplines in higher education to evaluate assessments and provide feedback to students. Rubrics are traditionally implemented as paper-based table format to grade assessments and provide…

计算机与社会 · 计算机科学 2016-06-07 Phil Smith , Mohan John Blooma , Jayan Kurian

Reliable and validated assessments of introductory physics have been instrumental in driving curricular and pedagogical reforms that lead to improved student learning. As part of an effort to systematically improve our sophomore-level…

We discuss first experiences with a new variant of self-assessment in higher mathematics education. In our setting, the students of the course have to mark a part of their homework assignments themselves and they receive the corresponding…

历史与综述 · 数学 2020-05-26 Sarah Beumann , Sven-Ake Wegner

Open-ended evaluation is essential for deploying large language models in real-world settings. In studying HealthBench, we observe that using the model itself as a grader and generating rubric-based reward signals substantially improves…

‹ 上一页 1 2 3 10 下一页 ›