RISS 학술연구정보서비스

검색

인기 검색어

    다국어 입력

    http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.

    변환된 중국어를 복사하여 사용하시면 됩니다.

    예시)
    • 中文 을 입력하시려면 zhongwen을 입력하시고 space를누르시면됩니다.
    • 北京 을 입력하시려면 beijing을 입력하시고 space를 누르시면 됩니다.
    닫기

    Comparability and linking in direct writing assessment: Benchmarks, discourse mode, and grade level.

    한글로보기

    https://www.riss.kr/link?id=T10565283

    • 0

      상세조회
    • 0

      다운로드
    서지정보 열기
    • 내보내기
    • 내책장담기
    • 공유하기
    • 오류접수

    부가정보

    다국어 초록 (Multilingual Abstract) kakao i 다국어 번역

    Increasingly, direct assessments of writing performance are being included in large-scale testing programs despite concerns regarding reliability and validity. Issues regarding assessing student writing across discourse modes and measuring growth across grade level have generated interest as well as concern. The purpose of this study was to examine the effects of: (a) different scoring benchmarks on scores for the same papers, (b) discourse modes on scores for papers by the same students, (c) grade level on scores for papers written in a single discourse mode, and (d) grade level on scores for papers written in different discourse modes. Raters scored writing samples from students in Grades 3, 5, and 8 against a common rubric. Raw ratings were analyzed using multi-facet Rasch models. Raw ratings and Rasch-estimated student abilities, trait difficulties, and rater leniency-severity parameters were examined. Ratings of the same essays differed in magnitude and relative rank when scored against different sets of benchmarks. Ratings of papers written in different discourse modes by the same students had similar features such as similarly rank-ordered analytic trait difficulties. However, ratings for different modes led to substantial inconsistencies in how students were classified based on various performance standards. Ratings of student writing in a single mode increased with grade level. Comparisons of writing ability on a common task appear to be possible across grade levels, given benchmarks chosen from the multi-grade set of sample papers. The validity of comparing ratings of student writing in different modes across grade levels remains questionable. Results indicate that directly adjusting for discourse mode may be a promising approach to assess general writing quality across modes, but adjusting for mode may not be sufficient to allow for successful linking across grade levels. The benchmark papers used to operationalize the rubric score points strongly influenced the ratings of students' papers as well. Results of this work add to growing cautions and concerns regarding the use and interpretation of large-scale writing assessment scores and suggest the need for careful research on the nature of benchmark papers and the processes used to select them.
    번역하기

    Increasingly, direct assessments of writing performance are being included in large-scale testing programs despite concerns regarding reliability and validity. Issues regarding assessing student writing across discourse modes and measuring growth acr...

    Increasingly, direct assessments of writing performance are being included in large-scale testing programs despite concerns regarding reliability and validity. Issues regarding assessing student writing across discourse modes and measuring growth across grade level have generated interest as well as concern. The purpose of this study was to examine the effects of: (a) different scoring benchmarks on scores for the same papers, (b) discourse modes on scores for papers by the same students, (c) grade level on scores for papers written in a single discourse mode, and (d) grade level on scores for papers written in different discourse modes. Raters scored writing samples from students in Grades 3, 5, and 8 against a common rubric. Raw ratings were analyzed using multi-facet Rasch models. Raw ratings and Rasch-estimated student abilities, trait difficulties, and rater leniency-severity parameters were examined. Ratings of the same essays differed in magnitude and relative rank when scored against different sets of benchmarks. Ratings of papers written in different discourse modes by the same students had similar features such as similarly rank-ordered analytic trait difficulties. However, ratings for different modes led to substantial inconsistencies in how students were classified based on various performance standards. Ratings of student writing in a single mode increased with grade level. Comparisons of writing ability on a common task appear to be possible across grade levels, given benchmarks chosen from the multi-grade set of sample papers. The validity of comparing ratings of student writing in different modes across grade levels remains questionable. Results indicate that directly adjusting for discourse mode may be a promising approach to assess general writing quality across modes, but adjusting for mode may not be sufficient to allow for successful linking across grade levels. The benchmark papers used to operationalize the rubric score points strongly influenced the ratings of students' papers as well. Results of this work add to growing cautions and concerns regarding the use and interpretation of large-scale writing assessment scores and suggest the need for careful research on the nature of benchmark papers and the processes used to select them.

    더보기

    분석정보

    View

    상세정보조회

    0

    Usage

    원문다운로드

    0

    대출신청

    0

    복사신청

    0

    EDDS신청

    0

    동일 주제 내 활용도 TOP

    더보기

    주제

    연도별 연구동향

    연도별 활용동향

    연관논문

    연구자 네트워크맵

    공동연구자 (7)

    유사연구자 (20) 활용도상위20명

    이 자료와 함께 이용한 RISS 자료

    나만을 위한 추천자료

    해외이동버튼