CHI '24

Critical AI Writing Assessment Workbench

Operationalizing Shibani et al. Empirical Framework across 5 Dimensions

Presets:

1. Turn-by-Turn AI Dialogue Classifier

Classify each prompt/response against the 5 CHI '24 dimensions. Toggle depth to observe metric shifts.

2. 5-Question Intellectual Contribution Scaffolder

CHI '24 Assessment Design guidelines to verify human agency, critical scrutiny, and hallucination containment.

Audit Telemetry & CHI '24 Benchmark

Empirical baseline of 49 graduate student assignments

5
Total Turns
2
Deep Turns
3
Shallow Turns
Planning & Generating Ideas Shallow: 100% (Cohort: 84%)
Finding & Evaluating Info Shallow: 100% (Cohort: 90%)
Writing & Revising Shallow: 100% (Cohort: 100%, 0% Deep)
Dialogue & Conversational Probing Deep: 100% (Cohort: 8% Deep, 91% Shallow)
Reflecting on AI Learning Deep: 100% (Cohort: Deepest Area)
Student Transcript Shibani '24 Baseline
Assessment Diagnostic Verdict
Critical Depth in Reflection & Dialogue; Shallow Planning and Revision typical of baseline cohort
Explicitly Documented with Independent Cross-Verification