Despite AI reshaping writing assessment, fairness remains a matter of human judgment as well as careful measurement. This study uses cross-classified modeling of an operational L2 writing placement test to examine prompt comparability and argues that writing teachers’ interpretive and ethical roles remain central in identifying and addressing inequitable task effects.