Independent evaluation of an oral reading fluency model
Before a speech-based literacy model is deployed at scale, its accuracy and fairness need to be independently assessed. APLYD evaluated the system without involvement in its development.
Operational Context
Oral reading fluency tools use speech models to assess children’s reading performance. At scale, differences in model performance across learner groups can affect the reliability and fairness of the resulting assessments.
Institutional Delivery
We independently evaluated the model across approximately 8,000 Hindi and Gujarati recordings, assessing accuracy and bias across grades, dialects and gender groups before deployment.
Public Value & Lasting Impact
Independent evaluation provides an evidence base for deployment decisions by separating assessment from development. In this case, the model was assessed for accuracy and bias before wider use, providing evidence on its performance across different learner groups.




