Brett Reynolds

Brett Reynolds

Professor of English for Academic Purposes and TESL, Humber Polytechnic. Adjunct professor of linguistics, University of Toronto.

I study when labels, scores, and categories support reliable inference, and when they don’t. The cases come from AI evaluation and assurance, English grammar, and the philosophy of science.


Current work

In development

Two companion frameworks: delegation assurance, for tool-using systems that act under delegated authority, and evidentiary assurance, for audit, challenge, and remediation. Two completed manuscripts, being prepared for public release.

What does a label license you to infer?

Labels don’t need essences to be useful. They need a record of supporting the right inferences for a stated field and purpose. I study how that warrant is produced, how far it projects, and how it fails, in grammatical categories, benchmark scores, and evaluator judgments.

Each row is one form of that question. Each column is a field that has to settle it on different evidence. Some work answers a row in more than one field at once, and appears in more than one cell; that overlap is the point rather than an error in the filing.

Trace one field: AI evaluation · philosophy of science · English grammar · show all

In AI evaluationIn philosophy of scienceIn English grammar
What does membership let you predict?
In AI evaluation Adversarial Pragmatics Public preprint — Adversarial Pragmatics · What a safety-benchmark score licenses In English grammar Interjection as a lexical category Public preprint — Interjection as a lexical category · What membership lets us predict
Who is competent to judge, and in what role?
In AI evaluation The adjudication protocol Public preprint — The adjudication protocol · LLM-judge validation, open to inspection In philosophy of science Truth-tracking profiles Public preprint — Truth-tracking profiles · What large language models participate in In English grammar Expert grammaticality judges Public preprint — Expert grammaticality judges · Evaluators, not participants. The same argument constrains who may adjudicate an AI evaluation
Where does the inference stop?
In AI evaluation Delegation assurance In development · Tool-using systems acting under delegated authority In philosophy of science Effective without warrant Completed manuscript — Effective without warrant · Status that is effective without being warranted In English grammar The homeostatic maintenance of English countability Completed manuscript — The homeostatic maintenance of English countability · The cluster dissociates in a constrained order
What keeps a category standing when nothing essential holds it together?
In AI evaluation Evidentiary assurance In development · Audit, challenge, and remediation In philosophy of science Not every stable cluster is homeostatic Public preprint — Not every stable cluster is homeostatic · Separating achievements the homeostatic component conflates In English grammar Grammaticality de-idealized Public preprint — Grammaticality de-idealized · Grammaticality as conditioned stability
General framework
Words That Won’t Hold Still: How Linguistic Categories WorkCompleted manuscript · Book-length statement.
Kinds as projectibility profiles: support grades and demotion rulesPublic preprint — Kinds as projectibility profiles: support grades and demotion rules · How a category earns, keeps, or loses its standing

Open questions

Four things I don’t know, and what would make me give up each position.

If you work on any of these, I’d like to hear from you: brett.reynolds@humber.ca

Selected publications

Full publication list, 1998 to present.

Grammar, teaching, and resources