NIST CAISI
Katherine Elkins and Jon Chun lead the Modern Language Association’s consortium team, contributing humanities and language expertise to federal AI standards.
- Agent security. How agents interpret natural-language instructions rather than simply executing them.
- Interpretive tractability. Whether human overseers can meaningfully evaluate an agent’s actions and reported reasoning at the pace agents operate.
- Benchmarks. Evaluation must account for the languages, cultural assumptions, and prompt formulations through which capabilities are measured.
Syntactic fragility: small structural changes in otherwise equivalent prompts can reverse a model’s recommendation.
