AI philosopher systems invite unusually strong claims about voice, history, sources, and intellectual authority. Publishing evaluation boundaries makes those claims easier to scrutinize and gives educators, philosophers, researchers, and users a clearer basis for deciding what the product should and should not be trusted to do.
These pages are implementation evidence, not peer-reviewed academic research. They can complement scholarly work on Socratic dialogue, AI tutoring, public philosophy, and human-AI interaction, but they should be cited for what they actually document rather than treated as independent academic validation.