Skip to content
Merged
Show file tree
Hide file tree
Changes from all commits
Commits
File filter

Filter by extension

Filter by extension

Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
9 changes: 8 additions & 1 deletion Docs/Teaching/committee.en.md
Original file line number Diff line number Diff line change
Expand Up @@ -93,9 +93,16 @@ not, in the finding itself:

> A writing-analysis tool was used to decide which sections to examine. Its output is not evidence of
> authorship and was not treated as such. It publishes a false-positive rate on writing known to be
> human of under 4.1% overall, with no threshold supported for either language on its own — the
> human, measured for the language this work is written in and printed on the report itself — the
> findings below rest on the source verification and the interview, not on that tool.

**Copy the figure from the report in front of you, not from here.** The tool measures a separate rate
per language and the report prints the one that applies to the work being judged. Quoting the pooled
figure instead is the single easiest way to overstate this tool at a hearing: on the current corpus
the pooled rate is under 4.1%, while Spanish on its own supports only 13.3% — three times worse. A
committee handed the flattering number about a Spanish essay has been given a better tool than the
one that was actually used, and the difference is the sort of thing that surfaces on appeal.

A committee that writes this sentence is in a stronger position than one that omits it, because the
sentence is going to be raised at appeal whether or not it appears.

Expand Down
14 changes: 11 additions & 3 deletions Docs/Teaching/committee.es.md
Original file line number Diff line number Diff line change
Expand Up @@ -96,9 +96,17 @@ resolución qué es y qué no es:

> Se utilizó una herramienta de análisis de escritura para decidir qué secciones examinar. Su
> resultado no es prueba de autoría y no se ha tratado como tal. Publica una tasa de falsos positivos
> sobre escritura conocidamente humana inferior al 4,1% en agregado, sin que ningún idioma por
> separado respalde un umbral propio; las conclusiones siguientes se apoyan en la verificación de
> fuentes y en la entrevista, no en esa herramienta.
> sobre escritura conocidamente humana, medida para el idioma en que está escrito este trabajo e
> impresa en el propio informe; las conclusiones siguientes se apoyan en la verificación de fuentes y
> en la entrevista, no en esa herramienta.

**Copie la cifra del informe que tiene delante, no de aquí.** La herramienta mide una tasa distinta
por idioma y el informe imprime la que corresponde al trabajo que se está juzgando. Citar la cifra
agregada es la forma más fácil de exagerar esta herramienta en una audiencia: en el corpus actual la
agregada queda por debajo del 4,1%, mientras que el español por separado solo respalda un 13,3% —
tres veces peor. A un comité al que se le entrega el número halagador sobre un ensayo en español se
le ha dado una herramienta mejor que la que realmente se usó, y esa diferencia es de las que salen a
la luz en una apelación.

Un comité que escribe esa frase queda en mejor posición que uno que la omite, porque la cuestión se
va a plantear en apelación aparezca o no.
Expand Down
7 changes: 6 additions & 1 deletion Docs/Teaching/student-sheet.en.md
Original file line number Diff line number Diff line change
Expand Up @@ -24,7 +24,12 @@ patterns. Software notices patterns in innocent work constantly.
It publishes how often it is wrong about writing known to be human — a figure almost no tool in this
category will state about itself. On its test corpus of ninety texts published before AI writing
existed, it flagged **none of them** at its recommended setting. But that is ninety texts, and the
honest reading is the range around it, not the zero: somewhere **under about 4 in 100**.
honest reading is the range around it, not the zero.

Of those ninety, sixty-five were in English. This sheet quotes the **English** figure, which is the
one that applies to you: somewhere **under about 6 in 100**. The pooled figure that mixes both
languages is more flattering, and using it here would credit the tool with a precision nobody
measured on writing in your language.

It also does something most tools refuse to do: **below its supported threshold it prints no verdict
at all.** A low score is not a certificate that a person wrote something, and it does not claim to be
Expand Down
8 changes: 6 additions & 2 deletions Docs/Teaching/student-sheet.es.md
Original file line number Diff line number Diff line change
Expand Up @@ -24,8 +24,12 @@ programas detectan patrones en trabajos inocentes todo el tiempo.
Publica con qué frecuencia se equivoca sobre escritura que se sabe humana — una cifra que casi
ninguna herramienta de esta categoría dice sobre sí misma. En su corpus de noventa textos publicados
antes de que existiera la escritura con IA, no marcó **ninguno** en su ajuste recomendado. Pero son
noventa textos, y la lectura honesta es el rango alrededor de ese cero, no el cero: algo **por debajo
de 4 de cada 100**.
noventa textos, y la lectura honesta es el rango alrededor de ese cero, no el cero.

Y de esos noventa, solo veinticinco estaban en español. Esta hoja cita la cifra del **español**, que
es la que te corresponde: algo **por debajo de 13 de cada 100**. La cifra agregada que mezcla los dos
idiomas es más favorable, y usarla contigo sería atribuirle a la herramienta una precisión que nadie
midió sobre escritura en tu idioma.

También hace algo que la mayoría se niega a hacer: **por debajo de su umbral respaldado no imprime
ningún veredicto.** Una puntuación baja no es un certificado de que lo escribió una persona, y no
Expand Down
8 changes: 4 additions & 4 deletions Docs/Teaching/syllabus.en.md
Original file line number Diff line number Diff line change
Expand Up @@ -40,10 +40,10 @@ never the reason for a decision about a student.** Everything else is negotiable
> Being asked is not an accusation and carries no penalty. It is part of how work is assessed in
> this course, and I may ask anyone.
>
> I may run submitted work through an offline writing-analysis tool. It produces no verdict, and no
> score from it will ever be the basis of a decision about you. Its purpose is to tell me where to
> read more carefully — the same thing my own eyes do, applied evenly to everyone's work rather than
> only to the students I happen to wonder about.
> I may run submitted work through an offline writing-analysis tool. It reaches no conclusion about
> who wrote anything, and nothing it produces will ever be the basis of a decision about you. Its
> purpose is to tell me where to read more carefully — the same thing my own eyes do, applied evenly
> to everyone's work rather than only to the students I happen to wonder about.

## Option C — Not permitted

Expand Down
5 changes: 3 additions & 2 deletions Docs/Teaching/syllabus.es.md
Original file line number Diff line number Diff line change
Expand Up @@ -39,8 +39,9 @@ detector nunca es el motivo de una decisión sobre un estudiante.** Todo lo dem
> Que se lo pida no es una acusación ni conlleva ninguna penalización. Es parte de cómo se evalúa en
> esta asignatura, y puedo pedírselo a cualquiera.
>
> Puedo analizar los trabajos con una herramienta que funciona sin conexión. No emite veredictos, y
> ninguna puntuación suya será la base de una decisión sobre usted. Sirve para indicarme dónde leer
> Puedo analizar los trabajos con una herramienta que funciona sin conexión. No concluye nada sobre
> quién escribió qué, y nada de lo que produzca será la base de una decisión sobre usted. Sirve para
> indicarme dónde leer
> con más atención — lo mismo que hacen mis propios ojos, aplicado por igual a todos los trabajos y
> no solo a los estudiantes sobre los que casualmente me pregunte algo.

Expand Down
Loading