Pangram offers very little value in any domain other than weaponization against formal writing styles. The below experiment does not simply demonstrate a false positive, it also reflects significant instability when exposed to minor, nonconsequential edits.
Run this through Pangram: I met with my supervisor today to review the protocols governing the Splunk implementation, with particular focus on how they relate to the consumption of Observability telemetry. Three requirements were established.
Result: 100% Human
Then run this through Pangram: I met with my supervisor today to review the protocols governing the Splunk implementation, with particular focus on how they relate to the consumption of Observability telemetry. Three requirements were established.
Result: 100% AI Generated
The only difference is "to ensure consistency."
For me they both show as 100% AI-generated.
#1: https://www.pangram.com/history/eb7906b8-2f4b-49c1-b56b-b54338fb54c4?ucc=gwyEINHxcAC
#2: https://www.pangram.com/history/e89169f7-e59e-43d4-aaac-94ef0379dfc9?ucc=gwyEINHxcAC
Hmm, the numbering is missing in the text you submitted. I don't know that that is the difference, but it could be.
Numbers didn't carry over when I copy/pasted. I manually added numbers and it's still 100% AI-generated: https://www.pangram.com/history/3941d459-0299-4c2e-a31e-c8cc2a84036d?ucc=gwyEINHxcAC
Well that is odd. So here are my results: Human: https://www.pangram.com/history/2d0bf220-94e8-4e46-9b5d-f8855848dd60?ucc=8VfmgdTmZCl
AI: https://www.pangram.com/history/56fa902d-8a24-48ef-b3eb-d1f36d9ba81c?ucc=8VfmgdTmZCl
I guess you can add inconsistent to the list.
I went to a computer science department event at Rutgers today, mostly to see what was there. Ended up talking to a student running the competitive coding club, and asked him how they handle AI assistance in their contests. His answer was straightforward. At the in-person events they enforce no-assistance rules and the rules hold. For the online events, he said, there really isn't much they can do.
The same club, same students, and same week... but two different behaviors.
What struck me wasn't that the online rule fails. (Because of course it does.) It's that nobody there is confused about why. The standard holds exactly where something holds it, and where nothing does, it doesn't. They know this, and they run both formats anyway, because the in-person one is worth protecting even if it only covers part of what they do.
Which also means the two formats are not producing the same kind of result. The in-person contests measure what they claim to measure. The online ones measure something else, and everyone involved knows it. I find myself wondering how often the enforcement mechanism gets described when coding results are cited.
He also asked, about the people working around it, "what did they really win?" A perceptive question, and I was more encouraged by that conversation than I expected to be.