Alasdair Stewart
@abestew
Sociologist | Socialist | Neurodivergent | Linux & FLOSS advocate | he/they
The demo is also reliant upon info from TeachFX (see screenshots), a concerning AI product that records lessons & creates summary, being fed to Claude. So 'Claude for Teachers' basically ingests this summary & spits out a verbose version with a suggested next lesson plan & risk of hallucinations.
Anthropic released 'Claude for Teachers' & even the promo video shows issues with it. Image 1 = 1st response that was generated, Images 2&3 = 2nd response generated. Some aspects are the same, but also contain critical differences - such as the educational standards. www.youtube.com/watch?v=V-OO...
I wonder why OpenAI is joining Anthropic in suggesting a slow down in the area that bleeds them the most money... 🤔
Chapter 8 also looks to be OK. Chapter 1 similarly aside from one segment. So again, if remove those chapters, what is the AI score for chapters 2, 3, 5, 7, and 9. - 37% AI generated with minimal human input. - 3% moderatley AI assisted. - 3% lightly AI assisted.
With chapters 4 & 6 removed, Pangram gives the following for the main text of the Milburn interim report: - 23% AI generated with minimal human input or revision. - 7% moderatley AI assisted. - 3% lightly AI assisted.
Pangram's graph of potential AI usage across the Milburn interim report. The two long dips map directly onto Chapters 4 & 6. So, whoever wrote those chapters, well-done - your colleagues clearly let you down. Which obviously leaves the question, how much is AI if remove those two chapters? 23%...
Some more (highly likely) egregious AI slop within the Milburn interim report.
AI for some reason often paraphrases percentages needlessly and in way that makes the figures vaguer, such as replacing 24% with "over 23%". See screenshot from report. Actual figures in cited text: lazy and work shy - 23% overly sensitive - 34% entitled - 27% rejected solely on basis of age - 9%
OK, it looks like there is a lot of AI slop in the Milburn interim report. Signed up for Pangram 7 day free trial. Of the main text: - 13% is high confidence AI generated - 3% is moderately AI-assisted - 1% is lightly AI-assisted 17% accounts for **10,451 words** of the main text.
The Milburn led "Young people and work: interim report" reads like genAI had a heavy hand in 'writing' it? Pangram suggests 429-438 - where the main attack on neurodivergents begins - is 100% AI-generated. Even the subtitles - "configured for treatment, not participation" - scream genAI.
Note - whilst you can manually turn off the AI bullshit in DuckDuckGo, it's also possible to add a 'No AI bullshit' version of DuckDuckGo as your browser's default search engine as well.
Google Search replacing dictionary definitions with AI bullshit was final straw, switched to duckduckgo.
Looks like chick reached safety. All quiet at the back, and the magpies - including short-tail - have returned to their usual hangout spot round the front.
Now it's becoming crows v magpies AND seagulls, the latter clearly assuming the crows are after the seagull chicks rather than the crows merely following their own along the ground near them.
Crow taking issue with me getting photo of chick near window is doing show of strength by attempting to fuck up my satellite dish - little does it know its not been in use for over five years.
Surprise development - the near two hour battle seems to be crows defending a chick thats fallen from the nest.
Spoke too soon, crows retreated merely to plan some bizarre ground offensive - remaining as clearly pissed off as before.
Just when I thought they had won the crows have retreated. No surprise to see the first magpie to claim victory in the backgarden is short-tail, quickly joined by their partner.
Battle seems to be coming to end, though crows still taking their aggression out on the trees.
This crow did not like my attempt to get a photo and immediately turned to swoop in at the window. Got to hand it to the magpies, I wouldn't fancy my odds against one of these riled up angry bastards never mind trying to take on two of them.
Also to demonstrate it's flexibility, I included within the prototype a series of checks for potential genAI misuse in data analysis projects. (Screenshots are using a dummy project created by using ChatGPT to create code + interpretation for an assessment on module I co-convene.)
I didn't get as far in the prototype with this aspect - but if your institution requires marking up the submission & matching sources, it has a tool with original submission on left, tabs for matching sources on right, and streamlines adding highlighting - auto-picking unique colours for the sources
The app does not 'automate' plagiarism investigation, nothing goes into the report that a user doesn't confirm. It flags things for attention, streamlines manual checks academics are already doing, enables writing comments as go, and outputs a neat and consistently formatted report.
A year ago, I made a prototype plagiarism investigation & referral writing tool - SherlockCite. For over a year now, I have struggled to find way to get it approved for use at my institution. So - on my todo list is refactoring the main code and redoing the UI in QML then open-sourcing it.
Initial quick sketch for an alternative to the usual AGENTS / CLAUDE md files.
NetHack (thegreatestgameyouwilleverplay.com) version 5.0.0 released with... "over 3100 fixes and changes". nethack.org/v500/release...
"The Boy That Cried Mythos: Verification is Collapsing Trust in Anthropic" www.flyingpenguin.com/the-boy-that...