

it’s fundamentally unsolvable, you can only mitigate it, mostly by using classical means to constrain the deterministic (i.e. non-AI) tools the chatbot is allowed access to, and constantly asking the user for confirmation.
With yolo/auto mode (no user confirmation required) and training LLMs on known vulnerabilities things will inevitably get more complicated.














The university actually cleared him of the plagiarism charges early on, they had hearings and landed on it being within accepted limits for this sort of thing.
It’s insane how the whole thing went from the initial coverage implying he’s a complete and total fraud that Cambridge didn’t do due diligence on because of woke to actually, even his more far fetched claims about endurance running seem to check out.
It would be hilarious if it wasn’t murder.