When models are hitting new highs on benchmarks, are they good in general or good at getting good results on benchmarks? Björn Mattsson and Pat Walter's preprint on data leakage is available on biorxiv:
https://www.biorxiv.org/content/10.64898/2026.06.29.735309v1
#bio #ai
stian
stian@nearhood.co.uk
npub1zzrs...9jem
nostr projects:
https://agentlist.com
https://nearhood.co.uk
Soon the question won't be a model's cost or "accuracy", it'll be who made it. The lab behind a model shapes how it nudges and persuades. We already see that people have "a high tendency to follow the potentially harmful medical advice" from model output (1) - imagine the power these providers can have, not just in intelligence but in persuasion.
1:
MIT Media Lab
People Overtrust AI-Generated Medical Advice despite Low Accuracy – MIT Media Lab
We conducted a study in which a total of 300 participants gave evaluations for medical responses that were either written by a medical doctor on an...
nostr is an amazing platform to build on!
View quoted note →
The greatest things to come from the AI boom is that people will meet each other more and, ironically, since we have an explosion in new apps and websites being built these days, will spend less time using them.


Event reminders working as intended 😀


Great work @Satoshi Consult🔑 First time the staff member had accepted btc at Oslo Airport, but still went very smoothly


#ProofOfWalk #London @BitcoinWalk

