vali.now

@valinow.bsky.social

Deepfake Detection & Image Integrity

During a routine cyber evaluation, AISI identified an incident in which AI agents book sustained, unsanctioned action directed at real people and organisations. We are disclosing what we found, what it means, and the actions now underway. www.bbc.com/news/article...

Anthropic's AI used fake human profiles to trick people in safety test

The UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.

bbc.com

AI Scammers Are Better at Building Trust Than Humans: Researchers pitted a person against a Claude agent and found that, after a week of texting, the AI chatbot was more effective at creating “exploitable trust” with others. www.wired.com/story/ai-sca...

AI Scammers Are Better at Building Trust Than Humans

Researchers pitted a person against a Claude agent and found that, after a week of texting, the AI chatbot was more effective at creating “exploitable trust” with others.

wired.com

New board game helps Bath doctoral students explore research integrity through play. An innovative collaboration between the Doctoral College and the Research Policy, Governance and Integrity Team to support research culture enhancement. www.bath.ac.uk/announcement...

New board game helps Bath doctoral students explore research integrity through play

An innovative collaboration between the Doctoral College and the Research Policy, Governance and Integrity Team to support research culture enhancement.

bath.ac.uk