Researchers surprised that with AI, toxicity is harder to fake than intelligence

The next time you encounter an unusually polite reply on social media, you might want to check twice. It could be an AI model trying (and failing) to blend in with the crowd.

On Wednesday, researchers from the University of Zurich, University of Amsterdam, Duke University, and NYU released a study revealing that AI models remain easily distinguishable from humans in social media conversations, with overly friendly emotional tone serving as the most persistent giveaway. The research, which tested nine open-weight models across Twitter/X, Bluesky, and Reddit, found that classifiers developed by the researchers detected AI-generated replies with 70 to 80 percent accuracy.

The study introduces what the authors call a “computational Turing test” to assess how closely AI models approximate human language. Instead of relying on subjective human judgment about whether text sounds authentic, the framework uses automated classifiers and linguistic analysis to identify specific features that distinguish machine-generated from human-authored content.

Read full article

Comments

Researchers surprised that with AI, toxicity is harder to fake than intelligence

Leave a Reply Cancel reply

James Watson, who helped unravel DNA’s double-helix, has died

Researchers surprised that with AI, toxicity is harder to fake than intelligence

Commercial spyware “Landfall” ran rampant on Samsung phones for almost a year

FBI subpoena tries to unmask mysterious founder of Archive.today

You Missed

James Watson, who helped unravel DNA’s double-helix, has died

Researchers surprised that with AI, toxicity is harder to fake than intelligence

Commercial spyware “Landfall” ran rampant on Samsung phones for almost a year

FBI subpoena tries to unmask mysterious founder of Archive.today

Researchers surprised that with AI, toxicity is harder to fake than intelligence

Related Post

Leave a Reply Cancel reply

You Missed