Friday, November 7, 2025

Researchers surprised that with AI, toxicity is harder to fake than intelligence


The next time you encounter an unusually polite reply on social media, you might want to check twice. It could be an AI model trying (and failing) to blend in with the crowd.

On Wednesday, researchers from the University of Zurich, University of Amsterdam, Duke University, and NYU released a study revealing that AI models remain easily distinguishable from humans in social media conversations, with overly friendly emotional tone serving as the most persistent giveaway. The research, which tested nine open-weight models across Twitter/X, Bluesky, and Reddit, found that classifiers developed by the researchers detected AI-generated replies with 70 to 80 percent accuracy.

The study introduces what the authors call a “computational Turing test” to assess how closely AI models approximate human language. Instead of relying on subjective human judgment about whether text sounds authentic, the framework uses automated classifiers and linguistic analysis to identify specific features that distinguish machine-generated from human-authored content.

Read full article

Comments

Reference : https://ift.tt/dmqjg07

No comments:

Post a Comment

Low-Vision Programmers Can Now Design 3D Models Independently

Most 3D design software requires visual dragging and rotating—posing a challenge for blind and low-vision users. As a result, a range of ...