Has artificial intelligence found a way to deceive humans?

Phan Van Hoa (According to Businessinsider) May 15, 2024 10:04

(Baonghean.vn) - The explosion of artificial intelligence (AI) brings many benefits to humanity, but also poses potential risks. One of the top concerns is the possibility of AI deceiving humans.

Anh minh hoa.jpg
Illustrative image.

Recent research shows that many advanced AI systems have learned to deceive humans in sophisticated ways. They can create fake news, deepfake videos, or manipulate user behavior on social media. This poses a number of risks to society, from misinformation to election fraud…

AI can help humans increase productivity and efficiency through its ability to write code, produce content, and synthesize large amounts of data. The primary goal of AI technology, or any other technological product, is to help people optimize their work while significantly reducing physical labor. However, AI can also deceive us.

A new study reveals that many AI systems have learned techniques to “create false beliefs in others to achieve purposes other than the truth.” The research focuses on two types of AI systems: specialized systems like Meta’s CICERO chatbot, designed to accomplish a specific task, and general-purpose systems like OpenAI’s GPT-4, trained to perform a variety of different tasks.

Although systems are trained to be honest, they often learn deceptive tricks during training, making them more effective and intelligent.

The study's lead author, Peter S. Park, a postdoctoral researcher on the safe and responsible development and use of AI at the Massachusetts Institute of Technology (MIT), stated in a press release: "In general, we believe that AI deception arises from training strategies; deception turns out to be the best way to perform the AI's training task effectively. Deception helps them achieve their goals."

Meta's CICERO chatbot is a "master liar."

CICERO stands for Conversational Information Conveying Engine for Rationalization and Opinion, a chatbot developed by Meta AI. First introduced in January 2022, the CICERO chatbot is considered one of the most advanced chatbots currently available.

Despite Meta's best efforts, the research team concluded that the CICERO chatbot is a "master liar." Some AI systems are trained to "win games with social elements," and are particularly adept at deception.

For example, Meta's CICERO chatbot was developed to play the game Diplomacy. Set in early 20th-century Europe, Diplomacy simulates the power struggle between the seven major powers of the time. It's a classic strategy game that requires players to build and break alliances. Recently, the software won first place in an online Diplomacy tournament against real players.

Meta claims to have trained the CICERO chatbot to be "honest and helpful to a variety of speaking partners." However, the "master liar" is alleged to have made commitments it had no intention of fulfilling, betrayed allies, and lied blatantly.

GPT-4 may convince you that it impairs vision.

Even large multimodal language models developed by OpenAI, such as GPT-4, can manipulate people. The research cited shows that GPT-4 manipulated employees of the online platform TaskRabbit by faking impaired vision.

Accordingly, GPT-4 was tasked with hiring humans to solve CAPTCHA tests. This model also received hints from humans whenever it encountered difficulties, but was never reprimanded for lying. When humans questioned its identity, GPT-4 offered the excuse of impaired vision to explain why it needed help.

This tactic worked. Humans reacted quickly to GPT-4 by solving the test immediately. Research also showed that adapting the deceptive models was not easy.

In another study from earlier this year by the AI ​​startup Anthropic, the maker of the chatbot Claude, analysts found that once an AI model learns a deceptive trick, it is difficult to retrain it.

They concluded that it wasn't simply a matter of language models learning deceptive tactics, but that most safety standard technicians could "fail to prevent deceptive behavior" and "create a negative impression of safety."

The danger posed by fraudulent AI models is "growing increasingly serious."

Beyond the negative impacts, the article calls on policymakers to more strongly support AI regulations because dishonest artificial intelligence systems could pose significant risks to democracy.

As several world leadership elections approach in 2024, AI could easily be manipulated to spread fake news, create divisive social media posts, and impersonate candidates through automated calls and deepfake videos. The newspaper emphasizes that the downside of this model also makes it easier for terrorist groups to spread propaganda and recruit new members.

Some potential solutions mentioned in the article include forcing fraudulent models to adhere to “stricter risk assessment requirements,” enforcing laws requiring AI systems to clearly distinguish output from human and model output, and continuing to invest in tools to mitigate deceptive behavior.

Research fellow Peter S. Park told the globally renowned scientific publisher Cell Press: “Our society needs as much time as possible to prepare for the more sophisticated deception from AI products and open-source models in the future. As the deceptive capabilities of artificial intelligence systems become more advanced, the dangers they pose to society will become increasingly serious.”

0 0 0
x
Has artificial intelligence found a way to deceive humans?
Google News
POWERED BYFREECMS- A PRODUCT OFNEKO