中文
相关论文

相关论文: Evaluating Commercial AI Chatbots as News Intermed…

200 篇论文

Artificial Intelligence (AI) chatbots leveraging Large Language Models (LLMs) are gaining traction in healthcare for their potential to automate patient interactions and aid clinical decision-making. This study examines the reliability of…

人工智能 · 计算机科学 2024-05-24 Ayesha Siddika Nipu , K M Sajjadul Islam , Praveen Madiraju

The purpose of this study is to assess how large language models (LLMs) can be used for fact-checking and contribute to the broader debate on the use of automated means for veracity identification. To achieve this purpose, we use AI…

AI chatbots are increasingly used by students as study tools in physics, raising practical questions about their reliability on conceptual tasks. Existing evaluations of large language models (LLMs) on physics concept inventories rely…

物理教育 · 物理学 2026-05-12 Eugenio Tufino , Caterina Giovanzana , Andrea Zamboni , Pasquale Onorato , Stefano Oss

This article presents a comparative analysis of the ability of two large language model (LLM)-based chatbots, ChatGPT and Bing Chat, recently rebranded to Microsoft Copilot, to detect veracity of political information. We use AI auditing…

In the digital age, the prevalence of misleading news headlines poses a significant challenge to information integrity, necessitating robust detection mechanisms. This study explores the efficacy of Large Language Models (LLMs) in…

计算与语言 · 计算机科学 2024-05-07 Md Main Uddin Rony , Md Mahfuzul Haque , Mohammad Ali , Ahmed Shatil Alam , Naeemul Hassan

Interactive chat systems that build on artificial intelligence frameworks are increasingly ubiquitous and embedded into search engines, Web browsers, and operating systems, or are available on websites and apps. Researcher efforts have…

计算机与社会 · 计算机科学 2026-05-26 Katherine M. FitzGerald , Michelle Riedlinger , Axel Bruns , Stephen Harrington , Timothy Graham , Daniel Angus

This study examines the impact of AI on human false memories -- recollections of events that did not occur or deviate from actual occurrences. It explores false memory induction through suggestive questioning in Human-AI interactions,…

计算与语言 · 计算机科学 2024-08-12 Samantha Chan , Pat Pataranutaporn , Aditya Suri , Wazeer Zulfikar , Pattie Maes , Elizabeth F. Loftus

Conversational systems or chatbots are an example of AI-Infused Applications (AIIA). Chatbots are especially important as they are often the first interaction of clients with a business and are the entry point of a business into the AI…

Generative artificial intelligence (AI) holds enormous potential to revolutionize decision-making processes, from everyday to high-stake scenarios. By leveraging generative AI, humans can benefit from data-driven insights and predictions,…

综合经济学 · 经济学 2024-02-19 Valerio Capraro , Roberto Di Paolo , Veronica Pizziol

The ability to discern between true and false information is essential to making sound decisions. However, with the recent increase in AI-based disinformation campaigns, it has become critical to understand the influence of deceptive…

计算机与社会 · 计算机科学 2022-10-18 Valdemar Danry , Pat Pataranutaporn , Ziv Epstein , Matthew Groh , Pattie Maes

Although large conversational AI models such as OpenAI's ChatGPT have demonstrated great potential, we question whether such models can guarantee factual accuracy. Recently, technology companies such as Microsoft and Google have announced…

计算与语言 · 计算机科学 2023-04-24 Ruochen Zhao , Xingxuan Li , Yew Ken Chia , Bosheng Ding , Lidong Bing

We perform a mixed-method frame semantics-based analysis on a dataset of more than 49,000 sentences collected from 5846 news articles that mention AI. The dataset covers the twelve-month period centred around the launch of OpenAI's chatbot…

计算与语言 · 计算机科学 2024-11-26 Igor Ryazanov , Carl Öhman , Johanna Björklund

This study analyzes the performance of eight generative artificial intelligence chatbots -- ChatGPT, Claude, Copilot, DeepSeek, Gemini, Grok, Le Chat, and Perplexity -- in their free versions, in the task of generating academic…

信息检索 · 计算机科学 2026-05-20 Álvaro Cabezas-Clavijo , Pavel Sidorenko-Bautista

A good open-domain chatbot should avoid presenting contradictory responses about facts or opinions in a conversational session, known as its consistency capacity. However, evaluating the consistency capacity of a chatbot is still…

计算与语言 · 计算机科学 2021-06-07 Zekang Li , Jinchao Zhang , Zhengcong Fei , Yang Feng , Jie Zhou

Large Language Models (LLMs) have made significant progress in recent years, achieving remarkable results in question-answering tasks (QA). However, they still face two major challenges: hallucination and outdated information after the…

Automated benchmarks dominate the evaluation of large language models, yet no systematic study has compared user satisfaction, adoption motivations, and frustrations across competing platforms using a consistent instrument. We address this…

人机交互 · 计算机科学 2026-03-27 Moiz Sadiq Awan , Muhammad Haris Noor , Muhammad Salman Munaf

AI chatbots are becoming a primary interface for seeking information. As their popularity grows, chatbot providers are starting to deploy advertising and analytics. Despite this, tracking on AI chatbots has not been systematically studied.…

密码学与安全 · 计算机科学 2026-05-14 Muhammad Jazlan , Ethan Wang , Yash Vekaria , Zubair Shafiq

Generative AI paraphrased text can be used for copyright infringement and the AI paraphrased content can deprive substantial revenue from original content creators. Despite this recent surge of malicious use of generative AI, there are few…

This study aimed to evaluate the proficiency of prominent Large Language Models (LLMs), namely OpenAI's ChatGPT 3.5 and 4.0, Google's Bard(LaMDA), and Microsoft's Bing AI in discerning the truthfulness of news items using black box testing.…

计算与语言 · 计算机科学 2023-07-03 Kevin Matthe Caramancion

This article explores the phenomenon of confirmation bias in generative AI chatbots, a relatively underexamined aspect of AI-human interaction. Drawing on cognitive psychology and computational linguistics, it examines how confirmation…

人机交互 · 计算机科学 2025-04-15 Yiran Du
‹ 上一页 1 2 3 10 下一页 ›