中文
相关论文

相关论文: Prakriti200: A Questionnaire-Based Dataset of 200 …

200 篇论文

The identification of Prakriti types for the human body is a long-lost medical practice in finding the harmony between the nature of human beings and their behaviour. There are 3 fundamental Prakriti types of individuals. A person can…

机器学习 · 计算机科学 2023-10-05 Pranav Bidve , Shalini Mishra , Annapurna J

Automatically detecting personality traits can aid several applications, such as mental health recognition and human resource management. Most datasets introduced for personality detection so far have analyzed these traits for each…

人机交互 · 计算机科学 2020-09-01 Shahid Nawaz Khan , Maitree Leekha , Jainendra Shukla , Rajiv Ratn Shah

Sanskrit is a classical language with about 30 million extant manuscripts fit for digitisation, available in written, printed or scannedimage forms. However, it is still considered to be a low-resource language when it comes to available…

计算与语言 · 计算机科学 2022-11-16 Ayush Maheshwari , Nikhil Singh , Amrith Krishna , Ganesh Ramakrishnan

Language Models (LMs) are indispensable tools shaping modern workflows, but their global effectiveness depends on understanding local socio-cultural contexts. To address this, we introduce SANSKRITI, a benchmark designed to evaluate…

计算与语言 · 计算机科学 2025-10-29 Arijit Maji , Raghvendra Kumar , Akash Ghosh , Anushka , Sriparna Saha

The study presents a comprehensive benchmark for retrieving Sanskrit documents using English queries, focusing on the chapters of the Srimadbhagavatam. It employs a tripartite approach: Direct Retrieval (DR), Translation-based Retrieval…

计算与语言 · 计算机科学 2025-05-27 Manoj Balaji Jagadeeshan , Prince Raj , Pawan Goyal

Improving human health and well-being requires an accurate and effective understanding of an individual's physical and mental state throughout daily life. To support this goal, we utilized smartphones, smartwatches, and sleep sensors to…

信号处理 · 电气工程与系统科学 2025-08-07 Se Won Oh , Hyuntae Jeong , Seungeun Chung , Jeong Mook Lim , Kyoung Ju Noh , Sunkyung Lee , Gyuwon Jung

India's linguistic landscape, spanning 22 scheduled languages and hundreds of marginalized dialects, has driven rapid growth in NLP datasets, benchmarks, and pretrained models. However, no dedicated survey consolidates resources developed…

计算与语言 · 计算机科学 2026-04-21 Raghvendra Kumar , Devankar Raj , Sriparna Saha

Great research interests have been attracted to devise AI services that are able to provide mental health support. However, the lack of corpora is a main obstacle to this research, particularly in Chinese language. In this paper, we propose…

计算与语言 · 计算机科学 2021-06-04 Hao Sun , Zhenru Lin , Chujie Zheng , Siyang Liu , Minlie Huang

The goal of the present paper is to develop and validate a questionnaire to assess AI literacy. In particular, the questionnaire should be deeply grounded in the existing literature on AI literacy, should be modular (i.e., including…

人工智能 · 计算机科学 2023-02-21 Astrid Carolus , Martin Koch , Samantha Straka , Marc Erich Latoschik , Carolin Wienrich

The recent advances in deep-learning have led to the development of highly sophisticated systems with an unquenchable appetite for data. On the other hand, building good deep-learning models for low-resource languages remains a challenging…

计算与语言 · 计算机科学 2024-02-20 Maithili Sabane , Onkar Litake , Aman Chadha

Knowledge bases (KB) are an important resource in a number of natural language processing (NLP) and information retrieval (IR) tasks, such as semantic search, automated question-answering etc. They are also useful for researchers trying to…

信息检索 · 计算机科学 2023-10-13 Hrishikesh Terdalkar , Arnab Bhattacharya , Madhulika Dubey , Ramamurthy S , Bhavna Naneria Singh

Effective driving style analysis is critical to developing human-centered intelligent driving systems that consider drivers' preferences. However, the approaches and conclusions of most related studies are diverse and inconsistent because…

机器人学 · 计算机科学 2024-06-13 Chaopeng Zhang , Wenshuo Wang , Zhaokun Chen , Junqiang Xi

A distinct feature of Hindu religious and philosophical text is that they come from a library of texts rather than single source. The Upanishads is known as one of the oldest philosophical texts in the world that forms the foundation of…

计算与语言 · 计算机科学 2022-10-12 Rohitash Chandra , Mukul Ranjan

This study aims to develop a semi-automatically labelled prosody database for Hindi, for enhancing the intonation component in ASR and TTS systems, which is also helpful for building Speech to Speech Machine Translation systems. Although no…

计算与语言 · 计算机科学 2021-12-14 Esha Banerjee , Atul Kr. Ojha , Girish Nath Jha

The effectiveness of Large Language Models (LLMs) depends heavily on the availability of high-quality post-training data, particularly instruction-tuning and preference-based examples. Existing open-source datasets, however, often lack…

We introduce a dataset for classifying wellness dimensions in social media user posts, covering six key aspects: physical, emotional, social, intellectual, spiritual, and vocational. The dataset is designed to capture these dimensions in…

机器学习 · 计算机科学 2025-07-18 Heba Shakeel , Tanvir Ahmad , Chandni Saxena

This study demonstrates how hybrid neural-symbolic methods can yield significant new insights into the evolution of a morphologically rich, low-resource language. We challenge the naive assumption that linguistic change is simplification by…

计算与语言 · 计算机科学 2025-12-08 Ananth Hariharan , David Mortensen

Developing benchmark datasets for low-resource languages poses significant challenges, primarily due to the limited availability of native linguistic experts and the substantial time and cost involved in annotation. Given these challenges,…

计算与语言 · 计算机科学 2025-10-28 Rahul Ranjan , Mahendra Kumar Gurve , Anuj , Nitin , Yamuna Prasad

Sanskrit, an ancient language with a rich linguistic heritage, presents unique challenges for automatic speech recognition (ASR) due to its phonemic complexity and the phonetic transformations that occur at word junctures, similar to the…

计算与语言 · 计算机科学 2025-06-03 Sujeet Kumar , Pretam Ray , Abhinay Beerukuri , Shrey Kamoji , Manoj Balaji Jagadeeshan , Pawan Goyal

We introduce MentalChat16K, an English benchmark dataset combining a synthetic mental health counseling dataset and a dataset of anonymized transcripts from interventions between Behavioral Health Coaches and Caregivers of patients in…

‹ 上一页 1 2 3 10 下一页 ›