SOTAVerified

Chatbot

Chatbot or conversational AI is a language model designed and implemented to have conversations with humans.

Source: Open Data Chatbot

Image source

Papers

Showing 351–400 of 971 papers

TitleStatusHype
ChatGPT Alternative Solutions: Large Language Models Survey—0
Facilitating Pornographic Text Detection for Open-Domain Dialogue Systems via Knowledge Distillation of Large Language ModelsCode0
Characteristic AI Agents via Large Language ModelsCode1
Empowering Air Travelers: A Chatbot for Canadian Air Passenger Rights—0
Interpretable User Satisfaction Estimation for Conversational Systems with Large Language Models—0
ClimateQ&A: Bridging the gap between climate scientists and the general public—0
Tur[k]ingBench: A Challenge Benchmark for Web Agents—0
Regulating Chatbot Output via Inter-Informational Competition—0
Large language model-powered chatbots for internationalizing student support in higher education—0
AI on AI: Exploring the Utility of GPT as an Expert Annotator of AI Publications—0
CuentosIE: can a chatbot about "tales with a message" help to teach emotional intelligence?—0
DeepSeek-VL: Towards Real-World Vision-Language UnderstandingCode7
Alto: Orchestrating Distributed Compound AI Systems with Nested Ancestry—0
Yi: Open Foundation Models by 01.AICode9
Chatbot Arena: An Open Platform for Evaluating LLMs by Human PreferenceCode14
Explaining Genetic Programming Trees using Large Language Models—0
AI Insights: A Case Study on Utilizing ChatGPT Intelligence for Research Paper Analysis—0
Breeze-7B Technical Report—0
A General and Flexible Multi-concept Parsing Framework for Multilingual Semantic Matching—0
Large Language Models in Fire Engineering: An Examination of Technical Questions Against Domain Knowledge—0
The Heterogeneous Productivity Effects of Generative AI—0
The Ink Splotch Effect: A Case Study on ChatGPT as a Co-Creative Game Designer—0
Ever-Evolving Memory by Blending and Refining the Past—0
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition—0
Impact of Decentralized Learning on Player Utilities in Stackelberg Games—0
MedAide: Leveraging Large Language Models for On-Premise Medical Assistance on Edge Devices—0
Do Large Language Models Mirror Cognitive Language Processing?—0
Making Them Ask and Answer: Jailbreaking Large Language Models in Few Queries via Disguise and ReconstructionCode2
KoDialogBench: Evaluating Conversational Understanding of Language Models with Korean Dialogue BenchmarkCode1
Prediction-Powered Ranking of Large Language ModelsCode0
A Piece of Theatre: Investigating How Teachers Design LLM Chatbots to Assist Adolescent Cyberbullying Education—0
A Fine-tuning Enhanced RAG System with Quantized Influence Measure as AI Judge—0
Long Dialog Summarization: An Analysis—0
ASEM: Enhancing Empathy in Chatbot through Attention-based Sentiment and Emotion ModelingCode0
HypoTermQA: Hypothetical Terms Dataset for Benchmarking Hallucination Tendency of LLMsCode0
Citation-Enhanced Generation for LLM-based ChatbotsCode1
A Study on the Vulnerability of Test Questions against ChatGPT-based Cheating—0
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation—0
SPML: A DSL for Defending Language Models Against Prompt Attacks—0
Compress to Impress: Unleashing the Potential of Compressive Memory in Real-World Long-Term ConversationsCode1
Understanding the Impact of Long-Term Memory on Self-Disclosure with Large Language Model-Driven Chatbots for Public Health Intervention—0
AbuseGPT: Abuse of Generative AI ChatBots to Create Smishing Campaigns—0
Fine-tuning Large Language Model (LLM) Artificial Intelligence Chatbots in Ophthalmology and LLM-based evaluation using GPT-4—0
SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware DecodingCode2
Measuring and Controlling Instruction (In)Stability in Language Model DialogsCode1
Making a prototype of Seoul historical sites chatbot using Langchain—0
LLaVA-Docent: Instruction Tuning with Multimodal Large Language Model to Support Art Appreciation Education—0
CataractBot: An LLM-Powered Expert-in-the-Loop Chatbot for Cataract PatientsCode1
Hydragen: High-Throughput LLM Inference with Shared PrefixesCode1
DFA-RAG: Conversational Semantic Router for Large Language Model with Definite Finite Automaton—0
Show:102550
← PrevPage 8 of 20Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Yi 34B ChatAverage win rate27.2—Unverified