LLMs Struggle with Real Conversations: Microsoft and Salesforce Researchers Reveal a 39% Performance Drop in Multi-Turn Underspecified Tasks
5月 18, 2025admin NUAI,Committee,ニュース,Uncategorized(0)Conversational artificial intelligence is centered on enabling large language models (LLMs) to engage in dynamic...

LLMs for Customized Marketing Content Generation and Evaluation at Scale
6月 24, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2506.17863v1 Announce Type: new Abstract: Offsite marketing is essential in e-commerce, enabling businesses to reach...
LLMs Can Also Do Well! Breaking Barriers in Semantic Role Labeling via Large Language Models
6月 9, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2506.05385v1 Announce Type: new Abstract: Semantic role labeling (SRL) is a crucial task of natural...
LLMs are stuck in a groupthink groove. This startup is trying to get them out.
7月 1, 2026admin NUAI,Committee,ニュース,Uncategorized(0)Let’s start with a game. Open up your chatbot of choice—Claude, ChatGPT, Gemini—and type “Give...
LLM-Pruning Collection: A JAX Based Repo For Structured And Unstructured LLM Compression
1月 5, 2026admin NUAI,Committee,ニュース,Uncategorized(0)Zlab Princeton researchers have released LLM-Pruning Collection, a JAX based repository that consolidates major pruning...
LLM-Independent Adaptive RAG: Let the Question Speak for Itself
5月 8, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2505.04253v1 Announce Type: new Abstract: Large Language Models~(LLMs) are prone to hallucinations, and Retrieval-Augmented Generation...
LLM-Based Instance-Driven Heuristic Bias In the Context of a Biased Random Key Genetic Algorithm
9月 15, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2509.09707v1 Announce Type: cross Abstract: Integrating Large Language Models (LLMs) within metaheuristics opens a novel...
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
6月 27, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2505.00753v4 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing...
LLM-as-a-Judge: Where Do Its Signals Break, When Do They Hold, and What Should “Evaluation” Mean?
9月 21, 2025admin NUAI,Committee,ニュース,Uncategorized(0)What exactly is being measured when a judge LLM assigns a 1–5 (or pairwise) score?...