LLMs Struggle with Real Conversations: Microsoft and Salesforce Researchers Reveal a 39% Performance Drop in Multi-Turn Underspecified Tasks
5 月 18, 2025admin NUAI,Committee,新闻,Uncategorized(0)Conversational artificial intelligence is centered on enabling large language models (LLMs) to engage in dynamic...

LLMs for Customized Marketing Content Generation and Evaluation at Scale
6 月 24, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2506.17863v1 Announce Type: new Abstract: Offsite marketing is essential in e-commerce, enabling businesses to reach...
LLMs Can Also Do Well! Breaking Barriers in Semantic Role Labeling via Large Language Models
6 月 9, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2506.05385v1 Announce Type: new Abstract: Semantic role labeling (SRL) is a crucial task of natural...
LLMs are stuck in a groupthink groove. This startup is trying to get them out.
7 月 1, 2026admin NUAI,Committee,新闻,Uncategorized(0)Let’s start with a game. Open up your chatbot of choice—Claude, ChatGPT, Gemini—and type “Give...
LLM-Pruning Collection: A JAX Based Repo For Structured And Unstructured LLM Compression
1 月 5, 2026admin NUAI,Committee,新闻,Uncategorized(0)Zlab Princeton researchers have released LLM-Pruning Collection, a JAX based repository that consolidates major pruning...
LLM-Independent Adaptive RAG: Let the Question Speak for Itself
5 月 8, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2505.04253v1 Announce Type: new Abstract: Large Language Models~(LLMs) are prone to hallucinations, and Retrieval-Augmented Generation...
LLM-Based Instance-Driven Heuristic Bias In the Context of a Biased Random Key Genetic Algorithm
9 月 15, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2509.09707v1 Announce Type: cross Abstract: Integrating Large Language Models (LLMs) within metaheuristics opens a novel...
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
6 月 27, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2505.00753v4 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing...
LLM-as-a-Judge: Where Do Its Signals Break, When Do They Hold, and What Should “Evaluation” Mean?
9 月 21, 2025admin NUAI,Committee,新闻,Uncategorized(0)What exactly is being measured when a judge LLM assigns a 1–5 (or pairwise) score?...