LLMs Can Also Do Well! Breaking Barriers in Semantic Role Labeling via Large Language Models
6 月 9, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2506.05385v1 Announce Type: new Abstract: Semantic role labeling (SRL) is a crucial task of natural...
LLMs are stuck in a groupthink groove. This startup is trying to get them out.
7 月 1, 2026admin NUAI,Committee,新闻,Uncategorized(0)Let’s start with a game. Open up your chatbot of choice—Claude, ChatGPT, Gemini—and type “Give...
LLM-Pruning Collection: A JAX Based Repo For Structured And Unstructured LLM Compression
1 月 5, 2026admin NUAI,Committee,新闻,Uncategorized(0)Zlab Princeton researchers have released LLM-Pruning Collection, a JAX based repository that consolidates major pruning...
LLM-Independent Adaptive RAG: Let the Question Speak for Itself
5 月 8, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2505.04253v1 Announce Type: new Abstract: Large Language Models~(LLMs) are prone to hallucinations, and Retrieval-Augmented Generation...
LLM-Based Instance-Driven Heuristic Bias In the Context of a Biased Random Key Genetic Algorithm
9 月 15, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2509.09707v1 Announce Type: cross Abstract: Integrating Large Language Models (LLMs) within metaheuristics opens a novel...
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
6 月 27, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2505.00753v4 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing...
LLM-as-a-Judge: Where Do Its Signals Break, When Do They Hold, and What Should “Evaluation” Mean?
9 月 21, 2025admin NUAI,Committee,新闻,Uncategorized(0)What exactly is being measured when a judge LLM assigns a 1–5 (or pairwise) score?...
LLM-as-a-Judge: Can Language Models Be Trusted to Evaluate Other Models?
4 月 30, 2025admin NUAI,Committee,新闻,Uncategorized(0)Exploring the promise, pitfalls, and practical applications of using LLMs to automate AI evaluation — from synthetic...
LLM-as-a-Judge for Reference-less Automatic Code Validation and Refinement for Natural Language to Bash in IT Automation
6 月 16, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2506.11237v1 Announce Type: cross Abstract: In an effort to automatically evaluate and select the best...