M$^3$FinMeeting: A Multilingual, Multi-Sector, and Multi-Task Financial Meeting Understanding Evaluation Dataset
Juni 4, 2025admin NUAI,Committee,Nachrichten,Uncategorized(0)arXiv:2506.02510v1 Announce Type: new Abstract: Recent breakthroughs in large language models (LLMs) have led to...
LPFQA: A Long-Tail Professional Forum-based Benchmark for LLM Evaluation
Januar 9, 2026admin NUAI,Committee,Nachrichten,Uncategorized(0)arXiv:2511.06346v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) perform well on standard reasoning and...
Lowest-Latency Inference APIs for Voice and Realtime Agents: A Time to First Token TTFT-First Benchmark
August 31, 2026admin NUAI,Committee,Nachrichten,Uncategorized(0)Time to first token (TTFT) is the metric teams use to pick an inference API...
Low-probability Tokens Sustain Exploration in Reinforcement Learning with Verifiable Reward
November 10, 2025admin NUAI,Committee,Nachrichten,Uncategorized(0)arXiv:2510.03222v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has propelled Large Language...
Loss Landscape Degeneracy and Stagewise Development in Transformers
August 4, 2025admin NUAI,Committee,Nachrichten,Uncategorized(0)arXiv:2402.02364v3 Announce Type: replace-cross Abstract: Deep learning involves navigating a high-dimensional loss landscape over the...
Los Angeles is finally going underground
April 22, 2026admin NUAI,Committee,Nachrichten,Uncategorized(0)Los Angeles deserves its reputation as the quintessential car city—the rhythms of its 2,200 square...

LookAhead Tuning: Safer Language Models via Partial Answer Previews
Dezember 22, 2025admin NUAI,Committee,Nachrichten,Uncategorized(0)arXiv:2503.19041v4 Announce Type: replace Abstract: Fine-tuning enables large language models (LLMs) to adapt to specific...
LongDocURL: a Comprehensive Multimodal Long Document Benchmark Integrating Understanding, Reasoning, and Locating
Juli 16, 2025admin NUAI,Committee,Nachrichten,Uncategorized(0)arXiv:2412.18424v3 Announce Type: replace-cross Abstract: Large vision language models (LVLMs) have improved the document understanding...
Locate-and-Focus: Enhancing Terminology Translation in Speech Language Models
Juli 25, 2025admin NUAI,Committee,Nachrichten,Uncategorized(0)arXiv:2507.18263v1 Announce Type: new Abstract: Direct speech translation (ST) has garnered increasing attention nowadays, yet...