What Don’t You Understand? Using Large Language Models to Identify and Characterize Student Misconceptions About Challenging Topics
5月 4, 2026admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2605.00294v1 Announce Type: new Abstract: This study presents a systematic approach to identifying and characterizing...
What Does Neuro Mean to Cardio? Investigating the Role of Clinical Specialty Data in Medical LLMs
10月 16, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2505.10113v3 Announce Type: replace Abstract: In this paper, we introduce S-MedQA, an English medical question-answering...
What do Speech Foundation Models Learn? Analysis and Applications
8月 19, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2508.12255v1 Announce Type: new Abstract: Speech foundation models (SFMs) are designed to serve as general-purpose...
What could possibly go wrong if an enterprise replaces all its engineers with AI?
11月 9, 2025admin NUAI,Committee,ニュース,Uncategorized(0)AI coding, vibe coding and agentic swarm have made a dramatic and astonishing recent market...

What Are They Talking About? A Benchmark of Knowledge-Grounded Discussion Summarization
11月 7, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2505.12474v3 Announce Type: replace Abstract: Traditional dialogue summarization primarily focuses on dialogue content, assuming it...
What Anthropic’s latest AI discovery does—and doesn’t—show
7月 13, 2026admin NUAI,Committee,ニュース,Uncategorized(0)This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories...
Welcome to the dark side of crypto’s permissionless dream
2月 18, 2026admin NUAI,Committee,ニュース,Uncategorized(0)“We’re out of airspace now. We can do whatever we want,” Jean-Paul Thorbjornsen tells me...

WebThinker: Empowering Large Reasoning Models with Deep Research Capability
5月 1, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2504.21776v1 Announce Type: new Abstract: Large reasoning models (LRMs), such as OpenAI-o1 and DeepSeek-R1, demonstrate...
Weak-for-Strong (W4S): A Novel Reinforcement Learning Algorithm that Trains a weak Meta Agent to Design Agentic Workflows with Stronger LLMs
10月 19, 2025admin NUAI,Committee,ニュース,Uncategorized(0)Researchers from Stanford, EPFL, and UNC introduce Weak-for-Strong Harnessing, W4S, a new Reinforcement Learning RL...
