Is It Thinking or Cheating? Detecting Implicit Reward Hacking by Measuring Reasoning Effort
10月 8, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2510.01367v3 Announce Type: replace-cross Abstract: Reward hacking, where a reasoning model exploits loopholes in a...
Is In-Context Learning Learning?
9月 16, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2509.10414v2 Announce Type: replace Abstract: In-context learning (ICL) allows some autoregressive models to solve tasks...
Is fake grass a bad idea? The AstroTurf wars are far from over.
4月 9, 2026admin NUAI,Committee,ニュース,Uncategorized(0)A rare warm spell in January melted enough snow to uncover Cornell University’s newest athletic...

Investigating Gender Stereotypes in Large Language Models via Social Determinants of Health
3月 11, 2026admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2603.09416v1 Announce Type: new Abstract: Large Language Models (LLMs) excel in Natural Language Processing (NLP)...
Investigating Factuality in Long-Form Text Generation: The Roles of Self-Known and Self-Unknown
9月 26, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2411.15993v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated strong capabilities in text...
Introspective Growth: Automatically Advancing LLM Expertise in Technology Judgment
6月 10, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2505.12452v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly demonstrate signs of conceptual understanding...
Introduction to Small Language Models: The Complete Guide for 2026
2月 24, 2026admin NUAI,Committee,ニュース,Uncategorized(0) AI deployment is changing...

Introducing OmniGEC: A Silver Multilingual Dataset for Grammatical Error Correction
9月 19, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2509.14504v1 Announce Type: new Abstract: In this paper, we introduce OmniGEC, a collection of multilingual...
Interpreting the Latent Structure of Operator Precedence in Language Models
10月 17, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2510.13908v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated impressive reasoning capabilities but...