HiMATE: A Hierarchical Multi-Agent Framework for Machine Translation Evaluation
9 月 16, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2505.16281v2 Announce Type: replace Abstract: The advancement of Large Language Models (LLMs) enables flexible and...
Highlighted at CVPR 2025: Google DeepMind’s ‘Motion Prompting’ Paper Unlocks Granular Video Control
6 月 14, 2025admin NUAI,Committee,新闻,Uncategorized(0)Key Takeaways: Researchers from Google DeepMind, the University of Michigan & Brown university have developed...

High-Entropy Token Selection in Reinforcement Learning with Verifiable Rewards (RLVR) Improves Accuracy and Reduces Training Cost for LLMs
6 月 9, 2025admin NUAI,Committee,新闻,Uncategorized(0)Large Language Models (LLMs) generate step-by-step responses known as Chain-of-Thoughts (CoTs), where each token contributes...

High Accuracy, Less Talk (HALT): Reliable LLMs through Capability-Aligned Finetuning
6 月 5, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2506.04051v1 Announce Type: new Abstract: Large Language Models (LLMs) currently respond to every prompt. However...
Hierarchical Memory for High-Efficiency Long-Term Reasoning in LLM Agents
8 月 1, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2507.22925v1 Announce Type: new Abstract: Long-term memory is one of the key factors influencing the...
Hierarchical Chain-of-Thought Prompting: Enhancing LLM Reasoning Performance and Efficiency
4 月 3, 2026admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2604.00130v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has significantly improved the reasoning capabilities of...
HFS: Holistic Query-Aware Frame Selection for Efficient Video Reasoning
12 月 15, 2025admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2512.11534v1 Announce Type: cross Abstract: Key frame selection in video understanding presents significant challenges. Traditional...
HeuriGym: An Agentic Benchmark for LLM-Crafted Heuristics in Combinatorial Optimization
1 月 29, 2026admin NUAI,Committee,新闻,Uncategorized(0)arXiv:2506.07972v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have demonstrated significant advancements in...
Hermes Agent Ships Tool Search for MCP: Anthropic Evals Show 49% to 74% Accuracy Gain on Opus 4
5 月 30, 2026admin NUAI,Committee,新闻,Uncategorized(0)Nous Research’s open-source Hermes Agent now ships a Tool Search feature. It directly addresses a...
