From Individuals to Interactions: Benchmarking Gender Bias in Multimodal Large Language Models from the Lens of Social Relationship
juillet 1, 2025admin NUAI,Committee,Actualités,Uncategorized(0)arXiv:2506.23101v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities across...
From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs
janvier 8, 2026admin NUAI,Committee,Actualités,Uncategorized(0)arXiv:2601.03682v1 Announce Type: new Abstract: Recent studies reveal that large language models (LLMs) exhibit limited...
From hallucinations to hardware: Lessons from a real-world computer vision project gone sideways
juin 29, 2025admin NUAI,Committee,Actualités,Uncategorized(0)What we tried, what didn’t work and how a combination of approaches eventually helped us...

From Gemma 3 270M to FunctionGemma, How Google AI Built a Compact Function Calling Specialist for Edge Workloads
décembre 27, 2025admin NUAI,Committee,Actualités,Uncategorized(0)Google has released FunctionGemma, a specialized version of the Gemma 3 270M model that is...
From Essence to Defense: Adaptive Semantic-aware Watermarking for Embedding-as-a-Service Copyright Protection
décembre 19, 2025admin NUAI,Committee,Actualités,Uncategorized(0)arXiv:2512.16439v1 Announce Type: cross Abstract: Benefiting from the superior capabilities of large language models in...
From Entity Mentions to Tone: An LLM-Based Pipeline for Media Bias Analysis
août 20, 2026admin NUAI,Committee,Actualités,Uncategorized(0)arXiv:2608.17454v1 Announce Type: new Abstract: This paper presents a pipeline for analyzing media bias and...
From disruption to reinvention: How knowledge workers can thrive after AI
mai 27, 2025admin NUAI,Committee,Actualités,Uncategorized(0)We are beginning a cognitive migration: Away from what AI now does well, and toward...

From Correction to Mastery: Reinforced Distillation of Large Language Model Agents
octobre 10, 2025admin NUAI,Committee,Actualités,Uncategorized(0)arXiv:2509.14257v2 Announce Type: replace Abstract: Large Language Model agents excel at solving complex tasks through...
From Clicking to Reasoning: WebChoreArena Benchmark Challenges Agents with Memory-Heavy and Multi-Page Tasks
juin 6, 2025admin NUAI,Committee,Actualités,Uncategorized(0)Web automation agents have become a growing focus in artificial intelligence, particularly due to their...
