Notizie
Notizie
MapFormer: Self-Supervised Learning of Cognitive Maps with Input-Dependent Positional Embeddings
arXiv:2511.19279v3 Announce Type: replace-cross Abstract: A cognitive map is an internal model which encodes the...
MANZANO: A Simple and Scalable Unified Multimodal Model with a Hybrid Vision Tokenizer
arXiv:2509.16197v1 Announce Type: cross Abstract: Unified multimodal Large Language Models (LLMs) that can both understand...
Manus has kick-started an AI agent boom in China
Last year, China saw a boom in foundation models, the do-everything large language models that...
MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized Collaboration
arXiv:2506.19835v1 Announce Type: new Abstract: Recent advancements in medical Large Language Models (LLMs) have showcased...
MALLM: Multi-Agent Large Language Models Framework
arXiv:2509.11656v1 Announce Type: cross Abstract: Multi-agent debate (MAD) has demonstrated the ability to augment collective...
Making AI operational in constrained public sector environments
The AI boom has hit across industries, and public sector organizations are facing pressure to...
Make It Hard to Hear, Easy to Learn: Long-Form Bengali ASR and Speaker Diarization via Extreme Augmentation and Perfect Alignment
arXiv:2602.23070v1 Announce Type: cross Abstract: Although Automatic Speech Recognition (ASR) in Bengali has seen significant...
MacroBench: A Novel Testbed for Web Automation Scripts via Large Language Models
arXiv:2510.04363v2 Announce Type: replace-cross Abstract: We introduce MacroBench, a code-first benchmark that evaluates whether LLMs...
MAC: A Multi-Agent Framework for Interactive User Clarification in Multi-turn Conversations
arXiv:2512.13154v1 Announce Type: cross Abstract: Conversational agents often encounter ambiguous user requests, requiring an effective...
M$^3$FinMeeting: A Multilingual, Multi-Sector, and Multi-Task Financial Meeting Understanding Evaluation Dataset
arXiv:2506.02510v1 Announce Type: new Abstract: Recent breakthroughs in large language models (LLMs) have led to...
LPFQA: A Long-Tail Professional Forum-based Benchmark for LLM Evaluation
arXiv:2511.06346v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) perform well on standard reasoning and...
Low-probability Tokens Sustain Exploration in Reinforcement Learning with Verifiable Reward
arXiv:2510.03222v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has propelled Large Language...
