Notizie
Notizie
Moonshot AI Releases Kimi K2.7-Code: a Coding Model Reporting +21.8% on Kimi Code Bench v2 Over K2.6
This week, Moonshot AI released Kimi K2.7-Code. It is a coding-focused, agentic model. The model...
Moonshot AI Releases Kimi K2: A Trillion-Parameter MoE Model Focused on Long Context, Code, Reasoning, and Agentic Behavior
Kimi K2, launched by Moonshot AI in July 2025, is a purpose-built, open-source Mixture-of-Experts (MoE)...
Moonshot AI Releases Kimi Code CLI: A Terminal AI Coding Agent Built in TypeScript for Next-Gen Agents
Moonshot AI has released Kimi Code CLI, an open-source coding agent that runs in the...
Moonshot AI Releases 𝑨𝒕𝒕𝒆𝒏𝒕𝒊𝒐𝒏 𝑹𝒆𝒔𝒊𝒅𝒖𝒂𝒍𝒔 to Replace Fixed Residual Mixing with Depth-Wise Attention for Better Scaling in Transformers
Residual connections are one of the least questioned parts of modern Transformer design. In PreNorm...
Moonshot AI Launches Kimi Work, a Local Desktop Agent Reportedly Running on Kimi K2.6 With a 300-Sub-Agent Agent Swarm
Moonshot AI has introduced Kimi Work, an AI agent that runs on your own desktop...
MoonMath AI Open-Sources a HIP Attention Kernel for AMD MI300X That Beats AITER v3 on Every Shape and Rounding Mode
MoonMath AI team has released a bf16 forward attention kernel for AMD’s MI300X GPU. It...
Montana’s plan to become an experimental medical hub just pushed forward
As of this week in Montana, any biotech company with an experimental drug has a...
Monadic Context Engineering
arXiv:2512.22431v2 Announce Type: replace-cross Abstract: The proliferation of Large Language Models (LLMs) has catalyzed a...
Moltbook was peak AI theater
For a few days this week the hottest new hangout on the internet was a...
MolLangBench: A Comprehensive Benchmark for Language-Prompted Molecular Structure Recognition, Editing, and Generation
arXiv:2505.15054v3 Announce Type: replace Abstract: Precise recognition, editing, and generation of molecules are essential prerequisites...
MoE Architecture Comparison: Qwen3 30B-A3B vs. GPT-OSS 20B
This article provides a technical comparison between two recently released Mixture-of-Experts (MoE) transformer models: Alibaba’s...
Modeling and Predicting Multi-Turn Answer Instability in Large Language Models
arXiv:2511.10688v1 Announce Type: new Abstract: As large language models (LLMs) are adopted in an increasingly...

