Stop guessing why your LLMs break: Anthropic’s new tool shows you exactly what goes wrong
June 5, 2025admin NUAI,Committee,News,Uncategorized(0)Anthropic’s open-source circuit tracing tool can help developers debug, optimize, and control AI for reliable...

StepFun Releases Step 3.7 Flash: A 198B MoE Vision-Language Model for Coding Agents and Search Workflows
May 30, 2026admin NUAI,Committee,News,Uncategorized(0)StepFun today released Step 3.7 Flash, a multimodal Mixture-of-Experts model targeting agentic use cases. It...
StepFun AI Releases Step-Audio-R1: A New Audio LLM that Finally Benefits from Test Time Compute Scaling
November 30, 2025admin NUAI,Committee,News,Uncategorized(0)Why do current audio AI models often perform worse when they generate longer reasoning instead...

StepFun AI Releases Step-Audio 2 Mini: An Open-Source 8B Speech-to-Speech AI Model that Surpasses GPT-4o-Audio
September 1, 2025admin NUAI,Committee,News,Uncategorized(0)The StepFun AI team has released Step-Audio 2 Mini, an 8B parameter speech-to-speech large audio...

STEPER: Step-wise Knowledge Distillation for Enhancing Reasoning Ability in Multi-Step Retrieval-Augmented Language Models
October 10, 2025admin NUAI,Committee,News,Uncategorized(0)arXiv:2510.07923v1 Announce Type: new Abstract: Answering complex real-world questions requires step-by-step retrieval and integration of...
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models
September 10, 2025admin NUAI,Committee,News,Uncategorized(0)arXiv:2507.15512v3 Announce Type: replace Abstract: Test-Time Scaling (TTS) is a promising approach to progressively elicit...
Step by Step Guide to Build an End-to-End Model Optimization Pipeline with NVIDIA Model Optimizer Using FastNAS Pruning and Fine-Tuning
April 3, 2026admin NUAI,Committee,News,Uncategorized(0)In this tutorial, we build a complete end-to-end pipeline using NVIDIA Model Optimizer to train...
Steering Language Models with Weight Arithmetic
November 10, 2025admin NUAI,Committee,News,Uncategorized(0)arXiv:2511.05408v1 Announce Type: new Abstract: Providing high-quality feedback to Large Language Models (LLMs) on a...
Static vs. Dynamic vs. Continuous Batching in LLM Inference
August 4, 2026admin NUAI,Committee,News,Uncategorized(0)In this article, you will learn how static, dynamic, and continuous batching work in LLM...
