DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains
8月 1, 2026admin NUAI,Committee,ニュース,Uncategorized(0)DeepSeek published DeepSeek-V4-Flash-0731 on Hugging Face and moved the official V4-Flash API into public beta...
DeepSeek Researchers Apply a 1967 Matrix Normalization Algorithm to Fix Instability in Hyper Connections
1月 4, 2026admin NUAI,Committee,ニュース,Uncategorized(0)DeepSeek researchers are trying to solve a precise issue in large language model training. Residual...

DeepSeek Releases DSpark, a Speculative Decoding Framework That Accelerates DeepSeek-V4 Per-User Generation 60–85% Over MTP-1
6月 27, 2026admin NUAI,Committee,ニュース,Uncategorized(0)DeepSeek released DSpark, a speculative decoding framework, with open-source checkpoints and training code. It is...
DeepSeek R1-0528 arrives in powerful open source challenge to OpenAI o3 and Google Gemini 2.5 Pro
5月 30, 2025admin NUAI,Committee,ニュース,Uncategorized(0)Additionally, the model’s hallucination rate has been reduced, contributing to more reliable and consistent output.Read...

DeepSeek AI Researchers Introduce Engram: A Conditional Memory Axis For Sparse LLMs
1月 15, 2026admin NUAI,Committee,ニュース,Uncategorized(0)Transformers use attention and Mixture-of-Experts to scale computation, but they still lack a native way...

DeepSeek AI Releases DeepSeek-OCR 2 with Causal Visual Flow Encoder for Layout Aware Document Understanding
1月 30, 2026admin NUAI,Committee,ニュース,Uncategorized(0)DeepSeek AI released DeepSeek-OCR 2, an open source document OCR and understanding system that restructures...

DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
6月 16, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2506.11763v1 Announce Type: new Abstract: Deep Research Agents are a prominent category of LLM-based agents...
DeepReinforce Team Introduces CUDA-L1: An Automated Reinforcement Learning (RL) Framework for CUDA Optimization Unlocking 3x More Power from GPUs
8月 3, 2025admin NUAI,Committee,ニュース,Uncategorized(0)Estimated reading time: 6 minutes Table of contents The Breakthrough: Contrastive Reinforcement Learning (Contrastive-RL) How...

DeepReinforce Releases Ornith-1.0: An Open-Source Coding Model Family That Learns Its Own RL Scaffolds
6月 25, 2026admin NUAI,Committee,ニュース,Uncategorized(0)DeepReinforce has released Ornith-1.0, an open-source model family built for agentic coding. The lineup spans...