Nachrichten
Nachrichten
Deploying a 1-Bit Bonsai-27B Model with PrismML llama.cpp and OpenAI-Compatible Local Inference Workflows
In this tutorial, we deploy the 1-bit Bonsai-27B language model using the PrismML fork of...
Demystifying ChatGPT: How It Masters Genre Recognition
arXiv:2507.03875v1 Announce Type: new Abstract: The introduction of ChatGPT has garnered significant attention within the...
Demographic Biases and Gaps in the Perception of Sexism in Large Language Models
arXiv:2508.18245v1 Announce Type: new Abstract: The use of Large Language Models (LLMs) has proven to...
DeMeVa at LeWiDi-2025: Modeling Perspectives with In-Context Learning and Label Distribution Learning
arXiv:2509.09524v1 Announce Type: new Abstract: This system paper presents the DeMeVa team’s approaches to the...
Delving into Multilingual Ethical Bias: The MSQAD with Statistical Hypothesis Tests for Large Language Models
arXiv:2505.19121v2 Announce Type: replace Abstract: Despite the recent strides in large language models, studies have...
DelvePO: Direction-Guided Self-Evolving Framework for Flexible Prompt Optimization
arXiv:2510.18257v1 Announce Type: new Abstract: Prompt Optimization has emerged as a crucial approach due to...
DefenderBench: A Toolkit for Evaluating Language Agents in Cybersecurity Environments
arXiv:2506.00739v3 Announce Type: replace Abstract: Large language model (LLM) agents have shown impressive capabilities in...
DEER: Disentangled Mixture of Experts with Instance-Adaptive Routing for Generalizable Machine-Generated Text Detection
arXiv:2511.01192v1 Announce Type: new Abstract: Detecting machine-generated text (MGT) has emerged as a critical challenge...
DEER: A Benchmark for Evaluating Deep Research Agents on Expert Report Generation
arXiv:2512.17776v2 Announce Type: replace Abstract: As large language models advance, deep research systems capable of...
DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains
DeepSeek published DeepSeek-V4-Flash-0731 on Hugging Face and moved the official V4-Flash API into public beta...
DeepSeek Researchers Apply a 1967 Matrix Normalization Algorithm to Fix Instability in Hyper Connections
DeepSeek researchers are trying to solve a precise issue in large language model training. Residual...
DeepSeek Releases DSpark, a Speculative Decoding Framework That Accelerates DeepSeek-V4 Per-User Generation 60–85% Over MTP-1
DeepSeek released DSpark, a speculative decoding framework, with open-source checkpoints and training code. It is...
