YouZum

Nachrichten

Nachrichten

Deploying a 1-Bit Bonsai-27B Model with PrismML llama.cpp and OpenAI-Compatible Local Inference Workflows

In this tutorial, we deploy the 1-bit Bonsai-27B language model using the PrismML fork of...

Demystifying ChatGPT: How It Masters Genre Recognition

arXiv:2507.03875v1 Announce Type: new Abstract: The introduction of ChatGPT has garnered significant attention within the...

Demographic Biases and Gaps in the Perception of Sexism in Large Language Models

arXiv:2508.18245v1 Announce Type: new Abstract: The use of Large Language Models (LLMs) has proven to...

DeMeVa at LeWiDi-2025: Modeling Perspectives with In-Context Learning and Label Distribution Learning

arXiv:2509.09524v1 Announce Type: new Abstract: This system paper presents the DeMeVa team’s approaches to the...

Delving into Multilingual Ethical Bias: The MSQAD with Statistical Hypothesis Tests for Large Language Models

arXiv:2505.19121v2 Announce Type: replace Abstract: Despite the recent strides in large language models, studies have...

DelvePO: Direction-Guided Self-Evolving Framework for Flexible Prompt Optimization

arXiv:2510.18257v1 Announce Type: new Abstract: Prompt Optimization has emerged as a crucial approach due to...

DefenderBench: A Toolkit for Evaluating Language Agents in Cybersecurity Environments

arXiv:2506.00739v3 Announce Type: replace Abstract: Large language model (LLM) agents have shown impressive capabilities in...

DEER: Disentangled Mixture of Experts with Instance-Adaptive Routing for Generalizable Machine-Generated Text Detection

arXiv:2511.01192v1 Announce Type: new Abstract: Detecting machine-generated text (MGT) has emerged as a critical challenge...

DEER: A Benchmark for Evaluating Deep Research Agents on Expert Report Generation

arXiv:2512.17776v2 Announce Type: replace Abstract: As large language models advance, deep research systems capable of...

DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains

DeepSeek published DeepSeek-V4-Flash-0731 on Hugging Face and moved the official V4-Flash API into public beta...

DeepSeek Researchers Apply a 1967 Matrix Normalization Algorithm to Fix Instability in Hyper Connections

DeepSeek researchers are trying to solve a precise issue in large language model training. Residual...

DeepSeek Releases DSpark, a Speculative Decoding Framework That Accelerates DeepSeek-V4 Per-User Generation 60–85% Over MTP-1

DeepSeek released DSpark, a speculative decoding framework, with open-source checkpoints and training code. It is...

We use cookies to improve your experience and performance on our website. You can learn more at Datenschutzrichtlinie and manage your privacy settings by clicking Settings.

Privacy Preferences

You can choose your cookie settings by turning on/off each type of cookie as you wish, except for essential cookies.

Allow All
Manage Consent Preferences
  • Always Active

Save
de_DE