YouZum

Noticias

Noticias

One-Token Rollout: Guiding Supervised Fine-Tuning of LLMs with Policy Gradient

arXiv:2509.26313v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is the predominant method for adapting large...

One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models

arXiv:2505.07167v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been extensively used across diverse...

One Token Is Enough: Improving Diffusion Language Models with a Sink Token

arXiv:2601.19657v3 Announce Type: replace Abstract: Diffusion Language Models (DLMs) have emerged as a compelling alternative...

One Joke to Rule them All? On the (Im)possibility of Generalizing Humor

arXiv:2508.19402v1 Announce Type: new Abstract: Humor is a broad and complex form of communication that...

On the Reliability of Large Language Models for Causal Discovery

arXiv:2407.19638v2 Announce Type: replace Abstract: This study investigates the efficacy of Large Language Models (LLMs)...

On the generalization of language models from in-context learning and finetuning: a controlled study

arXiv:2505.00661v2 Announce Type: replace Abstract: Large language models exhibit exciting capabilities, yet can show surprisingly...

On the Effectiveness of Membership Inference in Targeted Data Extraction from Large Language Models

arXiv:2512.13352v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are prone to memorizing training data...

OmniGAIA: Towards Native Omni-Modal AI Agents

arXiv:2602.22897v1 Announce Type: cross Abstract: Human intelligence naturally intertwines omni-modal perception — spanning vision, audio...

OMIND: Framework for Knowledge Grounded Finetuning and Multi-Turn Dialogue Benchmark for Mental Health LLMs

arXiv:2603.25105v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown remarkable capabilities for complex...

Olica: Efficient Structured Pruning of Large Language Models without Retraining

arXiv:2506.08436v1 Announce Type: new Abstract: Most existing structured pruning methods for Large Language Models (LLMs)...

OCRmyPDF Tutorial: Convert Scanned Documents into Searchable PDF/A Files with Sidecar Text Extraction and Batch Processing

In this tutorial, we build an advanced, self-contained OCRmyPDF workflow. We start by installing the...

OceanBase Releases seekdb: An Open Source AI Native Hybrid Search Database for Multi-model RAG and AI Agents

AI applications rarely deal with one clean table. They mix user profiles, chat logs, JSON...

We use cookies to improve your experience and performance on our website. You can learn more at Política de privacidad and manage your privacy settings by clicking Settings.

Privacy Preferences

You can choose your cookie settings by turning on/off each type of cookie as you wish, except for essential cookies.

Allow All
Manage Consent Preferences
  • Always Active

Save
es_ES