Nachrichten
Nachrichten
Five things you need to know about AI
At SXSW London last week I gave a talk called “Five things you need to...
FIT: Defying Catastrophic Forgetting in Continual LLM Unlearning
arXiv:2601.21682v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate impressive capabilities across diverse tasks...
Fish Audio Releases Fish Audio S2: A New Generation of Expressive Text-to-Speech (TTS) with Absurdly Controllable Emotion
The landscape of Text-to-Speech (TTS) is moving away from modular pipelines toward integrated Large Audio...
FireRedTeam Releases FireRed-OCR-2B Utilizing GRPO to Solve Structural Hallucinations in Tables and LaTeX for Software Developers
Document digitization has long been a multi-stage problem: first detect the layout, then extract the...
FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design
arXiv:2506.13066v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) demonstrate significant cross-modal reasoning capabilities. However...
Fingerprint Vector: Enabling Scalable and Efficient Model Fingerprint Transfer via Vector Addition
arXiv:2409.08846v3 Announce Type: replace-cross Abstract: Backdoor-based fingerprinting has emerged as an effective technique for tracing...
FineWeb2: One Pipeline to Scale Them All — Adapting Pre-Training Data Processing to Every Language
arXiv:2506.20920v1 Announce Type: new Abstract: Pre-training state-of-the-art large language models (LLMs) requires vast amounts of...
Fine-tuning vs. in-context learning: New research guides better LLM customization for real-world tasks
By combining fine-tuning and in-context learning, you get LLMs that can learn tasks that would...
Fine-Tuning Qwen3 with LoRA Using NVIDIA NeMo AutoModel: A Complete Single-GPU Google Colab Workflow Tutorial
In this tutorial, we build an end-to-end NVIDIA NeMo AutoModel workflow in Google Colab and...
Fine-Tuning a BERT Model
This article is divided into two parts; they are: • Fine-tuning a BERT Model for...
Finding My Voice: Generative Reconstruction of Disordered Speech for Automated Clinical Evaluation
arXiv:2509.19231v1 Announce Type: cross Abstract: We present ChiReSSD, a speech reconstruction framework that preserves children...
FinCoT: Grounding Chain-of-Thought in Expert Financial Reasoning
arXiv:2506.16123v2 Announce Type: replace Abstract: This paper presents FinCoT, a structured chain-of-thought (CoT) prompting framework...

