Actualités
Actualités
Forensic deepfake audio detection using segmental speech features
arXiv:2505.13847v1 Announce Type: cross Abstract: This study explores the potential of using acoustic features of...
Forcing LLMs to be evil during training can make them nicer in the long run
A new study from Anthropic suggests that traits such as sycophancy or evilness are associated...
FMBench: Adaptive Large Language Model Output Formatting
arXiv:2602.06384v1 Announce Type: new Abstract: Producing outputs that satisfy both semantic intent and format constraints...
FLUX.1 Kontext enables in-context image generation for enterprise AI pipelines
FLUX.1 Kontext from Black Forest Labs aims to let users edit images multiple times through...
FluoroSAM: A Language-promptable Foundation Model for Flexible X-ray Image Segmentation
arXiv:2403.08059v3 Announce Type: replace-cross Abstract: Language promptable X-ray image segmentation would enable greater flexibility for...
Five Years of SciCap: What We Learned and Future Directions for Scientific Figure Captioning
arXiv:2512.21789v1 Announce Type: new Abstract: Between 2021 and 2025, the SciCap project grew from a...
Five things you need to know about AI
At SXSW London last week I gave a talk called “Five things you need to...
FIT: Defying Catastrophic Forgetting in Continual LLM Unlearning
arXiv:2601.21682v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate impressive capabilities across diverse tasks...
Fish Audio Releases Fish Audio S2: A New Generation of Expressive Text-to-Speech (TTS) with Absurdly Controllable Emotion
The landscape of Text-to-Speech (TTS) is moving away from modular pipelines toward integrated Large Audio...
FireRedTeam Releases FireRed-OCR-2B Utilizing GRPO to Solve Structural Hallucinations in Tables and LaTeX for Software Developers
Document digitization has long been a multi-stage problem: first detect the layout, then extract the...
FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design
arXiv:2506.13066v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) demonstrate significant cross-modal reasoning capabilities. However...
Fingerprint Vector: Enabling Scalable and Efficient Model Fingerprint Transfer via Vector Addition
arXiv:2409.08846v3 Announce Type: replace-cross Abstract: Backdoor-based fingerprinting has emerged as an effective technique for tracing...
