ข่าว
ข่าว
Format-Adapter: Improving Reasoning Capability of LLMs by Adapting Suitable Format
arXiv:2506.23133v1 Announce Type: new Abstract: Generating and voting multiple answers is an effective method to...
Forget the hype — real AI agents solve bounded problems, not open-world fantasies
Event-driven multi-agent systems are a practical architecture for working with imperfect tools in a structured...
Forensic deepfake audio detection using segmental speech features
arXiv:2505.13847v1 Announce Type: cross Abstract: This study explores the potential of using acoustic features of...
Forcing LLMs to be evil during training can make them nicer in the long run
A new study from Anthropic suggests that traits such as sycophancy or evilness are associated...
FMBench: Adaptive Large Language Model Output Formatting
arXiv:2602.06384v1 Announce Type: new Abstract: Producing outputs that satisfy both semantic intent and format constraints...
FLUX.1 Kontext enables in-context image generation for enterprise AI pipelines
FLUX.1 Kontext from Black Forest Labs aims to let users edit images multiple times through...
FluoroSAM: A Language-promptable Foundation Model for Flexible X-ray Image Segmentation
arXiv:2403.08059v3 Announce Type: replace-cross Abstract: Language promptable X-ray image segmentation would enable greater flexibility for...
Five Years of SciCap: What We Learned and Future Directions for Scientific Figure Captioning
arXiv:2512.21789v1 Announce Type: new Abstract: Between 2021 and 2025, the SciCap project grew from a...
Five things you need to know about AI
At SXSW London last week I gave a talk called “Five things you need to...
FIT: Defying Catastrophic Forgetting in Continual LLM Unlearning
arXiv:2601.21682v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate impressive capabilities across diverse tasks...
Fish Audio Releases Fish Audio S2: A New Generation of Expressive Text-to-Speech (TTS) with Absurdly Controllable Emotion
The landscape of Text-to-Speech (TTS) is moving away from modular pipelines toward integrated Large Audio...
FireRedTeam Releases FireRed-OCR-2B Utilizing GRPO to Solve Structural Hallucinations in Tables and LaTeX for Software Developers
Document digitization has long been a multi-stage problem: first detect the layout, then extract the...

