YouZum

Gradients Must Earn Their Influence: Unifying SFT with Generalized Entropic Objectives

arXiv:2602.11424v1 Announce Type: new Abstract: Standard negative log-likelihood (NLL) for Supervised Fine-Tuning (SFT) applies uniform...

Gradient Descent:The Engine of Machine Learning Optimization

Editor’s note: This article is a part of our series on visualizing the foundations of...

GPZ: A Next-Generation GPU-Accelerated Lossy Compressor for Large-Scale Particle Data

Particle-based simulations and point-cloud applications are driving a massive expansion in the size and complexity...

GPTopic: Dynamic and Interactive Topic Representations

arXiv:2403.03628v3 Announce Type: replace Abstract: Topic modeling seems to be almost synonymous with generating lists...

GPT-4o Understands Text, But Does It See Clearly? A Benchmarking Study of MFMs on Vision Tasks

Multimodal foundation models (MFMs) like GPT-4o, Gemini, and Claude have shown rapid progress recently, especially...

GPT and Prejudice: A Sparse Approach to Understanding Learned Representations in Large Language Models

arXiv:2510.01252v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly trained on massive...

Governance-Aware Hybrid Fine-Tuning for Multilingual Large Language Models

arXiv:2512.17344v1 Announce Type: new Abstract: We present a governance-aware hybrid fine-tuning framework for multilingual, low-resource...

Google’s new AI training method helps small models tackle complex reasoning

Researchers at Google Cloud and UCLA have proposed a new reinforcement learning framework that significantly...

We use cookies to improve your experience and performance on our website. You can learn more at นโยบายความเป็นส่วนตัว and manage your privacy settings by clicking Settings.

ตั้งค่าความเป็นส่วนตัว

You can choose your cookie settings by turning on/off each type of cookie as you wish, except for essential cookies.

ยอมรับทั้งหมด
จัดการความเป็นส่วนตัว
  • เปิดใช้งานตลอด

บันทึกการตั้งค่า
th