← Back to home

HuggingFace

307 items in the index — today's curated AI papers from HuggingFace Daily. Everything opens right here on the site — you never leave.

Index updated 3h ago · refreshes hourly

← HighlightsStrict order
HuggingFace Daily Papers

BeaconKV: Key-Value Cache Compression Guided by Beacon Queries for Efficient Large Reasoning Model Inference

Large Reasoning Models (LRMs) achieve superior problem-solving through extended Chain-of-Thought (CoT) generation, but the resulting key-value (KV) cache grows linearly with sequence length and creates severe memory bottlenecks, often exceeding GPU capacity for long reasoning tra

Janghyeon Kim, Minsoo Kim, Kyuhong Shim, Jungwook Choi · Sep 4, 202625
32 likes
HuggingFace Daily Papers

Beneath the Surface of Chains-of-Thought: A Mechanistic Interpretation of Reasoning Operations in LLMs

Reasoning in large language models unfolds through diverse functional operations, such as problem formulation, goal decomposition, and deduction. Although these operations are explicitly distinguished in text, little is known about how they are geometrically organized in represen

Seogyeong Jeong, Jaehui Hwang, Dongyoon Han, Geonmo Gu · Sep 4, 202628
14 likes
HuggingFace Daily Papers

Safety for Whom? Boundary-Aware Self-Distillation for Controlled LLM Safety Refusal

Safety alignment is usually posed as a topic-level question: is this subject harmful? Deployments ask a narrower one. A civics tutor and a public-sector assistant may share a base model yet need different boundaries inside the same topic, refusing targeted political manipulation

Alejo López-Ávila, Iker García-Ferrero, Jezabel Garcia, Antonio Tiene · Sep 3, 202624
5 likes
HuggingFace Daily Papers

Iris: Climbing to the Search Frontier

We present Iris-mini and Iris-pro, two search agents trained at the 35B-A3B and 397B-A17B scales, together with the data pipeline and training recipe behind them. Tasks are reverse-constructed from the hyperlink structure of a web corpus: we author multi-hop chains over an entity

Ziyuan Liu, Hengqi Liu, Zichuan Wang, Yang Qin · Sep 3, 202626
54 likes
HuggingFace Daily Papers

One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing

Video editing spans diverse editing paradigms, yet achieving high-quality instruction-guided and subject-guided editing within a single unified framework remains challenging. We introduce EditVid, a training-free framework combining sparse causal memory for local coherence, corre

Adheesh Sunil Juvekar, Onkar Kishor Susladkar, Kiet A. Nguyen, Muntasir Wahed · Sep 3, 202626
6 likes
HuggingFace Daily Papers

Last Translation Benchmark

For scientific progress, we need benchmarks that test the limits of state-of-the-art models, and evaluation methods that inform us about failure cases. As models get stronger, standard benchmarks for machine translation are approaching saturation. Further, automatic translation m

Vilém Zouhar, Niyati Bafna, Mukund Choudhary, Maike Züfle · Sep 3, 202626
33 likes