← Back to home

HuggingFace

307 items in the index — today's curated AI papers from HuggingFace Daily. Everything opens right here on the site — you never leave.

Index updated 1h ago · refreshes hourly

← HighlightsStrict order
HuggingFace Daily Papers

SynthGait-19K: A Physically Grounded Synthetic Video Dataset for Gait Parameter Estimation

Accurate estimation of clinically meaningful gait parameters from monocular video is important for scalable mobility assessment, yet progress is limited by the small scale, restricted viewpoints, and limited visual diversity of existing datasets. We introduce SynthGait-19k, a phy

Soroush Mehraban, Xin Lei Lin, Vida Adeli, Majid Mirmehdi · Sep 8, 202632
16 likes
HuggingFace Daily Papers

Scores Alone Do Not Prove Discovery: The Discovery Certification Protocol for Auditing AI Research Agents

AI research agents combine prior knowledge, public sources, and experimental feedback to produce useful results. The Discovery Certification Protocol (DCP) turns claims about these results into executable recovery and feedback tests. Gate 1 validates useful improvement on sealed

Jingjie Ning, Shanshan Zhong, Xiaochuan Li, Ji Zeng · Sep 7, 202624
19 likes
HuggingFace Daily Papers

Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection

The growing realism and accessibility of manipulated and generated faces threaten the trustworthiness of digital media. To detect such forgeries, deepfake detectors based on vision foundation models have shown promising performance, but they typically rely on a single pretrained

Xuechao Zou, Yi Zhou, Kai Li, Shun Zhang · Sep 7, 202630
15 likes
HuggingFace Daily Papers

CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements

Transferring human hand demonstrations to robotic grippers has recently emerged as a cost-effective solution for robot learning. However, existing methods are largely confined to simple, planar tasks and fail to handle complex spatial movements (e.g., intricate trajectories invol

Hongxiang Zhao, Mutian Xu, Zeyu Jin, Yiming Hao · Sep 7, 202632
25 likes
HuggingFace Daily Papers

Measuring Language Transfer in Robot Policies: Adding Greek to a Cosmos3 Vision-Language-Action Policy

Robot foundation models are trained and evaluated predominantly in English, and robot demonstration corpora do not exist for most languages. We study the addition of Greek to an open vision-language-action stack using only machine-rephrased instructions and no architecture change

Ayoub Kirouane, Georgios Giaples, Christos Petrocheilos · Sep 7, 202630
19 likes
HuggingFace Daily Papers

DF26: We Cannot Tell Fake From Real Anymore

We introduce DF26, a novel benchmark for detecting AI-generated videos containing fully synthetic clips produced by recent text-to-video and image-to-video models. The videos capture single-person public-speaking scenarios, spanning direct-to-camera recordings, official statement

Severyn Shykula, Andrii Yermakov, Ivan Samarskyi, Dmytro Mishkin · Sep 7, 202624
5 likes