搜索:HuggingFace Daily Papers

共命中 50 条(服务端检索)
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation
Benchmark researchers and developers of large language models (LLMs) and other AI systems need to find relevant evaluati…
研究前沿 HuggingFace Daily Papers 6天前 阅读 1 · 访客 0
IdeaAMBIG: Benchmarking Implementation-Critical Gaps in Research-Idea Specifications
A research idea may be novel, coherent, and scientifically plausible, yet its proposed method may remain insufficiently …
研究前沿 HuggingFace Daily Papers 9-9 阅读 2 · 访客 0
WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data
Recent advances in wearable sensing enable continuous monitoring of physiological and behavioral signals, yet existing b…
研究前沿 HuggingFace Daily Papers 9-4 阅读 2 · 访客 0
RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments
Digital agents must often adapt to new environments whose interfaces, tools, and failure modes are not fully captured by…
智能体 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models
We present PhysBrain 1.5, a unified model for understanding physical environments, generating actions, and predicting fu…
行业动态 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training
Training prompts in online reinforcement learning (RL) differ substantially in how informative they are for the current …
行业动态 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus
Orthrus is a hybrid autoregressive-diffusion architecture that accelerates autoregressive language-model inference by ge…
行业动态 HuggingFace Daily Papers 2天前 阅读 1 · 访客 0
Kaininja: Extending Native 3D Generators to the Part Level
Native 3D generators turn one image into a single mesh. TRELLIS.2 and its peers deliver high-fidelity non-watertight geo…
行业动态 HuggingFace Daily Papers 2天前 阅读 1 · 访客 0
BVB: Benchmarking Agentic Video Understanding via Programmatic Reconstruction in Blender
Multimodal agents can create complex videos in software such as Blender by coding without relying on diffusion models. Y…
智能体 HuggingFace Daily Papers 2天前 阅读 1 · 访客 0
Discovery Foundation Models: Toward Open-Ended Discovery Intelligence
Foundation models have progressed from learning and reasoning over existing knowledge, to increasingly learning through …
研究前沿 HuggingFace Daily Papers 2天前 阅读 1 · 访客 0
Pick Your Poison: Learning to Select Poison Sets for Stronger LLM Backdoor Attacks
Backdoor poisoning attacks add poisoned examples to otherwise-clean finetuning data, pairing a trigger with a target beh…
大模型 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
ModaLens: Measuring Image Sensitivity in Report-Conditioned Medical VLMs
A radiology report can already answer a clinical question, so it is hard to tell whether a vision-language model also us…
行业动态 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
Enabling Creative Exploration for Vibe Design Agents
Vibe design agents turn natural-language briefs into rendered interfaces and frontend code. Yet a useful design agent sh…
智能体 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
When Agents Slow Down: Understanding LLM Agents' Test-Time Strategies via Elo-per-token Analysis
Large language model (LLM) agents allocate test-time compute adaptively as they revise solutions, use tools, explore alt…
智能体 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows
Video diffusion models are stochastic and hard to control: precise content often requires repeated sampling without guar…
智能体 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
Dream-RSI: Recursive Self-Improvement through Evolving Worlds
Recursive self-improvement is becoming increasingly vital for autonomous AI agents, where progress hinges on discovering…
智能体 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
HazardAuditor: From Executable Threats to Safer Computer-Use Agents
Computer-use agents increasingly interact with browsers, terminals, file systems, and external services, introducing saf…
智能体 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
Omni-Streaming Thinking
Streaming omni-modal models must decide what and when to answer from the video chunks and synchronized audio observed so…
行业动态 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
Atria Dawn: The Dawn of Agentic Superintelligence
As AI agents become participants in the development of their successors, they reshape both the production of intelligenc…
智能体 HuggingFace Daily Papers 2天前 阅读 0 · 访客 0
AlayaVista: Streaming World Modeling from Panoramic States to Perspective Video
Interactive video world models must maintain broad scene context under camera motion while producing high-fidelity obser…
行业动态 HuggingFace Daily Papers 3天前 阅读 0 · 访客 0
E2A-Bench: Benchmarking Evidence-to-Action Reliability in Financial Chart Reasoning
Can financial vision-language models (VLMs) turn chart evidence into reliable action recommendations? Existing hallucina…
研究前沿 HuggingFace Daily Papers 3天前 阅读 0 · 访客 0
Thought without systematicity? Evaluating reasoning models on rule induction tasks
A central tenet of human cognition is systematicity, the principle that understanding one concept is inherently tied to …
研究前沿 HuggingFace Daily Papers 4天前 阅读 0 · 访客 0
Dynin-Robotics: Omnimodal Unified Diffusion Vision-Language-Action Model
Visual goal and dynamics prediction can provide language-conditioned robot policies with both a target outcome and a rep…
智能体 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
Breaking the Vision-Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models
Robot foundation models achieve strong in-distribution performance but often degrade under visual distribution shifts. W…
智能体 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
Expert-Space Exploration in MoE Reinforcement Learning
Reinforcement learning (RL) has become central to post-training of large language models. Recent advances in RL for Mixt…
行业动态 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
In this work, we present ZGCM-1, a fully open 7B dense foundation model trained from scratch with extreme data, system, …
智能体 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
StepAudio 3 Gen Technical Report
We introduce StepAudio 3 Gen, a general-purpose audio generation model that supports zero-shot text-to-speech (TTS), voi…
行业动态 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
Root-Cause Attribution Is a Search Problem: Continual Search for Long-Horizon Agent Failures
The increasing deployment of AI agents in long-horizon tasks yields massive execution logs. Diagnosing failures within t…
智能体 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
MInTRL: Off-policy Intervention can boost On-policy RL
Reinforcement learning with verifiable rewards is typically performed on-policy, keeping training data close to the curr…
行业动态 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking
Post-training attention sparsification reduces the quadratic cumulative attention cost of pretrained Transformers by sel…
行业动态 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
Agent as Policy for Robotic Manipulation
We demonstrate that a general-purpose agent can directly drive a physical robot throughout task execution without any ta…
智能体 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
Building a Production Greek-English Speech Recognizer
We report a multi-month engineering program to build Sophea, a production bilingual Greek-English automatic speech recog…
行业动态 HuggingFace Daily Papers 5天前 阅读 0 · 访客 0
SNAP3D: Physically Grounded 3D Parts for Assembly from a Single Image
Part-aware 3D asset generation enables applications such as editing, articulation, simulation, and fabrication, yet exis…
行业动态 HuggingFace Daily Papers 5天前 阅读 2 · 访客 0
Recursive Code World Models: Building Complex Worlds through Recursive Scene Programs
Code world models represent worlds as executable programs, but this representation alone does not determine how to const…
行业动态 HuggingFace Daily Papers 6天前 阅读 6 · 访客 0
Ambient @ EgoProactive 2026 : Proactive Egocentric Assistance with Visually Grounded Supervision
We present our submission to the EgoProactive track of the ECCV 2026 Wearable AI Challenge, which ranked first in the la…
行业动态 HuggingFace Daily Papers 6天前 阅读 0 · 访客 0
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation
We present Vidu S2, which comprises Vidu S2-Avatar, a real-time interactive digital-character model, and Vidu S2-Editing…
行业动态 HuggingFace Daily Papers 6天前 阅读 0 · 访客 0
Ambient @ EgoLongQA 2026: Distilling Long-Video perception into a Sub-2B Model
We describe our entry to the EgoLongQA track of the Wearable-AI Challenge in ECCV 2026, which placed first in the
行业动态 HuggingFace Daily Papers 6天前 阅读 1 · 访客 0
Attention-DP3: Spatially Object-aware 3D Diffusion Policy via Geometry-aligned Attentional Conditioning
3D point-cloud observations are inherently ambiguous in complex, cluttered manipulation scenes, where target objects may…
行业动态 HuggingFace Daily Papers 6天前 阅读 0 · 访客 0
FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow Estimation
Optical flow methods typically rely on task-specific inductive biases, such as correlation volumes, feature warping, and…
行业动态 HuggingFace Daily Papers 6天前 阅读 2 · 访客 0
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization
Large language model (LLM) agents can benefit from reusable skills distilled from prior task experience, yet existing sk…
智能体 HuggingFace Daily Papers 6天前 阅读 0 · 访客 0
Feature Recovery for Object Understanding After Irreversible Fire Damage
Objects in post-fire environments often undergo irreversible physical transformations that change their geometry, materi…
行业动态 HuggingFace Daily Papers 6天前 阅读 0 · 访客 0
Mi-Ripple: Restoring Images Degraded by Iterative AI Editing
Iterative reference-conditioned image editing can introduce grid-like and granular textures, commonly described as digit…
行业动态 HuggingFace Daily Papers 6天前 阅读 4 · 访客 0
Competence-Gated Pooling of Language Models and Priors for Event Forecasting
In hybrid forecasting, a language model is often one of several available signals. A system may already have a market, c…
行业动态 HuggingFace Daily Papers 6天前 阅读 0 · 访客 0
DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-Agent Reinforcement Learning for Cooperative Air Combat
Multi-Agent Reinforcement Learning (MARL) has emerged as a pivotal paradigm for complex decision-making in autonomous sy…
智能体 HuggingFace Daily Papers 6天前 阅读 5 · 访客 0
X-AuT: Progressive Audio-Encoder Compression for Speech LLMs with Cross-Scale Distillation
Reducing audio-encoder depth lowers the inference cost of speech large language models, but removing complete blocks per…
大模型 HuggingFace Daily Papers 6天前 阅读 3 · 访客 0
Beyond Solver Verdicts: Generative Reward Models for Autoformalization
Neurosymbolic systems rely on mathematical solvers to guarantee reasoning correctness, yet solvers are fundamentally bli…
研究前沿 HuggingFace Daily Papers 6天前 阅读 4 · 访客 0
Pelican-Sim 1.0: A General World Model Simulator for Embodied Intelligence
In this technical report, we propose Pelican-Sim 1.0, a general world model simulator for embodied intelligence that pre…
智能体 HuggingFace Daily Papers 6天前 阅读 0 · 访客 0
Memory as Plans: World-Action Modeling with Memory-Grounded Planning
Mainstream robotic policies often adopt a Markovian formulation, but many complex real-world manipulation tasks are inhe…
智能体 HuggingFace Daily Papers 6天前 阅读 4 · 访客 0
World in World: Explore the World with World Models
Autoregressive video world models enable interactive, long-horizon exploration, but flexible control remains challenging…
行业动态 HuggingFace Daily Papers 6天前 阅读 3 · 访客 0
Negative Self-Distillation: Learning to Reason by Avoiding Flaws
On-Policy Self-Distillation (OPSD) has emerged as a popular paradigm for large language model (LLM) self-improvement, al…
大模型 HuggingFace Daily Papers 6天前 阅读 6 · 访客 0