AI Research Papers — Page 5

Archive, newest first. Filter by topic, impact and company →

Notableresearch

Behavioral Foundation Models for Quality Diversity

Problem The paper addresses a gap in the capability of existing Quality-Diversity (QD) methods to effectively search for behaviorally diverse and high-performing policies. The authors highlight the limitations of traditional…

arXivcodeNazim Bendib +2
Notableresearch

Minute-scale training for microrobot navigation

Problem Current deep reinforcement learning approaches for microrobot navigation exhibit limited learning efficiency and effectiveness. This paper addresses these shortcomings by proposing a framework that enables effective navigation policy training…

Notableresearch

RAPID: Robot Agentic Programming from Demonstrations

Problem The paper addresses the challenge of generating, verifying, and refining robot programs based on a single visual demonstration from a human. This capability is crucial for enhancing robot autonomy…

arXivcodeYuyao Liu +4
Notableresearch

Rolling-WAM: World Action Models with Rolling Imagination

Problem The paper addresses the latency issues associated with the joint video-action denoising process, which delays action updates and limits the responsiveness of closed-loop systems. This is particularly critical in…

arXivcodeYinghua Zhou +10
Notableresearch

PoEM: Predicting RL Outcomes from Existing Policies

Problem The paper addresses the challenge of predicting reinforcement learning (RL) outcomes for new reward functions without the need to run RL algorithms on these functions. This is particularly relevant…

arXivcodeKimia Hamidieh +2
Notableresearch

Minimally Invasive Steering of Language Models

Problem Unregularized reward optimization can significantly alter the output distribution of language models, leading to a degradation in generation quality. This paper addresses this gap by proposing a method that…

arXivcodeTaha Entesari +3
Notableresearch

Jev-Mobile: Jev as an Executor for Mobile GUI Agents

{'Problem': 'Existing systems for mobile GUI agents typically rely on Vision-Language Models (VLM) for both planning and action grounding, which results in increased latency and higher model-serving costs. This paper…

arXivcodeLinghua Zhang