Research Feed

Search source-linked summaries of recent AI and machine-learning papers by topic, by date, or by whether they include code or a diagram.

Filter papers All papers

Browse by date

Resource filters

Sort options

Research results

Robotics / Multimodal By Zaibin Zhang 2026-08-26
MA-VLA: Multi-Arm Vision-Language-Action Model for Collaboration and Compositional Generalization

The researchers developed a vision-language-action model designed to improve how multiple robotic arms collaborate on complex tasks by using techniques that enforce role-agnostic instruction following.

Robotics / Reinforcement Learning By Lehong Wu 2026-08-26
$R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning

The paper introduces a method that uses vision-language models to perform reasoning that guides robot manipulation policies, improving performance on long-horizon tasks.

Robotics / Efficiency & Inference By Zhe Liu 2026-08-26
StreamPI: Streaming Multimodal Temporal Modeling for Vision-Language-Action Models

StreamPI adds historical context to vision-language-action models to improve robotic task performance without increasing the model parameter count.

Computer Vision / Robotics By Yueen Ma 2026-08-26
4DGS-WAM: Bridging Past and Future with an Object-Centric World Action Model based on 4D Gaussian Splatting

The 4DGS-WAM model enables future video prediction by explicitly decomposing scenes into dynamic objects and a static background using 4D Gaussian Splatting.

Robotics / Efficiency & Inference By Xiang Li 2026-08-25
Latent Action as Intention Enables Efficient Future Imagination for World Action Models

The LAWA architecture optimizes robot action planning by using latent intentions to reduce inference latency while maintaining high success rates across robotics benchmarks.

Reinforcement Learning / Robotics By Zihao Wu 2026-08-25 40
WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation

WarpSAC is a scalable reinforcement learning framework that adapts its architecture based on available compute resources to accelerate training and improve deployment success.

Agents / Robotics By Alperen Avan 2026-08-24
OptiSight: Bridging Semantic Reasoning and Geometric Control for Embodied Navigation

OptiSight combines semantic object identification with geometric control to enable efficient robot navigation while minimizing reliance on high-frequency language model inference.

Robotics / Multimodal By Siyuan Ma 2026-08-20
DECOWAM: Decoupled Whole-Body World-Action Model for Legged Mobile Manipulation

DECOWAM is a new model architecture that optimizes how legged robots coordinate whole body actions with visual environment predictions.

Robotics / Training & Fine-Tuning By Varun Giridhar 2026-08-21
Beyond Imitation: Self-Improving Robot Policies via Off-Policy Q-Planning

The paper presents a method that enables robot policies to self-improve through iterative deployment without the need to modify the original policy weights.

Robotics / Efficiency & Inference By Zhuoyuan Li 2026-08-21
Just Noticeable Difference Modeling for Token Compression in Vision-Language-Action Models

The authors introduce a method to compress token data in vision-language-action models by identifying and prioritizing information that has the least impact on physical robot movements.

Robotics / Multimodal By Yaowei Guo 2026-08-19
RoboEdit: Turning Human Manipulation Videos into Scalable Robot Experience

The authors introduce a pipeline to automatically reconstruct and retarget 3D human interaction data into a large-scale dataset for training diverse robotic embodiments.

Robotics / Reinforcement Learning By Bhavya Sukhija 2026-08-20 10
EXIMO: VLM Guided Exploration of VLA Policies

EXIMO leverages a vision-language model to decompose complex robotic tasks into smaller steps, improving the efficiency of training vision-language-action policies.

Robotics / Efficiency & Inference By Shaoxuan Wang 2026-08-20
RoMAN-Flow: Taming Autoregressive Normalizing Flows for Offline Reinforcement Learning in Robotic Manipulation

RoMAN-Flow introduces post-training optimization and distillation techniques to eliminate the sequential sampling latency inherent in autoregressive normalizing flows for robotic control.

Agents / Robotics By Hongyan Feng 2026-08-18 29
Embodied-Navigator: Point, Think, Memorize, and Align for Efficient Navigation

TAMP-Nav improves embodied navigation by combining efficient 3D spatial grounding with selective reasoning and a multi-level reward training approach.

Robotics / Efficiency & Inference By Yuxuan Chen 2026-08-14
Reflex: Enabling Fast and Predictive Vision-Language-Action Models for Reaction-Critical Manipulation

The paper introduces ReflexVLA, a vision-language-action model architecture that uses future prediction and optimized inference to improve performance in time-sensitive robotics tasks.

Robotics / Safety & Alignment By Alexei Odinokov 2026-08-14
Ensuring Safe Physical AI in Urban Mobility via Hazard-Informed Synthesized Envelopes

The paper introduces a framework to ensure safe robotic operation in complex urban environments by defining a dynamic safety envelope rather than using static constraints.

Agents / Robotics By Ross D. King 2026-08-14
The Past and Future of AI Scientists

The paper explores the development of autonomous AI systems capable of performing scientific research by integrating neural learning, robotics, and formal reasoning.

Robotics / Efficiency & Inference By Ann-Kathrin Schwehn 2026-08-14
Control-Informed Constraint Adaptation in Minimum-Time Trajectory Planning for Autonomous Racing

The paper introduces a feedback loop between the tracking controller and trajectory planner that adjusts spatial constraints to prevent sub-optimal performance caused by model mismatches.

Robotics / Benchmarks & Evals By Yuyang Liu 2026-08-14
PRM-as-a-Judge 1.5: A Toolkit for Robot Process Assessment

The paper introduces a toolkit that assesses robotic task execution by analyzing continuous progress curves rather than relying on binary success rates.

Robotics / Efficiency & Inference By Gang Zhang 2026-08-13
Capstan-driven Continuum Surgical Robot: Design, Modeling, and Perception

The researchers developed a sensing framework that enables surgical robots to estimate cable tension and contact location in real time using a parallelized computation model.