News / #robotics Tag Robotics 419 articles archived under #robotics · RSS Sign in to follow Hugging Face Daily Papers research 1mo ago From Foundation to Application: Improving VLA Models in Practice Abstract LingBot-VLA 2.0 enhances generalization across tasks and embodiments through expanded data preprocessing and training on diverse robot configurations, extends action space to include whole-body degrees of freedom for complex manipulation tasks, and incorporates… 32 Hugging Face Daily Papers research 1mo ago 3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance Abstract 3D HAMSTER framework enhances robot manipulation by integrating a vision-language model with depth encoding to generate metrically accurate 3D trajectories for point cloud-based control policies. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Hierarchical… 27 arXiv — Machine Learning research 1mo ago Deep Reinforcement Learning for Dynamic Battery Management of Autonomous Order Pickers arXiv:2607.05683v1 Announce Type: new Abstract: Battery charging of Autonomous Mobile Robots (AMRs) in warehouses is a critical operational challenge that heavily impacts both order processing times and throughput. In this study, we address the dynamic AMR charging problem under… 21 Hugging Face Daily Papers research 1mo ago GaP: A Graph-as-Policy Multi-Agent Self-Learning Harness For Variational Automation Tasks Abstract Graph-as-Policy system combines modular robot skills with multi-agent coding to improve reliability in variable automation tasks through parallel simulation refinement. Generated by Qwen/Qwen2.5-Coder-32B-Instruct For robots to work reliably in commercial and industrial… 32 NVIDIA Developer Blog official-blog 1mo ago Develop Humanoid Robot Policies End-to-End with NVIDIA Isaac GR00T As more teams move from humanoid robot bring-up to task-specific skill development, the need for repeatable development workflows is growing. Building humanoids... 23 Ars Technica — AI news-outlet 1mo ago Robot workers rising: How AI may drive general-purpose autonomy in robotics Top robotics researchers and founders explain how robot autonomy is evolving. 20 Ars Technica — AI news-outlet 1mo ago How AI could enable autonomous robot workers in workplaces—and maybe homes Top robotics researchers and founders explain how robot autonomy is evolving. 7 Hugging Face Daily Papers research 1mo ago InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization Abstract InternVLA-A1.5 integrates pretrained vision-language models with future prediction in latent space to enable efficient robot manipulation with preserved semantics and long-horizon execution. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Unified models for robot… 15 Hugging Face Daily Papers research 1mo ago EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots Abstract EVA-Client is an open-source framework that unifies real-robot policy deployment, data collection, and evaluation through a component-decoupled architecture with inspectable execution workflows. Generated by Qwen/Qwen2.5-Coder-32B-Instruct We present EVA-Client, an… 22 Hugging Face Daily Papers research 1mo ago Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models Abstract A large-scale visuotactile dataset called Deform360 is introduced to study deformable object dynamics, enabling comparison between 2D video and 3D particle world models for robotic manipulation tasks. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Predicting object… 23 Hugging Face Daily Papers research 1mo ago Vision Pretraining for Dense Spatial Perception Abstract Boundary modeling enables dense spatial perception by learning sub-pixel representations that enhance depth estimation and support embodied AI applications. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Dense spatial perception is essential for physical intelligence,… 13 Hugging Face Daily Papers research 1mo ago GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation Abstract World models for robotic policy evaluation are systematically studied through a new benchmark, revealing that long-horizon rollout consistency and robot-specific controllability are more important than short-term visual realism for reliable policy assessment. Generated… 22 Hugging Face official-blog 1mo ago LeRobot v0.6.0: Imagine, Evaluate, Improve Back to Articles a]:hidden"> LeRobot v0.6.0: Imagine, Evaluate, Improve Published July 7, 2026 Update on GitHub Upvote 1 Steven Palma imstevenpmwork Pepijn Kooijmans pepijn223 Caroline Pascal CarolinePascal Khalil Meftah lilkm Martino Russi nepyope Nikodem Bartnik nikodembartnik… 26 r/LocalLLaMA community 1mo ago GitHub - kallewoof/tftf: Transforming Transformers -- ultra light-weight pipeline for enormous transformer model manipulation with minimal overhead Working with large models that don't fit in your VRAM+RAM is extremely annoying when you want to do things like LoRA merging or converting between formats, so I started the tftf (transforming transformers) project. The idea is simple: do all operations on a per tensor level.… 20 Hugging Face Daily Papers research 1mo ago VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon Abstract VLA-Corrector addresses limitations of action chunking in vision-language-action models by introducing a lightweight latent-space vision monitor that enables adaptive corrective replanning, improving robustness in contact-rich manipulation tasks. Generated by… 26 Hugging Face Daily Papers research 1mo ago Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots Abstract Embodied.cpp is a portable C++ runtime that enables efficient deployment of vision-language-action and world-action models across heterogeneous edge devices through modular execution layers and optimized inference. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Embodied… 7 r/MachineLearning community 1mo ago Question regarding Xournal++ and software 4 taking university notes during class [D] Hi. I have a question, could this plan and pipeline work?. I will be attending university master's classes on AI (thankfully got accepted a few days ago) and computers in a few months. There will be university lectures on machine learning, computer vision, robotics, video games… 9 Hugging Face Daily Papers research 1mo ago Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs Abstract Task-Agnostic Pretraining framework trains robotic models using self-supervised inverse dynamics on unlabeled data followed by lightweight language grounding, achieving superior performance with minimal expert demonstrations. Generated by Qwen/Qwen2.5-Coder-32B-Instruct… 28 arXiv — Machine Learning research 1mo ago Gaming Consensus: Coordinated Manipulation in Crowdsourced Fact-Checking arXiv:2607.01824v1 Announce Type: new Abstract: Crowdsourced fact-checking systems have been adopted by major social media companies such as X, Meta, TikTok and Google with the aim of combating misleading information at scale without relying on centralized editorial control.… 15 arXiv — Machine Learning research 1mo ago Privacy-Preserving and Verifiable Approximate Distributed Coded Computing arXiv:2607.02187v1 Announce Type: new Abstract: Distributed machine learning enables collaborative model training without centralizing data, but it also exposes learning processes to privacy leakage and malicious manipulation. Existing defenses typically address these threats in… 13 arXiv — Machine Learning research 1mo ago QFedAgent: Quantum-Enhanced Personalized Federated Learning for Multi-Agent Activity Recognition arXiv:2607.02426v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data, making it suitable for privacy-sensitive robotic sensing applications. However, multi-agent systems generate… 8 arXiv — NLP / Computation & Language research 1mo ago PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation arXiv:2607.01938v1 Announce Type: cross Abstract: Manipulating fast and dynamically moving targets in unstructured 3D environments remains challenging for embodied AI. Existing visual-language-action models and world models struggle with accurate 3D geometry and physically… 35 Hugging Face Daily Papers research 1mo ago ASPIRE: Agentic /Skills Discovery for Robotics Abstract ASPIRE is a continual learning system that autonomously develops and refines robot control programs through iterative exploration, achieving superior performance and zero-shot generalization in manipulation and household tasks while enabling sim-to-real transfer.… 12 arXiv — Machine Learning research 1mo ago HydraCollab: Adaptive Collaborative-Perception for Distributed Autonomous Systems arXiv:2607.00191v1 Announce Type: cross Abstract: Collaborative-perception enables multi-robot systems to enhance situational awareness by sharing perceptual information. Existing collaborative-perception systems face an inherent trade-off between communication bandwidth… 22 Hugging Face Daily Papers research 1mo ago ABot-M0.5: Unified Mobility-and-Manipulation World Action Model Abstract ABot-M0.5 is a World Action Model for mobile manipulation that improves performance through temporal granularity alignment, action space disentanglement, and train-test consistency in autoregressive prediction. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Mobile… 16 Hacker News — AI on Front Page community 1mo ago Oomwoo, an open-source robot vacuum you build yourself Article URL: https://makerspet.com/blog/building-an-open-source-robot-vacuum-meet-oomwoo/ Comments URL: https://news.ycombinator.com/item?id=48755005 Points: 241 # Comments: 41 37 Hacker News — AI on Front Page community 1mo ago Weave Robotics launches Isaac 1, a $7,999 home robot with Fall 2026 deliveries https://runtimewire.com/article/weave-robotics-isaac-1-home-... Comments URL: https://news.ycombinator.com/item?id=48750989 Points: 203 # Comments: 288 8 Hugging Face Daily Papers research 1mo ago Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly? Abstract A reinforcement learning framework called Play2Perfect enables sample-efficient robotic assembly tasks by first learning general manipulation skills through playful interaction with diverse objects, then adapting these skills for precise assembly through fine-tuning.… 34 Hugging Face Daily Papers research 2mo ago Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing Abstract A large-scale video editing dataset and model are introduced that support multi-task and structural manipulations through advanced data synthesis and network architectures. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Existing instruction-based video editing datasets… 38 Hugging Face Daily Papers research 2mo ago Scenes as Objects, Not Primitives: Instance-Structured 3D Tokenization from Unposed Views Abstract A feed-forward framework decomposes 3D scenes into instance-structured token groups from multi-view images, enabling direct object-level reconstruction, segmentation, and manipulation without 3D annotations. Generated by Qwen/Qwen2.5-Coder-32B-Instruct A 3D scene is… 38 arXiv — Machine Learning research 2mo ago Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation arXiv:2606.31043v1 Announce Type: new Abstract: Residual reinforcement learning adapts a pretrained robot policy by learning an additive correction to its actions. While effective when adaptation amounts to shifting the base policy's action distribution, additive corrections… 26 arXiv — NLP / Computation & Language research 2mo ago ViTL: Temporal Logic-Guided Zero-Shot Natural Language Navigation via Vision-Language Models arXiv:2606.30696v1 Announce Type: cross Abstract: Enabling robots to follow natural language commands to complete zero-shot long-horizon tasks remains challenging. It requires extracting implicit temporal and logical constraints from natural language commands and executing… 4 arXiv — NLP / Computation & Language research 2mo ago RCT: A Robot-Collected Touch-Vision-Language Dataset for Tactile Generalization arXiv:2606.31694v1 Announce Type: cross Abstract: For robots manipulating open-world objects, tactile representations must generalize to unseen materials. We introduce RCT (Robotic Contact Tactile), a robot-collected touch-vision-language dataset with 29,279 tactile frames from… 18 Hugging Face Daily Papers research 2mo ago Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models? Abstract Research reveals that language backbones in Vision-Language-Action models are highly redundant for robotic manipulation tasks, while vision and action pathways are more critical, suggesting need for deliberate capacity allocation in future architectures. Generated by… 11 Hugging Face Daily Papers research 2mo ago Learning Transferable Dynamics Priors from Action to World Modeling Abstract Action-conditioned world modeling enables transferable dynamics priors for robot learning through pretraining on large-scale manipulation data, supporting both simulator-based policy evaluation and video-action prediction. Generated by Qwen/Qwen2.5-Coder-32B-Instruct We… 27 arXiv — Machine Learning research 2mo ago A Linear Matching Bandit Approach to Online Multi-Human Multi-Robot Teaming arXiv:2606.29221v1 Announce Type: new Abstract: We address the problem of online multi-human multi-robot teaming through the lens of a linear matching bandit framework, where a learner assigns robots with unknown features from a fixed pool to distinct sets of human agents over… 15 Ars Technica — AI news-outlet 2mo ago South Korea to spend $1T on more memory chip production and humanoid robots South Korea targets physical AI lead and commercial humanoid robots by 2028. 9 r/MachineLearning community 2mo ago I do historical swordfighting and noticed AI struggles to track it. I’m building an open dataset to help fix this. Does my schema make sense? [P] Hi everyone, I’m a historical swordfighter (HEMA practitioner), and while I’m not a computer vision engineer or a roboticist, I’ve been reading a lot about the current bottlenecks in embodied AI, specifically around the Sim2Real gap and thin-object tracking. It occurred to me… 18 TechCrunch — AI news-outlet 2mo ago Robot hand company settles Tesla trade secret suit and announces $11M raise Jay Li doesn’t recommend getting sued by Tesla if you’re trying to get a startup off the ground. But he does think his company, Proception, might be better off for having endured the experience. “I think it’s kind of like a resilience test, or pressure… 15 Import AI (Jack Clark) community 2mo ago Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era What eras bookend our interregnum? 36 arXiv — Machine Learning research 2mo ago Support-Constrained RL Enables Real-World Policy Improvement without Real-World Experience arXiv:2606.27475v1 Announce Type: cross Abstract: Robots trained on real world data tend to be imprecise, slow, and brittle to perturbations. Improving these policies with reinforcement learning (RL) is an appealing alternative, but this process often requires expensive training… 28 arXiv — Machine Learning research 2mo ago Physics-Guided Robotic Radiation Source Localization along Arbitrary Measurement Paths in Unstructured Environments arXiv:2606.27624v1 Announce Type: cross Abstract: Using robots to estimate the location of the radiation source is an effective way to improve efficiency and safety. Existing methods focus on planning the robot's path to achieve precise estimation, typically approaching the… 19 MIT News — AI research 2mo ago LLMs help robots understand vague instructions and focus on key details To help robots do chores in places like homes and factories, a new approach from MIT uses one language model to clarify users’ instructions, then another to ignore irrelevant info. 19 arXiv — Machine Learning research 2mo ago Revisiting Action Factorization for Complex Action Spaces arXiv:2606.26574v1 Announce Type: new Abstract: Many real-world control problems involve hybrid discrete-continuous action spaces. For example, steering and signaling in autonomous driving, and aiming and firing in robotics or video-games. Despite real-world hybrid factorization… 10 arXiv — NLP / Computation & Language research 2mo ago Charting the Growth of Social-Physical HRI (spHRI): A Systematic Review Pipeline Augmented by Small Language Models arXiv:2606.26382v1 Announce Type: new Abstract: Social-physical human-robot interaction (spHRI) has grown rapidly across robotics, human-computer interaction, human-robot interaction, and haptics. Yet, fragmented terminology and inconsistent methodologies make systematic… 35 Hugging Face Daily Papers research 2mo ago In-Context World Modeling for Robotic Control Abstract ICWM enables robot policies to infer system variables from self-generated interactions, allowing adaptation to novel configurations without parameter updates by treating system identification as an in-context adaptation problem. Generated by… 8 arXiv — NLP / Computation & Language research 2mo ago RAVEN: Long-Horizon Reasoning & Navigation with a Visuo-Spatio-Temporal Memory arXiv:2606.25206v1 Announce Type: cross Abstract: Long-term robot deployment requires a compact and scalable memory that preserves fine-grained visual semantics, grounds observations in space and time, and enables efficient storage and retrieval. In this paper, we propose RAVEN,… 21 Hugging Face Daily Papers research 2mo ago EBench: Elemental Diagnosis of Generalist Mobile Manipulation Policies Abstract EBench is a comprehensive simulation benchmark for evaluating generalist mobile manipulation policies across diverse tasks and dimensions, revealing distinct capability profiles and generalization patterns among state-of-the-art models. Generated by… 18 Hugging Face Daily Papers research 2mo ago InSight: Self-Guided Skill Acquisition via Steerable VLAs Abstract InSight enables autonomous skill acquisition for vision-language-action models through primitive-action level steerability and automated demonstration generation. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Vision-language-action (VLA) models can learn manipulation… 19 TechCrunch — AI news-outlet 2mo ago Agility Robotics plans to go public via SPAC in a $2.5B deal Agility Robotics, the humanoid robotics startup that spun out of Oregon State University in 2015, expects to generate $620 million in proceeds. 13 Page 5 of 9 · 419 articles ← Newer Older →