News / #robotics Tag Robotics 419 articles archived under #robotics · RSS Sign in to follow Ars Technica — AI news-outlet 12d ago Former SpaceX engineers are building a robotic factory for making steel parts “We're not necessarily building in a dogmatic fashion towards full autonomy.” 36 MIT Technology Review — AI news-outlet 13d ago What happens when a kid’s robot best friend dies? When Xander first met Moxie, she taught him that when he was anxious, he could calm down by exhaling through his lips so that he buzzed like a bee. They practiced breathing like dragons to manage feeling mad and sniffing like bunnies to boost his energy. But in the six years… 17 Hugging Face Daily Papers research 13d ago PRM-as-a-Judge 1.5: A Toolkit for Robot Process Assessment Abstract PRM-as-a-Judge 1.5 provides fine-grained process metrics and reliability tools to evaluate embodied robotic models beyond binary success rates. Generated by thinkingmachines/Inkling-Small Fine-grained robotic evaluation matters for understanding embodied models, going… 33 arXiv — Machine Learning research 13d ago hint$^2$: Hierarchical World Models for Inference-Time Temporal Logic Guidance arXiv:2608.13678v1 Announce Type: cross Abstract: A central goal of robot learning is to enable robots to execute rich instructions specified at runtime. Large-scale language-conditioned policies have made substantial progress toward this goal, yet still struggle with temporal… 4 arXiv — NLP / Computation & Language research 13d ago Agentic Transaction: Towards ACID-Compliant Agent Systems arXiv:2608.13900v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from conversational assistants into autonomous systems that execute long-horizon tasks through reasoning, tool use, code generation, and workspace manipulation. As agents… 20 Hugging Face Daily Papers research 13d ago HumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmark Abstract HumanTracker introduces a large-scale benchmark and preference-aligned metric to evaluate humanoid motion tracking based on perceptual quality and physical contact stability. Generated by thinkingmachines/Inkling-Small Humanoid motion tracking is central to… 5 r/MachineLearning community 13d ago [Career Advice] Final-year in Physical AI / Robotics. How is the market & global hiring for freshers? [D] Hi everyone, I am heading into my final year of my BTech at a tier 1 college in India and just wrapped up a Physical AI internship at a MNC, working heavily with NVIDIA Isaac Sim and OpenFOAM. My background is fully focused on robotics and autonomy. My tech stack includes:… 35 Hugging Face Daily Papers research 16d ago H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models Abstract H2R-Bench evaluates video generation models on transforming human manipulation videos into robot-centric demonstrations across embodiment constraints and interaction fidelity. Generated by thinkingmachines/Inkling-Small Large-scale manipulation data is essential for… 22 arXiv — Machine Learning research 16d ago Towards Socially Compliant Navigation in Deep Reinforcement Learning via Proxemics-Based Reward Modeling arXiv:2608.12917v1 Announce Type: new Abstract: Developing effective robot navigation methods in crowded environments is essential for real-world applications. Although recent deep reinforcement learning (DRL) methods have improved navigation performance in crowded environments,… 31 Hugging Face Daily Papers research 16d ago DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation Abstract DreamX-Phi 1.0 is an action-conditioned video world model for robotic manipulation that uses geometric attention encoding, depth estimation, object masks with a frozen teacher, and distillation to generate faithful future observations. Generated by… 23 Hugging Face official-blog 16d ago Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets Back to Articles a]:hidden"> Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets Enterprise Article Published August 13, 2026 Upvote 4 Sundar Raghavan rsundaraws amazon Steven Palma imstevenpmwork amazon Cagatay Cali cagataydev… 34 Hugging Face Daily Papers research 17d ago AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Abstract AtlasVLA improves embodied AI by replacing reactive control with proactive reasoning via persistent world-ego memory, enabling robust long-horizon manipulation from a single wrist camera. Generated by thinkingmachines/Inkling-Small While Vision-Language-Action (VLA)… 36 arXiv — Machine Learning research 17d ago MOON: Multi-Objective OrthoNormalized Updates for Multitask Learning arXiv:2608.11749v1 Announce Type: new Abstract: Multi-objective optimization (MOO) has demonstrated significant success in multi-task learning by mitigating task conflicts through gradient manipulation. However, most existing methods flatten model parameters into vectors and… 8 NVIDIA Developer Blog official-blog 18d ago NVIDIA JetPack 7.2.1 Adds Agentic Video Skills and T3000 Emulation Video is a core data path across NVIDIA Jetson applications, from robotics and intelligent video analytics to industrial automation, healthcare, media... 5 Hugging Face Daily Papers research 19d ago RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance Abstract RynnValue is a scalable open-source value foundation model for robot manipulation that uses temporal distance instead of preferences or progress to learn generalizable value predictions and improve real-world policy success. Generated by thinkingmachines/Inkling-Small… 28 arXiv — Machine Learning research 19d ago LUCID: Latent-Skill Unified Control via Imagined Dynamics for Long-Horizon Humanoid Loco-Manipulation arXiv:2608.07746v1 Announce Type: new Abstract: Long-horizon humanoid loco-manipulation requires composing versatile whole-body skills and reliable high-level decision making. Existing methods often coordinate pretrained skills with scripted planners, finite-state machines or… 30 arXiv — Machine Learning research 19d ago V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control arXiv:2608.07870v1 Announce Type: new Abstract: Improving sample efficiency remains a core challenge in reinforcement learning (RL), especially in real-world settings like robotics, where data collection is costly. This challenge is pronounced in visual RL, where… 14 Hacker News — AI on Front Page community 19d ago Show HN: Needle2: 14MB agentic LLM for phones, wearables, smart home and robots Hey HN, Henry from Cactus here! We previously released Cactus Needle, a 14MB agentic LLM for tool call, device use, and structured extraction for phones, wearables, smart homes, small robots and microcontrollers. We got really great feedback here, and have now incorporated the… 17 r/LocalLLaMA community 19d ago Needle 2: 14MB agentic LLM for phones, wearables, smart home and robots. Hey LocalLlaMa, Henry from Cactus here! We previously released Cactus Needle, a 14MB agentic LLM for tool call, device use, and structured extraction for phones, wearables, smart homes, small robots and microcontrollers. We got really great feedback here, and have now… 15 Hugging Face Daily Papers research 20d ago Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence Abstract Vision-language models are increasingly serving as the reasoning core of embodied agents. Robot execution is inherently iterative: each action reshapes the scene and physical state, continually renewing what must be perceived, reasoned about, and verified. Meeting these… 25 r/LocalLLaMA community 20d ago omlab/VLX-Seek-1.5-10B · Hugging Face VLX-Seek-1.5-10B VLX-Seek-1.5-10B is the open-source 10B model in the VLX-Seek 1.5 family, designed for fine-grained perception and visual grounding in embodied scenarios. It targets practical settings such as drones, robots, robotic dogs, surveillance cameras, inspection… 27 arXiv — Machine Learning research 20d ago Fairis: Fairness-Aware Aggregation with Provable Influence Containment against Fairness Poisoning Attacks in Collaborative Machine Learning arXiv:2608.06469v1 Announce Type: cross Abstract: Collaborative machine learning among financial institutions must be both group-fair and robust against deliberate adversarial manipulation. Existing fairness-aware aggregation methods remain formally vulnerable to fairness… 21 arXiv — NLP / Computation & Language research 20d ago How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots arXiv:2608.06898v1 Announce Type: cross Abstract: Researchers who seek to build social robot applications on foundation models are faced with a difficult question: how should we pick a model? Public leaderboards offer little guidance: the demands of real-time, embodied social… 14 r/MachineLearning community 20d ago Non-Physical Intelligence Has A Ceiling [D] Reasoning alone cannot predict the chaotic physical world. Without a sensory and motor interface to reality, non-physical AI will not deliver the scientific and technological breakthroughs we expect.   submitted by   /u/dontkry4me [link]   [comments] 6 Hugging Face Daily Papers research 22d ago Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Abstract Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that write and refine their own executable skills as code. This survey organises the field around that axis of weights… 38 Hugging Face Daily Papers research 22d ago DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation Abstract Vision-Language-Action (VLA) models have become a powerful paradigm for robot manipulation, but training a single generalist policy for heterogeneous robot embodiments remains an open problem. Existing methods have two main limitations. First, they underuse dynamics… 38 arXiv — Machine Learning research 23d ago Failing Gracefully: Mitigating Impact of Inevitable Robot Failures arXiv:2608.05313v1 Announce Type: cross Abstract: Service robots operate in household environments shared with humans, pets, and everyday objects, where they are highly susceptible to failures such as software crashes, hardware degradation, or unpredictable interactions. While… 25 arXiv — Machine Learning research 23d ago Velocity- and Regime-Aware Detection of Intraday Options Market Manipulation, with Explainable Attribution arXiv:2608.05373v1 Announce Type: cross Abstract: Intraday market manipulation is hard to detect because its footprint is brief, buried in millions of quotes, and statistically similar to ordinary volatility. Detectors reach high recall only by flagging so many other days that… 6 Hugging Face Daily Papers research 23d ago World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation Abstract Vision-language-action (VLA) models often treat main-view and wrist-view observations as parallel visual inputs, overlooking their distinct roles in robot manipulation. Fine-grained manipulation, however, benefits from anticipating how wrist-local interactions may… 38 Hugging Face Daily Papers research 24d ago DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack Abstract Flow-matching vision-language-action (VLA) models such as pi0 generate robot actions by integrating a learned denoising velocity field, and have been reported to resist adversarial perturbations that readily fool autoregressive VLAs. We show that this robustness is… 25 arXiv — Machine Learning research 24d ago Manipulation-Proof Oblivious Audits against Deceptive Model Providers arXiv:2608.04365v1 Announce Type: new Abstract: Audits have emerged as a critical instrument for algorithmic governance, providing a mechanism for external scrutiny and governance of machine learning models. However, ensuring the integrity of such assessments remains a… 19 Hugging Face Daily Papers research 24d ago Ego2Robot: Scalable Robot Data Synthesis from Egocentric Human Data Abstract Learning generalizable robot manipulation policies requires large-scale and diverse demonstration data. Egocentric human manipulation videos offer rich scene and task diversity, and prior work has shown that retargeting and rendering such videos into robot-format data… 12 Hugging Face Daily Papers research 24d ago BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Abstract Leveraging pre-trained vision-language models (VLMs) to construct vision-language-action (VLA) models has emerged as a promising paradigm for 3D robot manipulation. However, existing 3D VLA methods remain data-hungry, exhibit limited generalization under distribution… 4 Marcus on AI community 24d ago Elon Musk’s preposterous and possibly harmful prediction about robotic surgery Some predictions are off. This one is way off. 13 r/LocalLLaMA community 24d ago Xiaomi-Robotics-1: New robotics model released Xiaomi-Robotics-1 is a robot foundation model trained on over 100K hours of real-world manipulation trajectories. It is a Vision-Language-Action (VLA) model engineered for out-of-the-box mobile manipulation in unseen environments and efficient adaptation to new tasks. XR-1… 14 TechCrunch — AI news-outlet 24d ago TechCrunch Disrupt 2026’s Real World AI Stage features robots, automated factories, and extinct animals On our new Real World AI stage, we’ll be focusing on the intersection between the digital and physical, and all the ways we’ll continue to see a blending of the two. 24 Hugging Face Daily Papers research 25d ago Push-Wiper: Toward General-Purpose Robotic Cleaning across Varied Stains and Surfaces with Segmented Pushing Trajectories Abstract Viscous stains, characterized by high viscosity and complex rheological properties, remain a major challenge for robotic surface cleaning. Conventional wiping often spreads the stain, while scrubbing provides stronger friction but risks damaging the surface. In this… 12 Hugging Face Daily Papers research 25d ago ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts Abstract World Action Models (WAMs) have emerged as a promising paradigm by jointly modeling robot actions and future visual dynamics. However, their reliance on pixel-generative future supervision can entangle action-relevant state transitions with task-irrelevant visual… 5 NVIDIA Developer Blog official-blog 25d ago Beyond VLAs: How World Action Models Reshape Robot Manipulation A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene... 16 TechCrunch — AI news-outlet 25d ago Elon Musk spends half his time talking robots and AI on Tesla earnings calls An analysis of the last seven years of Tesla earnings calls shows just little attention Musk pays to Tesla's car business. 30 Hugging Face Daily Papers research 26d ago DreamTraj: Generating 6-DoF Object Trajectories by Reading Unrendered Video Diffusion Latents Abstract Accurate prediction of object trajectories during manipulation is essential for closing the perception-action loop. Progress is limited on two fronts: available datasets lack fine-grained language-to-motion annotations, and existing predictors either rely on privileged… 19 arXiv — Machine Learning research 26d ago Fairness Auditing: Lower Bounds on Company Manipulation arXiv:2608.00568v1 Announce Type: new Abstract: Fairness audits are increasingly mandated in high-stakes applications such as hiring, lending, and automated decision-making. Recent work has established fundamental impossibility results for black-box fairness auditing, showing… 32 arXiv — NLP / Computation & Language research 26d ago DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text arXiv:2608.01046v1 Announce Type: new Abstract: The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation, and risks to vulnerable online communities. Existing transformer-based… 7 Hugging Face Daily Papers research 26d ago WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Abstract Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods, critic-based approaches rely on a value estimator that predominantly operates on single-frame observations or… 24 MIT Technology Review — AI news-outlet 26d ago Trump’s AI protectionism has come for robotics This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Humanoid robots usually elicit more cringe than awe: They stumble, kick children, and despite advances are still worse at using their hands… 8 Hugging Face Daily Papers research 27d ago One Future, Every Robot: Label-Efficient Collective-State Prediction with Decentralized JEPA Abstract Can every robot in a swarm predict the same future collective state from only local observations and bandwidth-limited messages? We formulate this as decentralized shared-state prediction and introduce Collective-State JEPA (CS-JEPA), a recurrent joint-embedding… 29 Hugging Face Daily Papers research 27d ago N_0-TWAM: Scaling Tactile-Native World-Action Model for Contact-Rich Manipulation Abstract We present N_0-TWAM, a tactile-native world-action model for contact-rich manipulation that predicts both future vision and future contact. To our knowledge, it is the first tactile world-action model trained at large scale, and it shows strong capability on… 31 arXiv — Machine Learning research 27d ago When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning arXiv:2607.29617v1 Announce Type: new Abstract: Imitation learning (IL)---training an agent to replicate expert behavior from demonstrations---underpins applications from robotics to language model training. Standard approaches such as Behavior Cloning (BC) are known to suffer… 24 arXiv — NLP / Computation & Language research 27d ago WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning arXiv:2607.29613v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods, critic-based approaches rely on a value estimator that predominantly operates… 34 Hugging Face Daily Papers research 27d ago N_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens Abstract We present N_0-VTLA, a vision-tactile-language-action (VTLA) foundation model capable of (1) fine-grained contact-rich manipulation with tactile perception and tactile-feedback control, and (2) offline policy improvement from stored deployment data. Building on current… 24 Page 2 of 9 · 419 articles ← Newer Older →