News / #robotics Tag Robotics 360 articles archived under #robotics · RSS Sign in to follow Hugging Face Daily Papers research 1mo ago Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models Abstract A large-scale visuotactile dataset called Deform360 is introduced to study deformable object dynamics, enabling comparison between 2D video and 3D particle world models for robotic manipulation tasks. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Predicting object… 23 Hugging Face Daily Papers research 1mo ago Vision Pretraining for Dense Spatial Perception Abstract Boundary modeling enables dense spatial perception by learning sub-pixel representations that enhance depth estimation and support embodied AI applications. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Dense spatial perception is essential for physical intelligence,… 13 Hugging Face Daily Papers research 1mo ago GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation Abstract World models for robotic policy evaluation are systematically studied through a new benchmark, revealing that long-horizon rollout consistency and robot-specific controllability are more important than short-term visual realism for reliable policy assessment. Generated… 22 Hugging Face official-blog 1mo ago LeRobot v0.6.0: Imagine, Evaluate, Improve Back to Articles a]:hidden"> LeRobot v0.6.0: Imagine, Evaluate, Improve Published July 7, 2026 Update on GitHub Upvote 1 Steven Palma imstevenpmwork Pepijn Kooijmans pepijn223 Caroline Pascal CarolinePascal Khalil Meftah lilkm Martino Russi nepyope Nikodem Bartnik nikodembartnik… 26 r/LocalLLaMA community 1mo ago GitHub - kallewoof/tftf: Transforming Transformers -- ultra light-weight pipeline for enormous transformer model manipulation with minimal overhead Working with large models that don't fit in your VRAM+RAM is extremely annoying when you want to do things like LoRA merging or converting between formats, so I started the tftf (transforming transformers) project. The idea is simple: do all operations on a per tensor level.… 20 Hugging Face Daily Papers research 1mo ago VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon Abstract VLA-Corrector addresses limitations of action chunking in vision-language-action models by introducing a lightweight latent-space vision monitor that enables adaptive corrective replanning, improving robustness in contact-rich manipulation tasks. Generated by… 26 Hugging Face Daily Papers research 1mo ago Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots Abstract Embodied.cpp is a portable C++ runtime that enables efficient deployment of vision-language-action and world-action models across heterogeneous edge devices through modular execution layers and optimized inference. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Embodied… 7 r/MachineLearning community 1mo ago Question regarding Xournal++ and software 4 taking university notes during class [D] Hi. I have a question, could this plan and pipeline work?. I will be attending university master's classes on AI (thankfully got accepted a few days ago) and computers in a few months. There will be university lectures on machine learning, computer vision, robotics, video games… 9 Hugging Face Daily Papers research 1mo ago Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs Abstract Task-Agnostic Pretraining framework trains robotic models using self-supervised inverse dynamics on unlabeled data followed by lightweight language grounding, achieving superior performance with minimal expert demonstrations. Generated by Qwen/Qwen2.5-Coder-32B-Instruct… 28 arXiv — Machine Learning research 1mo ago Gaming Consensus: Coordinated Manipulation in Crowdsourced Fact-Checking arXiv:2607.01824v1 Announce Type: new Abstract: Crowdsourced fact-checking systems have been adopted by major social media companies such as X, Meta, TikTok and Google with the aim of combating misleading information at scale without relying on centralized editorial control.… 15 arXiv — Machine Learning research 1mo ago Privacy-Preserving and Verifiable Approximate Distributed Coded Computing arXiv:2607.02187v1 Announce Type: new Abstract: Distributed machine learning enables collaborative model training without centralizing data, but it also exposes learning processes to privacy leakage and malicious manipulation. Existing defenses typically address these threats in… 13 arXiv — Machine Learning research 1mo ago QFedAgent: Quantum-Enhanced Personalized Federated Learning for Multi-Agent Activity Recognition arXiv:2607.02426v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data, making it suitable for privacy-sensitive robotic sensing applications. However, multi-agent systems generate… 8 arXiv — NLP / Computation & Language research 1mo ago PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation arXiv:2607.01938v1 Announce Type: cross Abstract: Manipulating fast and dynamically moving targets in unstructured 3D environments remains challenging for embodied AI. Existing visual-language-action models and world models struggle with accurate 3D geometry and physically… 35 Hugging Face Daily Papers research 1mo ago ASPIRE: Agentic /Skills Discovery for Robotics Abstract ASPIRE is a continual learning system that autonomously develops and refines robot control programs through iterative exploration, achieving superior performance and zero-shot generalization in manipulation and household tasks while enabling sim-to-real transfer.… 12 arXiv — Machine Learning research 1mo ago HydraCollab: Adaptive Collaborative-Perception for Distributed Autonomous Systems arXiv:2607.00191v1 Announce Type: cross Abstract: Collaborative-perception enables multi-robot systems to enhance situational awareness by sharing perceptual information. Existing collaborative-perception systems face an inherent trade-off between communication bandwidth… 22 Hugging Face Daily Papers research 1mo ago ABot-M0.5: Unified Mobility-and-Manipulation World Action Model Abstract ABot-M0.5 is a World Action Model for mobile manipulation that improves performance through temporal granularity alignment, action space disentanglement, and train-test consistency in autoregressive prediction. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Mobile… 16 Hacker News — AI on Front Page community 1mo ago Oomwoo, an open-source robot vacuum you build yourself Article URL: https://makerspet.com/blog/building-an-open-source-robot-vacuum-meet-oomwoo/ Comments URL: https://news.ycombinator.com/item?id=48755005 Points: 241 # Comments: 41 37 Hacker News — AI on Front Page community 1mo ago Weave Robotics launches Isaac 1, a $7,999 home robot with Fall 2026 deliveries https://runtimewire.com/article/weave-robotics-isaac-1-home-... Comments URL: https://news.ycombinator.com/item?id=48750989 Points: 203 # Comments: 288 8 Hugging Face Daily Papers research 1mo ago Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly? Abstract A reinforcement learning framework called Play2Perfect enables sample-efficient robotic assembly tasks by first learning general manipulation skills through playful interaction with diverse objects, then adapting these skills for precise assembly through fine-tuning.… 34 Hugging Face Daily Papers research 1mo ago Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing Abstract A large-scale video editing dataset and model are introduced that support multi-task and structural manipulations through advanced data synthesis and network architectures. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Existing instruction-based video editing datasets… 38 Hugging Face Daily Papers research 1mo ago Scenes as Objects, Not Primitives: Instance-Structured 3D Tokenization from Unposed Views Abstract A feed-forward framework decomposes 3D scenes into instance-structured token groups from multi-view images, enabling direct object-level reconstruction, segmentation, and manipulation without 3D annotations. Generated by Qwen/Qwen2.5-Coder-32B-Instruct A 3D scene is… 38 arXiv — Machine Learning research 1mo ago Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation arXiv:2606.31043v1 Announce Type: new Abstract: Residual reinforcement learning adapts a pretrained robot policy by learning an additive correction to its actions. While effective when adaptation amounts to shifting the base policy's action distribution, additive corrections… 26 arXiv — NLP / Computation & Language research 1mo ago ViTL: Temporal Logic-Guided Zero-Shot Natural Language Navigation via Vision-Language Models arXiv:2606.30696v1 Announce Type: cross Abstract: Enabling robots to follow natural language commands to complete zero-shot long-horizon tasks remains challenging. It requires extracting implicit temporal and logical constraints from natural language commands and executing… 4 arXiv — NLP / Computation & Language research 1mo ago RCT: A Robot-Collected Touch-Vision-Language Dataset for Tactile Generalization arXiv:2606.31694v1 Announce Type: cross Abstract: For robots manipulating open-world objects, tactile representations must generalize to unseen materials. We introduce RCT (Robotic Contact Tactile), a robot-collected touch-vision-language dataset with 29,279 tactile frames from… 18 Hugging Face Daily Papers research 1mo ago Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models? Abstract Research reveals that language backbones in Vision-Language-Action models are highly redundant for robotic manipulation tasks, while vision and action pathways are more critical, suggesting need for deliberate capacity allocation in future architectures. Generated by… 11 Hugging Face Daily Papers research 1mo ago Learning Transferable Dynamics Priors from Action to World Modeling Abstract Action-conditioned world modeling enables transferable dynamics priors for robot learning through pretraining on large-scale manipulation data, supporting both simulator-based policy evaluation and video-action prediction. Generated by Qwen/Qwen2.5-Coder-32B-Instruct We… 27 arXiv — Machine Learning research 1mo ago A Linear Matching Bandit Approach to Online Multi-Human Multi-Robot Teaming arXiv:2606.29221v1 Announce Type: new Abstract: We address the problem of online multi-human multi-robot teaming through the lens of a linear matching bandit framework, where a learner assigns robots with unknown features from a fixed pool to distinct sets of human agents over… 15 Ars Technica — AI news-outlet 1mo ago South Korea to spend $1T on more memory chip production and humanoid robots South Korea targets physical AI lead and commercial humanoid robots by 2028. 9 r/MachineLearning community 1mo ago I do historical swordfighting and noticed AI struggles to track it. I’m building an open dataset to help fix this. Does my schema make sense? [P] Hi everyone, I’m a historical swordfighter (HEMA practitioner), and while I’m not a computer vision engineer or a roboticist, I’ve been reading a lot about the current bottlenecks in embodied AI, specifically around the Sim2Real gap and thin-object tracking. It occurred to me… 18 TechCrunch — AI news-outlet 1mo ago Robot hand company settles Tesla trade secret suit and announces $11M raise Jay Li doesn’t recommend getting sued by Tesla if you’re trying to get a startup off the ground. But he does think his company, Proception, might be better off for having endured the experience. “I think it’s kind of like a resilience test, or pressure… 15 Import AI (Jack Clark) community 1mo ago Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era What eras bookend our interregnum? 36 arXiv — Machine Learning research 1mo ago Support-Constrained RL Enables Real-World Policy Improvement without Real-World Experience arXiv:2606.27475v1 Announce Type: cross Abstract: Robots trained on real world data tend to be imprecise, slow, and brittle to perturbations. Improving these policies with reinforcement learning (RL) is an appealing alternative, but this process often requires expensive training… 28 arXiv — Machine Learning research 1mo ago Physics-Guided Robotic Radiation Source Localization along Arbitrary Measurement Paths in Unstructured Environments arXiv:2606.27624v1 Announce Type: cross Abstract: Using robots to estimate the location of the radiation source is an effective way to improve efficiency and safety. Existing methods focus on planning the robot's path to achieve precise estimation, typically approaching the… 19 MIT News — AI research 1mo ago LLMs help robots understand vague instructions and focus on key details To help robots do chores in places like homes and factories, a new approach from MIT uses one language model to clarify users’ instructions, then another to ignore irrelevant info. 19 arXiv — Machine Learning research 1mo ago Revisiting Action Factorization for Complex Action Spaces arXiv:2606.26574v1 Announce Type: new Abstract: Many real-world control problems involve hybrid discrete-continuous action spaces. For example, steering and signaling in autonomous driving, and aiming and firing in robotics or video-games. Despite real-world hybrid factorization… 10 arXiv — NLP / Computation & Language research 1mo ago Charting the Growth of Social-Physical HRI (spHRI): A Systematic Review Pipeline Augmented by Small Language Models arXiv:2606.26382v1 Announce Type: new Abstract: Social-physical human-robot interaction (spHRI) has grown rapidly across robotics, human-computer interaction, human-robot interaction, and haptics. Yet, fragmented terminology and inconsistent methodologies make systematic… 35 Hugging Face Daily Papers research 1mo ago In-Context World Modeling for Robotic Control Abstract ICWM enables robot policies to infer system variables from self-generated interactions, allowing adaptation to novel configurations without parameter updates by treating system identification as an in-context adaptation problem. Generated by… 8 arXiv — NLP / Computation & Language research 1mo ago RAVEN: Long-Horizon Reasoning & Navigation with a Visuo-Spatio-Temporal Memory arXiv:2606.25206v1 Announce Type: cross Abstract: Long-term robot deployment requires a compact and scalable memory that preserves fine-grained visual semantics, grounds observations in space and time, and enables efficient storage and retrieval. In this paper, we propose RAVEN,… 21 Hugging Face Daily Papers research 1mo ago EBench: Elemental Diagnosis of Generalist Mobile Manipulation Policies Abstract EBench is a comprehensive simulation benchmark for evaluating generalist mobile manipulation policies across diverse tasks and dimensions, revealing distinct capability profiles and generalization patterns among state-of-the-art models. Generated by… 18 Hugging Face Daily Papers research 1mo ago InSight: Self-Guided Skill Acquisition via Steerable VLAs Abstract InSight enables autonomous skill acquisition for vision-language-action models through primitive-action level steerability and automated demonstration generation. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Vision-language-action (VLA) models can learn manipulation… 19 TechCrunch — AI news-outlet 1mo ago Agility Robotics plans to go public via SPAC in a $2.5B deal Agility Robotics, the humanoid robotics startup that spun out of Oregon State University in 2015, expects to generate $620 million in proceeds. 13 NVIDIA Developer Blog official-blog 1mo ago Accelerating BEV Pooling on NVIDIA GPUs for Physical AI Applications An increasingly common design pattern for autonomous vehicles (AVs), robotics, and spatial AI systems is bird's-eye-view (BEV) perception. BEV models project... 31 Hugging Face Daily Papers research 1mo ago EventVLA: Event-Driven Visual Evidence Memory for Long-Horizon Vision-Language-Action Policies Abstract EventVLA addresses long-horizon robotic manipulation challenges by introducing a sparse visual evidence memory framework with visual anchors and dynamic Keyframe Evidence Memory module for improved task performance. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Memory… 23 Hugging Face Daily Papers research 1mo ago World Value Models for Robotic Manipulation Abstract World Value Model combines world models with value estimation to provide accurate task progression assessment and improve robotic policy learning from mixed-quality data. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Generalist value models play a pivotal role in scaling… 6 arXiv — Machine Learning research 1mo ago Verifiable Foundation Models for Robot Safety arXiv:2606.23754v1 Announce Type: cross Abstract: Deploying foundation models for robot control raises a central challenge: the expressive power that enables rich, multimodal perception also makes these models opaque and difficult to analyze formally, rendering them intractable… 4 arXiv — Machine Learning research 1mo ago RE4: Transformation-aware Imitation of Object Interactions Using Manipulation Modes arXiv:2606.24403v1 Announce Type: cross Abstract: Object interaction tasks have been a focus of advances in imitation learning. End-to-end methods, dominated by diffusion and flow-based variants have shown leaps in performance while sacrificing interpretability. Object-centric… 23 Hugging Face Daily Papers research 1mo ago ShotcreteDepth: A Bi-modal Dataset for Robust Robotic Depth Perception in Shotcrete Construction Environments Abstract A bi-modal construction domain dataset combining stereo RGB and LiDAR data under challenging environmental conditions is introduced for autonomous system perception research. Generated by Qwen/Qwen2.5-Coder-32B-Instruct We introduce ShotcreteDepth, a bi-modal dataset… 22 Hugging Face Daily Papers research 1mo ago Foresight: Failure Detection for Long-Horizon Robotic Manipulation with Action-Conditioned World Model Latents Abstract A failure detection framework for long-horizon robotic tasks uses action-conditioned world models and functional conformal prediction to monitor manipulation trajectories with only final task labels. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Long-horizon tasks are… 8 Hugging Face Daily Papers research 1mo ago PoLAR: Factorizing Extent and Mode in Latent Actions for Robot Policy Learning Abstract PoLAR introduces a geometrically structured latent action representation in hyperbolic space that separates transition extent from transition mode, improving robotic policy learning performance. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Latent action pretraining… 12 MIT News — AI research 1mo ago New chip could help tiny robots traverse complex environments Researchers combined an efficient algorithm with dedicated hardware to rapidly generate 3D maps for navigation using minimal memory and power. 10 Page 4 of 8 · 360 articles ← Newer Older →