News / #robotics Tag Robotics 500 articles archived under #robotics · RSS Sign in to follow Hugging Face Daily Papers research 19d ago OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining Abstract OpenWAM factorizes world-action pretraining into modular components to identify key design principles, yielding a scalable open model with strong simulation and real-robot performance. Generated by thinkingmachines/Inkling-Small World-Action Models inherit world… 34 Hugging Face Daily Papers research 19d ago Measuring Language Transfer in Robot Policies: Adding Greek to a Cosmos3 Vision-Language-Action Policy Abstract Adding Greek to a robot vision-language-action model via machine-translated instructions reveals measurement pitfalls and shows bilingual training improves performance over monolingual baselines, though it remains far below English levels. Generated by… 35 Hugging Face Daily Papers research 19d ago RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks? Abstract RoboSPA is a large-scale robotic manipulation benchmark that evaluates vision-language-action models on fine-grained spatial reasoning and long-horizon procedural planning across progressively harder task variants. Generated by thinkingmachines/Inkling-Small… 25 Hugging Face Daily Papers research 19d ago CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements Abstract A two-stage framework predicts sparse gripper keyframes and continuous actions to transfer complex hand demonstrations to robotic grippers while reducing drift via kinematic optimization. Generated by thinkingmachines/Inkling-Small Transferring human hand demonstrations… 23 Hugging Face Daily Papers research 19d ago GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation Abstract GE-Act 2.0 is a world-action model trained from scratch with a control-oriented autoencoder, single-step visual planner, and inverse dynamics model, using knowledge-aligned selective optimization to enable scalable zero-shot robot manipulation across diverse skills and… 32 The Information — AI news-outlet 19d ago China Curbs Humanoid IPOs After Unitree’s Volatile Debut Chinese regulators are tightening the approval for humanoid startups seeking to go public after a volatile debut by industry leader Unitree Robotics, as they aim to cool down a sizzling sector rife with copycats but little real technological breakthroughs. The Chinese Securities… 32 The Information — AI news-outlet 19d ago China Curbs Humanoid IPOs After Unitree’s Volatile Debut Chinese regulators are tightening the approval for humanoid startups seeking to go public after a volatile debut by industry leader Unitree Robotics, as they aim to cool down a sizzling sector rife with copycats but little real technological breakthroughs. The Chinese Securities… 27 The Information — AI news-outlet 19d ago Hugging Face Is Making A Big Robotics Push Nvidia is paying a steep price for Hugging Face , a repository of open source models, primarily for its strategic position in the center of the open source AI world. But it turns out Nvidia is also getting a budding robot business that could help CEO Jensen Huang fulfill the… 15 Hugging Face Daily Papers research 20d ago EmbodiedSkills: A Unified Framework for Orchestrating, Training, and Deploying VLA Agents Abstract EmbodiedSkills proposes a unified framework that validates and verifies robot skill executions through a fixed interface, enabling closed-loop embodied agents with adaptable low-level vision-language-action policies. Generated by thinkingmachines/Inkling-Small… 8 r/LocalLLaMA community 20d ago I REALLY hope the new gemma 5 family sticks to the "chat model first" philsophy and doesn't fall into the Qwen trap It just seems every local 30b class model is just trying so hard to be the next Qwen that they all just kinda blend into a mass of code focused models. I really like how gemma 4 31b turned out with it feeling a lot less robotic and more creative than other models even knowing… 33 r/MachineLearning community 21d ago Roboticists working in Learning-from-Demonstrations and Behavioral Cloning : What is going on in your field these days? [D] Is LfD and BC research being effected by recent advances in (so-called) Frontier LLMs? Or is research in LfD and BC sort of going along in an independent direction from these? Are you seeing any use from ViTs or VLAs? Any other recent advances you would like to bring up?  … 28 arXiv — Machine Learning research 21d ago VLA-Precision: Asymmetric Co-Bootstrapping for Efficient Real-World Online RL of Vision-Language-Action Models arXiv:2609.04355v1 Announce Type: cross Abstract: Pretrained vision-language-action (VLA) models enable broad manipulation but remain unreliable in tasks demanding precision and repeatability. Applying real-world online reinforcement learning (RL) to VLA post-training enables… 19 arXiv — NLP / Computation & Language research 21d ago Same Trajectory, Contradictory Rewards (ROBORMBENCH): Paraphrase Fragility in Vision Language Reward Models arXiv:2609.05401v1 Announce Type: cross Abstract: Vision-language models are increasingly used as reward functions for robotic learning, but this role requires paraphrase invariance: the same trajectory should receive the same reward under semantically equivalent goal… 19 r/MachineLearning community 21d ago Reproducibility seems to be headed towards irrelevance in ML research. Is it too late? [D] I feel that reproducibility is now a lost cause in machine learning research for three reasons: Many research is moving towards the physical AI territory, where you need expensive hardwares or even entire laboratories with high-speed cameras, in order to perform an experiment.… 28 TechCrunch — AI news-outlet 21d ago Travis Kalanick’s Atoms might be getting into the robotaxi business The Uber founder has said that Atoms will allow him to complete "unfinished business." 21 Hacker News — AI on Front Page community 22d ago GPT-6 Astra on robot arms Article URL: https://openai.robocurve.org/gpt-6-astra/ Comments URL: https://news.ycombinator.com/item?id=49582582 Points: 204 # Comments: 155 30 Hugging Face Daily Papers research 23d ago RoboTok: An Internet-Scale Data Engine for Human Demonstration Retrieval and Dexterous Manipulation Learning Abstract RoboTok retrieves relevant human manipulation videos from the web using a latent motion space derived from 3D hand trajectories to improve robot policy training. Generated by thinkingmachines/Inkling-Small Robot learning increasingly depends on broad and diverse… 6 TechCrunch — AI news-outlet 23d ago XDOF, just three months out of stealth, is in talks for a Series B at a $1.2B valuation The round is being raised just months after the robot data startup exited from stealth. 6 Hugging Face Daily Papers research 25d ago Does Imitation Learning Preserve Temporal Robustness in Dexterous Manipulation? An Expert-Learner Comparison Across Task Execution Speeds Abstract Imitation-learned dexterous manipulation policies degrade more sharply than expert policies when execution speed increases, with insertion misalignment being the primary failure mode. Generated by thinkingmachines/Inkling-Small Dexterous manipulation policies learned by… 21 TechCrunch — AI news-outlet 25d ago TechCrunch Disrupt 2026’s new Real World AI Stage features Nvidia, robots, and extinct animals On our new Real World AI stage, we’ll be focusing on the intersection between the digital and physical, and all the ways we’ll continue to see a blending of the two. 23 arXiv — NLP / Computation & Language research 26d ago Exploring Collaboration between a language and a non-language agent arXiv:2609.00474v1 Announce Type: new Abstract: LLMs are increasingly deployed as orchestrators that coordinate specialized subagents to solve complex tasks through natural language. However, in many important domains like game playing and robotics, the strongest available… 17 Hugging Face Daily Papers research 26d ago ZimaBlue: Evolving Generalizable World Action Models through Scalable Video Pre-training Abstract ZimaBlue learns generalizable world action models from large-scale egocentric video via a three-stage curriculum and a slow-fast architecture, substantially improving zero-shot robotic manipulation. Generated by thinkingmachines/Inkling-Small Robotic manipulation faces… 12 Hugging Face Daily Papers research 27d ago LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation Abstract LightNav-0 is a compact generalist navigation model that leverages a pretrained vision-language model’s spatial reasoning via unified pointing tokens and action tokenization to achieve state-of-the-art embodied navigation across diverse tasks and robots. Generated by… 34 arXiv — Machine Learning research 27d ago Fully Distributed GNE Algorithms for Multi-Robot Placement without Consensus on Multipliers arXiv:2608.29388v1 Announce Type: new Abstract: Recent machine learning research has increasingly focused on equilibrium analysis in non-cooperative games rather than solely on optimal solutions. Many such problems involve shared constraints and can be formulated as Generalized… 8 arXiv — Machine Learning research 27d ago Asynchronous Cooperative Online Learning for Multi-Robot Control under Computational Delays arXiv:2608.29562v1 Announce Type: new Abstract: Ensuring the safe operation of multi-agent systems (MASs) under uncertain environments is crucial for cooperative robotic, where external disturbances and inaccurate dynamic models can significantly compromise performance and… 23 The Information — AI news-outlet 27d ago FTC Sues Amazon, Alleging Ad Price Manipulation The Federal Trade Commission and attorneys general from 22 states sued Amazon for overcharging its advertising customers by allegedly manipulating ad auctions. Advertisers promoting their products on Amazon’s marketplace buy ad spots where the price is set in auctions. In the… 35 The Information — AI news-outlet 27d ago Exclusive: Reframe Raises Funds to Bring Amazon Robotics Know-How to Home Building For all the talk of robots that look like humans, the next robots to move into people’s houses could look more like ducks. Sales of such a robot, made by open-source model platform Hugging Face , reached more than $2.5 million on its first day on Thursday. Microduck , which… 21 Hugging Face Daily Papers research 28d ago Act with Intent: Distilling Behavior Intent for Vision-Language-Action Models Abstract Vision-Language-Action (VLA) models can turn multimodal context into robot actions, but their action decoders are still trained largely by behavior cloning. This supervises which motor command was demonstrated while leaving implicit the local objective served by the… 29 arXiv — NLP / Computation & Language research 28d ago When Robots Mishear Us: Mapping the Safety Risks of Voice-Controlled Embodied AI arXiv:2608.28518v1 Announce Type: cross Abstract: We investigate whether automatic speech recognition (ASR) errors in user input can lead to unsafe outputs from Embodied AI (EAI) models. We find that ASR errors can lead to harmful instructions being accepted and executed by EAI… 38 TechCrunch — AI news-outlet 28d ago The U.S. is building barriers around drones and robots, but China has scale to get around them The U.S. is shutting out more foreign-made drones and robots. China’s scale means the global competition may simply move elsewhere. 11 Hugging Face Daily Papers research 28d ago Beyond Data Scaling: Representation-Centric Continued Pre-training for Vision-Language-Action Models Abstract VLAct improves vision-language-action model performance by pre-training on diverse robot data with preserved vision-language priors and shared action semantics, achieving strong results across simulations and unseen embodiments with limited compute. Generated by… 32 Hugging Face Daily Papers research 28d ago PonderPounce: A Pretrained MLLM as an Episode Context Engine for Robot Control Abstract PonderPounce leverages native causal context in multimodal language models as robot episode memory, jointly training a reasoning System2 module and a fast System1 action model to improve long-horizon policy performance without dedicated memory architectures. Generated… 14 r/LocalLLaMA community 28d ago We used HFlow to evaluate the latest open weights VLMs for processing egocentric data We used HFlow to evaluate the latest open weights VLMs for processing egocentric data. This was based on Build AI's Egocentric-10k evaluation , which used Gemini 2.5 Flash to measure hand visibility and active manipulation. We kept the same prompts and the same dataset, only… 17 Ars Technica — AI news-outlet 29d ago Inside Meta’s push to put robots to work in data centers The company is testing robots on tasks that can performed by technicians. 28 r/MachineLearning community 1mo ago PhD Internship in smaller lab [D] How much of a disadvantage is it if your only internship is not at one of the big frontier labs when it comes to post-phd opportunities in robotics/ML? My PhD is at a top university (UK) and my internship is interesting and relevant but the team itself is smaller and it's no… 30 The Information — AI news-outlet 1mo ago Andreessen Horowitz Raises $1.1 Billion for AI Hardware Fund Andreessen Horowitz raised $1.1 billion for a fund that will invest in physical AI and infrastructure startups, the firm announced Friday. The money will be used to back companies building chips, memory, networking, storage and other companies involved in the AI buildout.… 15 Hugging Face Daily Papers research 1mo ago TacForcing: Streaming Action Generation with Execution-Time Tactile Feedback Abstract TacForcing is a streaming action-generation framework that integrates real-time tactile feedback during execution via a streaming action expert and execution-aware tactile attention, improving contact-rich manipulation. Generated by thinkingmachines/Inkling-Small… 9 arXiv — Machine Learning research 1mo ago Ultra Low-Power, Lightweight, Probabilistic RSS-Based Path Reconstruction: A System for Landscape-Scale Bee Tracking arXiv:2608.27152v1 Announce Type: new Abstract: Applications in fields such as movement ecology, Internet of Things or robotics share the need for systems that localize devices that are too small and power constrained to implement GNSS (Global Navigation Satellite Systems).… 37 arXiv — Machine Learning research 1mo ago Diffusion Policies for Short-Horizon Planning in Robot Crowd Navigation arXiv:2608.27158v1 Announce Type: new Abstract: Robot crowd navigation requires safe and efficient decision-making under dense, dynamic, and multimodal human--robot interactions. Existing reinforcement-learning methods typically output a single reactive action at each timestep,… 10 arXiv — Machine Learning research 1mo ago Making Latent Evolution Explicit: Operator-Structured Transitions for World Action Models arXiv:2608.27259v1 Announce Type: new Abstract: World Action Models (WAMs) augment robot policies by predicting how task-relevant scene states may evolve under interaction. Recent WAMs increasingly perform such prediction in latent representation spaces, avoiding full… 30 Hugging Face Daily Papers research 1mo ago Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Abstract Zero-WAM enables robotic manipulation of unseen tasks by conditioning a causal video-action model on in-context human video guidance, supported by an automatically generated dataset and a future-chunk prediction objective. Generated by thinkingmachines/Inkling-Small… 21 TechCrunch — AI news-outlet 1mo ago Hugging Face is selling a cute $399 open-source duck robot, Microduck Hugging Face is taking orders for the Microduck, a $399 tiny open-source duck robot that developers can train at home out of the box. 37 r/LocalLLaMA community 1mo ago Microduck by Pollen Robotics & Hugging Face Pollen Robotics and Hugging Face are releasing an open-source bipedal robot that comes with reinforcement learning software. It looks like it has a speaker, camera + LiDAR, NFC, Wifi, Bluetooth, etc. And roller-skates, because that's just the cutest thing ever. They announced it… 19 Hacker News — AI on Front Page community 1mo ago Pollen Robotics (Hugging Face) Microduck Article URL: https://pollen-robotics.com/microduck/ Comments URL: https://news.ycombinator.com/item?id=49462763 Points: 205 # Comments: 72 21 The Information — AI news-outlet 1mo ago SoftBank Explores Buying Majority Stake in 1X Humanoid Maker SoftBank is in talks to buy a majority stake in 1X Technologies, an OpenAI-backed humanoid robot developer, The Information reported late Wednesday . The investment would support SoftBank’s robotics ambition and give 1X more runway to put its soft-bodied bots in customers’… 20 arXiv — Machine Learning research 1mo ago Simultaneous inference of environmental and interaction forces in collective dynamics arXiv:2608.25181v1 Announce Type: new Abstract: Collective dynamics arise in a wide range of physical, biological, and engineering applications. Examples include cell migration, swarm robotics, social dynamics, and animal behavior. A defining characteristic of these systems is… 15 Hugging Face Daily Papers research 1mo ago StreamPI: Streaming Multimodal Temporal Modeling for Vision-Language-Action Models Abstract StreamPI enhances single-frame vision-language-action models with streaming temporal reasoning via instruction-anchored attention and randomized interval training, improving robot manipulation without extra parameters. Generated by thinkingmachines/Inkling-Small… 6 The Information — AI news-outlet 1mo ago SoftBank in Talks to Buy Majority Stake in Humanoid Maker 1X at $6 Billion Valuation SoftBank is in talks to buy a majority stake in 1X Technologies, an OpenAI-backed humanoid robot developer, in a deal that would value the startup at about $6 billion, according to people with knowledge of the deal. The investment would buttress SoftBank’s robotics ambition. The… 21 NVIDIA Developer Blog official-blog 1mo ago How to Train a Cross-Embodiment Robot Navigation Policy with AI Agents Navigation enables a robot to turn perception and motion into purposeful autonomy. Unlike locomotion, which produces stable movement, navigation must be used to... 7 TechCrunch — AI news-outlet 1mo ago Bill Gates wants to see a robot tax and ‘Human Reserved’ jobs to mitigate harms from AI Gates is mostly in the Responsible AI camp, but there are a few ideas in here we hadn't heard before. 30 Page 2 of 10 · 500 articles ← Newer Older →