News / #robotics Tag Robotics 500 articles archived under #robotics · RSS Sign in to follow arXiv — NLP / Computation & Language research 7h ago Sorry Robot, Happy Human: Vision-Language Models Read Only One of Two Legible Typographic Layers arXiv:2609.31403v1 Announce Type: new Abstract: Vision-language models (VLMs), despite their success in optical character recognition (OCR) tasks, are vulnerable to typographic attacks and have a fragile structure for images with multiple text layers. In this study, the… 24 Ars Technica — AI news-outlet 2d ago Tesla workers balk at training Optimus humanoid robots as replacements Despite challenges, Tesla aims for 1,000 Optimus robots per week by end of 2026. 35 The Information — AI news-outlet 2d ago Tesla’s Optimus Hits Snags in Hands, Suppliers as Scale-Up Begins Tesla has ramped up production of its Optimus humanoid robot roughly tenfold in recent months, but the company is still struggling to manufacture the machines reliably at scale. Its production lines are grappling with problems involving the robot’s intricate hands, automated… 22 arXiv — Machine Learning research 3d ago Learning from Mixed-Quality Deployment Experience for Robot Manipulation arXiv:2609.29000v1 Announce Type: new Abstract: Robot policies deployed in real environments naturally accumulate mixed-quality experience, including successful executions, partial progress, and failures. Although these rollouts provide valuable information for further learning,… 21 arXiv — NLP / Computation & Language research 3d ago Design and Evaluation of LLM Chaining-Based Task Planning for General Purpose Service Robots arXiv:2609.29043v1 Announce Type: cross Abstract: General Purpose Service Robot (GPSR) tasks, as defined in the RoboCup@Home benchmark, require robots to interpret diverse natural language commands and generate multi-step action sequences in real home environments. Conventional… 6 arXiv — Machine Learning research 4d ago Data-driven discrete-time deep recurrent neural network-based modeling for dissipative systems arXiv:2609.27186v1 Announce Type: new Abstract: Physical AI has gained increasing attention for its role in developing AI systems that better understand, predict, and control real-world dynamics. Achieving this requires AI models that not only achieve high prediction accuracy… 22 arXiv — NLP / Computation & Language research 4d ago Psychoacoustically Aligned Latent Smoothing for Adversarial Robustness of Full-Duplex Speech-to-Speech Dialogue Models arXiv:2609.27378v1 Announce Type: cross Abstract: End-to-end speech-to-speech dialogue models listen and speak simultaneously, so a continuously open acoustic channel is exposed to adversarial manipulation. We formalize imperceptible attacks on full-duplex agents as optimization… 22 Hugging Face official-blog 4d ago How to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning Workflows Back to Articles a]:hidden"> How to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning Workflows Enterprise + Article Published September 23, 2026 Upvote - Johnny Nuñez Cano johnnynv nvidia Asier Arranz asiernvidia nvidia Rishabh Chadha rchadha-nv nvidia… 9 r/LocalLLaMA community 4d ago BFL releases FLUX 3 Action: a 7B robot model read more: https://bfl.ai/models/flux-3-action   submitted by   /u/paf1138 [link]   [comments] 36 llama.cpp releases dev-tools 5d ago b11117 HIP : optimize IQ2/IQ3 ( __vsub4 __vcmpne4 ) using SWAR ( #27962 ) HIP : use bit manipulation for __vcmpne4 HIP : use bit manipulation for __vsub4 Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/49401157 macOS/iOS: macOS Apple Silicon… 11 r/LocalLLaMA community 5d ago To my surprise I found gemma4 much better at tool-calling than Qwen I've had fairly good luck getting off the ground coding at home with both qwen3.6 35B a3b, and also qwen3.8 27B. However once I switched from a chat window (where the robot wrote code blocks that I could copy/paste into a text editor) to a simple agentic loop, things got funny.… 4 Ars Technica — AI news-outlet 5d ago Toyota orders workers to train humanoid robots but says humans won't be replaced Toyota's push comes as automakers race to develop and deploy humanoid robots. 24 TechCrunch — AI news-outlet 5d ago TechCrunch Disrupt 2026: Aaron Edsinger brings Hello Robot’s Stretch 4 to life onstage Hello Robot CEO and co-founder Aaron Edsinger will bring Stretch 4 for a live demo on the Real World AI Stage at TechCrunch Disrupt 2026. Register before September 25 to save up to $200, plus get a second pass at 50% off. 9 NVIDIA Developer Blog official-blog 5d ago Accelerating a ROS 2 Node with an AI Agent and NVIDIA Isaac ROS GPU acceleration can speed up compute-intensive robotics workloads, but a fast CUDA kernel alone does not guarantee a fast ROS 2 graph. As messages move between... 28 arXiv — Machine Learning research 6d ago Prioritized Rollouts for Efficient World Model-based Vision-Language-Action Policy Optimization arXiv:2609.22879v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for embodied intelligence, but fine-tuning them with reinforcement learning (RL) remains constrained by the cost of real-world robot interaction. Model-based… 38 arXiv — NLP / Computation & Language research 6d ago The Role of AI in Online Reviews arXiv:2609.22198v1 Announce Type: new Abstract: The rapid adoption of large language models (LLMs) creates new opportunities for strategic content generation on online platforms, including potentially harmful forms of manipulation that may undermine platform effectiveness and… 11 arXiv — NLP / Computation & Language research 7d ago From Generation to Detection: Exploration of Discourse Driven Scenario based LLM Generated Fake News arXiv:2609.20838v1 Announce Type: new Abstract: In this study, we examine how modern LLMs generate and detect fake news under controlled settings across four manipulation scenarios. These are open-ended generation, rewriting, manipulation prompts and attribute based prompts… 11 arXiv — NLP / Computation & Language research 7d ago Do Personality-Tuned LLMs Make Better Social Agents? arXiv:2609.21857v1 Announce Type: new Abstract: LLMs are increasingly used in social simulations for socially interactive agents and robots, offering more flexibility than rule-based systems. However, even though they mimic human behaviour very well, there is a persistent… 18 TechCrunch — AI news-outlet 9d ago A startup that builds other startups raised $100M, and is all-in on physical AI UP.Labs, now doing business under the name Vantora, is building startups for industrial corporations. 35 r/LocalLLaMA community 9d ago NGL, I’m hyped to see if Qwen3.8 27b can make me a sandwich. Instant buy for me. Saw this little dude in a Forbes article ( https://www.forbes.com/sites/johnkoetsier/2026/08/18/american-humanoid-robot-launches-for-just-1688-delivery-this-fall-made-in-san-francisco/ ) This is definitely for the DIY researcher crowd who want to dip their toes into the robotics… 14 r/MachineLearning community 10d ago Future of general LLM work (interp/inference/alignment) vs agentic/physical AI (VLA, multimodal) for career [D] I'm at a crossroads with two grad school options that would take me in somewhat different research directions, and I wanted some general advice on these fields, their growth, and industry alignment. I'm leaving out the specifics of the programs since I'm tryna compare the… 5 Hugging Face Daily Papers research 11d ago In-Context Robot Learning with VLM Agents Abstract Enabling robots to adapt to unfamiliar environments as readily as humans remains a moonshot goal of embodied AI. No finite collection of demonstrations can cover every task and situation a robot will encounter, making the ability to learn from context at deployment… 24 TechCrunch — AI news-outlet 11d ago Iceland-based Treble raises $18 million for its voice simulation platform Treble's voice simulation platform is used by voice AI model developers, AI wearable, and robotics companies 34 arXiv — Machine Learning research 11d ago Changepoint-Aware World Models: Detecting Dynamics Shifts and Recovering by Forgetting Stale Replay in Model-Based RL arXiv:2609.18950v1 Announce Type: new Abstract: A robot's learned model of its own dynamics is only valid until those dynamics change: actuators wear, payloads shift, and joints stiffen. A model-based agent that keeps training as if nothing happened adapts slowly, dragged back… 26 Hugging Face Daily Papers research 11d ago EventEgoHands++: Event-based Egocentric 3D Hand Mesh Reconstruction with Real Dataset Abstract 3D hand mesh reconstruction is a challenging yet essential task for downstream applications, including human-robot interaction and AR/VR. Although conventional cameras have been widely adopted for this task, methods that rely on them struggle in low-light environments… 22 NVIDIA Developer Blog official-blog 11d ago How to Use AI Agents to Prepare 3D Scenes for Simulation Agentic AI workflows can be used to prepare and validate digital twins for physical AI systems. Agents can inspect 3D scenes, author simulation-relevant data in... 16 NVIDIA Developer Blog official-blog 11d ago TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor AI agents are moving from cloud data centers to vehicles, robots, and other edge devices. Unlike a chatbot that answers a single prompt, an agent works through... 14 TechCrunch — AI news-outlet 11d ago Robots are waiting for a ChatGPT moment: Nvidia’s Les Karpas explains why at TechCrunch Disrupt 2026 The robotics industry is still waiting for their breakthrough into day-to-day life. Nvidia's Les Karpas has an answer as to why at TechCrunch Disrupt 2026. Register before September 25 to save up to $200 on your pass. 9 arXiv — Machine Learning research 12d ago Autonomous Droplet Navigation via Model-Based Reinforcement Learning arXiv:2609.16369v1 Announce Type: new Abstract: Precise manipulation of liquid droplets underpins lab-on-a-chip platforms for diagnostics, chemical synthesis, and biological assays. Yet autonomous droplet transport through confined geometries of varying complexity remains an… 18 arXiv — Machine Learning research 12d ago MyoFlow: Anchor-Tied Rectified Flow for HD-sEMG Gesture Recognition Across Sessions and Subjects arXiv:2609.17194v1 Announce Type: new Abstract: High-density surface electromyography (HD-sEMG) gesture recognition supports prosthetic control, assistive robotics, and rehabilitation, but electrode re-donning and physiological variability cause distribution shifts that degrade… 21 Ars Technica — AI news-outlet 12d ago Agility’s new humanoid robot will stop, squat to avoid harming human coworkers Robots can start working outside physical cages and without safety barriers. 8 TechCrunch — AI news-outlet 12d ago Discover how to take your startup from prototype to production at TechCrunch Disrupt 2026 Learn how to scale your startup breakthrough from prototype to production at TechCrunch Disrupt 2026 with scaling leaders, Adrian Macneil (Foxglove), John Mackey (MBRYONICS), and Boris Sofman (Bedrock Robotics. Register before September 25 to save up to $200 on your pass. 4 Hugging Face Daily Papers research 13d ago Agent as Policy for Robotic Manipulation Abstract A general-purpose agent directly controls a physical robot by interpreting visuals, writing executable programs, and revising actions based on physical feedback across diverse manipulation tasks. Generated by thinkingmachines/Inkling-Small We demonstrate that a… 36 Ars Technica — AI news-outlet 13d ago Founder’s cost-cutting obsession drove Unitree lead in cheap humanoid robots Wang Xingxing micromanaged Unitree to success—will his leadership style scale? 9 arXiv — NLP / Computation & Language research 14d ago Agent as Policy for Robotic Manipulation arXiv:2609.12541v1 Announce Type: new Abstract: We demonstrate that a general-purpose agent can directly drive a physical robot throughout task execution without any task-specific or environment-specific training. We introduce Agent as Policy (AGP), which places task planning… 24 Hugging Face Daily Papers research 14d ago Breaking the Vision-Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models Abstract LIT improves robot action generalization by first training pose-conditioned action priors without images, then constraining visual inputs through a pose-supervised latent interface that preserves spatial goal information. Generated by thinkingmachines/Inkling-Small… 20 Simon Willison community 15d ago Quoting Paul Ford For a while, I must admit, it looked as if software developer roles like mine were done for. How could we fight against tireless robots? But our industry is slowly realizing that making truly cutting-edge software still requires humans to think and work together, to maximize… 23 Ars Technica — AI news-outlet 16d ago I spent $4,000 on a robot dog from China Unitree might be the world’s most important robotics company. 18 TechCrunch — AI news-outlet 16d ago Mecka AI nears $500M valuation in Sequoia-led deal amid rush for robot training data The round for the two-year-old startup is coming together months after Mecka announced its Series A. 32 Hacker News — AI on Front Page community 16d ago I spent $220 on Google app ads and 60% of the installs were robots Article URL: https://dayzlegame.com/blog/google-ads-bot-farm/ Comments URL: https://news.ycombinator.com/item?id=49662990 Points: 257 # Comments: 142 22 Hugging Face Daily Papers research 16d ago Adaptive Bridge: A Proxy-Based Decoupling Layer for Mitigating DDS Backpressure in ROS 2 Abstract A proxy layer isolates critical ROS 2 subscribers from degraded ones via topic splitting and dynamic rate control to eliminate DDS backpressure. Generated by thinkingmachines/Inkling-Small In systems built on Robot Operating System 2 (ROS 2) and using Data Distribution… 26 Hugging Face Daily Papers research 16d ago Memory as Plans: World-Action Modeling with Memory-Grounded Planning Abstract MaP-WAM improves non-Markovian robotic manipulation by separating memory-grounded planning from plan-conditioned execution, using compact episodic segment records and progress-calibrated action chunks to maintain fixed inference latency. Generated by… 35 arXiv — Machine Learning research 17d ago HuRo: Robotizing Human Videos for Scalable VLA Pretraining arXiv:2609.10706v1 Announce Type: cross Abstract: Human video datasets have emerged as a compelling alternative to expensive real-robot data, offering rich diversity at scale. To bridge the human-to-robot embodiment gap, existing approaches either robotize videos in task-matched… 5 Hugging Face Daily Papers research 17d ago SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem Abstract Large vision-language models trained on synthetic block-manipulation tasks improve 3D spatial reasoning and generalize to real-world visual tasks. Generated by thinkingmachines/Inkling-Small Large Vision-Language Models (LVLMs) have achieved strong performance on… 25 Hacker News — AI on Front Page community 17d ago Show HN: Bodily Oddities When I was about 11 years old, my best friend and I were playing during recess at school, and I was carrying him around on my back, presumably pretending to be a multipart attack robot. All of a sudden, my heart started hurting, and I collapsed to my knees, and the robot was no… 7 TechCrunch — AI news-outlet 17d ago Maven Robotics wants to steal your robot deployment deal Maven Robotics emerged from stealth today with a $100 million Series A and active deployments. 29 The Information — AI news-outlet 18d ago Chinese CEO Laments Many Robotics Firms Fabricate Revenue The co-founder and CEO of Mech-Mind Robotics, a Beijing-based robotics firm that recently went public, said in a post on WeChat on Thursday that many Chinese embodied AI companies are “creating false and unsustainable revenue,” in response to The Information’s scoop on Chinese… 34 Hugging Face Daily Papers research 18d ago SyncWorld: Visual Calibration Enables World Models as Zero-Shot Simulators Abstract SyncWorld is an action-conditioned world model that uses visual calibration episodes to learn environment-specific action-to-visual mappings, enabling zero-shot simulation and test-time policy improvement across unseen robotic settings. Generated by… 6 Hugging Face Daily Papers research 18d ago Show-Harness: Just a VLM Agent Can Play Robots Abstract Show-Harness links vision-language models to robot control via discrete semantic actions interpreted by embodiment-specific modules, enabling zero-shot and efficient fine-tuned deployment across robots and GUIs. Generated by thinkingmachines/Inkling-Small Foundation… 6 Hugging Face Daily Papers research 18d ago TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model Abstract TANGO is a vision-language framework that predicts whole-body joint actions for humanoid robots to navigate cluttered indoor environments using only simulated training data. Generated by thinkingmachines/Inkling-Small We study the problem of navigating cluttered indoor… 36 Page 1 of 10 · 500 articles Older →