News / #robotics Tag Robotics 362 articles archived under #robotics · RSS Sign in to follow r/LocalLLaMA community 3mo ago AllenAI has been iterating on their MolmoAct2 models for robotics r/AllenAI is cooking with MolmoAct2, a 5B vision-language-action model for robot control. They keep releasing new fine-tunes on different kinds of robotics datasets, including (but not limited to, and they keep releasing new ones): https://huggingface.co/allenai/MolmoAct2-LIBERO… 31 r/LocalLLaMA community 3mo ago Built a fully offline suitcase robot around a Jetson Orin NX SUPER 16GB. Gemma 4 E4B, ~200ms cached TTFT, 30+ sensors, no WiFi/BT/cellular. He has opinions. Sparky runs entirely on the Jetson. Gemma 4 E4B at Q4_K_M via llama.cpp with q8_0 KV cache and flash attention. 12K context, native system role, sampler defaults from the model card. Cached TTFT around 200ms, sustained 14-15 tok/s. SenseVoiceSmall for STT, Piper for TTS with… 21 Hugging Face Daily Papers research 3mo ago Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding Abstract Multi-agent pathfinding solver enhanced with learnable communication module improves coordination and performance while maintaining scalability. AI-generated summary Multi-agent pathfinding (MAPF) is a widely used abstraction for multi-robot trajectory planning… 37 arXiv — Machine Learning research 3mo ago WarmPrior: Straightening Flow-Matching Policies with Temporal Priors arXiv:2605.13959v1 Announce Type: new Abstract: Generative policies based on diffusion and flow matching have become a dominant paradigm for visuomotor robotic control. We show that replacing the standard Gaussian source distribution with WarmPrior, a simple temporally grounded… 19 arXiv — Machine Learning research 3mo ago R2R2: Robust Representation for Intensive Experience Reuse via Redundancy Reduction in Self-Predictive Learning arXiv:2605.14026v1 Announce Type: new Abstract: For reinforcement learning in data-scarce domains like real-world robotics, intensive data reuse enhances efficiency but induces overfitting. While prior works focus on critic bias, representation-level instability in… 4 arXiv — NLP / Computation & Language research 3mo ago IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation arXiv:2605.14712v1 Announce Type: cross Abstract: Robot imitation data are often multimodal: similar visual-language observations may be followed by different action chunks because human demonstrators act with different short-horizon intents, task phases, or recent context.… 37 Hugging Face Daily Papers research 3mo ago IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation Abstract IntentVLA is a history-conditioned visual-language action framework that improves robot imitation learning stability by encoding short-horizon intents from visual observations, addressing challenges from partial observability and ambiguous observations. AI-generated… 21 Hugging Face Daily Papers research 3mo ago From Pixels to Concepts: Do Segmentation Models Understand What They Segment? Abstract CAFE is a new benchmark for evaluating concept-faithful segmentation in promptable models through attribute-level counterfactual manipulation, revealing that accurate mask prediction does not guarantee semantic grounding. AI-generated summary Segmentation is a… 18 arXiv — Machine Learning research 3mo ago Ergodic Trajectory Design by Learned Pushforward Maps: Provable Coverage via Conditional Flow Matching arXiv:2605.13063v1 Announce Type: new Abstract: Designing continuous trajectories whose time-averaged occupancy provably matches a prescribed spatial density (the \emph{ergodic coverage} problem) is central to UAV-assisted data collection and sensing, robotic exploration, and… 21 Hugging Face Daily Papers research 3mo ago RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data Abstract RoboEvolve combines vision-language and video generation models in a co-evolutionary framework to enable scalable robotic manipulation with improved data efficiency and continuous learning capabilities. AI-generated summary The scalability of robotic manipulation is… 29 Hugging Face Daily Papers research 3mo ago World Action Models: The Next Frontier in Embodied AI Abstract World Action Models unify predictive state modeling with action generation for embodied policy learning, forming a cohesive framework for understanding environment dynamics and action prediction. AI-generated summary Vision-Language-Action (VLA) models have achieved… 15 Hugging Face Daily Papers research 3mo ago World Model for Robot Learning: A Comprehensive Survey Abstract World models as predictive representations of environmental dynamics have become essential for robot learning, supporting policy learning, planning, and simulation across various embodied applications. AI-generated summary World models, which are predictive… 12 Page 8 of 8 · 362 articles ← Newer