ComBodied Agents: a New Paradigm of Human-Centric Agentic AI Paper • 2608.10915 • Published 5 days ago • 188
Uncertainty-Aware World Model for Aerial Image-Goal Navigation Paper • 2608.05597 • Published 10 days ago • 10
YOLO-PEFT: Parameter-Efficient Fine-Tuning on YOLO Family Paper • 2608.07051 • Published 9 days ago • 20
World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation Paper • 2608.05369 • Published 11 days ago • 26
SKILL-KD: Contrastive Skill Distillation for LLM Agents Paper • 2607.28048 • Published 12 days ago • 14
OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents Paper • 2608.05013 • Published 12 days ago • 35
SmartMage: Dynamic Modality Orchestration for 3D Scene Understanding Paper • 2608.05137 • Published 6 days ago • 27
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published 13 days ago • 171
InfiniSplat: Implicit Gaussian Decoding for Large-Baseline Monocular View Synthesis Paper • 2608.02437 • Published 13 days ago • 67
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published 21 days ago • 105
QQWorld: Quantile-Quantile Matching for World Model Regularization Paper • 2607.28415 • Published 17 days ago • 30
SULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based SUrface LANDmine Detection Under Domain Shift Paper • 2607.28996 • Published 16 days ago • 6
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 17 days ago • 302
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing Paper • 2607.19064 • Published 26 days ago • 77
UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models Paper • 2607.23373 • Published 22 days ago • 7
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Paper • 2607.24904 • Published 20 days ago • 37
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published 25 days ago • 193
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 25 days ago • 32