Skip to content

DAILY INTELLIGENCE · AI APP + VLA

研究日報研究日报

AI App 精選與 VLA 論文評級的每日合輯AI App 精选与 VLA 论文评级的每日合辑下方匯總近 7 天的信號總覽,歸檔索引可按領域過濾下方汇总近 7 天的信号总览,归档索引可按领域过滤

近 7 日匯總近 7 日汇总 08-30 → 09-10
43 AI精選AI精选
155 VLA論文VLA论文
⚡ 0 突破
🔧 19 工具/技術工具/技术
📖 136 背景/觀點背景/观点
近 3 天內容近 3 天内容 09-08 → 09-10
2026-09-10 最新 41
GPT-6 Astra 登陆 Amazon Bedrock AI GPT-Image-2.5 发布:多轮指令遵循 + 参考图主体保留 AI Vercel Eve 推出 Persistent Memory AI Vercel 全计划免费启用 Production Protect AI OpenAI 宣布攻克纳维-斯托克斯千禧年难题 AI Physics filtering favors the generalization of robot learning VLA Aerial tactile perching via an anthropomorphic hand with embodied soft tactile receptors VLA ZEST: Zero-shot embodied skill transfer for athletic robot control VLA SONIC: Supersizing motion tracking for natural humanoid whole-body control VLA Learning vision-driven reactive soccer skills for humanoid robots VLA Robot in a crib: How a playing robot helps us understand sensorimotor contingency learning VLA BeyondMimic: From motion tracking to versatile humanoid control via guided diffusion VLA Convergent binocular stereo: Depth perception for humanoid robot vision VLA DeCAL: Towards Physically-Grounded Dexterous Vision-Language-Action Models via Contact-Aware Latent Co-Imagination VLA 3DWay: Generalizing Robot Manipulation via 3D Consistent Waypoints VLA GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation VLA CR-VLA-Force: Learning Control-aware Compliance VLA Model for Robust Contact-rich Robotic Manipulation VLA A4A: Cross-Embodiment Transfer of Action-Oriented 4D Affordances from Human Demonstrations VLA GIF: Agentic Generation of Interactive and Functional Object Compositions for Robot Learning VLA A Brain-inspired Hierarchical Framework for Zero-Shot Robot Task Reasoning and Execution VLA Where Success Breaks: Failure-Boundary Learning for Robust Vision-Language-Action Models VLA RefGuard: Identity-Aware Language-Guided Robot Manipulation via Joint Target-Anchor-Frame Grounding VLA GloVLA: Let Geometry Move and Local VLA Interact for Robust Object-Centric Manipulation in Unstructured Environments VLA IM-ENGINE: Image Editing for Embodied Data Generation VLA VLA-Corrector: Stage-Aware Observable State Understanding for Prompt-Based Closed-Loop Recovery of Vision-Language-Action Policies VLA MemCorr-DP: Counterfactual Correspondence Conditioning for a Diffusion Policy Guided by a Reference VLA ContextFlow: In-Context Flow Matching for Robot Manipulation VLA MEMOBench: A Process Level Memory Benchmark for Robotic Manipulation VLA Large Discrete Policy: Advancing Explicit Behavior Modeling with Stochastic Iterative Scoring VLA RoboDreamer: Anticipatory Humanoid Locomotion with Predictive State-Space Models VLA Measuring Language Transfer in Robot Policies: Adding Greek to a Cosmos3 Vision-Language-Action Policy VLA CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements VLA ICI-VLA: In-Context Imitation with Spatiotemporally Aligned Demonstrations for Vision-Language-Action Models VLA ComVLA: Communication-Aware Split Inference for VLA Models in 6G-Connected Robotics VLA M3-Tele: A Unified Multimodal Teleoperational Framework for Compliant Whole-Body Mobile Manipulation VLA Monkey See, Can Monkey Do? A Benchmark for Evaluating Robot Skill Learning by Observation VLA RoboCousin: Build Your Own Simulation Playground for Robust Bimanual Robotic Manipulation VLA Localized Visual Feature Aggregation via Focus Pooling for Visuomotor Policies VLA CASD: Chunk-Aligned Semantic Distillation for Multi-StageRobot Manipulation VLA BIFTA: Brain-Inspired Few-Shot Tactile Adaptation for Unknown Sensors VLA TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model VLA
2026-09-08 2日前 41
英伟达拟斥 30 亿美元投穆拉蒂创办的 Thinking Machines Lab AI HuggingFace 发布 funes:给 Coding Agent 装上持久记忆层 AI Simon Willison:OpenAI 研究加速的内部视角 AI OpenAI「AI研究实习生」正式入职,黄仁勋:AGI已来 AI GPT-6 Astra 零失误通关「我不是机器人」验证码 AI 一个人,4 个岗位,20 天:Cursor + Codex 上线微信小游戏 AI Physics filtering favors the generalization of robot learning VLA Aerial tactile perching via an anthropomorphic hand with embodied soft tactile receptors VLA ZEST: Zero-shot embodied skill transfer for athletic robot control VLA SONIC: Supersizing motion tracking for natural humanoid whole-body control VLA Learning vision-driven reactive soccer skills for humanoid robots VLA Robot in a crib: How a playing robot helps us understand sensorimotor contingency learning VLA BeyondMimic: From motion tracking to versatile humanoid control via guided diffusion VLA Convergent binocular stereo: Depth perception for humanoid robot vision VLA Evolution of humanoid locomotion control VLA From acrobatics to generality: Humanoid robots at an inflection point VLA FailureSpot: Label-Efficient Timestamp-Level Failure Detection for Vision-Language-Action Models VLA VLA-Precision: Asymmetric Co-Bootstrapping for Efficient Real-World Online RL of Vision-Language-Action Models VLA Continual Field-Adaptive Models (CFAMs) for Post-Deployment Physical AI VLA Dressing in Motion: A Human Motion-Aware Diffusion Policy for Robot-Assisted Dressing VLA Reasoning Without Inference Cost: Latent Semantic Scaffolding for Robot VLA Policies VLA LIBERO-RECOVER: Beyond Task Success Towards Failure Recovery in Robotic Manipulation Models VLA Temporal Tactile Encoding and Compliance for Intent-Aware Robot-to-Human Bimanual Handover VLA RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks? VLA Towards Neuro-Symbolic Procedural Reasoning for Long-Horizon Vision-Language-Action Manipulation VLA What Matters, When? Diagnosing and Improving Conditional Visual Grounding in Visuomotor Imitation Policies VLA CoFreeVLA: Short-Horizon Collision-Free Dual-Arm Manipulation via Vision-Language-Action Model and Risk Estimation VLA RedVLA: Physical Red Teaming for Vision-Language-Action Models VLA Rapid On-Robot Learning for Dynamic Manipulation Skills: Robot Juggling VLA Knowing When to Stop: Adaptive Action Chunking via Internal Cross-Attention Dynamics in VLAs VLA FWBC-VLA: Force-Aware Whole-Body Compensation for Contact-Rich Loco-Manipulation VLA Evaluating Uncertainty and Quality of Vision-Language-Action-enabled Robots VLA SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-Based Humanoid Control VLA RL-VLA$^3$: A Flexible and Asynchronous Reinforcement Learning Framework for VLA Training VLA CF-VLA: Efficient Coarse-to-Fine Action Generation for Vision-Language-Action Policies VLA TacPAC: Tactile Prediction and Real-Time Action Correction in World-Action Models for Contact-Rich Manipulation VLA Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics VLA Persistent Robot World Models: Stabilizing Multi-Step Rollouts via Reinforcement Learning VLA Improving Weak World Models Behind Strong Agents in Atari Pong VLA TourPhysics: Bringing Physics to World Models for Exploration and Manipulation from a Single Image VLA Spectral-Target Physical Latent Structuring for JEPA-Style World Models VLA

歸檔索引归档索引

83 期
— 2026 年 9 月 —
AI 最新
GPT-6 Astra 登陆 Amazon Bedrock
5 篇
VLA 最新
🔧 5 📖 31 Physics filtering favors the generalization of robot learning
36 篇
AI 昨日
OpenAI 用未发布模型攻克纳维-斯托克斯千禧年难题
6 篇
VLA 昨日
📖 9 Physics filtering favors the generalization of robot learning
9 篇
AI 2日前
英伟达拟斥 30 亿美元投穆拉蒂创办的 Thinking Machines Lab
6 篇
VLA 2日前
🔧 5 📖 30 Physics filtering favors the generalization of robot learning
35 篇
AI
索尼华纳集体起诉Anthropic,AI版权战进入总攻阶段
6 篇
VLA
📖 10 Physics filtering favors the generalization of robot learning
10 篇
AI
Altman 自曝超级果粉,为苹果起诉 OpenAI 感到难过
6 篇
VLA
📖 10 Physics filtering favors the generalization of robot learning
10 篇
AI
GPT-6 Astra 正式发布:OpenAI 新一代旗舰,全面超越 Claude Fable 5.1
8 篇
VLA
🔧 3 📖 37 Physics filtering favors the generalization of robot learning
40 篇
VLA
🔧 6 📖 9 GIFT: Guided Intermediate Feature Training via Action-Oriented Structural Supervision for Robotic Manipulation
15 篇
VLA
🔧 2 📖 7 RoboTok: An Internet-Scale Data Engine for Human Demonstration Retrieval and Dexterous Manipulation Learning
9 篇
VLA
🔧 2 📖 8 Does Imitation Learning Preserve Temporal Robustness in Dexterous Manipulation? An Expert-Learner Comparison Across Task Execution Speeds
10 篇
VLA
🔧 2 📖 13 IMPACT: Attention Is the Interaction Map for Scalable Interaction-Aware World Model Training
15 篇
— 2026 年 8 月 —
AI
OpenAI 宣布断供 Cursor:SpaceX 收购触发控制权变更条款
6 篇
VLA
📖 5 ZEST: Zero-shot embodied skill transfer for athletic robot control
5 篇
AI
智谱 GLM-5.3-Flash 发布:320B MoE 价格仅为 DeepSeek 1/40,终结其 56 天霸榜
6 篇
VLA
🔧 5 📖 25 ZEST: Zero-shot embodied skill transfer for athletic robot control
30 篇
AI
Anthropic 发布 Model Hardware Standard (MHS) 研究预览:AI Agent 标准化操控物理设备
5 篇
VLA
🔧 5 📖 23 $R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning
28 篇
AI
Paul Dix:AI 写了 100 万行代码并持续打磨到可靠,正在百万开发者机器上运行
6 篇
VLA
🔧 5 📖 33 ZEST: Zero-shot embodied skill transfer for athletic robot control
38 篇
AI
OpenAI 首颗自研推理芯片 Jalapeño 实测:吞吐+能效领先
6 篇
VLA
🔧 5 📖 25 ZEST: Zero-shot embodied skill transfer for athletic robot control
30 篇
AI
LLM 可通过推理引擎漏洞控制宿主机
6 篇
VLA
🔧 5 📖 15 Logic-VLA: A Temporal Logic Conditioned Vision-Language-Action Model
20 篇
AI
Anthropic Opus 5 降智风波:跑分向上、体感向下
8 篇
AI
Linus Torvalds 公开承认 AI 帮助调试 Linux 内核驱动
4 篇
VLA
🔧 1 📖 2 Aerial tactile perching via an anthropomorphic hand with embodied soft tactile receptors
3 篇
AI
OpenAI 全面开源 Codex Harness:Agent 循环引擎 Apache 2.0 许可
5 篇
VLA
🔧 9 📖 22 SONIC: Supersizing motion tracking for natural humanoid whole-body control
31 篇
AI
Mojo🔥 1.0 正式开源,编译器与工具链 Apache 2.0 许可
7 篇
VLA
⚡ 1 🔧 10 📖 24 SONIC: Supersizing motion tracking for natural humanoid whole-body control
35 篇
AI
Claude 自主设计蛋白质:15 靶标命中 14 个,成功率超人类专家 2-3 倍
7 篇
VLA
⚡ 1 🔧 24 📖 13 SONIC: Supersizing motion tracking for natural humanoid whole-body control
38 篇
AI
Mojo 1.0 正式发布并全面开源(Apache 2.0)
6 篇
VLA
⚡ 1 🔧 19 📖 11 SONIC: Supersizing motion tracking for natural humanoid whole-body control
31 篇
AI
Anthropic 年化营收超650亿美元,2万亿美元估值冲刺IPO
6 篇
VLA
🔧 8 📖 11 SONIC: Supersizing motion tracking for natural humanoid whole-body control
19 篇
AI
OpenAI IPO前大换血:12位高管离职,安全团队撤编
6 篇
VLA
📖 2 SONIC: Supersizing motion tracking for natural humanoid whole-body control
2 篇
AI
Anthropic Risk Report 曝光内部最强模型 Model 2,AI 研发速度逼近人类研究员
5 篇
VLA
🔧 1 📖 2 Learning contact representations in real-world clutter for universal robotic grasping
3 篇
AI
Qwen3.8-27B 正式发布:原生视觉+Agent增强,27B紧凑密度模型
5 篇
VLA
⚡ 1 🔧 10 📖 21 RoboSynChallenge: Mastering Real-World Dexterity via Generalizing Synthesized Manipulation Skills
32 篇
VLA
⚡ 1 🔧 7 📖 21 Redistribution-based Cost Inference Improves Sparse Safe Offline RL
29 篇
AI
DeepSeek V4 Pro 0813 正式发布,支持 Responses API
6 篇
VLA
🔧 7 📖 37 Real-World Cooperative Bimanual Dexterous Grasp of Large Objects from Single-View Observations
44 篇
AI
Claude 挑战黎曼猜想:从 41.6% 推进到 67.2% 临界线下界
5 篇
VLA
🔧 18 📖 8 SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning
26 篇
AI
Meta 发布 Muse Glimmer 30B:Apache 2.0 本地多模态 Agent 模型
7 篇
VLA
⚡ 1 🔧 10 📖 19 Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection
30 篇
VLA
📖 1 Energy-Guided Flow Matching
1 篇
VLA
⚡ 1 🔧 19 📖 18 VLAff: Vision-Language-Affordance Model for Unified Actionable Affordances
38 篇
VLA
🔧 15 📖 9 Kitchen Robotic Manipulation utilizing Foundation Models
24 篇
VLA
🔧 11 📖 14 ValueFormer: A Causal Transformer Value Function with Stage-Aware Labels for Semi-Autonomous Vision-Language-Action Policies
25 篇
VLA
🔧 19 📖 8 Towards General Language-Conditioned Latent Safety Filters
27 篇
VLA
🔧 5 📖 14 ActFovea: Runtime Safeguarding for VLA Policies via Spatiotemporal Visual-Action Consistency
19 篇
VLA
🔧 1 📖 2 Security of World-Model-Based Embodied AI: A Lifecycle of Threats, Defenses, and Evaluation
3 篇
VLA
⚡ 1 🔧 8 📖 28 It's Not Just More Demos: Counterfactual Action Sensitivity Coverage for Data-Efficient Robust Robot Imitation
37 篇
— 2026 年 7 月 —
VLA
🔧 10 📖 14 CG-World: A Large-Scale World-State Dataset and Protocol for World Models
24 篇
VLA
🔧 9 📖 16 DeVA: Decoupled Video-Action Model with physical guidance for robot policy learning
25 篇
VLA
⚡ 1 🔧 9 📖 16 DeVA: Decoupled Video-Action Model with physical guidance for robot policy learning
26 篇
VLA
⚡ 1 🔧 8 📖 16 Agile perceptive multiskill locomotion for quadrupedal robots in the wild
25 篇
VLA
📖 1 Offline RL with Hierarchical Action Chunking
1 篇
VLA
🔧 6 📖 6 PhysCoRe: Physics-Corrected Residual World Models for Material-Aware Deformable Dynamics
12 篇
VLA
⚡ 1 🔧 12 📖 11 NavVerse: Benchmarking Indoor-to-Outdoor Embodied Navigation in Continuous Robot Simulation
24 篇
VLA
🔧 13 📖 11 STeP: Signal Temporal Logic for Precise Specifications for Action Generation with Vision Language Models
24 篇
VLA
🔧 10 📖 16 ConceptTree: Bringing Semantic Transparency to Black-Box Decision Making for Robotic Manipulation
26 篇
VLA
⚡ 1 🔧 4 📖 17 Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories
22 篇
VLA
⚡ 1 🔧 9 📖 17 Open-AoE: An Open Egocentric Manipulation Dataset and Toolchain for Embodied Learning
27 篇
VLA
🔧 15 📖 17 HRIBench: Benchmarking Interaction-Centric Human-Robot Collaboration
32 篇
VLA
🔧 9 📖 17 FlowWAM: Optical Flow as a Unified Action Representation for World Action Models
26 篇
VLA
🔧 10 📖 19 EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos
29 篇
VLA
🔧 10 📖 15 FlowDAgger: Human-in-the-Loop Adaptation of Generative Robot Policies in Latent Space
25 篇
VLA
⚡ 1 🔧 11 📖 18 LEEVLA: Seeing What Matters in Latent Environment Evolution for Vision-Language-Action
30 篇
VLA
⚡ 1 🔧 11 📖 18 GeoProp: Grounding Robot State in Vision for Generalist Manipulation
30 篇
VLA
🔧 9 📖 23 Learning 4D Geometric Priors for Inference-Efficient World Action Models
32 篇
VLA
🔧 8 📖 20 Simple-to-Complex Structured Demonstrations for Vision-Language-Action Learning
28 篇
VLA
📖 1 High-resolution real-time mechanochromic tactile sensors
1 篇
VLA
📖 1 High-resolution real-time mechanochromic tactile sensors
1 篇