EMBODIED BRAIN

PhiAgent

把人类第一视角视频,转化为规模化具身训练数据

PhiAgent 是灵生科技的具身数据引擎,利用大规模人类第一视角操作视频,将人类操作示范转化为机器人运动、仿真轨迹与可视化视频,实现跨本体、跨动作的规模化合成数据生成,为具身模型训练提供可扩展的行为学习信号。

商务咨询 →

Like what you see?

Support PhiAgent with a GitHub Star. It is the project's public, verifiable like.

Like on GitHubOpen repository → click StarGitHub

Opens the matching GitHub repository in a new tab.

Action-conditioned bowl grasp comparison

Input: The same real first frame with two action conditions.
Output: A lift-and-carry-right future and a lift-up future.

Shadow hand and forearm replacement

Input: A 621-frame continuous human-hand gesture.
Output: A 24-DOF Shadow Hand and forearm retargeting.

Robot-Arm RGB-to-Depth Comparison

Left: Original RGB.
Right: PhiAgent relative depth.

Human-Hand RGB-to-Depth Comparison

Left: Original RGB.
Right: PhiAgent relative depth.

Three-hand confidence-routed comparison

Input: One human-hand source video.
Output: Frame-aligned human, silver robot-hand, and graphite robot-hand views.

Vendor-hand comparison

Input: The same source motion and scene.
Output: Sharpa, Wonik Allegro, and Shadow Hand variants.

Short human-hand to robot-arm conversion

Input: An 89-frame human-hand motion video.
Output: Synchronized silver and graphite robot-arm conversions.

Robot appearance comparison

Input: The same human motion video.
Output: Silver, graphite, and Sudo R1-style robot variants.

View source and reproducibility details on GitHub