VLA
Step2 Vision language Action (VLA) model - architecture
VLM에 무엇을 붙이면 robot policy가 되는가? RT-2의 action tokenization, OpenVLA의 구조와 디멘션, Behavior Cloning 데이터 모양, π₀의 Action Expert와 flow matching까지 VLA 구조를 한 흐름으로 정리한다.
Step1 Vision Language Action (VLA) Model - basic concepts
로봇이 커피를 옮기는 예시로 policy·dynamics model·planner, 보상과 가치, feedback loop, observation과 state, 입력·출력, 모방학습과 강화학습 데이터, 토큰·임베딩, π₀ 구조, WAM·WFM의 차이와 현재 업계 지도까지 VLA의 ...
No posts match your search.