WorldVLA: Towards Autoregressive Action World Model

24 pointsposted 2 days ago
by chrsw

2 Comments

blovescoffee

2 days ago

More details on the actual arch would be nice. It seems like this autoregressive action world model space is a local max until the JEPA work takes over.

gunalx

2 days ago

Dont see how we could not hack on jepa instead of any other input encoders on a transformers llm .