Large models as the agent's brain
Our aim is to give agents a cognitive core that connects perception, language, reasoning, and decisions. Our current work develops two ingredients: efficient multimodal inference and spatially grounded understanding.


























