Theo Lang
theolang97
AI & ML interests
Multimodal learning, vision-language models, image captioning, cross-modal retrieval
Recent Activity
liked a model about 23 hours ago
era-temporary/eb_alfred_sft_single_step_reasoning_dataset_success_episodes_multimodal-lr1e-5-full-e1-bs-16 upvoted a paper about 23 hours ago
Show-Harness: Just a VLM Agent Can Play Robots upvoted a paper 2 days ago
Reason Through the Latent! Making Latent Visual Reasoning NecessaryOrganizations
None yet