Rosenberg, G., Stadhard, S., Hansen, B. C., & Greene, M. R. (2026). The Limits of Learning from Pictures and Text: Vision-Language Models and Embodied Scene Understanding.
Chicago Style (17th ed.) CitationRosenberg, Gillian, Skylar Stadhard, Bruce C. Hansen, and Michelle R. Greene. The Limits of Learning from Pictures and Text: Vision-Language Models and Embodied Scene Understanding. 2026.
MLA (9th ed.) CitationRosenberg, Gillian, et al. The Limits of Learning from Pictures and Text: Vision-Language Models and Embodied Scene Understanding. 2026.
Warning: These citations may not always be 100% accurate.