Preprint Attention Is All You Need Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, Illia Polosukhin 2025 Read Paper Attention Is All You Need Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, Illia Polosukhin 2025 Open access Multimodal Machine Learning Applications Natural Language Processing Techniques Topic Modeling
Preprint SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training arXiv (Cornell University) 2025 Read Paper SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training Chu, Tianzhe, Yuexiang Zhai, Jihan Yang, Shengbang Tong, Saining Xie, Dale Schuurmans, Quoc V. Le, Sergey Levine, Yi Ma arXiv (Cornell University) 2025 Open access Domain Adaptation and Few-Shot Learning Multimodal Machine Learning Applications Topic Modeling
Journal article The Illusion of Thinking SuperIntelligence - Robotics - Safety & Alignment 2025 Read Paper The Illusion of Thinking Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh, Maxwell Horton, Samy Bengio, Mehrdad Farajtabar SuperIntelligence - Robotics - Safety & Alignment 2025 Open access Multimodal Machine Learning Applications Natural Language Processing Techniques Topic Modeling
Preprint Tree of Thoughts: Deliberate Problem Solving with Large Language Models arXiv (Cornell University) 2023 Read Paper Tree of Thoughts: Deliberate Problem Solving with Large Language Models Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Thomas L. Griffiths, Yuan Cao, Karthik Narasimhan arXiv (Cornell University) 2023 Open access Multimodal Machine Learning Applications Natural Language Processing Techniques Topic Modeling
Preprint VL-JEPA: Joint Embedding Predictive Architecture for Vision-language arXiv (Cornell University) 2025 Read Paper VL-JEPA: Joint Embedding Predictive Architecture for Vision-language Delong Chen, Mustafa Shukor, Théo Moutakanni, Willy Chung, J. M. Yu, Tejaswi Kasarla, Bang, Yejin, Allen Bolourchi, Yann LeCun, Pascale Fung arXiv (Cornell University) 2025 Open access Advanced Image and Video Retrieval Techniques Domain Adaptation and Few-Shot Learning Multimodal Machine Learning Applications