![]() | Monter d'un niveau |
Ahmadi, S., Awal, R., Sikarwar, A., Kazemnejad, A., Luo, G. Y., Rodriguez, J., Rajeswar Mudumba, S., Reddy, S., Pal, C. J., Krojer, B., & Agrawal, A. (décembre 2025). The Promise of RL for Autoregressive Image Editing [Communication écrite]. 39th Conference on Neural Information Processing Systems (NeurIPS 2025), San Diego, CA, USA. Publié dans Advances in Neural Information Processing Systems, 38. Lien externe
Masry, A., Rodriguez, J., Zhang, T., Wang, S., Wang, C., Feizi, A., Kalkunte Suresh, A., Puri, A., Jian, X., Noël, P.-A., Tejaswi Madhusudhan, S., Pedersoli, M., Liu, B., Chapados, N., Bengio, Y., Hoque, E., Pal, C. J., Laradji, I., Vázquez, D., ... Rajeswar Mudumba, S. (décembre 2025). AlignVLM: Bridging Vision and Language Latent Spaces for Multimodal Document Understanding [Communication écrite]. 39th Conference on Neural Information Processing Systems (NeurIPS 2025), San Diego, CA, USA. Publié dans Advances in Neural Information Processing Systems, 38. Lien externe
Sharma, A., Dalmia, A., Kazemi, M., Zouaq, A., & Pal, C. J. (avril 2025). GeoCoder: Solving Geometry Problems by Generating Modular Code through Vision-Language Models [Communication écrite]. Findings of the Association for Computational Linguistics: NAACL 2025, Albuquerque, New Mexico. Lien externe