Publications (122)
Poutine: Vision-Language-Trajectory Pre-Training and Reinforcement Learning Post-Training Enable Robust End-to-End Autonomous Driving
Luke Rowe, Rodrigue de Schaetzen, Roger Girgis +2
Recombinator Networks: Learning Coarse-to-Fine Feature Aggregation
Sina Honari, Jason Yosinski, Pascal Vincent +1
Using Graph Algorithms to Pretrain Graph Completion Transformers
Jonathan Pilault, Michael Galkin, Bahare Fatemi +3
AR-DAE: Towards Unbiased Neural Entropy Gradient Estimation
Jae Hyun Lim, Aaron Courville, Christopher Pal +1
Exploring validation metrics for offline model-based optimisation with diffusion models
Christopher Beckham, Alexandre Piche, David Vazquez +1
Parallel-mentoring for Offline Model-based Optimization
Can Chen, Christopher Beckham, Zixuan Liu +2
Milo, a Fully Autonomous Indoor/Outdoor Robotic Guide Dog
Florian Golemo, Joanna Wolski, Joel Ruben Antony Moniz +1
Simple Video Generation using Neural ODEs
David Kanaa, Vikram Voleti, Samira Ebrahimi Kahou +1
CLAREL: Classification via retrieval loss for zero-shot learning
Boris N. Oreshkin, Negar Rostamzadeh, Pedro O. Pinheiro +1
Learning to Guide and to Be Guided in the Architect-Builder Problem
Paul Barde, Tristan Karch, Derek Nowrouzezahrai +3
On the impressive performance of randomly weighted encoders in summarization tasks
Jonathan Pilault, Jaehong Park, Christopher Pal
IntentGPT: Few-shot Intent Discovery with Large Language Models
Juan A. Rodriguez, Nicholas Botzer, David Vazquez +3
Movie Description
Anna Rohrbach, Atousa Torabi, Marcus Rohrbach +5
AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone?
Shambhavi Mishra, Gaurav Sahu, Marco Pedersoli +3
Overcoming challenges in leveraging GANs for few-shot data augmentation
Christopher Beckham, Issam Laradji, Pau Rodriguez +3
Multi-Resolution Continuous Normalizing Flows
Vikram Voleti, Chris Finlay, Adam Oberman +1
ExtremeWeather: A large-scale climate dataset for semi-supervised detection, localization, and understanding of extreme weather events
Evan Racah, Christopher Beckham, Tegan Maharaj +3
Generative Floor Plan Design with LLMs via Reinforcement Learning with Verifiable Rewards
Luis Lara, Aristides Milios, Zhi Hao Luo +5
InsightBench: Evaluating Business Analytics Agents Through Multi-Step Insight Generation
Gaurav Sahu, Abhay Puri, Juan Rodriguez +11
Do Neural Dialog Systems Use the Conversation History Effectively? An Empirical Study
Chinnadhurai Sankar, Sandeep Subramanian, Christopher Pal +2
XC-Cache: Cross-Attending to Cached Context for Efficient LLM Inference
João Monteiro, Ãtienne Marcotte, Pierre-André Noël +5
Does Entity Abstraction Help Generative Transformers Reason?
Nicolas Gontier, Siva Reddy, Christopher Pal
Neural Multisensory Scene Inference
Jae Hyun Lim, Pedro O. Pinheiro, Negar Rostamzadeh +2
SMPL-IK: Learned Morphology-Aware Inverse Kinematics for AI Driven Artistic Workflows
Vikram Voleti, Boris N. Oreshkin, Florent Bocquelet +3
Interactive Language Learning by Question Answering
Xingdi Yuan, Marc-Alexandre Cote, Jie Fu +4
Recurrent Transition Networks for Character Locomotion
Félix G. Harvey, Christopher Pal
Robust Motion In-betweening
Félix G. Harvey, Mike Yurick, Derek Nowrouzezahrai +1
The Alien Space of Science: Sampling Coherent but Cognitively Unavailable Research Directions
Alejandro H. Artiles, Martin Weiss, Levin Brinkmann +6
AgentAda: Skill-Adaptive Data Analytics for Tailored Insight Discovery
Amirhossein Abaskohi, Amrutha Varshini Ramesh, Shailesh Nanisetty +6
Convolutional Residual Memory Networks
Joel Moniz, Christopher Pal
AlignVLM: Bridging Vision and Language Latent Spaces for Multimodal Document Understanding
Ahmed Masry, Juan A. Rodriguez, Tianyu Zhang +19
Ctrl-Crash: Controllable Diffusion for Realistic Car Crashes
Anthony Gosselin, Ge Ya Luo, Luis Lara +5
From Machine Learning to Robotics: Challenges and Opportunities for Embodied Intelligence
Nicholas Roy, Ingmar Posner, Tim Barfoot +17
Generative Point Tracking with Flow Matching
Mattie Tesfaldet, Adam W. Harley, Konstantinos G. Derpanis +2
Improving GUI Grounding with Explicit Position-to-Coordinate Mapping
Suyuchen Wang, Tianyu Zhang, Ahmed Masry +4
Receptive Field Refinement for Convolutional Neural Networks Reliably Improves Predictive Performance
Mats L. Richter, Christopher Pal
Seq-VCR: Preventing Collapse in Intermediate Transformer Representations for Enhanced Reasoning
Md Rifat Arefin, Gopeshh Subbaraj, Nicolas Gontier +4
Rethinking Literature Search Evaluation: Deep Research Helps, and Human Citation Lists Are Not a Ground Truth
Gaurav Sahu, Laurent Charlin, Christopher Pal
Real-Time Reinforcement Learning
Simon Ramstedt, Christopher Pal
Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments
Luke Rowe, Roger Girgis, Anthony Gosselin +3
LitLLMs, LLMs for Literature Review: Are we there yet?
Shubham Agarwal, Gaurav Sahu, Abhay Puri +5
CtRL-Sim: Reactive and Controllable Driving Agents with Offline Reinforcement Learning
Luke Rowe, Roger Girgis, Anthony Gosselin +5
UI-Vision: A Desktop-centric GUI Benchmark for Visual Perception and Interaction
Shravan Nayak, Xiangru Jian, Kevin Qinghong Lin +11
Language Decision Transformers with Exponential Tilt for Interactive Text Environments
Nicolas Gontier, Pau Rodriguez, Issam Laradji +2
Are Diffusion Models Vision-And-Language Reasoners?
Benno Krojer, Elinor Poole-Dayan, Vikram Voleti +2
Unsupervised Depth Estimation, 3D Face Rotation and Replacement
Joel Ruben Antony Moniz, Christopher Beckham, Simon Rajotte +2
Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics
Jishnu Sethumadhavan Nair, Patrice Bechard, Rishabh Maheshwary +14
Neural Coherence : Find higher performance to out-of-distribution tasks from few samples
Simon Guiroy, Mats Richter, Sarath Chandar +1
Robust Guided Diffusion for Offline Black-Box Optimization
Can Sam Chen, Christopher Beckham, Zixuan Liu +2
Capture the Flag: Uncovering Data Insights with Large Language Models
Issam Laradji, Perouz Taslakian, Sai Rajeswar +6
Reinforcement Learning with Random Delays
Simon Ramstedt, Yann Bouteiller, Giovanni Beltrame +2
Supervise Thyself: Examining Self-Supervised Representations in Interactive Environments
Evan Racah, Christopher Pal
Conditionally Adaptive Multi-Task Learning: Improving Transfer Learning in NLP Using Fewer Parameters & Less Data
Jonathan Pilault, Amine Elhattami, Christopher Pal
Constrained Group Relative Policy Optimization
Roger Girgis, Rodrigue de Schaetzen, Luke Rowe +3
Improving Generalization in Task-oriented Dialogues with Workflows and Action Plans
Stefania Raimondo, Christopher Pal, Xiaotian Liu +2
ParetoFlow: Guided Flows in Multi-Objective Optimization
Ye Yuan, Can Chen, Christopher Pal +1
Improving Meta-Learning Generalization with Activation-Based Early-Stopping
Simon Guiroy, Christopher Pal, Gonçalo Mordido +1
StarVector: Generating Scalable Vector Graphics Code from Images and Text
Juan A. Rodriguez, Abhay Puri, Shubham Agarwal +6
Goal-conditioned GFlowNets for Controllable Multi-Objective Molecular Design
Julien Roy, Pierre-Luc Bacon, Christopher Pal +1
ReviewerToo: Should AI Join The Program Committee? A Look At The Future of Peer Review
Gaurav Sahu, Hugo Larochelle, Laurent Charlin +1
WebMMU: A Benchmark for Multimodal Multilingual Website Understanding and Code Generation
Rabiul Awal, Mahsa Massoud, Aarash Feizi +10
Diffusion Large Language Models for Black-Box Optimization
Ye Yuan, Can, Chen +4
EmoNets: Multimodal deep learning approaches for emotion recognition in video
Samira Ebrahimi Kahou, Xavier Bouthillier, Pascal Lamblin +15
Learning Action and Reasoning-Centric Image Editing from Videos and Simulations
Benno Krojer, Dheeraj Vattikonda, Luis Lara +4
Learning to Summarize Long Texts with Memory Compression and Transfer
Jaehong Park, Jonathan Pilault, Christopher Pal
Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning
Julien Roy, Paul Barde, Félix G. Harvey +2
Rendering-Aware Reinforcement Learning for Vector Graphics Generation
Juan A. Rodriguez, Haotian Zhang, Abhay Puri +12
DStruct2Design: Data and Benchmarks for Data Structure Driven Generative Floor Plan Design
Zhi Hao Luo, Luis Lara, Ge Ya Luo +3
MCVD: Masked Conditional Video Diffusion for Prediction, Generation, and Interpolation
Vikram Voleti, Alexia Jolicoeur-Martineau, Christopher Pal
SelvaBox: A high-resolution dataset for tropical tree crown detection
Hugo Baudchon, Arthur Ouaknine, Martin Weiss +5
A Meta-Transfer Objective for Learning to Disentangle Causal Mechanisms
Yoshua Bengio, Tristan Deleu, Nasim Rahaman +5
BigCharts-R1: Enhanced Chart Reasoning with Visual Reinforcement Finetuning
Ahmed Masry, Abhay Puri, Masoud Hashemi +13
Score-based Denoising Diffusion with Non-Isotropic Gaussian Noise Models
Vikram Voleti, Christopher Pal, Adam Oberman
Theano: A Python framework for fast computation of mathematical expressions
The Theano Development Team, Rami Al-Rfou, Guillaume Alain +110
Direct Behavior Specification via Constrained Reinforcement Learning
Julien Roy, Roger Girgis, Joshua Romoff +2
Interactive Machine Comprehension with Information Seeking Agents
Xingdi Yuan, Jie Fu, Marc-Alexandre Cote +3
Bridging the Gap Between Target Networks and Functional Regularization
Alexandre Piché, Valentin Thomas, Rafael Pardinas +4
Conservative objective models are a special kind of contrastive divergence-based energy model
Christopher Beckham, Christopher Pal
Unimodal probability distributions for deep ordinal classification
Christopher Beckham, Christopher Pal
Action-Based Representation Learning for Autonomous Driving
Yi Xiao, Felipe Codevilla, Christopher Pal +1
Describing Videos by Exploiting Temporal Structure
Li Yao, Atousa Torabi, Kyunghyun Cho +4
Towards Learning to Imitate from a Single Video Demonstration
Glen Berseth, Florian Golemo, Christopher Pal
Latent Variable Sequential Set Transformers For Joint Multi-Agent Motion Prediction
Roger Girgis, Florian Golemo, Felipe Codevilla +5
On Extractive and Abstractive Neural Document Summarization with Transformer Language Models
Sandeep Subramanian, Raymond Li, Jonathan Pilault +1
BigDocs: An Open Dataset for Training Multimodal Models on Document and Code Tasks
Juan Rodriguez, Xiangru Jian, Siba Smarak Panigrahi +40
The Promise of RL for Autoregressive Image Editing
Saba Ahmadi, Rabiul Awal, Ankur Sikarwar +8
DRBench: A Realistic Benchmark for Enterprise Deep Research
Amirhossein Abaskohi, Tianyi Chen, Miguel Muñoz-Mármol +11
Visual Question Answering From Another Perspective: CLEVR Mental Rotation Tests
Christopher Beckham, Martin Weiss, Florian Golemo +3
RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content
Joao Monteiro, Pierre-Andre Noel, Etienne Marcotte +6
Adversarial Soft Advantage Fitting: Imitation Learning without Policy Optimization
Paul Barde, Julien Roy, Wonseok Jeon +3
On Adversarial Mixup Resynthesis
Christopher Beckham, Sina Honari, Vikas Verma +5
Distilling Specialized Orders for Visual Generation
Rishav Pramanik, Amin Sghaier, Masih Aminbeidokhti +6
Score-based Diffusion Models in Function Space
Jae Hyun Lim, Nikola B. Kovachki, Ricardo Baptista +10
Mem-: Adaptive Memory through Learning When and What to Generate
Xiaoqiang Wang, Chao Wang, Hadi Nekoei +5
VectorGym: A Multitask Benchmark for SVG Code Generation, Sketching, and Editing
Juan Rodriguez, Haotian Zhang, Abhay Puri +13
The Liver Tumor Segmentation Benchmark (LiTS)
Patrick Bilic, Patrick Christ, Hongwei Bran Li +106
Recurrent Semi-supervised Classification and Constrained Adversarial Generation with Motion Capture Data
Félix G. Harvey, Julien Roy, David Kanaa +1
Reinforcement Learning for Blind Stair Climbing with Legged and Wheeled-Legged Robots
Simon Chamorro, Victor Klemm, Miguel de la Iglesia Valls +2
Grounding Computer Use Agents on Human Demonstrations
Aarash Feizi, Shravan Nayak, Xiangru Jian +14
A simple squared-error reformulation for ordinal classification
Christopher Beckham, Christopher Pal