NewEvery arXiv paper, its researchers & institutions — mapped.
papers

Publications (27)

cs.LG2018

Rapid Adaptation with Conditionally Shifted Neurons

Tsendsuren Munkhdalai, Xingdi Yuan, Soroush Mehri +1

cs.CL2017

Reasoning with Memory Augmented Neural Networks for Language Comprehension

Tsendsuren Munkhdalai, Hong Yu

cs.CL2021

Diverse Distributions of Self-Supervised Tasks for Meta-Learning in NLP

Trapit Bansal, Karthick Gunasekaran, Tong Wang +2

cs.CL2024

Deferred NAM: Low-latency Top-K Context Injection via Deferred Context Encoding for Non-Streaming ASR

Zelin Wu, Gan Song, Christopher Li +9

cs.CL2017

Neural Tree Indexers for Text Understanding

Tsendsuren Munkhdalai, Hong Yu

cs.CL2024

Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Tsendsuren Munkhdalai, Manaal Faruqui, Siddharth Gopal

cs.LG2021

Learning Associative Inference Using Fast Weight Memory

Imanol Schlag, Tsendsuren Munkhdalai, Jürgen Schmidhuber

cs.NE2019

Metalearned Neural Memory

Tsendsuren Munkhdalai, Alessandro Sordoni, Tong Wang +1

cs.CL2020

Self-Supervised Meta-Learning for Few-Shot Natural Language Classification Tasks

Trapit Bansal, Rishikesh Jha, Tsendsuren Munkhdalai +1

eess.AS2023

Improving Speech Recognition for African American English With Audio Classification

Shefali Garg, Zhouyuan Huo, Khe Chai Sim +11

cs.CL2024

Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Gemini Team, Petko Georgiev, Ving Ian Lei +1132

cs.NE2020

Sparse Meta Networks for Sequential Adaptation and its Application to Adaptive Language Modelling

Tsendsuren Munkhdalai

cs.NE2018

Metalearning with Hebbian Fast Weights

Tsendsuren Munkhdalai, Adam Trischler

cs.CL2023

Contextual Biasing with the Knuth-Morris-Pratt Matching Algorithm

Weiran Wang, Zelin Wu, Diamantino Caseiro +10

cs.CL2020

Exploring and Predicting Transferability across NLP Tasks

Tu Vu, Tong Wang, Tsendsuren Munkhdalai +5

cs.LG2017

Meta Networks

Tsendsuren Munkhdalai, Hong Yu

cs.CL2018

Building Dynamic Knowledge Graphs from Text using Machine Reading Comprehension

Rajarshi Das, Tsendsuren Munkhdalai, Xingdi Yuan +2

cs.CL2025

Multi-Turn Puzzles: Evaluating Interactive Reasoning and Strategic Dialogue in LLMs

Kartikeya Badola, Jonathan Simon, Arian Hosseini +7

cs.LG2024

What Matters for Model Merging at Scale?

Prateek Yadav, Tu Vu, Jonathan Lai +4

stat.ML2022

A Locally Adaptive Interpretable Regression

Lkhagvadorj Munkhdalai, Tsendsuren Munkhdalai, Keun Ho Ryu

eess.AS2021

Fast Contextual Adaptation with Neural Associative Memory for On-Device Personalized Speech Recognition

Tsendsuren Munkhdalai, Khe Chai Sim, Angad Chandorkar +4

cs.CL2025

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Gheorghe Comanici, Eric Bieber, Mike Schaekermann +3431

eess.AS2024

Hierarchical Recurrent Adapters for Efficient Multi-Task Adaptation of Large Speech Models

Tsendsuren Munkhdalai, Youzheng Chen, Khe Chai Sim +3

cs.CL2018

Understanding Deep Learning Performance through an Examination of Test Set Difficulty: A Psychometric Case Study

John P. Lalor, Hao Wu, Tsendsuren Munkhdalai +1

cs.CL2018

Sentence Simplification with Memory-Augmented Neural Networks

Tu Vu, Baotian Hu, Tsendsuren Munkhdalai +1

cs.LG2017

Neural Semantic Encoders

Tsendsuren Munkhdalai, Hong Yu

cs.CL2025

Do LLMs Really Need 10+ Thoughts for "Find the Time 1000 Days Later"? Towards Structural Understanding of LLM Overthinking

Xinliang Frederick Zhang, Anhad Mohananey, Alexandra Chronopoulou +5