Overview

To make use of diverse kinds of knowledge, AI needs to combine approaches suited to the nature of that knowledge: learning it during training, retrieving it from external sources when needed, and updating it as circumstances change.

We study large language models (LLMs) that flexibly handle diverse kinds of knowledge and search agents that autonomously find and integrate the information they need. Through this research, we aim to develop AI that draws on knowledge to reason, plan, and act autonomously.

Search Agents

We study search agents that autonomously alternate between search and reasoning, drawing on diverse sources such as the web, academic literature, and internal organizational documents. Our goal is to develop AI that searches for the information it needs based on the model’s knowledge and search results, integrates information while considering its reliability and recency, and provides answers grounded in evidence.

We also work on using information across languages and from sources containing figures, tables, and mathematical expressions, improving search efficiency, and developing evaluation datasets and methods to measure these capabilities. We aim to create environments where people and AI can collaborate on research and problem solving, drawing on human knowledge and judgment.

Selected Achievements

As foundational retrieval technologies, we proposed BPR, a retrieval model with reduced memory requirements, and KPR, a retrieval model that can incorporate additional entity knowledge without retraining. For the 2025 NeurIPS MMU-RAG Competition, we developed a search agent that repeatedly searches and reasons, integrating multiple sources to generate detailed answers. The agent won the static evaluation in the open-source division of the Text-to-Text track.

We also contribute to question answering evaluation and competition organization. Yamada co-organized the MIA Workshop at NAACL 2022, which hosted a competition on open-retrieval question answering across 16 languages. At the EMM-QA Workshop at ICML 2026, Yamada served as a co-organizer and also helped organize a competition on multimodal question answering using text and images.

Our results in international competitions include:

Photo from the award ceremony of the NeurIPS 2025 MMU-RAG Competition
NeurIPS 2025 MMU-RAG award ceremony
Quiz champions competing against AI at NIPS 2017
Quiz champions vs. AI at NIPS 2017

Related Papers

Large Language Models (LLMs)

We aim to develop large language models (LLMs) that can flexibly handle diverse kinds of knowledge, including advanced domain expertise, knowledge specific to organizations or individuals, and knowledge that requires frequent updates. Our research includes, for example, methods for learning knowledge efficiently, updating and correcting it while preserving existing capabilities, and applying learned knowledge to a variety of tasks and situations.

We also study the use and transfer of knowledge across languages. For example, we develop methods for applying knowledge learned in one language to another and analyze the factors behind performance differences across languages.

Selected Achievements

We have proposed LUKE, a language model that uses entity information; mLUKE, a multilingual model; and LEIA, a method for facilitating cross-lingual knowledge transfer in language models. The LUKE paper has been cited more than 1,000 times, and the model is also included in Hugging Face Transformers. Our foundational work on knowledge representations supporting these efforts includes Wikipedia2Vec, which learns vector representations of words and entities from Wikipedia; NTEE, which embeds texts and entities in a shared vector space; and EASE, which learns sentence representations using information about related entities.

Related Papers

Putting Research into Practice

We work on applications in academic research and industry to put our findings into practice. We aim to support activities such as reviewing academic literature, organizing claims and evidence across studies, and conducting technical investigations and R&D using organizational documents and advanced domain knowledge.

Alongside our papers, we release models, code, and evaluation datasets to make our research broadly accessible and reusable. Students participate in research projects throughout the full research process, from implementation, experimentation, and evaluation to publishing papers and releasing research outputs.