東京大学 · 情報科学
Masahiro Suzuki教授の研究室は、深層生成モデルを用いたマルチモーダルな情報処理に注力しています。特に、テキストと画像の間で双方向に変換可能な「共同マルチモーダル変分オートエンコーダー(JMVAE)」の開発を通じて、異種データ間の高次概念を統合的に捉える表現学習を追求しています。また、トランスファーラーニングにおける属性ベースの学習手法の改良や、嚥下機能の評価に向けた筋電図を用いた生理的指標の分析など、医療・バイオインフォマティクス分野への応用も進めています。
Figures are computed from collected data and may differ slightly.
We investigate deep generative models that can exchange multiple modalities bi-directionally, e.g., generating images from corresponding texts and vice versa. Recently, some studies handle multiple modalities on deep generative models, such as variational autoencoders (VAEs). However, these models typically assume that modalities are forced to have a conditioned relation, i.e., we can only generate modalities in one direction. To achieve our objective, we should extract a joint representation th
Multimodal learning is a framework for building models that make predictions based on different types of modalities. Important challenges in multimodal learning are the inference of shared representations from arbitrary modalities and cross-modal generation via these representations; however, achieving this requires taking the heterogeneous nature of multimodal data into account. In recent years, deep generative models, i.e. generative models in which distributions are parameterized by deep neur
The adenovirus vector (AdV) can carry two transgenes in its genome, the therapeutic gene and a reporter gene, for example. The E3 insertion site has often been used for the expression of the second transgene. A transgene can be inserted at six different sites/orientations: E1, E3 and E4 sites, and right and left orientations. However, the best combination of the insertion sites and orientations as for the titers and the expression levels has not sufficiently been studied. We attempted to constru
The ability to fine-tune the movement of swallowing-related organs and change the swallowing pattern to fit the volume of a bolus, texture and the physical properties of the food to be swallowed is referred to as the swallowing reserve. In other words, it is the response capability of food swallowing to avoid choking and aspiration. Herein, we focus on the coordination of the suprahyoid and infrahyoid muscles activities, which are closely related to swallowing movement, as a first step to develo
We investigate deep generative models that can exchange multiple modalities bi-directionally, e.g., generating images from corresponding texts and vice versa. Recently, some studies handle multiple modalities on deep generative models, such as variational autoencoders (VAEs). However, these models typically assume that modalities are forced to have a conditioned relation, i.e., we can only generate modalities in one direction. To achieve our objective, we should extract a joint representation th
Machine learning is the basis of important advances in artificial intelligence. Unlike the general methods of machine learning, which use the same tasks for training and testing, the method of transfer learning uses different tasks to learn a new task. Among the various transfer learning algorithms in the literature, we focus on the attribute-based transfer learning. This algorithm realizes transfer learning by introducing attributes and transferring the results of training to another task with
This paper proposes and analyzes a methodology of forecasting movements of the analysts’ net income estimates and those of stock prices. We achieve this by applying natural language processing and neural networks in the context of analyst reports. In the pre-experiment, we applied our method to extract opinion sentences from the analyst report while classifying the remaining parts as non-opinion sentences. Then, we performed two additional experiments. First, we employed our proposed method for
Open papers in the app to read, cite, and organize with AI.