国際会議
Satoshi Suzuki, Shin'ya Yamaguchi, Shoichiro Takeda, Taiga Yamane, Naoki Makishima, Naotaka Kawata, Mana Ihori, Tomohiro Tanaka, Shota Orihashi, Ryo Masumura, "Difference Vector Equalization for Robust Fine-tuning of Vision-Language Models", AAAI Conference on Artificial Intelligence (AAAI), 2026. (arXiv)
Ryo Masumura, Tomohiro Tanaka, Naoki Makishima, Mana Ihori, Shota Orihashi, Naotaka Kawata, Taiga Yamane, Satoshi Suzuki, Takafumi Moriya, "Phoneme Overlapping-Aware Pre-Training with External Text Resources for Multi-Talker ASR", IEEE Automatic Speech Recognition and Understanding Workshop (ASRU), 2025.
Mana Ihori, Taiga Yamane, Naotaka Kawata, Naoki Makishima, Tomohiro Tanaka, Satoshi Suzuki, Shota Orihashi, Ryo Masumura, "Few-shot Personalization via In-Context Learning for Speech Emotion Recognition based on Speech-Language Model", IEEE Automatic Speech Recognition and Understanding Workshop (ASRU), 2025.
Ryo Masumura, Shota Orihashi, Mana Ihori, Tomohiro Tanaka, Naoki Makishima, Taiga Yamane, Naotaka Kawata, Satoshi Suzuki, Taichi Katayama, "Joint Modeling of Big Five and HEXACO for Multimodal Apparent Personality-trait Recognition", Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), 2025.
Tomohiro Tanaka, Ryo Masumura, Naoki Makishima, Mana Ihori, Naotaka Kawata, Shota Orihashi, Satoshi Suzuki, Taiga Yamane, "Semi-Supervised End-to-End Speech-to-Text Translation with Joint Text-to-Text and Speech-to-Text Decoding", Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), 2025.
Taiga Yamane, Satoshi Suzuki, Ryo Masumura, Shota Orihashi, Tomohiro Tanaka, Mana Ihori, Naoki Makishima, Naotaka Kawata, "MSMVD: Exploiting Multi-scale Image Features via Multi-scale BEV Features for Multi-view Pedestrian Detection", The British Machine Vision Conference (BMVC), 2025. (arXiv)
Taiga Yamane, Ryo Masumura, Satoshi Suzuki, Shota Orihashi, "MVTrajecter: Multi-View Pedestrian Tracking with Trajectory Motion Cost and Trajectory Appearance Cost", International Conference on Computer Vision (ICCV), 2025. (arXiv)
Naoki Makishima, Naotaka Kawata, Taiga Yamane, Mana Ihori, Tomohiro Tanaka, Satoshi Suzuki, Shota Orihashi, Ryo Masumura, "SOMSRED-SVC: Sequential Output Modeling with Speaker Vector Constraints for Joint Multi-Talker Overlapped ASR and Speaker Diarization", Annual Conference of the International Speech Communication Association (INTERSPEECH), 2025.
Naoki Makishima, Naotaka Kawata, Taiga Yamane, Mana Ihori, Tomohiro Tanaka, Satoshi Suzuki, Shota Orihashi, Ryo Masumura, "Unified Audio-Visual Modeling for Recognizing Which Face Spoke When and What in Multi-Talker Overlapped Speech and Video", Annual Conference of the International Speech Communication Association (INTERSPEECH), 2025.
Naotaka Kawata, Shota Orihashi, Satoshi Suzuki, Tomohiro Tanaka, Mana Ihori, Naoki Maikishima, Taiga Yamane, Ryo Masumura, "Block Refinement Learning for Improving Early Exit in Autoregressive ASR", Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), 2024.
Taiga Yamane, Satoshi Suzuki, Ryo Masumura, Shotaro Tora, "MVAFormer: RGB-Based Multi-View Spatio-Temporal Action Recognition with Transformer", IEEE International Conference on Image Processing (ICIP), 2024. (arXiv) Best Industry Paper Award
Ryo Masumura, Naoki Makishima, Tomohiro Tanaka, Mana Ihori, Naotaka Kawata, Shota Orihashi, Kazutoshi Shinoda, Taiga Yamane, Saki Mizuno, Keita Suzuki, Satoshi Suzuki, Nobukatsu Hojo, Takafumi Moriya, Atsushi Ando, "Unified Multi-Talker ASR with and without Target-speaker Enrollment", Annual Conference of the International Speech Communication Association (INTERSPEECH), 2024.
Satoshi Suzuki, Taiga Yamane, Naoki Makishima, Keita Suzuki, Atsushi Ando, Ryo Masumura, "OnDA-DETR: Online Domain Adaptation for Detection Transformers with Self-Training Framework", IEEE International Conference on Image Processing (ICIP), 2023.
Ryo Masumura, Naoki Makishima, Taiga Yamane, Yoshihiko Yamazaki, Saki Mizuno, Mana Ihori, Mihiro Uchida, Keita Suzuki, Hiroshi Sato, Tomohiro Tanaka, Akihiko Takashima, Satoshi Suzuki, Takafumi Moriya, Nobukatsu Hojo, Atsushi Ando, "End-to-End Joint Target and Non-Target Speakers ASR", Annual Conference of the International Speech Communication Association (INTERSPEECH), 2023.
Tomohiro Tanaka, Ryo Masumura, Mana Ihori, Hiroshi Sato, Taiga Yamane, Takanori Ashihara, Kohei Matsuura, Takafumi Moriya, "Leveraging Language Embeddings for Cross-Lingual Self-Supervised Speech Representation Learning", IEEE International Conference on Acoustic, Speech, and Signal Processing (ICASSP), 2023.
国内会議
山根大河,増村亮,鈴木聡志,折橋翔太,"Trajectoryの外観情報とモーション情報を活用したエンドツーエンドマルチビュー歩行者追跡",情報論的学習理論ワークショップ(IBIS),2025.
山根大河,鈴木聡志,増村亮,折橋翔太,田中智大,庵愛,牧島直輝,河田尚孝,"マルチスケールBEV特徴を用いたマルチビュー歩行者検出",画像の認識・理解シンポジウム(MIRU),2025.
折橋翔太,山根大河,河田尚孝,牧島直輝,庵愛,鈴木聡志,田中智大,増村亮,"条件付きトークン系列生成を用いた静止画からのソーシャルグループアクティビティ認識",画像の認識・理解シンポジウム(MIRU),2025.
山根大河,鈴木聡志,増村亮,東羅翔太郎,"Same View AttentionとDifferent View Attentionを用いたマルチビュー時空間行動認識",画像の認識・理解シンポジウム(MIRU),2024.
折橋翔太,山根大河,河田尚孝,牧島直輝,庵愛,鈴木聡志,田中智大,増村亮,"タスク切り替えトークンを用いた顔画像に対する複数タスクのEnd-to-Endモデリング",画像の認識・理解シンポジウム(MIRU),2024.