Contents Science Lab

Nagoya University Graduate School of Informatics
Back to member list
Avatar image

Dr. Yasutomo Kawanishi

Visiting Professor (Ritsumeikan University)

E-Mail: yasutomo.kawanishi ( a t ) riken.jp

Personal website: https://yasutomo57jp.github.io/en/

Dr. Yasutomo Kawanishi received his BEng and MEng degrees in Engineering and a PhD degree in Informatics from Kyoto University, Japan, in 2006,2008, and 2012, respectively. He became an Post Doctoral Fellow at Kyoto University, Japan in 2012. He moved to Nagoya University, Japan as an Designated Assistant Professor in 2014. Since 2015, he has been an Assistant Professor at Nagoya University, Japan. His research interests are Pedestrian-centric Vision, which includes Pedestrian Detection, Tracking, and Retrieval, for surveillance and in-vehicle videos. He received the best paper award from SPC2009, and Young Researcher Award from IEEE ITS Society Nagoya Chapter. He is a member of IEICE and IEEE.

Recent publications

    MultiSensor-Home: Multi-modal multi-view dataset and benchmarks for action recognition in home environments
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    Pattern Recognition, 179(113810), pp.1-12, To be published in November 2026.
    MultiSensor-Home: Multi-modal multi-view dataset and benchmarks for action recognition in home environments
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    Pattern Recognition, 179(113810), pp.1-12, To be published in November 2026.
    TraRA: Trajectory-level Recognition Aggregation for Video Text Spotting in Urban Surveillance
    Duc Tri Tran, Trung Thanh Nguyen, Vijay John, Phi Le Nguyen, Yasutomo Kawanishi
    IEEE International Conference on Advanced Visual and Signal-Based Systems, pp.1-6, To be published in September 2026.
    TraRA: Trajectory-level Recognition Aggregation for Video Text Spotting in Urban Surveillance
    Duc Tri Tran, Trung Thanh Nguyen, Vijay John, Phi Le Nguyen, Yasutomo Kawanishi
    IEEE International Conference on Advanced Visual and Signal-Based Systems, pp.1-6, To be published in September 2026.
    Linear-time 3D Forest Point Cloud Segmentation with Geometry-guided Queries
    Trung Thanh Nguyen, Tuan-Anh Vu, Duc Viet Le, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide, Teja Katternborn
    Meeting on Image Recognition & Understanding, pp.1-5, To be published in August 2026.
    Linear-time 3D Forest Point Cloud Segmentation with Geometry-guided Queries
    Trung Thanh Nguyen, Tuan-Anh Vu, Duc Viet Le, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide, Teja Katternborn
    Meeting on Image Recognition & Understanding, pp.1-5, To be published in August 2026.
    Beyond Single-view: Cross-view Context Modeling for Multi-view Action Recognition
    Trung Thanh Nguyen, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    Meeting on Image Recognition & Understanding, pp.1-5, To be published in August 2026.
    Beyond Single-view: Cross-view Context Modeling for Multi-view Action Recognition
    Trung Thanh Nguyen, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    Meeting on Image Recognition & Understanding, pp.1-5, To be published in August 2026.
    View-aware Cross-modal Distillation for Multi-view Action Recognition
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    IEEE/CVF Winter Conference on Applications of Computer Vision, p.7769–7778, March 2026.
    View-aware Cross-modal Distillation for Multi-view Action Recognition
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    IEEE/CVF Winter Conference on Applications of Computer Vision, p.7769–7778, March 2026.
    Hierarchical Global-Local Fusion for One-stage Open-vocabulary Temporal Action Detection
    Trung Thanh Nguyen, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    ACM Transactions on Multimedia Computing, Communications, and Applications, 2(1), p.6_1–6_23, January 2026.
    Hierarchical Global-Local Fusion for One-stage Open-vocabulary Temporal Action Detection
    Trung Thanh Nguyen, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    ACM Transactions on Multimedia Computing, Communications, and Applications, 2(1), p.6_1–6_23, January 2026.
    Semantic Alignment on Action for Image Captioning
    Da Huo, Marc A Kastner, Takatsugu Hirayama, Takahiro Komamizu, Yasutomo Kawanishi, Ichiro Ide
    IEEE Access, 13(2025), p.199615–199629, November 2025.
    Semantic Alignment on Action for Image Captioning
    Da Huo, Marc A Kastner, Takatsugu Hirayama, Takahiro Komamizu, Yasutomo Kawanishi, Ichiro Ide
    IEEE Access, 13(2025), November 2025.
    Semantic Alignment on Action for Image Captioning
    Da Huo, Marc A Kastner, Takatsugu Hirayama, Takahiro Komamizu, Yasutomo Kawanishi, Ichiro Ide
    IEEE Access, 13(2025), p.199615–199629, November 2025.
    IntentVC 2025: The ACM Multimedia Grand Challenge on Intention-Oriented Controllable Video Captioning
    Takahiro Komamizu, Marc A. Kastner, Yasutomo Kawanishi, Trung Thanh Nguyen, Junan Chen
    ACM International Conference on Multimedia, p.13813–13814, October 2025.
    IntentVC 2025: The ACM Multimedia Grand Challenge on Intention-Oriented Controllable Video Captioning
    Takahiro Komamizu, Marc A. Kastner, Yasutomo Kawanishi, Trung Thanh Nguyen, Junan Chen
    ACM International Conference on Multimedia, p.13813–13814, October 2025.
    Analyzing the Visual Variety of Adjectives based on Clustering of Visual Features
    Yui Tanaka, Marc A. Kastner, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    MUWS ‘25: Proceedings of the 4th International Workshop on Multimodal Human Understanding for the Web and Social Media, October 2025.
    Analyzing the Visual Variety of Adjectives based on Clustering of Visual Features
    Yui Tanaka, Marc A. Kastner, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    MUWS ‘25: Proceedings of the 4th International Workshop on Multimodal Human Understanding for the Web and Social Media, October 2025.
    Exploring Unknown Image Generation for Zero Shot Learning via Diffusion Models
    Lei Xiang, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    Unknown Journal, IS3-079, August 2025.
    Exploring Unknown Image Generation for Zero Shot Learning via Diffusion Models
    Lei Xiang, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    Unknown Journal, IS3-079, August 2025.
    MultiSensor-Home: Benchmark for Multi-modal Multi-view Action Recognition in Home Environments
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    Unknown Journal, IS3-038, August 2025.
    MultiSensor-Home: Benchmark for Multi-modal Multi-view Action Recognition in Home Environments
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    Unknown Journal, IS3-038, August 2025.
    Image Retrieval based on Editable Scene Graph with Contrastive Representation Learning
    PHAM Dinh Duy, Itthisak PHUEAKSRI, Marc A. Kastner, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    Unknown Journal, IS2-097, July 2025.
    Image Retrieval based on Editable Scene Graph with Contrastive Representation Learning
    PHAM Dinh Duy, Itthisak PHUEAKSRI, Marc A. Kastner, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    Unknown Journal, IS2-097, July 2025.
    Action Selection Learning for Weakly Labeled Multi-modal Multi-view Action Recognition
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    ACM Transactions on Multimedia Computing, Communications, and Applications, June 2025.
    Action Selection Learning for Weakly Labeled Multi-modal Multi-view Action Recognition
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    ACM Transactions on Multimedia Computing, Communications, and Applications, June 2025.
    MultiSensor-Home: A Wide-area Multi-modal Multi-view Dataset for Action Recognition and Transformer-based Sensor Fusion
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    The 19th IEEE International Conference on Automatic Face and Gesture Recognition, No.137, pp.1-10, May 2025.
    MultiSensor-Home: A Wide-area Multi-modal Multi-view Dataset for Action Recognition and Transformer-based Sensor Fusion
    Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu, Ichiro Ide
    The 19th IEEE International Conference on Automatic Face and Gesture Recognition, No.137, pp.1-10, May 2025.
    Towards Visual Storytelling by Understanding Narrative Context through Scene-Graphs
    PHUEAKSRI Itthisak, Marc A. Kastner, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide
    International Conference on Multimedia Modeling (MMM), 15523 (4), pp.226-239, January 2025.


Last updated: 2026-07-28 00:17:04.402338309 +0000 UTC m=+3.163495743.