佐藤研究室/菅野研究室
佐藤研究室/菅野研究室
佐藤 (洋) 研究室
菅野研究室
ニュース
発表文献
連絡先
リソース
内部ページ
日本語
English
Paper-Conference
Affordance-Guided Diffusion Prior for 3D Hand Reconstruction
Naru Suzuki
,
Takehiko Ohkawa
,
Tatsuro Banno
,
Jihyun Lee
,
Ryosuke Furuta
,
Yoichi Sato
The N-Body Problem: Parallel Execution from Single-Person Egocentric Video
Zhifan Zhu
,
Yifei Huang
,
Yoichi Sato
,
Dima Damen
Vinci2: Providing Proactive Assistance in Continuous Egocentric Videos
Sitong Gong
,
Tianyu Yan
,
Caixin Kang
,
Bo Zheng
,
Xiang Ruan
,
Huchuan Lu
,
Kaipeng Zhang
,
Yoichi Sato
,
Yifei Huang
Leveraging RGB Images for Pre-Training of Event-Based Hand Pose Estimation
This paper presents RPEP: RGB Pre-training for Event-based hand Pose estimation, the first framework that uses labeled RGB images and …
Ruicong Liu
,
Takehiko Ohkawa
,
Tze Ho Elden Tse
,
Mingfang Zhang
,
Angela Yao
,
Yoichi Sato
PDF
引用
DOI
Embodied Interaction with Large Language Models: A Spatio-Physical Approach for Engaging Non-technical Users
Despite their growing capabilities, large language models (LLMs) remain unfamiliar to many non-technical users. The true potential of …
Hiroto Fukuda
,
Wataru Kawabe
,
Yusuke Sugano
PDF
引用
DOI
Constrained Rotation Optimization: Revisiting Crop-Based Gaze Estimation
Appearance-based gaze estimation typically relies on face normalization to reduce appearance variability, but this requires costly and …
Riccardo Santambrogio
,
Jiawei Qin
,
Matteo Matteucci
,
Yusuke Sugano
PDF
引用
ソースコード
Prosthesis-Aware 3D Human Pose Estimation: A Dataset and Benchmark for RSP Users
Recovering 3D human body motion from video is important for applications such as rehabilitation assessment and sports performance …
Yilin Wen
,
Kechuan Dong
,
Fumiya Suginaka
,
Ken Endo
,
Yusuke Sugano
引用
プロジェクト
BioVITA: Biological Dataset, Model, and Benchmark for Visual-Textual-Acoustic Alignment
Understanding animal species from multimodal data poses an emerging challenge at the intersection of computer vision and ecology. While …
Risa Shinoda
,
Kaede Shiohara
,
Nakamasa Inoue
,
Kuniaki Saito
,
Hiroaki Santo
,
Fumio Okura
PDF
引用
プロジェクト
CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering
Cause-and-effect reasoning in video is a significant challenge for Vision-Language Models (VLMs), as it requires going beyond …
Mingfang Zhang
,
Jingjing Pan
,
Ashutosh Kumar
,
Rajat Saini
,
Mustafa Erdogan
,
Hsuan-Kung Yang
,
Caixin Kang
,
Yifei Huang
,
Yoichi Sato
,
Quan Kong
PDF
引用
HanDyVQA: A Video QA Benchmark for Fine-Grained Hand-Object Interaction Dynamics
Hand-object interaction (HOI) involves dynamics where human manipulations produce spatio-temporal effects on objects. However, existing …
Masatoshi Tateno
,
Gido Kato
,
Hirokatsu Kataoka
,
Yoichi Sato
,
Takuma Yagi
PDF
引用
プロジェクト
»
引用
×