Long-Term Temporally Consistent Unpaired Video Translation from Simulated Surgical 3D Data

被引:7
|
作者
Rivoir, Dominik [1 ,2 ]
Pfeiffer, Micha [1 ]
Docea, Reuben [1 ]
Kolbinger, Fiona [1 ,3 ]
Riediger, Carina [3 ]
Weitz, Juergen [2 ,3 ]
Speidel, Stefanie [1 ,2 ]
机构
[1] NCT UCC Dresden, Dresden, Germany
[2] Tech Univ Dresden, CeTI, Dresden, Germany
[3] Univ Hosp Dresden, Dresden, Germany
关键词
SLAM;
D O I
10.1109/ICCV48922.2021.00333
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Research in unpaired video translation has mainly focused on short-term temporal consistency by conditioning on neighboring frames. However for transfer from simulated to photorealistic sequences, available information on the underlying geometry offers potential for achieving global consistency across views. We propose a novel approach which combines unpaired image translation with neural rendering to transfer simulated to photorealistic surgical abdominal scenes. By introducing global learnable textures and a lighting-invariant view-consistency loss, our method produces consistent translations of arbitrary views and thus enables long-term consistent video synthesis. We design and test our model to generate video sequences from minimally-invasive surgical abdominal scenes. Because labeled data is often limited in this domain, photorealistic data where ground truth information from the simulated domain is preserved is especially relevant. By extending existing image-based methods to view-consistent videos, we aim to impact the applicability of simulated training and evaluation environments for surgical applications. Code and data: http://opencas.dkfz.de/video-sim2real.
引用
收藏
页码:3323 / 3333
页数:11
相关论文
共 50 条
  • [1] Look Outside the Room Synthesizing A Consistent Long-Term 3D Scene Video from A Single Image
    Ren, Xuanchi
    Wang, Xiaolong
    2022 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2022), 2022, : 3553 - 3563
  • [2] Learning Temporally and Semantically Consistent Unpaired Video-to-video Translation Through Pseudo-Supervision From Synthetic Optical Flow
    Wang, Kaihong
    Akash, Kumar
    Misu, Teruhisa
    THIRTY-SIXTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE / THIRTY-FOURTH CONFERENCE ON INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE / THE TWELVETH SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2022, : 2477 - 2486
  • [3] Beyond Static Features for Temporally Consistent 3D Human Pose and Shape from a Video
    Choi, Hongsuk
    Moon, Gyeongsik
    Chang, Ju Yong
    Lee, Kyoung Mu
    2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021, 2021, : 1964 - 1973
  • [4] Temporally Consistent Depth Map Estimation for 3D Video Generation and Coding
    Lee, Sang-Beom
    Ho, Yo-Sung
    CHINA COMMUNICATIONS, 2013, 10 (05) : 39 - 49
  • [5] FutureHuman3D: Forecasting Complex Long-Term 3D Human Behavior from Video Observations
    Diller, Christian
    Funkhouser, Thomas
    Dai, Angela
    2024 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2024, : 19902 - 19914
  • [6] Combining FAIR principles and long-term archival of 3D data
    Quantin, Matthieu
    Tournon, Sarah
    Grimaud, Valentin
    Laroche, Florent
    Granier, Xavier
    28TH INTERNATIONAL CONFERENCE ON WEB3D TECHNOLOGY, WEB3D 2023, 2023,
  • [7] TEMPORALLY CONSISTENT HOLE FILLING METHOD BASED ON GLOBAL OPTIMIZATION WITH LABEL PROPAGATION FOR 3D VIDEO
    Kim, Hak Gu
    Yoon, Soo Sung
    Ro, Yong Man
    2015 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2015, : 3136 - 3140
  • [8] Organizing anatomical 3D data for long-term storage and frequent reuse
    Schlager, Stefan
    Engel, Felix
    AMERICAN JOURNAL OF BIOLOGICAL ANTHROPOLOGY, 2023, 180 : 157 - 157
  • [9] Long-term 3D Localization and Pose from Semantic Labellings
    Toft, Carl
    Olsson, Carl
    Kahl, Fredrik
    2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION WORKSHOPS (ICCVW 2017), 2017, : 650 - 659
  • [10] Make-It-4D: Synthesizing a Consistent Long-Term Dynamic Scene Video from a Single Image
    Shen, Liao
    Li, Xingyi
    Sun, Huiqiang
    Peng, Juewen
    Xian, Ke
    Cao, Zhiguo
    Lin, Guosheng
    PROCEEDINGS OF THE 31ST ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2023, 2023, : 8167 - 8175