000 06133nam a22005415i 4500
001 978-3-031-19833-5
003 DE-He213
005 20240730170904.0
007 cr nn 008mamaa
008 221103s2022 sz | s |||| 0|eng d
020 _a9783031198335
_9978-3-031-19833-5
024 7 _a10.1007/978-3-031-19833-5
_2doi
050 4 _aTA1634
072 7 _aUYQV
_2bicssc
072 7 _aCOM016000
_2bisacsh
072 7 _aUYQV
_2thema
082 0 4 _a006.37
_223
245 1 0 _aComputer Vision - ECCV 2022
_h[electronic resource] :
_b17th European Conference, Tel Aviv, Israel, October 23-27, 2022, Proceedings, Part XXXV /
_cedited by Shai Avidan, Gabriel Brostow, Moustapha Cissé, Giovanni Maria Farinella, Tal Hassner.
250 _a1st ed. 2022.
264 1 _aCham :
_bSpringer Nature Switzerland :
_bImprint: Springer,
_c2022.
300 _aLV, 745 p. 222 illus., 219 illus. in color.
_bonline resource.
336 _atext
_btxt
_2rdacontent
337 _acomputer
_bc
_2rdamedia
338 _aonline resource
_bcr
_2rdacarrier
347 _atext file
_bPDF
_2rda
490 1 _aLecture Notes in Computer Science,
_x1611-3349 ;
_v13695
505 0 _aEfficient One-Stage Video Object Detection by Exploiting Temporal Consistency -- Leveraging Action Affinity and Continuity for Semi-Supervised Temporal Action Segmentation -- Spotting Temporally Precise, Fine-Grained Events in Video -- Unified Fully and Timestamp Supervised Temporal Action Segmentation via Sequence to Sequence Translation -- Efficient Video Transformers with Spatial-Temporal Token Selection -- Long Movie Clip Classification with State-Space Video Models -- Prompting Visual-Language Models for Efficient Video Understanding -- Asymmetric Relation Consistency Reasoning for Video Relation Grounding -- Self-Supervised Social Relation Representation for Human Group Detection -- K-Centered Patch Sampling for Efficient Video Recognition -- A Deep Moving-Camera Background Model -- GraphVid: It Only Takes a Few Nodes to Understand a Video -- Delta Distillation for Efficient Video Processing -- MorphMLP: An Efficient MLP-Like Backbone for Spatial-Temporal Representation Learning -- COMPOSER: Compositional Reasoning of Group Activity in Videos with Keypoint-Only Modality -- E-NeRV: Expedite Neural Video Representation with Disentangled Spatial-Temporal Context -- TDViT: Temporal Dilated Video Transformer for Dense Video Tasks -- Semi-Supervised Learning of Optical Flow by Flow Supervisor -- Flow Graph to Video Grounding for Weakly-Supervised Multi-step Localization -- Deep 360° Optical Flow Estimation Based on Multi-Projection Fusion -- MaCLR: Motion-Aware Contrastive Learning of Representations for Videos -- Learning Long-Term Spatial-Temporal Graphs for Active Speaker Detection -- Frozen CLIP Models Are Efficient Video Learners -- PIP: Physical Interaction Prediction via Mental Simulation with Span Selection -- Panoramic Vision Transformer for Saliency Detection in 360° Videos -- Bayesian Tracking of Video Graphs Using Joint Kalman Smoothing and Registration -- Motion Sensitive Contrastive Learning for Self-Supervised Video Representation -- Dynamic Temporal Filtering In Video Models -- Tip-Adapter: Training-Free Adaption of CLIP for Few-Shot Classification -- Temporal Lift Pooling for Continuous Sign Language Recognition -- MORE: Multi-Order RElation Mining for Dense Captioning in 3D Scenes -- SiRi: A Simple Selective Retraining Mechanism for Transformer-Based Visual Grounding -- Cross-Modal Prototype Driven Network for Radiology Report Generation -- TM2T: Stochastic and Tokenized Modeling for the Reciprocal Generation of 3D Human Motions and Texts -- SeqTR: A Simple Yet Universal Network for Visual Grounding -- VTC: Improving Video-Text Retrieval with User Comments -- FashionViL: Fashion-Focused Vision-and-Language Representation Learning -- Weakly Supervised Grounding for VQA in Vision-Language Transformers -- Automatic Dense Annotation of Large-Vocabulary Sign Language Videos -- MILES: Visual BERT Pre-training with Injected Language Semantics for Video-Text Retrieval -- GEB+: A Benchmark for Generic Event Boundary Captioning, Grounding and Retrieval -- A Simple and Robust Correlation Filtering Method for Text-Based Person Search.
520 _aThe 39-volume set, comprising the LNCS books 13661 until 13699, constitutes the refereed proceedings of the 17th European Conference on Computer Vision, ECCV 2022, held in Tel Aviv, Israel, during October 23-27, 2022. The 1645 papers presented in these proceedings were carefully reviewed and selected from a total of 5804 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; object recognition; motion estimation.
650 0 _aComputer vision.
_995996
650 1 4 _aComputer Vision.
_995997
700 1 _aAvidan, Shai.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_995998
700 1 _aBrostow, Gabriel.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_995999
700 1 _aCissé, Moustapha.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_996001
700 1 _aFarinella, Giovanni Maria.
_eeditor.
_0(orcid)
_10000-0002-6034-0432
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_996003
700 1 _aHassner, Tal.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_996005
710 2 _aSpringerLink (Online service)
_996007
773 0 _tSpringer Nature eBook
776 0 8 _iPrinted edition:
_z9783031198328
776 0 8 _iPrinted edition:
_z9783031198342
830 0 _aLecture Notes in Computer Science,
_x1611-3349 ;
_v13695
_923263
856 4 0 _uhttps://doi.org/10.1007/978-3-031-19833-5
912 _aZDB-2-SCS
912 _aZDB-2-SXCS
912 _aZDB-2-LNC
942 _cELN
999 _c87248
_d87248