000 07663nam a22006735i 4500
001 978-3-030-98358-1
003 DE-He213
005 20240730180703.0
007 cr nn 008mamaa
008 220314s2022 sz | s |||| 0|eng d
020 _a9783030983581
_9978-3-030-98358-1
024 7 _a10.1007/978-3-030-98358-1
_2doi
050 4 _aTA1634
072 7 _aUYQV
_2bicssc
072 7 _aCOM016000
_2bisacsh
072 7 _aUYQV
_2thema
082 0 4 _a006.37
_223
245 1 0 _aMultiMedia Modeling
_h[electronic resource] :
_b28th International Conference, MMM 2022, Phu Quoc, Vietnam, June 6-10, 2022, Proceedings, Part I /
_cedited by Björn Þór Jónsson, Cathal Gurrin, Minh-Triet Tran, Duc-Tien Dang-Nguyen, Anita Min-Chun Hu, Binh Huynh Thi Thanh, Benoit Huet.
250 _a1st ed. 2022.
264 1 _aCham :
_bSpringer International Publishing :
_bImprint: Springer,
_c2022.
300 _aXXVI, 641 p. 188 illus., 179 illus. in color.
_bonline resource.
336 _atext
_btxt
_2rdacontent
337 _acomputer
_bc
_2rdamedia
338 _aonline resource
_bcr
_2rdacarrier
347 _atext file
_bPDF
_2rda
490 1 _aLecture Notes in Computer Science,
_x1611-3349 ;
_v13141
505 0 _aBEST PAPER SESSION -- Real-time detection of tiny objects based on a weighted bi-directional FPN -- Multi-Modal Fusion Network for Rumor Detection with Texts and Images -- PF-VTON: Toward High-Quality Parser-Free Virtual Try-On Network -- MF-GAN: Multi-conditional fusion Generative Adversarial Network for Text-to-Image Synthesis -- APPLICATIONS 1 -- Learning to classify weather conditions from single images without labels -- Learning Image Representation via Attribute-aware Attention Networks for Fashion Classification -- Toward Detail-Oriented Image-Based Virtual Try-On with Arbitrary Poses -- Parallel DBSCAN-Martingale estimation of the number of concepts for automatic satellite image clustering -- MULTIMEDIA APPLICATIONS - PERSPECTIVES, TOOLS & APPLICATIONS (Special Session) & BRAVE NEW IDEAS -- AI for the Media Industry: Application Potential and Automation Level -- Color the Word: Leveraging Web Images for Machine Translation of Untranslatable Words -- ACTIVITIES & EVENTS -- MGMP: Multimodal Graph Message Propagation Network for Event Detection -- Pose-Enhanced Relation Feature for Action Recognition in Still Images.-Prostate Segmentation of Ultrasound Images based on Interpretable-guided Mathematical Model -- Spatiotemporal Perturbation Based Dynamic Consistency for Semi-Supervised Temporal Action Detection -- MULTIMEDIA DATASETS FOR REPEATABLE EXPERIMENTATION (Special Session) -- A Task Category Space for User-Centric Comparative Multimedia Search Evaluations -- GPR1200: A Benchmark for General-Purpose Content-Based Image Retrieval -- LLQA - Lifelog Question Answering Dataset -- LEARNING -- Category-sensitive Incremental Learning For Image-based 3D Shape Reconstruction -- AdaConfigure: Reinforcement Learning-based Adaptive Configuration for Video Analytics Services -- Mining Minority-class Examples With Uncertainty Estimates -- Conditional Context-aware Feature Alignment for Domain Adaptive Detection Transformer -- MULTIMEDIA for MEDICAL APPLICATIONS (Special Session) -- Human activity recognition with IMU and vital signs feature fusion -- On Identifying Pareidolia Phenomenon by Emulating Patient Behavior -- Using Explainable AI to Identify Differences between Clinical and Experimental Pain Detection Models Based on Facial Expressions -- APPLICATIONS 2 -- Double Granularity Relation Network with Self-Criticism for Occluded Person Re-Identification -- A Complementary Fusion Strategy for RGB-D Face Recognition -- Multi-scale Cross-modal Transformer Network for RGB-D Object Detection -- Joint Re-Detection and Re-Identification for Multi-Object Tracking -- MULTIMEDIA ANALYTICS for CONTEXTUAL HUMAN UNDERSTANDING (Special Session) -- An Investigation into Keystroke Dynamics and Heart Rate Variability as Indicators of Stress -- Fall detection using multimodal data -- Prediction of Blood Glucose using Contextual LifeLog Data -- Multimodal Embedding for Lifelog Retrieval -- APPLICATIONS 3 -- A Multiple Positives Enhanced NCE Loss for Image-Text Retrieval -- SAM: Self Attention Mechanism for Scene Text Recognition based on Swin Transformer -- JVCSR: Video Compressive Sensing Reconstruction with Joint In-loop Reference Enhancement and Out-loop Super-resolution -- Point Cloud Upsampling via a Coarse-to-fine Network -- IMAGE ANALYTICS -- Arbitrary Style Transfer With Adaptive Channel Network -- Fast Single Image Dehazing Using Morphological Reconstruction and Saturation Compensation -- One-Stage Image Inpainting with Hybrid Attention -- Real-time FPGA Design for OMP Targeting 8K Image Reconstruction -- SPEECH & MUSIC -- Time-Frequency Attention For Speech Emotion Recognition With Squeeze-and-Excitation Blocks -- SPEECH INTELLIGIBILITY ENHANCEMENT BY NON-PARALLEL SPEECH STYLE CONVERSION USING CWT AND iMetricGAN BASED CycleGAN -- A-Muze-Net: Music Generation by Composing the Harmony based on the Generated Melody -- MULTIMODAL ANALYTICS -- Bi-attention modal separation network for multimodal video fusion -- Combining Knowledge and Multi-modal Fusion for Meme Classification -- Non-Uniform Attention Network for Multi-modal Sentiment Analysis -- Multimodal Unsupervised Image-to-Image Translation Without Independent Style Encoder.
520 _aThe two-volume set LNCS 13141 and LNCS 13142 constitutes the proceedings of the 28th International Conference on MultiMedia Modeling, MMM 2022, which took place in Phu Quoc, Vietnam, during June 6-10, 2022. The 107 papers presented in these proceedings were carefully reviewed and selected from a total of 212 submissions. They focus on topics related to multimedia content analysis; multimedia signal processing and communications; and multimedia applications and services.
650 0 _aComputer vision.
_9121375
650 0 _aEducation
_xData processing.
_982607
650 0 _aComputer engineering.
_910164
650 0 _aComputer networks .
_931572
650 0 _aMultimedia systems.
_911575
650 0 _aPattern recognition systems.
_93953
650 1 4 _aComputer Vision.
_9121376
650 2 4 _aComputers and Education.
_941129
650 2 4 _aComputer Engineering and Networks.
_9121377
650 2 4 _aMultimedia Information Systems.
_931575
650 2 4 _aAutomated Pattern Recognition.
_931568
700 1 _aÞór Jónsson, Björn.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_9121378
700 1 _aGurrin, Cathal.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_9121379
700 1 _aTran, Minh-Triet.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_9121380
700 1 _aDang-Nguyen, Duc-Tien.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_9121381
700 1 _aHu, Anita Min-Chun.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_9121382
700 1 _aHuynh Thi Thanh, Binh.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_9121383
700 1 _aHuet, Benoit.
_eeditor.
_4edt
_4http://id.loc.gov/vocabulary/relators/edt
_929834
710 2 _aSpringerLink (Online service)
_9121384
773 0 _tSpringer Nature eBook
776 0 8 _iPrinted edition:
_z9783030983574
776 0 8 _iPrinted edition:
_z9783030983598
830 0 _aLecture Notes in Computer Science,
_x1611-3349 ;
_v13141
_923263
856 4 0 _uhttps://doi.org/10.1007/978-3-030-98358-1
912 _aZDB-2-SCS
912 _aZDB-2-SXCS
912 _aZDB-2-LNC
942 _cELN
999 _c90453
_d90453