MARC View

000			07663nam a22006735i 4500
001			978-3-030-98358-1
003			DE-He213
005			20240730180703.0
007			cr nn 008mamaa
008			220314s2022 sz \| s \|\|\|\| 0\|eng d
020			_a9783030983581 _9978-3-030-98358-1
024	7		_a10.1007/978-3-030-98358-1 _2doi
050		4	_aTA1634
072		7	_aUYQV _2bicssc
072		7	_aCOM016000 _2bisacsh
072		7	_aUYQV _2thema
082	0	4	_a006.37 _223
245	1	0	_aMultiMedia Modeling _h[electronic resource] : _b28th International Conference, MMM 2022, Phu Quoc, Vietnam, June 6-10, 2022, Proceedings, Part I / _cedited by Björn Þór Jónsson, Cathal Gurrin, Minh-Triet Tran, Duc-Tien Dang-Nguyen, Anita Min-Chun Hu, Binh Huynh Thi Thanh, Benoit Huet.
250			_a1st ed. 2022.
264		1	_aCham : _bSpringer International Publishing : _bImprint: Springer, _c2022.
300			_aXXVI, 641 p. 188 illus., 179 illus. in color. _bonline resource.
336			_atext _btxt _2rdacontent
337			_acomputer _bc _2rdamedia
338			_aonline resource _bcr _2rdacarrier
347			_atext file _bPDF _2rda
490	1		_aLecture Notes in Computer Science, _x1611-3349 ; _v13141
505	0		_aBEST PAPER SESSION -- Real-time detection of tiny objects based on a weighted bi-directional FPN -- Multi-Modal Fusion Network for Rumor Detection with Texts and Images -- PF-VTON: Toward High-Quality Parser-Free Virtual Try-On Network -- MF-GAN: Multi-conditional fusion Generative Adversarial Network for Text-to-Image Synthesis -- APPLICATIONS 1 -- Learning to classify weather conditions from single images without labels -- Learning Image Representation via Attribute-aware Attention Networks for Fashion Classification -- Toward Detail-Oriented Image-Based Virtual Try-On with Arbitrary Poses -- Parallel DBSCAN-Martingale estimation of the number of concepts for automatic satellite image clustering -- MULTIMEDIA APPLICATIONS - PERSPECTIVES, TOOLS & APPLICATIONS (Special Session) & BRAVE NEW IDEAS -- AI for the Media Industry: Application Potential and Automation Level -- Color the Word: Leveraging Web Images for Machine Translation of Untranslatable Words -- ACTIVITIES & EVENTS -- MGMP: Multimodal Graph Message Propagation Network for Event Detection -- Pose-Enhanced Relation Feature for Action Recognition in Still Images.-Prostate Segmentation of Ultrasound Images based on Interpretable-guided Mathematical Model -- Spatiotemporal Perturbation Based Dynamic Consistency for Semi-Supervised Temporal Action Detection -- MULTIMEDIA DATASETS FOR REPEATABLE EXPERIMENTATION (Special Session) -- A Task Category Space for User-Centric Comparative Multimedia Search Evaluations -- GPR1200: A Benchmark for General-Purpose Content-Based Image Retrieval -- LLQA - Lifelog Question Answering Dataset -- LEARNING -- Category-sensitive Incremental Learning For Image-based 3D Shape Reconstruction -- AdaConfigure: Reinforcement Learning-based Adaptive Configuration for Video Analytics Services -- Mining Minority-class Examples With Uncertainty Estimates -- Conditional Context-aware Feature Alignment for Domain Adaptive Detection Transformer -- MULTIMEDIA for MEDICAL APPLICATIONS (Special Session) -- Human activity recognition with IMU and vital signs feature fusion -- On Identifying Pareidolia Phenomenon by Emulating Patient Behavior -- Using Explainable AI to Identify Differences between Clinical and Experimental Pain Detection Models Based on Facial Expressions -- APPLICATIONS 2 -- Double Granularity Relation Network with Self-Criticism for Occluded Person Re-Identification -- A Complementary Fusion Strategy for RGB-D Face Recognition -- Multi-scale Cross-modal Transformer Network for RGB-D Object Detection -- Joint Re-Detection and Re-Identification for Multi-Object Tracking -- MULTIMEDIA ANALYTICS for CONTEXTUAL HUMAN UNDERSTANDING (Special Session) -- An Investigation into Keystroke Dynamics and Heart Rate Variability as Indicators of Stress -- Fall detection using multimodal data -- Prediction of Blood Glucose using Contextual LifeLog Data -- Multimodal Embedding for Lifelog Retrieval -- APPLICATIONS 3 -- A Multiple Positives Enhanced NCE Loss for Image-Text Retrieval -- SAM: Self Attention Mechanism for Scene Text Recognition based on Swin Transformer -- JVCSR: Video Compressive Sensing Reconstruction with Joint In-loop Reference Enhancement and Out-loop Super-resolution -- Point Cloud Upsampling via a Coarse-to-fine Network -- IMAGE ANALYTICS -- Arbitrary Style Transfer With Adaptive Channel Network -- Fast Single Image Dehazing Using Morphological Reconstruction and Saturation Compensation -- One-Stage Image Inpainting with Hybrid Attention -- Real-time FPGA Design for OMP Targeting 8K Image Reconstruction -- SPEECH & MUSIC -- Time-Frequency Attention For Speech Emotion Recognition With Squeeze-and-Excitation Blocks -- SPEECH INTELLIGIBILITY ENHANCEMENT BY NON-PARALLEL SPEECH STYLE CONVERSION USING CWT AND iMetricGAN BASED CycleGAN -- A-Muze-Net: Music Generation by Composing the Harmony based on the Generated Melody -- MULTIMODAL ANALYTICS -- Bi-attention modal separation network for multimodal video fusion -- Combining Knowledge and Multi-modal Fusion for Meme Classification -- Non-Uniform Attention Network for Multi-modal Sentiment Analysis -- Multimodal Unsupervised Image-to-Image Translation Without Independent Style Encoder.
520			_aThe two-volume set LNCS 13141 and LNCS 13142 constitutes the proceedings of the 28th International Conference on MultiMedia Modeling, MMM 2022, which took place in Phu Quoc, Vietnam, during June 6-10, 2022. The 107 papers presented in these proceedings were carefully reviewed and selected from a total of 212 submissions. They focus on topics related to multimedia content analysis; multimedia signal processing and communications; and multimedia applications and services.
650		0	_aComputer vision. _9121375
650		0	_aEducation _xData processing. _982607
650		0	_aComputer engineering. _910164
650		0	_aComputer networks . _931572
650		0	_aMultimedia systems. _911575
650		0	_aPattern recognition systems. _93953
650	1	4	_aComputer Vision. _9121376
650	2	4	_aComputers and Education. _941129
650	2	4	_aComputer Engineering and Networks. _9121377
650	2	4	_aMultimedia Information Systems. _931575
650	2	4	_aAutomated Pattern Recognition. _931568
700	1		_aÞór Jónsson, Björn. _eeditor. _4edt _4http://id.loc.gov/vocabulary/relators/edt _9121378
700	1		_aGurrin, Cathal. _eeditor. _4edt _4http://id.loc.gov/vocabulary/relators/edt _9121379
700	1		_aTran, Minh-Triet. _eeditor. _4edt _4http://id.loc.gov/vocabulary/relators/edt _9121380
700	1		_aDang-Nguyen, Duc-Tien. _eeditor. _4edt _4http://id.loc.gov/vocabulary/relators/edt _9121381
700	1		_aHu, Anita Min-Chun. _eeditor. _4edt _4http://id.loc.gov/vocabulary/relators/edt _9121382
700	1		_aHuynh Thi Thanh, Binh. _eeditor. _4edt _4http://id.loc.gov/vocabulary/relators/edt _9121383
700	1		_aHuet, Benoit. _eeditor. _4edt _4http://id.loc.gov/vocabulary/relators/edt _929834
710	2		_aSpringerLink (Online service) _9121384
773	0		_tSpringer Nature eBook
776	0	8	_iPrinted edition: _z9783030983574
776	0	8	_iPrinted edition: _z9783030983598
830		0	_aLecture Notes in Computer Science, _x1611-3349 ; _v13141 _923263
856	4	0	_uhttps://doi.org/10.1007/978-3-030-98358-1
912			_aZDB-2-SCS
912			_aZDB-2-SXCS
912			_aZDB-2-LNC
942			_cELN
999			_c90453 _d90453