ShopSpell

MultiMedia Modeling: 30th International Conference, MMM 2024, Amsterdam, The Netherlands, January 29 February 2, 2024, Proceedings, Part III [Paperback]

$70.99     $89.99   21% Off      (Free Shipping)
100 available
  • Category: Books (Computers)
  • ISBN-10:  3031533100
  • ISBN-10:  3031533100
  • ISBN-13:  9783031533105
  • ISBN-13:  9783031533105
  • Publisher:  Springer
  • Publisher:  Springer
  • Binding:  Paperback
  • Binding:  Paperback
  • SKU:  3031533100-11-SPRI
  • SKU:  3031533100-11-SPRI
  • Pages:  535
  • Pages:  535
  • Item ID: 106799639
  • List Price: $89.99
  • Seller: ShopSpell
  • Ships in: 5 business days
  • Transit time: Up to 5 business days
  • Delivery by: Oct 16 to Oct 18
  • Notes: Brand New Item. Not shipped to AK, HI, APO, FPO, AE.

This book constitutes the refereed proceedings of the 30th International Conference on MultiMedia Modeling, MMM 2024, held in Amsterdam, The Netherlands, during January 29February 2, 2024.

The 112 full papers included in this volume were carefully reviewed and selected from 297 submissions. The MMM conference were organized in topics related to multimedia modelling, particularly: audio, image, video processing, coding and compression; multimodal analysis for retrieval applications, and multimedia fusion methods.
Global-to-Local Feature Mining Network for RGB-Infrared Person Re-Identification.- Semantic Transition Detection for Self-Supervised Vide Scene Segmentation.- Multi-Task Collaborative Network for Image-text Retrieval.- FGENet:Fine-Grained Extraction Network for Congested Crowd Counting.- MSMV-UNet : A 2.5D Stroke Lesion Segmentation Method based on Multi-slice Feature Fusion.- Non-Local Spatial-Wise and Global Channel-Wise Transformer for
Efficient Image Super-Resolution.- MobileViT-FocR: MobileViT with Fixed-One-Centre Loss and Gradient Reversal for Generalised Fake Face Detection.- ASF-Conformer: Audio Scoring Conformer with FFC for Speaker Verification in Noisy Environments.- Prior-Knowledge-Free Video Frame Interpolation with Bidirectional Regularized Implicit Neural Representations.- Two-Stage Reasoning Network with Modality Decomposition for Text
VQA.- Localization and Local Motion Magnification of Pulsatile Regions in Endoscopic Surgery Videos.- Co-speech Gesture Generation with Variational Auto Encoder.- Differentiable Neural Architecture Search Based on Efficient Architecture for Lightweight Image Super-Resolution.- Learning Collaborative Reinforcement Attention for 3D Face l#2