Skip to Content

Multimodal Understanding: From Document Image Analysis to Large Language Models

This special issue belongs to the section "Image and Video Processing".

Special Issue Information

Keywords

  • multimodal document intelligence
  • document visual question answering
  • large language models (LLMs)
  • vision–language foundation models
  • handwriting and signature verification
  • parameter-efficient fine-tuning
  • robust representation learning
  • table and scene text recognition
  • retrieval-augmented generation
  • trustworthy AI and security

Benefits of Publishing in a Special Issue

  • Ease of navigation: Grouping papers by topic helps scholars navigate broad scope journals more efficiently.
  • Greater discoverability: Special Issues support the reach and impact of scientific research. Articles in Special Issues are more discoverable and cited more frequently.
  • Expansion of research network: Special Issues facilitate connections among authors, fostering scientific collaborations.
  • External promotion: Articles in Special Issues are often promoted through the journal's social media, increasing their visibility.
  • Reprint: MDPI Books provides the opportunity to republish successful Special Issues in book format, both online and in print.

Published Papers

XFacebookLinkedIn
J. Imaging - ISSN 2313-433X