Scope & Topics

International Journal of Multimedia & Its Applications (IJMA) is a bi monthly open access peer-reviewed journal that publishes articles which contribute new results in all areas of the Multimedia & its applications. The journal focuses on all technical and practical aspects of Multimedia and its applications. The goal of this journal is to bring together researchers and practitioners from academia and industry to focus on understanding recent developments this arena, and establishing new collaborations in these areas.

Authors are solicited to contribute to the journal by submitting articles that illustrate research results, projects, surveying works and industrial experiences that describe significant advances in the areas of Multimedia & its applications

Topics of interest include but are not limited to, the following

    Multimodal Foundation Models and AI Systems
  • Vision-Language, Video-Language and Audio-Language Models
  • Large Multimodal Models (LMMs) and Foundation Models
  • Cross-modal representation learning and alignment
  • Multimodal reasoning, grounding and knowledge integration
  • Multimodal AI agents and Vision-Language-Action (VLA) systems
  • Embodied and interactive multimodal intelligence
  • Generative Multimedia AI and Creative Systems
  • Text-to-image, text-to-video and text-to-audio generation
  • Diffusion models, GANs and autoregressive generative models
  • Controllable and prompt-driven multimedia generation
  • Neural avatars, digital humans and synthetic identity systems
  • AI-driven video editing, storytelling and content creation
  • Real-time generative multimedia and co-creative systems
  • Video, Image, Audio and Multimodal Understanding
  • Video foundation models and Video-LLMs
  • Long-form and streaming video understanding
  • Egocentric vision and wearable multimedia systems
  • Image recognition, scene understanding and fine-grained analysis
  • Speech, audio event detection and audio-visual learning
  • Self-supervised, weakly supervised and continual learning
  • 3D Vision, Neural Rendering and Immersive Media
  • Neural Radiance Fields (NeRF) and 3D Gaussian Splatting
  • Neural rendering and dynamic scene reconstruction
  • 4D spatio-temporal scene modeling
  • Digital twins and real-time 3D simulation systems
  • XR (VR/AR/MR), spatial computing and immersive environments
  • Neural avatars and photorealistic digital humans
  • Multimedia Retrieval, RAG and Knowledge Systems
  • Cross-modal semantic retrieval (image, video, audio, text)
  • Embedding-based search and vector database systems
  • Retrieval-Augmented Generation (RAG) for multimedia
  • Agentic retrieval and multimodal memory systems
  • Multimedia knowledge graphs and semantic reasoning
  • Large-scale indexing, search and recommendation systems
  • Multimedia Systems, Edge AI and Scalable Infrastructure
  • Edge, cloud and distributed multimedia computing systems
  • Real-time and low-latency multimedia pipelines
  • 5G/6G-enabled multimedia communication systems
  • Adaptive streaming and QoE optimization
  • Mobile and IoT-based multimedia systems
  • Efficient deployment of multimodal AI at scale
  • Human-Centered and Interactive Multimedia Systems
  • Multimodal human-computer interaction (speech, vision, gesture, emotion)
  • Affective computing and emotion-aware systems
  • Social media content analysis and user behavior modeling
  • Immersive and interactive multimedia interfaces
  • Accessibility and assistive multimedia technologies
  • Multimodal conversational agents and co-creative systems
  • Multimedia Security, Trust and Responsible AI
  • Deepfake detection and multimedia forensics
  • Media authentication, watermarking and provenance tracking
  • Privacy-preserving and federated multimedia learning
  • Adversarial robustness in multimodal AI systems
  • Bias, fairness and ethical AI in multimedia
  • Multimodal hallucination detection and grounding
  • Misinformation and synthetic media detection
  • Efficient, Data-Centric and Sustainable Multimedia AI
  • Data-centric AI for multimedia systems
  • Synthetic data generation for vision, audio, video and 3D
  • Continual and lifelong multimodal learning
  • Efficient and compressed multimodal models
  • Green AI and energy-efficient multimedia computing
  • Edge inference and resource-aware deployment
  • Evaluation, Benchmarks and Applications
  • Benchmarks for multimodal reasoning and generation
  • Evaluation of hallucination, robustness and grounding
  • Human-aligned evaluation frameworks
  • Real-world deployment and system-level evaluation
  • Healthcare, education, smart cities and industrial applications
  • Autonomous systems, robotics and embodied intelligence
  • Entertainment, gaming and creative industries

Important Dates

  • Submission Deadline : July 19, 2026
  • Notification                   : August 11, 2026
  • Final Manuscript Due : August 18, 2026
  • Publication Date          : Determined by the Editor-in-Chief

Note : AIRCC's International Journal of Multimedia & Its Applications (IJMA) is dedicated to strengthen the field of Multimedia & its applications and publishes only good quality papers. IJMA is highly selective and maintains less than 20% acceptance rate. All accepted papers will be tested for plagiarism manually as well as by Docoloc . Papers published in IJMA has received enormous citations and has been regarded as one of the best Journal in the field of Multimedia.

Call for Papers

  • Special issue will be published for selected papers from SIPM 2023, Vancouver, Canada

ERA Indexed


  • IJMA is listed in ERA 2023 as per the Australian Research Council (ARC) Journal Ranking.New