Generative AI–Based Multilingual Multimodal Framework for Depression Detection

Clicks: 3
ID: 312426
2026
Article Quality & Performance Metrics
Overall Quality
Not rated
Combines reader engagement with the AI quality analysis. This article has not been analysed, so there is no overall score — reader engagement is measured and shown alongside.
AI Quality Assessment
Not analyzed
Readership in this journal
Emerging

Ranked #4 of 15 articles by views in Southern Journal of Computer Science

Most read Least read

Bar heights use a square-root scale.

Mint this article as an NFT
Not yet minted

Create a permanent, verifiable on-chain record of this article on the Scimatic Network. The NFT is held in your Journament account, and you can withdraw it to your own wallet at any time.

5 SUSD one-off · no wallet required
Abstract
Depression is a prevalent psychological condition that is not easy to detect at an early stage due to its multipolar nature in terms of clinical manifestations and subjective clinical diagnoses. Recent advancements in generative artificial intelligence models, deep learning, and machine learning techniques have made it feasible to digitally identify depression based on how the illness manifests itself in speech, facial expressions, text, and physiological and behavioral characteristics. The study examines and discusses the use of large language, multimodal, and unimodal models for digital depression identification in various multilingual contexts. The technology consistently outperforms unimodal systems, according to the results, with enhanced transformer and cross-attention architecture performance in cross-modal relationship capture, a crucial component of clinical decision support. Large language models have been shown to have potential applications in few-shot learning, multilingual analysis, transcript-based severity estimation, data generation for simulation, and transparent clinical decision support systems. The current review aims to provide a systematic overview of the current methodologies and identify the areas of future research; however, much work needs to be done to have such applications extensive in number, culture-independent, and reflecting the ordinal level severity.
Reference Key
imported_1776990044_69eab75c9464e Use this key to autocite in the manuscript while using SciMatic Manuscript Manager or Thesis Manager
Authors Uswa Ashraf, Hamid Ghous, Mubasher H. Malik, Majid Khawar
Journal Southern Journal of Computer Science
Year 2026
DOI
DOI not found
URL
Keywords Keywords not found

Citations

No citations found. To add a citation, contact the admin at info@scimatic.org

No comments yet. Be the first to comment on this article.