Proposal Closed Access

Medical Image Captioning Using Deep Learning

Agegnehu Teshome


JSON Export

{
  "owners": [
    81
  ], 
  "doi": "10.20372/nadre:23965", 
  "stats": {}, 
  "links": {
    "conceptbadge": "https://nadre.ethernet.edu.et/badge/doi/10.20372/nadre%3A23964.svg", 
    "doi": "https://doi.org/10.20372/nadre:23965", 
    "conceptdoi": "https://doi.org/10.20372/nadre:23964", 
    "latest_html": "https://nadre.ethernet.edu.et/record/23965", 
    "badge": "https://nadre.ethernet.edu.et/badge/doi/10.20372/nadre%3A23965.svg", 
    "html": "https://nadre.ethernet.edu.et/record/23965", 
    "latest": "https://nadre.ethernet.edu.et/api/records/23965"
  }, 
  "conceptdoi": "10.20372/nadre:23964", 
  "created": "2026-01-13T12:17:48.880801+00:00", 
  "updated": "2026-02-10T08:46:56.439491+00:00", 
  "conceptrecid": "23964", 
  "revision": 2, 
  "id": 23965, 
  "metadata": {
    "access_right_category": "danger", 
    "doi": "10.20372/nadre:23965", 
    "description": "<p>This research proposal presents a novel deep learning framework for automated medical image captioning, focusing specifically on chest X-ray analysis. The study aims to develop a robust and interpretable system that automatically generates accurate, clinically relevant textual descriptions by integrating object detection and caption generation. The proposed methodology optimizes the Faster R-CNN architecture through advanced techniques&mdash;including Feature Pyramid Network (FPN), Feature Reuse, and Optimized Anchor Generation&mdash;to enhance the detection of small and overlapping anatomical structures. Furthermore, an end-to-end learning approach is adopted to improve contextual understanding and caption quality directly from raw image data. By addressing key limitations in current methods, such as spatial context neglect and sequential processing bottlenecks, this research seeks to contribute to diagnostic accuracy, workflow efficiency, and patient communication in clinical practice.</p>", 
    "contributors": [
      {
        "affiliation": "Mekdela Amba University", 
        "type": "ResearchGroup", 
        "name": "Habtamu Shiferaw"
      }, 
      {
        "affiliation": "Mekdela Amba University", 
        "type": "ResearchGroup", 
        "name": "Getie Balew"
      }, 
      {
        "affiliation": "Mekdela Amba University", 
        "type": "ResearchGroup", 
        "name": "Melese Alemante"
      }
    ], 
    "title": "Medical Image Captioning Using Deep Learning", 
    "notes": "On Progress", 
    "relations": {
      "version": [
        {
          "count": 1, 
          "index": 0, 
          "parent": {
            "pid_type": "recid", 
            "pid_value": "23964"
          }, 
          "is_last": true, 
          "last_child": {
            "pid_type": "recid", 
            "pid_value": "23965"
          }
        }
      ]
    }, 
    "communities": [
      {
        "id": "mau-community"
      }
    ], 
    "keywords": [
      "Medical Image Captioning", 
      "Deep Learning", 
      "Chest X-ray Analysis", 
      "Faster R-CNN"
    ], 
    "publication_date": "2026-01-13", 
    "creators": [
      {
        "affiliation": "Mekdela Amba University", 
        "name": "Agegnehu Teshome"
      }
    ], 
    "access_right": "closed", 
    "resource_type": {
      "subtype": "proposal", 
      "type": "publication", 
      "title": "Proposal"
    }, 
    "related_identifiers": [
      {
        "scheme": "doi", 
        "identifier": "10.20372/nadre:23964", 
        "relation": "isVersionOf"
      }
    ]
  }
}
0
0
views
downloads
All versions This version
Views 00
Downloads 00
Data volume 0 Bytes0 Bytes
Unique views 00
Unique downloads 00

Share

Cite as