--> --- name: 'multimodal-medical-imaging' description: 'Analyzes medical images (X-ray, MRI, CT) using multimodal LLMs to identify anomalies and generate reports.' measurable_outcome: Execute skill workflow successfully with valid output within 15 minutes. allowed-tools: - read_file - run_shell_command --- The **Multimodal Medical Imaging Analysis Skill** leverages state-of-the-art Vision-Language Models (VLMs) like Gemini 1.5 Pro and GPT-4o to interpret medical imagery alongside clinical text.
Scanned 9/8/2026
Install to Claude Code
npx -y skills add mdbabumiamssm/AI-Agentic-Skills-by-Dr.-Mia --skill Multimodal_Analysis --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Multimodal Analysis?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/mdbabumiamssm-multimodal-analysis-ai-agentic-skills-by-dr-mia)More formats (shields.io, HTML) on the badges page.
<!--
# COPYRIGHT NOTICE
# This file is part of the "Universal AI Agentic Skills" project.
# Copyright (c) 2026 MD BABU MIA, PhD <md.babu.mia@mssm.edu>
# All Rights Reserved.
#
# This code is proprietary and confidential.
# Unauthorized copying of this file, via any medium is strictly prohibited.
#
# Provenance: Authenticated by MD BABU MIA
-->
---
name: 'multimodal-medical-imaging'
description: 'Analyzes medical images (X-ray, MRI, CT) using multimodal LLMs to identify anomalies and generate reports.'
measurable_outcome: Execute skill workflow successfully with valid output within 15 minutes.
allowed-tools:
- read_file
- run_shell_command
---
# Multimodal Medical Imaging Analysis
The **Multimodal Medical Imaging Analysis Skill** leverages state-of-the-art Vision-Language Models (VLMs) like Gemini 1.5 Pro and GPT-4o to interpret medical imagery alongside clinical text.
## When to Use This Skill
* When you need a preliminary screening of medical images.
* When correlating visual findings with textual clinical notes.
* To generate structured reports (DICOM-SR-like) from raw images.
## Core Capabilities
1. **Anomaly Detection**: Identify potential pathologies in X-rays, CTs, etc.
2. **Report Generation**: Draft radiology reports in standard formats.
3. **VQA (Visual Question Answering)**: Answer specific questions about an image (e.g., "Is there a fracture in the left femur?").
4. **Dermatology and Dermoscopy Use Case**: For suspected basal cell carcinoma and common mimickers, use modality-specific prompts for clinical skin photographs versus dermoscopic images; check image quality for focus, lighting, scale, lesion completeness, artifacts, and occlusion before analysis; record available lesion metadata without inferring missing values; include basal cell carcinoma mimic differential checks; calibrate confidence to image quality and clinical context; report uncertainty explicitly; frame outputs as triage support rather than a definitive diagnosis; do not treat chatbot image output as diagnostic without task-specific clinical validation; and require clinician or dermatologist review before diagnostic use or any clinical action.
5. **Closed-System Radiography Response Evaluation**: For closed-system LLM radiography-response workflows, evaluate answer correctness against the clinical or radiography question before use; define escalation thresholds for uncertain, incomplete, inconsistent, diagnostic, or safety-critical responses; keep generated technical guidance within radiology-technologist scope limits and route diagnosis or treatment decisions to licensed radiologist or clinician review; require qualified human review of generated explanations before patient-facing, educational, or operational use.
## Workflow
1. **Input**: Provide an image file path (JPG, PNG) and a specific clinical question or "generate report" instruction.
2. **Analyze**: The agent sends the image and prompt to the VLM.
3. **Output**: Returns a JSON object with findings, confidence scores, and reasoning.
## Example Usage
**User**: "Analyze this chest X-ray for pneumonia."
**Agent Action**:
```bash
python3 Skills/Clinical/Medical_Imaging/Multimodal_Analysis/multimodal_agent.py \
--image "/path/to/cxr.jpg" \
--prompt "Check for signs of pneumonia and consolidation."
```
## References
- https://pubmed.ncbi.nlm.nih.gov/41952838/
- https://pubmed.ncbi.nlm.nih.gov/42024724/
<!-- AUTHOR_SIGNATURE: 9a7f3c2e-MD-BABU-MIA-2026-MSSM-SECURE -->
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!