Research

Advanced research capabilities for comprehensive medical analysis and insights

Research stories

Our papers, told as stories

Accessible, evidence-faithful narratives of Arkangel AI research, with the original paper always one click away.

Browse research stories

Published in Intelligence-Based Medicine (Elsevier)

An open-book exam and an assistant that already knows where to look

83 students, identical clinical cases, and blinded specialists tested whether Arkangel AI improved the validity and efficiency of open clinical answers.

Read the story

Arkangel AI medical-audit architecture for Colombian claims

By Arkangel AI

6/22/2026

Research brief for Arkangel AI: 98 independent administrative, clinical and financial rules for Colombian medical-claims audit under Resolución 3047 de 2008.

Arkangel AI, OpenEvidence, ChatGPT, and Medisearch evaluated against medical standards

By Jose Zea

5/21/2026

Peer-reviewed 2026 study compares four conversational agents across 128 Q/A pairs and 1024 expert evaluations.

AI detects bipolar disorder in EHRs of patients with affective diagnoses with AUC 0.93

By Jose Zea

5/21/2026

medRxiv diagnostic accuracy preprint: EHR model reports AUC 0.93, sensitivity 96.4%, and specificity 84.4%.

From free text to SOFA score reconstructs sepsis severity from clinical notes

By Jose Zea

5/21/2026

medRxiv preprint: automated SOFA reconstruction improved respiratory variable recovery from 33% to 100%.

Recursive learning architecture improves zero-shot clinical coding F1 from 0.318 to 0.605

By Jose Zea

5/21/2026

JMIR preprint: recursive memory raises zero-shot clinical coding F1 from 0.318 to 0.605 after 20 iterations.

Arkangel AI real-time multi-LLM agent answers clinicians' medical questions with 90% accuracy

By Jose Zea

8/15/2025

Arkangel AI: multi-LLM retrieval system gives evidence-based medical answers with 90% accuracy.

PANDORA LLM Automates COPD Risk Detection with Near-Perfect Extraction and 94% PUMA Accuracy

By Jose Zea

8/15/2025

LLM auto-extracted ICU/outpatient (~100%) and applied PUMA: 94% scoring accuracy; 100% sensitivity

Minimal-resource AI Detects CKD in Type 2 Diabetes Patients Across Six LMICs with 90% Sensitivity

By Jose Zea

8/15/2025

Minimal-data ML screened CKD in T2D across 6 LMICs: 90% sensitivity, AUC 0.63.

Patients and clinicians: LLMs achieve high QA accuracy but require human evaluation for clinical safety

By Jose Zea

8/15/2025

Review: LLMs score highly on QA but need human-in-loop, real-world evaluation for safe clinical use.

Vitruvius conversational AI achieves 90% accuracy on USMLE-style clinical queries for patients across specialties

By Jose Zea

8/15/2025

Vitruvius: multi-LLM, retrieval-augmented chat answers USMLE-style queries with 90.3% accuracy.

PANDORA AI Extracts EHR Data, Identifies COPD Risk in Patients with 98% PUMA Accuracy

By Jose Zea

8/15/2025

PANDORA used GPT-4 to extract clinical notes and apply PUMA: >90% extraction, 95-98% COPD scoring.

Clinicians Using Arkangel AI Conversational Search Answer Questions 79% Faster, 34% Fewer Searches

By Jose Zea

8/15/2025

Arkangel AI: real‑time, evidence‑based AI cut clinicians' answer time 79% and searches 34% in pilot.

CNN Identifies Exudative Macular Disease on OCT in AMD and DME Patients with 97% Accuracy

By Jose Zea

8/15/2025

AI-CNN reads OCT to detect intraretinal fluid/macular edema with ~97% accuracy, aiding retina care.

GPT-4o Conversational Agent Achieves 100% Guideline Accuracy in Alzheimer’s Disease Care

By Jose Zea

8/15/2025

GPT-4o conversational agent, trained on 17 AD guidelines, achieved near-perfect accuracy.

AI Chest X‑Ray + Clinical Data Predicts ICU Admission (AUC 0.92) in Hospitalized COVID‑19 Patients

By Jose Zea

8/15/2025

In COVID inpatients, AI chest X-ray + clinical data predicted ICU AUC 0.92; death AUC 0.81.

Ensemble AI Detects CKD with 91% Sensitivity in Diabetics and 92.5% in Non‑Diabetics

By Jose Zea

8/15/2025

Latin America: ensemble AI using routine clinical data flags CKD - 91% sensitivity (T2D), 92% (NT2D)

Ensemble AI Detects CKD with 91% Sensitivity in Diabetics and 92.5% in Non‑Diabetics

By Jose Zea

8/15/2025

Latin America: ensemble AI using routine clinical data flags CKD - 91% sensitivity (T2D), 92% (NT2D)

Arkangel.AI predictive models reduce hospital admissions up to 45% across 68 million chronic disease patients

By Jose Zea

8/15/2025

AI models across 68M patients predicted CKD, HF, diabetes risks (AUC>0.85), cutting admissions ≈45%.

Automated AI Chest Imaging Analysis Increases Lung Nodule Detection Sensitivity 13% in Screening Patients

By Jose Zea

8/15/2025

Automated AI reading of chest X‑rays and low‑dose CT raised nodule sensitivity from 47% to 60% (+13%), cut false positives ~11%, and achieved ~94% diagnostic accuracy, aiding earlier detection.

No-code Hippocrates AutoML Builds Pediatric Leukemia AI Models Tenfold Faster

By Jose Zea

8/15/2025

Hippocrates AutoML lets clinicians build no-code, HIPAA-compliant AI (auto data prep, model selection) 10x faster, cutting costs and enabling prospective validation—bias monitoring advised.

Noninvasive AI Algorithm Boosts Early CKD Detection by Up to 90% in Latin American Patients

By Jose Zea

8/15/2025

Arkangel AI used non-invasive ML on routine clinical data in Latin America to boost early CKD detection - high sensitivity/specificity, identifying up to 90% more at-risk patients.

Arkangel AI Tool Improves Early Breast Cancer Detection in Latin American Women to AUC 0.90

By Jose Zea

8/15/2025

In Latin America, Arkangel AI's mammogram + clinical-data AI decision support improved detection (AUC up to 0.90) and cut radiologist reading load by up to 88%, enabling earlier, scalable screening.