Safeguard evaluation where scientific AI meets the lab: evidence across agent handoffs, calibration across interfaces, hazard under misleading labels.
JangKeun Kim
jang1563
AI & ML interests
None yet
Recent Activity
updated a dataset 5 days ago
jang1563/genelab-benchmark updated a dataset 5 days ago
jang1563/narrow-model-safety-eval updated a dataset 6 days ago
jang1563/clinical-trial-decision-benchmarkOrganizations
Biological AI Evaluation & Scientific Agents
Grounding, verification-aware scientific agents, mission-held-out space biology, and laboratory-agent evaluation.
Space Biology and Multi-Omics
Spaceflight biomedical benchmarks and omics / foundation-model resources across NASA, JAXA, and Inspiration4 missions.
AI Safety for Biological Research
Safeguard evaluation where scientific AI meets the lab: evidence across agent handoffs, calibration across interfaces, hazard under misleading labels.
Biological AI Evaluation & Scientific Agents
Grounding, verification-aware scientific agents, mission-held-out space biology, and laboratory-agent evaluation.
Scientific Agents & Drug Discovery Evaluation
Auditable benchmarks for scientific agents: drug development decisions, calibrated abstention, tool use, and specialist-model reliability.
Space Biology and Multi-Omics
Spaceflight biomedical benchmarks and omics / foundation-model resources across NASA, JAXA, and Inspiration4 missions.