Safeguard evaluation where scientific AI meets the lab: evidence across agent handoffs, calibration across interfaces, hazard under misleading labels.
JangKeun Kim
jang1563
AI & ML interests
None yet
Recent Activity
updated a dataset about 7 hours ago
jang1563/genelab-benchmark updated a dataset about 7 hours ago
jang1563/clinical-trial-decision-benchmark updated a dataset about 21 hours ago
jang1563/narrow-model-safety-evalOrganizations
AI Safety for Biological Research
Safeguard evaluation where scientific AI meets the lab: evidence across agent handoffs, calibration across interfaces, hazard under misleading labels.
Biological AI Evaluation & Scientific Agents
Grounding, verification-aware scientific agents, mission-held-out space biology, and laboratory-agent evaluation.
models 4
jang1563/constitutional-bioguard-v4
Text Classification • 0.2B • Updated • 8
jang1563/constitutional-bioguard-response
Text Classification • 0.2B • Updated
jang1563/constitutional-bioguard-deberta-v1
Text Classification • 0.2B • Updated • 13
jang1563/constitutional-bioguard-prompt
Text Classification • 0.2B • Updated
datasets 24
jang1563/genelab-benchmark
Updated • 298
jang1563/clinical-trial-decision-benchmark
Viewer • Updated • 2.48k • 67
jang1563/narrow-model-safety-eval
Viewer • Updated • 8 • 580
jang1563/llm-sfm-safety-eval
Viewer • Updated • 24.3k • 207
jang1563/sci-agent-verification-cascade
Viewer • Updated • 69 • 87
jang1563/biothreat-eval
Viewer • Updated • 72 • 59
jang1563/bio-overrefusal-v0.1
Viewer • Updated • 201 • 216
jang1563/SpaceOmicsBench-v3
Viewer • Updated • 26.9k • 102
jang1563/protein-structure-trust-benchmark
Viewer • Updated • 238 • 53
jang1563/evo2-spaceflight-vep
Viewer • Updated • 215k • 37