Grounding, verification-aware scientific agents, mission-held-out space biology, and laboratory-agent evaluation.
JangKeun Kim
jang1563
AI & ML interests
None yet
Recent Activity
updated a dataset 4 days ago
jang1563/narrow-model-safety-eval updated a dataset 12 days ago
jang1563/agentic-drug-discovery-system updated a dataset 14 days ago
jang1563/ambiguity-casebookOrganizations
Biological AI Evaluation & Scientific Agents
Grounding, verification-aware scientific agents, mission-held-out space biology, and laboratory-agent evaluation.
Scientific Agents & Drug Discovery Evaluation
Auditable benchmarks for scientific agents: drug development decisions, calibrated abstention, tool use, and specialist-model reliability.
models 4
jang1563/constitutional-bioguard-v4
Text Classification • 0.2B • Updated • 3
jang1563/constitutional-bioguard-response
Text Classification • 0.2B • Updated
jang1563/constitutional-bioguard-deberta-v1
Text Classification • 0.2B • Updated • 13
jang1563/constitutional-bioguard-prompt
Text Classification • 0.2B • Updated
datasets 24
jang1563/narrow-model-safety-eval
Viewer • Updated • 8 • 40
jang1563/agentic-drug-discovery-system
Updated • 1.12k
jang1563/ambiguity-casebook
Viewer • Updated • 35 • 82
jang1563/sci-agent-verification-cascade
Viewer • Updated • 69 • 54
jang1563/bioreview-bench
Viewer • Updated • 100k • 71
jang1563/SpaceOmicsBench
Updated • 124
jang1563/causalatlas-move1
Updated • 187
jang1563/clinical-trial-decision-benchmark
Viewer • Updated • 1.48k • 61
jang1563/verify-or-trust
Viewer • Updated • 4.01k • 33 • 1
jang1563/genelab-benchmark
Updated • 135