1,200 robot-view video scenarios testing whether VLM guards catch real hazards without over-alarming
Context-flip evaluation revealing models cling to safety rules even when the safe action becomes harmful
Country-grounded safety benchmark with 5,500 cases across 10 country-language pairs
Embodiment-agnostic action embeddings anchoring humanoid robots to a shared human motion manifold
Multi-turn adversarial red-teaming of LLM operator agents in a simulated nuclear control room
SODA framework revealing demographic bias in generated objects across 8,000 images
VLMs misclassify 31–96% of safe situations as dangerous in visual emergency recognition
A unified framework for evaluating the Korean capabilities of language models
First systematic black-box jailbreak for T2V models, achieving 70–84% success rates
Jailbreaking audio-language models with benign-sounding adversarial inputs
Framework for evaluating LLM compliance with enterprise-specific allow/deny policies
HAERAE-Vision benchmark revealing VLM failures on ambiguous and incomplete queries
A comprehensive benchmark for evaluating object hallucination across multiple images in vision-language models
Motion-aware approach for referring image segmentation with improved temporal understanding
Frontier AI evaluation benchmark with 3,000+ expert-level questions across disciplines
Conversational red-teaming to elicit misalignment through narrative immersion and emotional pressure
Compressing multi-turn jailbreak strategies into single-turn attacks
Manipulating LLM internal representations to reduce harmful outputs
Safety evaluation framework for patient-facing medical AI systems
Benchmark for evaluating toxicity across language-image multimodal models
LLM-guided evolution for automated multi-turn jailbreak template discovery
Benchmarking whether LLM judges can recover hidden objectives in jailbreak transcripts
Systematic security analysis of autonomous AI agent vulnerabilities · Industry Track
Federated learning for heterogeneous resource-constrained clients
* 심사 중 / 게재 목표 학회
DB25-0017-KR0
DB25-0016-KR0
DB25-0015-KR0
10-2025-0038904
10-2024-0124354
10-2024-0116863
공동 논문 발표부터 엔터프라이즈 보안 평가까지, AI를 더 안전하게 만드는 길을 함께 걷겠습니다.