Publications

Author names in bold indicate my contribution. See also Google Scholar and ORCID.

Preprints and papers under review

KidnapRAG: A Black-Box Attack for Hijacking Reasoning in Agentic Retrieval-Augmented Generation Systems Preprint

Industry-academia collaboration with Prof. Buru Chang's group (Korea University, lead author). I led the KT side of the project, am fourth author on the paper, and hold first-inventor priority on the registered Korean patent (Jul 2026). ACL ARR 2026 May submission.

ARS-Bench: Benchmarking Security Risks Across the Agentic RAG Attack Lifecycle Under review

A benchmark that evaluates security risks across the entire agentic RAG attack lifecycle, from adversarial document ingestion to final response generation, structured along five security-critical dimensions: document ingestion, retrieval, reasoning, final response, and efficiency. Tracing attack effects through the lifecycle enables fine-grained diagnosis of where vulnerabilities enter and how defences should be targeted.

Silent Failure of Moral Alignment in Large Language-Model Evaluators Under review

Peer-reviewed workshop papers

Moral Orientation and Calibration: Coupled in Human Annotators, Separable in Judge LLMs Accepted

Technical reports

A Hybrid Corpus-to-Taxonomy Pipeline for Physical AI Risk Cards

Documents the data collection, deduplication, and hierarchical assignment pipeline behind the Physical AI Risk Taxonomy, with reproducibility artefacts.

Journal articles

AI Technology Specialization and National Competitiveness

The Nexus of Artificial Intelligence and Green Innovation: A Cross-Density Analysis at the Country Level

The Moderating Role of Design Innovation in the Relationship Between Technology Complexity and Firm Performance

Manuscripts in preparation

When Fear Becomes Beauty: Genre as Context in Human and AI Aesthetic Judgment

Cross-institutional collaboration between the Singapore AI Safety Hub (SASH) and KT Frontier AI Lab.

Patents

Method and Apparatus for Automatically Constructing a Physical AI Risk Taxonomy Filed 2026

Korean patent application filed 3 August 2026, covering the automated construction of a risk classification system for Physical AI, arising from the Physical AI Risk Taxonomy project.

Method, Server and Computer Program for Dynamically Adjusting Value Information Filed 2025

A value-alignment method for generative AI combining multi-agent debate with reinforcement-learning-based reward modelling, so systems can track evolving ethical and social values without continuous human intervention. Received the Excellent Patent Award, KT R&D Center Invention Competition 2025.

Reasoning-Chain Hijacking Attack and Detection for Agentic AI Registered 2026

KT-funded and filed ahead of the corresponding paper, with priority held as first inventor.

Selected talks