Abstract

As adolescents increasingly use LLMs in everyday life, ensuring safe and developmentally appropriate responses has become essential. However, existing LLM guardrails primarily target explicit harmful content in isolated prompts or responses and are less effective at identifying implicit, context-dependent developmental risks. To address this limitation, we propose SaplingGuard, a plug-and-play, profile-aware and dialogue-aware guardrail that requires no modification to downstream model parameters. SaplingGuard decomposes adolescent safety intervention into three specialized agents for user profile construction, context-aware risk assessment, and intent-preserving prompt optimization. Together, these agents leverage the current prompt, preceding dialogue, and structured user characteristics to identify contextual risks and guide downstream response generation. We evaluate SaplingGuard on SaplingBench, which contains 276 three-turn dialogues spanning seven categories of developmental risk. Across ten adolescent profile conditions and nine open- and closed-source downstream LLMs, profile-aware retrieval improves the Major Hit rate from 50.8% to 63.7+/-1.1%. End-to-end intervention further reduces the average harmful response rate from 17.10% to 5.27% and increases the average safety score from 0.7017 to 1.0043. These results show that user-profile and dialogue context provide complementary signals for identifying implicit developmental risks, and that SaplingGuard can serve as an effective external safety layer for adolescent-LLM interaction.

Keywords

Subject

Publication details

Journal
Not available
Open access
Green open access

Cite this article

APA 7

Tan, J., Liu, Y., Lin, Y., Guo, X., Wang, Z., Zhao, X., Ma, L., Yao, X., & Wei, X. (2026). SaplingGuard: A Multidimensional-Profile-Aware Multi-Agent Guardrail for Developmentally Safe Adolescent-LLM Interaction. https://omanscience.com/en/articles/saplingguard-a-multidimensional-profile-aware-multi-agent-guardrail-for-developmentally-safe-adolescent-llm-interaction

MLA 9

Tan, Jing, et al. "SaplingGuard: A Multidimensional-Profile-Aware Multi-Agent Guardrail for Developmentally Safe Adolescent-LLM Interaction." https://omanscience.com/en/articles/saplingguard-a-multidimensional-profile-aware-multi-agent-guardrail-for-developmentally-safe-adolescent-llm-interaction.

Chicago (author–date)

Tan, Jing, Yifan Liu, Yi Lin, Xinwei Guo, Ziwei Wang, Xiangyu Zhao, Lei Ma, Xin Yao, and Xuetao Wei. 2026. "SaplingGuard: A Multidimensional-Profile-Aware Multi-Agent Guardrail for Developmentally Safe Adolescent-LLM Interaction." https://omanscience.com/en/articles/saplingguard-a-multidimensional-profile-aware-multi-agent-guardrail-for-developmentally-safe-adolescent-llm-interaction.

Harvard

Tan, J., Liu, Y., Lin, Y., Guo, X., Wang, Z., Zhao, X., Ma, L., Yao, X. and Wei, X. (2026) 'SaplingGuard: A Multidimensional-Profile-Aware Multi-Agent Guardrail for Developmentally Safe Adolescent-LLM Interaction', Available at: https://omanscience.com/en/articles/saplingguard-a-multidimensional-profile-aware-multi-agent-guardrail-for-developmentally-safe-adolescent-llm-interaction.

Vancouver

Tan J, Liu Y, Lin Y, Guo X, Wang Z, Zhao X, et al. SaplingGuard: A Multidimensional-Profile-Aware Multi-Agent Guardrail for Developmentally Safe Adolescent-LLM Interaction. https://omanscience.com/en/articles/saplingguard-a-multidimensional-profile-aware-multi-agent-guardrail-for-developmentally-safe-adolescent-llm-interaction

IEEE

J. Tan, Y. Liu, Y. Lin, X. Guo, Z. Wang, X. Zhao, L. Ma, X. Yao, and X. Wei, "SaplingGuard: A Multidimensional-Profile-Aware Multi-Agent Guardrail for Developmentally Safe Adolescent-LLM Interaction," https://omanscience.com/en/articles/saplingguard-a-multidimensional-profile-aware-multi-agent-guardrail-for-developmentally-safe-adolescent-llm-interaction.