SemGuard: A Triple-Anchor Semantic Security Gateway for Multilingual Prompt Attack Detection in Large Language Models
SemGuard: A Triple-Anchor Semantic Security Gateway for Multilingual Prompt Attack Detection in Large Language Models Source: https://www.linkedin.com/safety/go/?url=https%3A%2F%2Fgithub%2Ecom%2FAbdaullahAG%2FSemGuard&urlhash=xA2K&mt=_QOfz9VNuMJfYK8ZZrRXV8rIsAVagy7I1X-qxe_HjXjcc45042OqTHPqfZGl016z68OO3ia8qS_ImVoi2_PgkKHmTg8hKXLTj5UkRvd68W5yqFOYq0wzGK8d6UHbWyVYKrAO7yrvqpeNvpRJcsT9YaKXeQQwyVEz9iw&isSdui=true
SemGuard: A Triple-Anchor Semantic Security Gateway for Multilingual Prompt Attack Detection in Large Language Models
Description
SemGuard: A Triple-Anchor Semantic Security Gateway for Multilingual Prompt Attack Detection in Large Language Models Source: https://www.linkedin.com/safety/go/?url=https%3A%2F%2Fgithub%2Ecom%2FAbdaullahAG%2FSemGuard&urlhash=xA2K&mt=_QOfz9VNuMJfYK8ZZrRXV8rIsAVagy7I1X-qxe_HjXjcc45042OqTHPqfZGl016z68OO3ia8qS_ImVoi2_PgkKHmTg8hKXLTj5UkRvd68W5yqFOYq0wzGK8d6UHbWyVYKrAO7yrvqpeNvpRJcsT9YaKXeQQwyVEz9iw&isSdui=true
Reddit Discussion
For researchers working on LLM security, prompt injection detection, or Arabic NLP:
Our paper "SemGuard: A Triple-Anchor Semantic Security Gateway for Multilingual Prompt Attack Detection in Large Language Models" is now live on IEEE Xplore.
It introduces the first formally validated Arabic LLM security dataset (807 examples, 7 threat categories, Fleiss' κ = 0.839) and a novel Triple-Anchor framework for explainable, multilingual threat detection — achieving 0.989 F1 / 0.991 recall on Arabic prompt injection, outperforming English-only baselines.
If this overlaps with your work on LLM safety, multilingual NLP, or adversarial robustness, happy to discuss or collaborate. Citations and feedback welcome.
DOI: 10.1109/AEECT69724.2026.11657880
the project on Github : https://github.com/AbdaullahAG/SemGuard
Dataset : https://huggingface.co/datasets/AG-31625874/SemGuard-Dataset
#LLMSecurity #PromptInjection #ArabicNLP #AISecurity #ExplainableAI #NLP
Links cited in this discussion
Technical Details
- Source Type
- Subreddit
- cybersecurity
- Reddit Score
- 0
- Discussion Level
- minimal
- Content Source
- reddit_link_post
- Post Type
- link
- Domain
- null
- Newsworthiness Assessment
- {"score":27,"reasons":["external_link","established_author","very_recent"],"isNewsworthy":true,"foundNewsworthy":[],"foundNonNewsworthy":[]}
- Has External Source
- true
- Trusted Domain
- false
Threat ID: 6a8e9097acd9273b49877861
Added to database: 08/26/2026, 07:07:03 UTC
Last updated: 08/26/2026, 21:52:08 UTC
Views: 15
Community Reviews
0 reviewsCrowdsource mitigation strategies, share intel context, and vote on the most helpful responses. Sign in to add your voice and help keep defenders ahead.
Want to contribute mitigation steps or threat intel context? Sign in or create an account to join the community discussion.
Actions
External Links
Need more coverage?
Upgrade to Pro Console for AI refresh and higher limits.
For incident response and remediation, OffSeq services can help resolve threats faster.
Latest Threats
Check if your credentials are on the dark web
Instant breach scanning across billions of leaked records. Free tier available.