Easily find secrets in files, directories, and repositories. Stop leaking secrets using git hooks.
-
Updated
Jan 25, 2024 - JavaScript
Easily find secrets in files, directories, and repositories. Stop leaking secrets using git hooks.
Dataset for Training and Evaluating LLM-Based SOC Agents
Cyber Security AI Dashboard
Unified security training dataset (2,185 examples) covering OWASP Top 10 2021 and OWASP LLM Top 10 2025
2,899 real-world malware families categorized for security teams & incident response. Schema.org-ready dataset derived from EMBER 2018 with FAQ, MITRE ATT&CK, CISA advisory cross-refs, and per-family profiles. Apache-2.0 licensed.
Dataset of what websites impose insecure password limits, or crash on strong passwords
AI-ready pentest/bug-bounty scenario dataset: generalized condition/step/impact/remediation playbooks grounded in real, publicly disclosed reports and writeups.
Vulnerability dataset: 16,597 real CVEs with fix commits, CWE labels, root-cause analysis, taint paths and detection heuristics. Two-model cross-check with blind adjudication. Train/eval split for vulnerability detection and security ML research.
Analysis of various aerospace and security data
A structured collection of cybersecurity findings extracted from public penetration testing reports, designed for training and evaluating large language models on offensive (red team) and defensive (blue team) tasks.
Research dataset for IIoT and wireless sensor network cyber threat detection, supporting ML-based anomaly detection and security education.
Agent Skill 恶意检测评测数据集目录:来源可追溯、标签可复现、按许可证独立发布
A structured NLP dataset for detecting prompt injection attacks, jailbreak attempts, and malicious instruction manipulation in Large Language Models (LLMs). Includes annotated threat categories, risk classifications, and validation-ready samples for AI safety training, security evaluation, and adversarial robustness research.
A 100-incident AI security benchmark for evaluating SOC analyst reasoning while under uncertainty.
Security research, datasets, and reproducible findings from AgentGuard across AI agents, plugins, MCP ecosystems, and agent runtimes.
中文提示注入攻击语料库 —— 三项目共享的语料资产层,npm 包 + HuggingFace dataset 同步;代码 MIT / 数据 CC-BY 4.0。Chinese prompt-injection corpus as a versioned shared asset.
To associate your repository with the security-dataset topic, visit your repo's landing page and select "manage topics."