
OpenAI open-sources Codex Security CLI to help developers find and fix vulnerabilities from the command line
Quick Answer
OpenAI has launched the open-source Codex Security CLI, a command-line tool for developers to identify and fix vulnerabilities in code repositories.
Quick Take
This tool, currently in beta, supports bulk scans, CI/CD integration, and requires Node.js 22 and Python 3.10 or higher. It has already helped fix over 3,000 critical vulnerabilities since its preview release in March 2026.
Key Points
- Codex Security CLI is licensed under Apache 2.0 and available via npm.
- The tool can scan multiple repositories and verify fixes across runs.
- It was previously known as 'Aardvark' and launched in March 2026.
- Codex Security has already fixed over 3,000 critical vulnerabilities.
- It competes with Anthropic's Claude Security in the vulnerability scanning space.
📖 Reader Mode
~1 min readOpenAI has released Codex Security CLI. The open-source command-line tool, licensed under Apache 2.0, helps security and development teams automatically find, confirm, and fix vulnerabilities in code repositories. Codex Security CLI can scan repositories, compare results across multiple runs, verify fixes, and plug security checks into CI/CD pipelines. Bulk scans across multiple repositories are also supported. The tool requires Node.js 22 and Python 3.10 or higher, is currently in beta, and installs via npm. The documentation covers all commands and output formats.
Codex Security was previously known internally as "Aardvark" and launched in March 2026 as a research preview for ChatGPT Enterprise, Business, and Edu customers. By April 2026, the system had helped fix more than 3,000 critical vulnerabilities, according to OpenAI. Codex Security competes directly with Anthropic's Claude Security, which also scans codebases for vulnerabilities and suggests patches. Both tools reflect a growing reality: as AI models give attackers more automated capabilities, defenders need to keep up.
— Originally published at the-decoder.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

