Skip to main content

Lightweight static analysis for many languages. Find bug variants with patterns that look like source code.

Project description


Semgrep logo

Code scanning at ludicrous speed.

Homebrew PyPI Documentation Join Semgrep community Slack Issues welcome! Star Semgrep on GitHub Docker Pulls Follow @semgrep on Twitter


This repository contains the source code for Semgrep OSS (open-source software). Semgrep OSS is a fast, open-source, static analysis tool for searching code, finding bugs, and enforcing code standards at editor, commit, and CI time. Semgrep is a semantic grep for code: where grep "2" would only match the exact string 2, Semgrep would match x = 1; y = x + 1 when searching for 2. And it does this in 30+ languages! Semgrep rules look like the code you already write; no abstract syntax trees, regex wrestling, or painful DSLs: read more below.

For companies who need SAST, SCA, and Secret scanning, we provide a product suite on top of Semgrep OSS that scans code and package dependencies for known issues, software vulnerabilities, and finds secrets with high accuracy:

  • Semgrep Code to find bugs & vulnerabilities using high-accuracy Pro rules in addition to the community rules
  • Semgrep Supply Chain to find dependencies with known vulnerabilities function-level reachability analysis
  • Semgrep Secrets to find hard-coded credentials that shouldn't be checked into source code

Semgrep analyzes code locally on your computer or in your build environment: by default, code is never uploaded. Get started →.

Semgrep CLI image

Language support

Semgrep Code supports 30+ languages.

Category Languages
GA C# · Go · Java · JavaScript · JSX · JSON · PHP · Python · Ruby · Scala · Terraform · TypeScript · TSX
Beta Kotlin · Rust
Experimental Bash · C · C++ · Clojure · Dart · Dockerfile · Elixir · HTML · Julia · Jsonnet · Lisp · Lua · OCaml · R · Scheme · Solidity · Swift · YAML · XML · Generic (ERB, Jinja, etc.)

Semgrep Supply Chain supports 8 languages across 15 package managers.

Category Languages
GA Go (Go modules, go mod) · Javascript/Typescript (npm, Yarn, Yarn 2, Yarn 3, pnpm) · Python (pip, pip-tool, Pipenv, Poetry) · Ruby (RubyGems) · Java (Gradle, Maven)
Beta C# (NuGet)
Lock file-only Rust (Cargo) · PHP (Composer)

For more information, visit our supported languages page.

Getting started 🚀

  1. From the Semgrep Cloud Platform
  2. From the CLI

For new users, we recommend starting with the Semgrep Cloud Platform because it provides a visual interface, a demo project, result triaging and exploration workflows, and makes setup in CI/CD fast. Scans are still local and code isn't uploaded. Alternatively, you can also start with the CLI and navigate the terminal output to run one-off searches.

Option 1: Getting started from the Semgrep Cloud Platform (Recommended)

Semgrep platform image

  1. Register on semgrep.dev

  2. Explore the demo findings to learn how Semgrep works

  3. Scan your project by navigating to Projects > Scan New Project > Run scan in CI

  4. Select your version control system and follow the onboarding steps to add your project. After this setup, Semgrep will scan your project after every pull request.

  5. [Optional] If you want to run Semgrep locally, follow the steps in the CLI section.

Notes:

If there are any issues, please ask for help in the Semgrep Slack.

Option 2: Getting started from the CLI

  1. Install Semgrep CLI
# For macOS
$ brew install semgrep

# For Ubuntu/WSL/Linux/macOS
$ python3 -m pip install semgrep

# To try Semgrep without installation run via Docker
$ docker run -it -v "${PWD}:/src" returntocorp/semgrep semgrep login
$ docker run -e SEMGREP_APP_TOKEN=<TOKEN> --rm -v "${PWD}:/src" returntocorp/semgrep semgrep ci
  1. Run semgrep login to create your account and login to Semgrep.

Logging into Semgrep gets you access to:

  1. Go to your app's root directory and run semgrep ci. This will scan your project to check for vulnerabilities in your source code and its dependencies.

  2. Try writing your own query interactively with -e. For example, a check for Python == where the left and right hand sides are the same (potentially a bug): $ semgrep -e '$X == $X' --lang=py path/to/src

Semgrep Ecosystem

The Semgrep ecosystem includes the following products:

  • Semgrep Code - Scan your code with Semgrep's proprietary rules (written by our Security Research team) using our cross-file and cross-function analysis. Designed to find OWASP Top 10 vulnerabilities and protect against critical security risks. Semgrep Code is available on both free and paid tiers.
  • Semgrep Supply Chain (SSC) - A high-signal dependency scanner that detects reachable vulnerabilities in open source third-party libraries and functions across the software development life cycle (SDLC). Semgrep Supply Chain is available on both free and paid tiers.
  • Semgrep Secrets [NEW!] - Secrets detection that uses semantic analysis, improved entropy analysis, and validation together to accurately detect sensitive credentials in developer workflows. Book a demo to request early access to the product.
  • Semgrep Cloud Platform (SCP) - Deploy, manage, and monitor Semgrep at scale, with free and paid tiers. Integrates with continuous integration (CI) providers such as GitHub, GitLab, CircleCI, and more.
  • Semgrep OSS Engine - The open-source engine and community-contributed rules at the heart of everything (this project).

To learn more about Semgrep, visit:

  • Semgrep Playground - An online interactive tool for writing and sharing rules.
  • Semgrep Registry - 2,000+ community-driven rules covering security, correctness, and dependency vulnerabilities.

Join hundreds of thousands of other developers and security engineers already using Semgrep at companies like GitLab, Dropbox, Slack, Figma, Shopify, HashiCorp, Snowflake, and Trail of Bits.

Semgrep is developed and commercially supported by Semgrep, Inc., a software security company.

Semgrep Rules

Semgrep rules look like the code you already write; no abstract syntax trees, regex wrestling, or painful DSLs. Here's a quick rule for finding Python print() statements.

Run it online in Semgrep’s Playground by clicking here.

Semgrep rule example for finding Python print() statements

Examples

Visit Docs > Rule examples for use cases and ideas.

Use case Semgrep rule
Ban dangerous APIs Prevent use of exec
Search routes and authentication Extract Spring routes
Enforce the use secure defaults Securely set Flask cookies
Tainted data flowing into sinks ExpressJS dataflow into sandbox.run
Enforce project best-practices Use assertEqual for == checks, Always check subprocess calls
Codify project-specific knowledge Verify transactions before making them
Audit security hotspots Finding XSS in Apache Airflow, Hardcoded credentials
Audit configuration files Find S3 ARN uses
Migrate from deprecated APIs DES is deprecated, Deprecated Flask APIs, Deprecated Bokeh APIs
Apply automatic fixes Use listenAndServeTLS

Extensions

Visit Docs > Extensions to learn about using Semgrep in your editor or pre-commit. When integrated into CI and configured to scan pull requests, Semgrep will only report issues introduced by that pull request; this lets you start using Semgrep without fixing or ignoring pre-existing issues!

Documentation

Browse the full Semgrep documentation on the website. If you’re new to Semgrep, check out Docs > Getting started or the interactive tutorial.

Metrics

Using remote configuration from the Registry (like --config=p/ci) reports pseudonymous rule metrics to semgrep.dev.

Using configs from local files (like --config=xyz.yml) does not enable metrics.

To disable Registry rule metrics, use --metrics=off.

The Semgrep privacy policy describes the principles that guide data-collection decisions and the breakdown of the data that are and are not collected when the metrics are enabled.

More

Upgrading

To upgrade, run the command below associated with how you installed Semgrep:

# Using Homebrew
$ brew upgrade semgrep

# Using pip
$ python3 -m pip install --upgrade semgrep

# Using Docker
$ docker pull returntocorp/semgrep:latest

Project details


Release history Release notifications | RSS feed

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

semgrep-1.60.0.tar.gz (35.5 MB view details)

Uploaded Source

Built Distributions

semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-musllinux_1_0_aarch64.manylinux2014_aarch64.whl (38.5 MB view details)

Uploaded CPython 3.10 CPython 3.11 CPython 3.8 CPython 3.9 Python 3.10 Python 3.11 Python 3.7 Python 3.8 Python 3.9 musllinux: musl 1.0+ ARM64

semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-macosx_11_0_arm64.whl (35.2 MB view details)

Uploaded CPython 3.10 CPython 3.11 CPython 3.8 CPython 3.9 Python 3.10 Python 3.11 Python 3.7 Python 3.8 Python 3.9 macOS 11.0+ ARM64

semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-macosx_10_14_x86_64.whl (30.8 MB view details)

Uploaded CPython 3.10 CPython 3.11 CPython 3.8 CPython 3.9 Python 3.10 Python 3.11 Python 3.7 Python 3.8 Python 3.9 macOS 10.14+ x86-64

semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-any.whl (36.1 MB view details)

Uploaded CPython 3.10 CPython 3.11 CPython 3.8 CPython 3.9 Python 3.10 Python 3.11 Python 3.7 Python 3.8 Python 3.9

File details

Details for the file semgrep-1.60.0.tar.gz.

File metadata

  • Download URL: semgrep-1.60.0.tar.gz
  • Upload date:
  • Size: 35.5 MB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/4.0.2 CPython/3.11.8

File hashes

Hashes for semgrep-1.60.0.tar.gz
Algorithm Hash digest
SHA256 9971beb735a567758863e5bbce0e2efae01c3ca8d9d11a1d3b07ebfc7f3246b7
MD5 b1dea5cc7f39428aefc21b2bf34e41c8
BLAKE2b-256 e2f4b81a2701c92afac84132f7424d4abb45f84ee5a51121bc0e146a20e47d00

See more details on using hashes here.

File details

Details for the file semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-musllinux_1_0_aarch64.manylinux2014_aarch64.whl.

File metadata

File hashes

Hashes for semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-musllinux_1_0_aarch64.manylinux2014_aarch64.whl
Algorithm Hash digest
SHA256 b8cc631a5a85798f3f64e83377054abe3dfa361d94c9a27ece36b3cd1171c6e8
MD5 55eb60ee9610cdb4aded224ffa15c0a6
BLAKE2b-256 6057070546fd8907b2f8e8ba4871c59e66ea74e934775b994df9124481d2b5fb

See more details on using hashes here.

File details

Details for the file semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-macosx_11_0_arm64.whl.

File metadata

File hashes

Hashes for semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-macosx_11_0_arm64.whl
Algorithm Hash digest
SHA256 4ffd07d37b659af20c052a9b1789cd4b63f53643451cc89d92685fef8861f03a
MD5 9573f4e1694d3453309c373300a0087b
BLAKE2b-256 1eaa217dc121ab33456ae6caf9537ab0c8ff96f13063285aaadc42e4e25546ac

See more details on using hashes here.

File details

Details for the file semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-macosx_10_14_x86_64.whl.

File metadata

File hashes

Hashes for semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-macosx_10_14_x86_64.whl
Algorithm Hash digest
SHA256 d47b425f84e1990b7f200668928fa1b3651f41358eca1d35a0658fd11d6b0584
MD5 768eaa5e4353a8eddea041343e53aad9
BLAKE2b-256 0082ae0e65229cc1f85ce4f16245336ecf67fe7b8988144a60f4be5190145b2c

See more details on using hashes here.

File details

Details for the file semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-any.whl.

File metadata

File hashes

Hashes for semgrep-1.60.0-cp38.cp39.cp310.cp311.py37.py38.py39.py310.py311-none-any.whl
Algorithm Hash digest
SHA256 c72fce30465527f751123e23d9a75919983f95e38a7fb482e25363cb0e6d915f
MD5 e869f0dabc17cff4c8a6ed289de32000
BLAKE2b-256 233e7373fcd8c502d8648900f29b490dfab7e5b0c2fa3098df9894908e93047f

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page