Back to dataset: Replication Code for: "From Constrictor to Serpent: Investigating the Threat of Cache Poisoning in the Python Ecosystem"

pyproject.toml

Important usage condition: Use and redistribution are governed by MIT. Review and comply with the license before using the data.
Dataset
Replication Code for: "From Constrictor to Serpent: Investigating the Threat of Cache Poisoning in the Python Ecosystem"
Published version
1.0
Media type
application/toml
Size
195 bytes
Checksum
MD5 250b03649f41a107784886f04056b12b
License
MIT

Before downloading: Use and redistribution are governed by MIT. Review and comply with the license before using the data.

Download pyproject.toml

Expected crawler behaviour

Use a stable, truthful User-Agent with product/version and a working contact URL. Across all IP addresses and HTTP connections used by one crawler identity, allow no more than 5 requests in flight and wait at least 20 seconds between request starts. Crawl URLs listed in the catalog sitemap, including file pages and download URLs when they are published, use conditional requests, honor Retry-After, and apply exponential backoff after errors.

The welcome page may link to the interactive repository for human navigation. Automated clients must not treat that human link as a catalog crawl target.

Read the live machine-readable crawler policy before and during a crawl. Stop crawling when it reports CPU or memory utilization at or above 80% and 80% respectively.