mirror of
https://github.com/vinta/awesome-python.git
synced 2026-10-02 08:23:10 +08:00
Provides per-sitting download evidence for prune sweeps, per the shortlist-reform tooling plan. Shells out to the bq CLI against bigquery-public-data.pypi.file_downloads, parses entry names from README.md via readme_parser, and supports --dry-run and --names-file. Merges results into the gitignored cache at website/data/pypi_downloads.tsv. The table is clustered on file.project, so scanned bytes grow with the IN-list size: a dry run against the full README (~530 names) scanned 1.21 TB, past the 1 TB/month free tier. Per-sitting --names-file fetches are used instead of one big query. Co-Authored-By: Claude <noreply@anthropic.com>