--- name: paro-benchmark description: Run Paro engineering performance gates and exploratory cold/warm, cross-engine or operator comparisons. Default to lightweight diagnosis; use paro-evidence only for formal parity, release or non-inferiority claims. --- # Paro benchmark Default to exploration: answer the specific question with the smallest useful experiment. A pilot needs source/binary, SQL/data, settings, timer scope and raw samples, not a preregistered certification campaign. Label conclusions accordingly. ## Use the selected checkout and maintained runner Inspect HEAD/status, preserve user work, and read [benchmark README](../../../benchmark/README.md). Discover commands from `make -C benchmark help`, the intended CLI's `--help` and its implementation; don't copy stale flags. Inspect recursive Make recipes before a dry run. Reuse the declared Python environment and toolchain. - Engineering workloads/gates: `benchmark/runner.py`, `harness/`, `policies/`. - Corpus A/B and cold/warm: `corpora/tpcds_compare.py`; read [CORPORA](../../../benchmark/CORPORA.md) for data and result contracts. - Compile diagnosis: `corpora/cold_planning.py` and EXPLAIN COMPILE. - Actual operators: EXPLAIN ANALYZE and `corpora/d6_execution_profile.py`. - Known-cardinality checks: `corpora/plan_quality.py`. Use supported collectors/validators; extend their missing capability instead of building another timer or per-report parser. Full-corpus certification is not required for a targeted diagnostic run. Help must not start a server. ## Running the TPC-DS collector The `corpora/` scripts are not importable as a package or runnable by path alone: run them from `benchmark/` with `PYTHONPATH=.:corpora .venv/bin/python corpora/