chore(ruff): the acceptance was whatever the default happened to be [skip-docs]

`uv sync --frozen` resolved ruff 0.15.22 and the tree read clean. A loose
install resolves 0.16.6, under which the SAME untouched code reports 148
findings -- 4 more than round 9 counted, because this round added four files.
All of them are new rules rather than new defects: 0.16 widened the default
rule set to whole families (YTT, ASYNC, PL, ISC, C4, UP, B, SIM, FURB, ...).

(`[skip-docs]` is for CLAUDE.md, which a lint-configuration change does not
reach. README's developer section IS updated in this commit.)

THE DEFECT IS NOT THE 148, IT IS THAT NOBODY CHOSE THEM. `[tool.ruff]` set only
`line-length` and `target-version`, so the acceptance was ruff's default, and
the tree stayed green only as long as the lockfile froze an old ruff. `select`
is now written down: `E4`, `E7`, `E9`, `F` (the historical default), `I`
because this tree already keeps imports sorted, and `RUF100` so a `noqa` that
has stopped meaning anything is caught rather than left as decoration. Pin
`ruff>=0.9` -> `ruff>=0.16.6,<0.17`.

Per rule, before -> after: RUF100 50 -> 0, I001 20 -> 0, ISC004 19, PLW1510 8,
C408 8, EXE001 6, RUF007 5, PLE2515 4, UP031 3, B017 3, and fourteen more with
2 or fewer -- the families out of the declared set are 0 by selection, and 148
is the number to start from if they are adopted, which is a separate decision
and not one to take inside a version-pin commit. 57 were auto-fixed; one E402
was reintroduced by the import-sorting fix merging a block away from its
`noqa`, and got the directive back rather than a bare one.

`S` IS MEASURED OUT, NOT ASSUMED OUT: it reports 2657 `S101` on a suite whose
every assertion is an `assert`, and `S603` flags 19 subprocess calls of which
one was ever marked -- selecting it buys 18 suppressions and no defect. Two
`noqa` directives naming non-selected rules were dropped with that reason
recorded in the configuration instead.

THE TWO FILES 0.16 WOULD REFORMAT ARE MARKDOWN, NOT PYTHON: `README.md` and
`docs/2026-09-08-blindsone-below-k-k2.md`. 0.16 formats fenced Python inside
markdown, and both blocks are RECORDS -- the second is a quotation of
`COST_VOCABULARY` as it stood when that measurement was taken. Reformatting a
quotation makes it stop being one, so markdown is excluded from the formatter
and `ruff format --check .` stays in the acceptance over `.py`.

`tools/okf_consume_measure.py` is fenced by the order as run-not-edited, so its
three findings are exempted by path with the reason and the debt named, and its
bytes are untouched.

THE LOCKFILE TRAP IS CLOSED, NOT AVOIDED. `uv.lock` predated the `[ocr]` extra,
so any unlocked resolve wrote that extra's transitive tree back into it -- 681
insertions over 4 deletions, twice now, and round 9 recorded the cause as
`uv run` OUTSIDE the project when it is `uv run` without `--frozen` INSIDE it.
The relock is complete for every declared extra (703 insertions, 26 deletions),
and measured after it, an unfrozen `uv run` leaves the file alone.

`ruff check src tests tools`, `ruff format --check .` (0.16.6), `mypy src` over
21 files and 1535 tests, all green.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
This commit is contained in:
Kjell Tore Guttormsen 2026-09-09 23:14:51 +02:00
commit 36c201cc8a
37 changed files with 843 additions and 105 deletions

View file

@ -39,8 +39,8 @@ from typing import Any
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf.errors import SegmentationError # noqa: E402
from llm_ingestion_okf.segmentation import PLAN_FIELDS, parse_segmentation_plan # noqa: E402
from llm_ingestion_okf.errors import SegmentationError
from llm_ingestion_okf.segmentation import PLAN_FIELDS, parse_segmentation_plan
#: This tool's identity, written into the artifact so an operator reading a
#: verdict six months later can tell what produced it.

View file

@ -27,8 +27,8 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf.errors import ExtractionError # noqa: E402
from llm_ingestion_okf.extract import extract_text # noqa: E402
from llm_ingestion_okf.errors import ExtractionError
from llm_ingestion_okf.extract import extract_text
#: What "a CID glyph code" means here -- pdfminer.six's placeholder for a
#: character its font's encoding cannot map to Unicode. Not an upstream term:

View file

@ -27,7 +27,7 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf import consume as _impl # noqa: E402
from llm_ingestion_okf import consume as _impl
if __name__ == "__main__":
raise SystemExit(_impl.main())

View file

@ -20,7 +20,7 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf import contract_check as _impl # noqa: E402
from llm_ingestion_okf import contract_check as _impl
if __name__ == "__main__":
raise SystemExit(_impl.main())

View file

@ -19,7 +19,7 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf.corpus import main # noqa: E402
from llm_ingestion_okf.corpus import main
if __name__ == "__main__":
raise SystemExit(main())

View file

@ -35,7 +35,7 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf.extract import extract_text # noqa: E402
from llm_ingestion_okf.extract import extract_text
#: A source string shorter than this is not evidence either way -- a lone digit
#: or a stray bullet appears in almost any output by accident, so counting it

View file

@ -46,10 +46,10 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf.errors import ExtractionError # noqa: E402
from llm_ingestion_okf.extract import extract_text # noqa: E402
from llm_ingestion_okf.materialize import reduce_to_id_grammar # noqa: E402
from llm_ingestion_okf.propose import ( # noqa: E402
from llm_ingestion_okf.errors import ExtractionError
from llm_ingestion_okf.extract import extract_text
from llm_ingestion_okf.materialize import reduce_to_id_grammar
from llm_ingestion_okf.propose import (
RULE_OUTLINE,
Candidate,
_segment_path,

View file

@ -19,7 +19,7 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf.propose import main # noqa: E402
from llm_ingestion_okf.propose import main
if __name__ == "__main__":
raise SystemExit(main())

View file

@ -19,7 +19,7 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
from llm_ingestion_okf import skill as _impl # noqa: E402
from llm_ingestion_okf import skill as _impl
if __name__ == "__main__":
raise SystemExit(_impl.main())

View file

@ -44,15 +44,15 @@ from pathlib import Path
sys.path.insert(0, str(Path(__file__).resolve().parents[1] / "src"))
sys.path.insert(0, str(Path(__file__).resolve().parent))
from okf_outline_measure import draw_sample # noqa: E402
from okf_outline_measure import draw_sample
from llm_ingestion_okf.errors import ExtractionError # noqa: E402
from llm_ingestion_okf.extract import extract_text # noqa: E402
from llm_ingestion_okf.propose import ( # noqa: E402
RULE_TABLE_BLOCK,
RULE_TABLE_GRID,
from llm_ingestion_okf.errors import ExtractionError
from llm_ingestion_okf.extract import extract_text
from llm_ingestion_okf.propose import (
_GRID_RULE,
_TABLE_ROW,
RULE_TABLE_BLOCK,
RULE_TABLE_GRID,
find_candidates,
)