mirror of
https://github.com/ChuckBuilds/LEDMatrix.git
synced 2026-08-07 19:58:08 +00:00
* ci: run the whole test tree and make the plugin-safety job assert something real The unit-tests CI job ran an explicit 24-file allowlist that had rotted: 63 of 90 test files (display, vegas, store manager, web API, web_interface) never ran on a PR. The job now runs all of test/ (minus test/plugins, which the plugin-safety job owns) so new test files are enrolled by default and any exclusion needs a visible, commented --ignore. The plugin-safety job was a green no-op: plugins/ is empty in CI, so every test skipped with 'Manifest not found'. It now renders a bundled deterministic fixture plugin (test/fixtures/plugins/ci-fixture-plugin, golden images included for all 8 default sizes) via LEDMATRIX_PLUGINS_DIR, and sets LEDMATRIX_REQUIRE_PLUGINS=1 so discovering zero plugins fails loudly instead of skipping green. The per-plugin suites document that they target dev machines with real plugins installed. Coverage is now measured and enforced in exactly one place — the CI unit-tests step (--cov=src --cov=web_interface --cov-fail-under=45, from a measured 47% baseline). pytest.ini previously declared --cov-fail-under=30 but CI always passed --no-cov, so the gate had never run anywhere; local pytest is now coverage-free and fast. Enabling the 63 unenrolled files surfaced three cases of test rot, fixed here: test_display_controller_vegas_tick.py could not collect without the hardware rgbmatrix module (now uses the emulator convention), the state-reconciliation unrecoverable-cache tests broke when production added the is_plugin_uninstalled tombstone check (bare Mock returned truthy), and test_get_system_status assumed the optional psutil dependency (now installed via requirements-test.txt and guarded by importorskip). Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh * test: replace can't-fail tests with real assertions test_font_manager.py was 5 of 6 tests shaped as 'try: call(); assert True / except: assert True' — running in CI while unable to fail on any regression. Rewritten against the real FontManager API and the bundled assets/fonts: returned font types, cache-hit identity, distinct entries per size, default-font fallback for unknown families and corrupt files (recorded in failed_loads), BDF native-size reading, text measurement, and cache lifecycle. test_display_manager.py's test_draw_text ended in 'assert True'; it now renders onto a known-black canvas and asserts pixels were actually lit — which required un-breaking the fixture's freetype MagicMock so draw_text's isinstance check doesn't silently swallow the draw. test_display_controller.py carried a permanently-skipped test whose skip reason already declared it redundant; deleted. Both display test files now set EMULATOR=true before importing display_manager (the same convention as test_display_dirty_tracking.py) so they collect standalone instead of depending on which test module imports display_manager first. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh * test: cover the untested fragile logic (compatibility gate, secrets, config merges, durations, skin cards) New unit tests for pure or filesystem-only logic that previously had zero direct coverage: - test_compatibility.py: the semver install gate (parse_semver suffix handling, every range operator, TRUSTWORTHY_FLOOR behavior for cores reporting untrustworthy versions, 'more restrictive wins', and the malformed-manifest shapes that used to raise). - test/web_interface/test_secret_helpers.py: the canonical x-secret helpers — find/separate/mask/remove, array-item secrets, no input mutation, and a separate->recombine round-trip. - test/web_interface/test_api_v3_helpers.py: the module-level helpers behind the plugin config save endpoint (_is_plugin_update_available, _coerce_to_bool including the int==1 quirk, deep_merge including its shared-subtree shallowness, _parse_form_value, dotted-key-aware _get_schema_property/_set_nested_value). - test_base_plugin_duration.py: get_display_duration's full coercion ladder (instance attr -> config -> 15.0), including the bool-is-int quirk where display_duration=True means one second. - test_config_manager_secrets.py: the secrets round-trip — deep-merge on load, strip on save, group pruning, the load fast path — and two characterized sharp edges marked SUSPECTED BUG: an unreadable secrets file at save time writes secrets into config.json in plaintext, and a same-mtime-same-size content swap is served stale. - test_schema_manager_merge.py: merge_with_defaults branch behavior (None replacement vs falsey preservation, dict-vs-scalar mismatches, arrays replaced wholesale, defaults never mutated). - test_skin_system.py (extended): render_skin_card shares _render_game's 3-strike counter but never resets it on success — the asymmetry is pinned in both directions, along with card fallthrough and the disable interaction between the two paths. Suspected bugs are characterized, not fixed — each carries a comment so a future behavior change is deliberate rather than accidental. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh * test: add drift guards for cross-file contracts Three guard suites that pin contracts spanning multiple files, where one side changing unilaterally breaks the other silently: - test_version_comparison_consistency.py: the repo's four version comparators (compatibility.parse_semver, api_v3's packaging-based _is_plugin_update_available, store_manager update_plugin's raw string equality, skin_runtime._major) answer differently on the same inputs. A table pins each one's verdict; update_plugin is driven through its real code path to show the SUSPECTED BUGs: 'v1.2.0' vs '1.2.0' triggers a full reinstall the UI calls unnecessary, and a locally-ahead plugin gets downgraded. A pairwise-ordering check keeps parse_semver agreeing with packaging on plain X.Y.Z. - test/web_interface/test_secret_separation_parity.py: api_v3.py carries three inline copies of find_secret_fields/separate_secrets that lack the canonical module's array-item support. The copy count is asserted exact (it may only go down; new copies must import src/web_interface/secret_helpers), the missing-array-support gap is asserted so it can't grow silently, and the canonical behavior that migration will adopt is documented executably. - test_discovery_path_contract.py: the three 'where is plugin X' resolvers (PluginManager discovery, StoreManager._find_plugin_path, SchemaManager.get_schema_path) agree on the configured directory, and their divergent fallback chains are characterized. Also pins the .standalone-backup- naming contract shared by store rollback and discovery, and _resolve_skin_target's path-traversal rejection. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh * test: address review feedback — fixture lifecycle, test names, ClassVar - ci-fixture-plugin: call display_manager.clear() before rendering (per plugin guidelines — the fixture should model a well-behaved plugin), add a class docstring, and document why Pillow is deliberately not pinned in its requirements.txt (core dependency; harness installs nothing). - Rename two tests whose names contradicted their assertions: test_unparseable_core_version_is_compatible -> test_unparseable_core_with_high_floor_is_blocked, and test_unreadable_secrets_file... -> test_corrupt_secrets_file... - Annotate TestGetSchemaProperty.SCHEMA as ClassVar (RUF012). Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh * ci: allow manual test.yml runs via workflow_dispatch Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh * fix: unify version comparison, refuse secret-leaking saves, reset skin strikes on card success Fixes the three suspected bugs this PR's characterization tests pinned, flipping those tests to assert the corrected behavior: - plugins/store: ONE shared update comparator. New compatibility.is_update_available() (PEP 440 via packaging) is now used by both the web UI's update badge (api_v3._is_plugin_update_available is a thin alias) and store_manager.update_plugin's reinstall decision. Previously update_plugin used raw string equality: 'v1.2.0' vs '1.2.0' triggered a full reinstall the UI called unnecessary, and a locally- ahead plugin (2.0.0 installed, registry 1.9.0) was silently DOWNGRADED. Now equivalent spellings skip the reinstall and locally-ahead versions are never downgraded; unparseable versions still reconcile by reinstalling from the registry. - config: save_config and save_config_atomic now refuse (ConfigError) when config_secrets.json exists but cannot be loaded. Both previously proceeded without stripping, writing the merged secrets into config.json in plaintext. The shared _load_secrets_for_save() helper raises with an actionable message instead; a missing secrets file is still fine (nothing to strip), and _migrate_config's catch-all keeps boot resilient. - skins: render_skin_card resets _skin_failures on both success paths (vegas card returned, or mode renderer handled), mirroring _render_game. Transient card failures no longer accumulate across a session until they permanently disable a working skin. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh * fix: harden shared comparator edges from review - is_update_available: reject truthy non-string versions (a malformed manifest can carry a number; packaging raises TypeError on those) by surfacing the mismatch instead of raising. - store_manager.update_plugin: drop the truthiness gate around the comparator so a missing version on either side follows the shared 'no update' verdict, keeping the store consistent with the UI badge; a missing manifest still uses the reinstall recovery path. - config_manager._load_secrets_for_save: catch only expected read/parse failures (OSError/ValueError/RecursionError) so implementation bugs propagate as themselves, and log with traceback. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01NohXi78cwsAKtN1sCfxjUh --------- Co-authored-by: Claude <noreply@anthropic.com>
176 lines
7.4 KiB
Python
176 lines
7.4 KiB
Python
"""
|
|
Drift guard for version comparison.
|
|
|
|
There is now ONE shared "should this plugin update?" comparator —
|
|
`src.plugin_system.compatibility.is_update_available` — used by both the web
|
|
UI's update badge (`api_v3._is_plugin_update_available`) and the store's
|
|
`update_plugin` reinstall decision, so the badge and the actual reinstall
|
|
can never disagree. (Historically the store used raw string equality, which
|
|
reinstalled over cosmetic differences like "v1.2.0" vs "1.2.0" and even
|
|
DOWNGRADED locally-ahead plugins; this file's tests killed that.)
|
|
|
|
Two other version parsers legitimately remain and are pinned here so they
|
|
don't drift: `compatibility.parse_semver` (the install-compatibility gate,
|
|
range-spec oriented) and `skin_runtime._major` (skin API major gate).
|
|
"""
|
|
|
|
import json
|
|
from unittest.mock import patch
|
|
|
|
import pytest
|
|
from packaging.version import parse as pkg_parse
|
|
|
|
from src.plugin_system.compatibility import is_update_available, parse_semver
|
|
from src.skin_system.skin_runtime import _major
|
|
from src.plugin_system.store_manager import PluginStoreManager
|
|
from web_interface.blueprints.api_v3 import _is_plugin_update_available
|
|
|
|
|
|
# (installed, registry) -> update available?
|
|
CASES = [
|
|
(("1.2.0", "1.2.0"), False), # identical
|
|
(("v1.2.0", "1.2.0"), False), # cosmetic v-prefix, semantically equal
|
|
(("1.2", "1.2.0"), False), # short form, semantically equal
|
|
(("1.2.0", "1.2.0-rc1"), False), # rc of same release is not newer
|
|
(("1.2.0", "1.3.0"), True), # registry genuinely newer
|
|
(("2.0.0", "1.9.0"), False), # locally ahead — never downgrade
|
|
(("abc.def", "1.0.0"), True), # unparseable — surface the mismatch
|
|
(("", "1.0.0"), False), # missing either side — nothing to do
|
|
(("1.0.0", ""), False),
|
|
]
|
|
|
|
|
|
class TestSharedComparatorMalformedInputs:
|
|
def test_truthy_non_string_surfaces_mismatch(self):
|
|
# A malformed manifest can carry version as a number; packaging would
|
|
# raise TypeError on it. The comparator must not raise.
|
|
assert is_update_available(1.2, "1.2.0") is True
|
|
assert is_update_available("1.2.0", 1.3) is True
|
|
|
|
def test_falsy_non_string_means_nothing_to_do(self):
|
|
assert is_update_available(None, "1.0.0") is False
|
|
assert is_update_available("1.0.0", None) is False
|
|
assert is_update_available(0, "1.0.0") is False
|
|
|
|
|
|
class TestSharedComparator:
|
|
@pytest.mark.parametrize("pair,expected", CASES)
|
|
def test_is_update_available(self, pair, expected):
|
|
installed, latest = pair
|
|
assert is_update_available(installed, latest) is expected
|
|
|
|
@pytest.mark.parametrize("pair,expected", CASES)
|
|
def test_api_v3_helper_agrees(self, pair, expected):
|
|
# The UI badge helper must be a pure alias of the shared comparator.
|
|
installed, latest = pair
|
|
assert _is_plugin_update_available(installed, latest) is expected
|
|
|
|
|
|
class TestStoreManagerUsesSharedComparator:
|
|
"""Drive update_plugin's real code path to its version check."""
|
|
|
|
def _store(self, tmp_path, local_version, registry_version):
|
|
plugin_dir = tmp_path / "plugins" / "demo-plugin"
|
|
plugin_dir.mkdir(parents=True)
|
|
(plugin_dir / "manifest.json").write_text(json.dumps({
|
|
"id": "demo-plugin", "version": local_version,
|
|
}))
|
|
store = PluginStoreManager(
|
|
plugins_dir=str(tmp_path / "plugins"),
|
|
uninstalled_registry_path=str(tmp_path / "uninstalled.json"),
|
|
)
|
|
registry_info = {
|
|
"id": "demo-plugin",
|
|
"repo": "https://github.com/example/ledmatrix-plugins",
|
|
"latest_version": registry_version,
|
|
}
|
|
return store, registry_info
|
|
|
|
def _run_update(self, store, registry_info):
|
|
with patch.object(store, "fetch_registry", return_value={"plugins": [registry_info]}), \
|
|
patch.object(store, "get_plugin_info", return_value=registry_info), \
|
|
patch.object(store, "_reinstall_with_rollback", return_value=True) as reinstall:
|
|
result = store.update_plugin("demo-plugin")
|
|
return result, reinstall
|
|
|
|
def test_equal_strings_skip_reinstall(self, tmp_path):
|
|
store, info = self._store(tmp_path, "1.2.0", "1.2.0")
|
|
result, reinstall = self._run_update(store, info)
|
|
assert result is True
|
|
reinstall.assert_not_called()
|
|
|
|
def test_v_prefix_equivalent_skips_reinstall(self, tmp_path):
|
|
# "v1.2.0" == "1.2.0" semantically — no pointless reinstall.
|
|
store, info = self._store(tmp_path, "v1.2.0", "1.2.0")
|
|
result, reinstall = self._run_update(store, info)
|
|
assert result is True
|
|
reinstall.assert_not_called()
|
|
|
|
def test_locally_ahead_version_is_never_downgraded(self, tmp_path):
|
|
# A plugin ahead of the registry (local dev build) must not be
|
|
# "updated" — that would be a downgrade.
|
|
store, info = self._store(tmp_path, "2.0.0", "1.9.0")
|
|
result, reinstall = self._run_update(store, info)
|
|
assert result is True
|
|
reinstall.assert_not_called()
|
|
|
|
def test_registry_newer_triggers_reinstall(self, tmp_path):
|
|
store, info = self._store(tmp_path, "1.2.0", "1.3.0")
|
|
result, reinstall = self._run_update(store, info)
|
|
reinstall.assert_called_once()
|
|
assert result is True
|
|
|
|
def test_unparseable_version_surfaces_via_reinstall(self, tmp_path):
|
|
# Direction unknowable → reconcile by reinstalling from the registry.
|
|
store, info = self._store(tmp_path, "abc.def", "1.0.0")
|
|
result, reinstall = self._run_update(store, info)
|
|
reinstall.assert_called_once()
|
|
assert result is True
|
|
|
|
def test_empty_local_version_follows_comparator_no_reinstall(self, tmp_path):
|
|
# The comparator says "nothing to do" for a missing version, and the
|
|
# store must agree with the UI badge — no reinstall.
|
|
store, info = self._store(tmp_path, "", "1.0.0")
|
|
result, reinstall = self._run_update(store, info)
|
|
assert result is True
|
|
reinstall.assert_not_called()
|
|
|
|
def test_empty_registry_version_follows_comparator_no_reinstall(self, tmp_path):
|
|
store, info = self._store(tmp_path, "1.0.0", "")
|
|
result, reinstall = self._run_update(store, info)
|
|
assert result is True
|
|
reinstall.assert_not_called()
|
|
|
|
|
|
class TestSkinRuntimeMajor:
|
|
def test_plain_versions(self):
|
|
assert _major("1.0.0") == 1
|
|
assert _major("2.1") == 2
|
|
|
|
def test_int_input_tolerated(self):
|
|
assert _major(2) == 2
|
|
|
|
def test_garbage_returns_none(self):
|
|
assert _major("garbage") is None
|
|
assert _major(None) is None
|
|
|
|
def test_v_prefix_not_tolerated(self):
|
|
# Unlike parse_semver, _major does NOT strip a leading 'v' —
|
|
# a skin.json declaring "v1.0.0" fails the API gate. Characterized
|
|
# so a manifest-format loosening elsewhere doesn't silently diverge.
|
|
assert _major("v1.0.0") is None
|
|
|
|
|
|
class TestParseSemverAgreesWithPackaging:
|
|
"""parse_semver and packaging must agree on ordering for plain X.Y.Z —
|
|
the region where the two ecosystems overlap and must never diverge."""
|
|
|
|
PLAIN = ["0.1.0", "1.0.0", "1.2.0", "1.2.3", "1.10.0", "2.0.0", "10.0.1"]
|
|
|
|
def test_pairwise_ordering_matches(self):
|
|
for a in self.PLAIN:
|
|
for b in self.PLAIN:
|
|
ours = parse_semver(a) < parse_semver(b)
|
|
theirs = pkg_parse(a) < pkg_parse(b)
|
|
assert ours == theirs, f"ordering diverges on ({a}, {b})"
|