Run tests in mflux (fast/slow/full), preserve image outputs, and handle golden image diffs safely.
60
70%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Passed
No findings from the security scan
Fix and improve this skill with Tessl
tessl review fix ./.cursor/skills/mflux-testing/SKILL.mdThis repo uses pytest with image-producing tests. Always preserve outputs for inspection and never update reference images unless explicitly asked.
make test-fast (fast tests, no image generation)make test-slow (slow tests, image generation)make test (full suite)MFLUX_PRESERVE_TEST_OUTPUT=1 on test runs (already built into the Makefile test targets).Golden tests compare generated PNGs to tests/resources/reference_*.png (typically 15% mismatch threshold).
When to update (only with explicit user approval):
MFLUX_PRESERVE_TEST_OUTPUT=1mflux-debugging)Workflow:
tests/resources/output_*.png vs reference_*.pngtest(<model>): update golden images for local hardware)Important: Golden tests lock mflux-native sampling (mx.random + mflux schedulers), not diffusers pixel parity. A good diffusers side-by-side or injected-latent run builds confidence in the model code; the golden still reflects mflux’s full recipe on CI hardware.
Use when a change touches model config resolution, mflux-save, or the model’s generate CLI, or when a PR fixes local model-path handling for the model under investigation. Refer to the mflux-cli skill to find the correct generate command for the model you are testing.
mflux-cli skill to look up the correct command and flags.--help before running it.mflux-cli skill to find the generate command and required flags.--help before running it.4ac641e
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.