Run tests in mflux (fast/slow/full), preserve image outputs, and handle golden image diffs safely.
60
70%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Passed
No findings from the security scan
Fix and improve this skill with Tessl
tessl review fix ./.cursor/skills/mflux-testing/SKILL.mdThis repo uses pytest with image-producing tests. Always preserve outputs for inspection and never update reference images unless explicitly asked.
just test-fast (fast tests, no image generation)just test-slow (slow tests, image generation)just test (default selection, skips slow model tests)just test-all (all except high-memory tests; slow tests download model weights)MFLUX_PRESERVE_TEST_OUTPUT=1 on test runs (already built into the justfile test recipes).Golden tests compare generated PNGs to tests/resources/reference_*.png (typically 15% mismatch threshold).
When to update (only with explicit user approval):
MFLUX_PRESERVE_TEST_OUTPUT=1mflux-debugging)Workflow:
tests/resources/output_*.png vs reference_*.pngtest(<model>): update golden images for local hardware)Important: Golden tests lock mflux-native sampling (mx.random + mflux schedulers), not diffusers pixel parity. A good diffusers side-by-side or injected-latent run builds confidence in the model code; the golden still reflects mflux’s full recipe on CI hardware.
Use when a change touches model config resolution, mflux-save, or the model’s generate CLI, or when a PR fixes local model-path handling for the model under investigation. Refer to the mflux-cli skill to find the correct generate command for the model you are testing.
mflux-cli skill to look up the correct command and flags.--help before running it.mflux-cli skill to find the generate command and required flags.--help before running it.1430b27
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.