スキル一覧に戻る
mvillmow

test-diff-analyzer

by mvillmow

Training framework written in Mojo

11🍴 4📅 2026年1月24日
GitHubで見るManusで実行

SKILL.md


name: test-diff-analyzer description: "Analyze test differences between runs to identify flaky tests and consistency issues. Use to find tests that fail intermittently." category: testing mcp_fallback: none user-invocable: false

Analyze Test Differences Between Runs

Compare test results across multiple runs to identify flaky tests.

When to Use

  • Test passes locally but fails in CI
  • Test sometimes passes, sometimes fails (flaky test)
  • Need to understand test consistency issues
  • Comparing test results before/after code changes
  • Debugging intermittent test failures

Quick Reference

# Run tests and capture output
pixi run mojo test -I . tests/ > /tmp/test_run_1.log

# Compare two test runs
diff -u /tmp/test_run_1.log /tmp/test_run_2.log

# Extract failures from log
grep "FAILED" /tmp/test_run_*.log | sort | uniq -c

# Show tests that sometimes pass, sometimes fail
grep "FAILED\|PASSED" /tmp/test_run_*.log | cut -d: -f2 | sort | uniq -d

Analysis Workflow

  1. Collect baseline: Run tests locally N times
  2. Collect CI data: Get CI test results from recent runs
  3. Compare outputs: Diff between test runs
  4. Identify flaky tests: Tests with inconsistent results
  5. Find patterns: When does test fail vs pass
  6. Root cause: Timing, randomness, resource issues
  7. Remediation: Fix or isolate flaky test

Flaky Test Indicators

Timing Issues:

  • Test passes when run in isolation
  • Test fails when run with other tests
  • Timeout values too aggressive
  • Race conditions in setup/teardown

Randomness Issues:

  • Random seed not fixed
  • Hash ordering varies
  • Dictionary/set iteration order
  • Floating point precision

Resource Issues:

  • Test passes locally but fails in CI
  • Fails under resource constraints
  • Out of memory errors intermittently
  • Disk space dependent

Output Format

Report analysis with:

  1. Flaky Tests - Tests with inconsistent results
  2. Consistency Score - Pass rate across runs (e.g., 80% pass rate)
  3. Failure Patterns - When/how tests fail
  4. Impact - How many test runs affected
  5. Root Cause Hypothesis - What likely causes instability
  6. Recommendations - How to fix or isolate flaky test

Error Handling

ProblemSolution
Different environmentRun in controlled environment (docker)
Insufficient dataRun more iterations to get pattern
No failure infoEnable debug output, increase verbosity
External dependenciesMock or isolate external services
Timing-dependentAdd explicit waits or retry logic

References

  • See mojo-test-runner for test execution options
  • See extract-test-failures for failure analysis
  • See CLAUDE.md for test standards and TDD workflow

スコア

総合スコア

60/100

リポジトリの品質指標に基づく評価

SKILL.md

SKILL.mdファイルが含まれている

+20
LICENSE

ライセンスが設定されている

+10
説明文

100文字以上の説明がある

0/10
人気

GitHub Stars 100以上

0/15
最近の活動

3ヶ月以内に更新がある

0/10
フォーク

10回以上フォークされている

0/5
Issue管理

オープンIssueが50未満

+5
言語

プログラミング言語が設定されている

+5
タグ

1つ以上のタグが設定されている

0/5

レビュー

💬

レビュー機能は近日公開予定です