github.com/ArabelaTso/Skills-4-SE
| Skill | Added | Review |
|---|---|---|
interval-profiling-performance-analyzer skills/interval-profiling-performance-analyzer/SKILL.md Profile programs at the function/method level to identify performance hotspots, bottlenecks, and optimization opportunities. Records execution time, memory usage, and call frequency for each interval. Generates actionable recommendations and visualizations. Use when users need to (1) analyze program performance, (2) identify slow functions or bottlenecks, (3) optimize execution time or memory usage, (4) profile Python, Java, or C/C++ programs with test cases or workload scenarios, or (5) generate performance reports with flame graphs and recommendations. | 91 91 1.23x Agent success vs baseline Impact 99% 1.23xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
interval-guided-regression-test-update skills/interval-guided-regression-test-update/SKILL.md Automatically updates regression tests based on interval analysis to maintain coverage of key program intervals. Use when code changes affect value ranges, conditionals, or control flow, and existing tests need updating to maintain interval coverage. Analyzes interval information from updated code, identifies coverage gaps, adjusts test inputs and assertions, removes redundant tests, and generates new tests for uncovered intervals. Supports Python, Java, JavaScript, and C/C++ with various test frameworks (pytest, JUnit, Jest, Google Test). | 79 79 0.92x Agent success vs baseline Impact 88% 0.92xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
interval-difference-analyzer skills/interval-difference-analyzer/SKILL.md Analyze differences in program intervals between two versions of a program (old and new) to identify added, removed, or modified intervals. Use when comparing program versions, analyzing variable ranges, detecting behavioral changes in numeric computations, validating refactorings, or assessing migration impacts. Supports optional test suite integration to validate interval changes. Generates comprehensive reports highlighting intervals requiring further testing or verification. | 81 81 1.85x Agent success vs baseline Impact 100% 1.85xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
interface-contract-verifier skills/interface-contract-verifier/SKILL.md Verify that interface and class contracts (preconditions, postconditions, invariants) are preserved across program versions. Use when validating refactorings, checking API compatibility, verifying design-by-contract implementations, or ensuring behavioral contracts remain intact after code changes. Automatically detects contract violations, identifies affected methods and classes, and provides actionable guidance for resolving violations while maintaining program correctness. | 90 90 2.02x Agent success vs baseline Impact 95% 2.02xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
integration-test-generator skills/integration-test-generator/SKILL.md Generate integration tests for multiple interacting components in Python. Use when testing interactions between: (1) Multiple services or APIs (REST/GraphQL endpoints, microservices), (2) Database operations with repositories/ORMs (SQLAlchemy, Django ORM), (3) External services (payment gateways, email services, third-party APIs), (4) Message queues and event-driven systems, (5) Full stack workflows (API + database + business logic). Provides test structure templates, fixtures, test data builders, and patterns for pytest-based integration testing. | 89 89 1.02x Agent success vs baseline Impact 91% 1.02xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
incremental-python-programmer skills/incremental-python-programmer/SKILL.md Takes a Python repository and natural language feature description as input, implements the feature with proper code placement, generates comprehensive tests, and ensures all tests pass. Use when Claude needs to: (1) Add new features to existing Python projects, (2) Implement functions, classes, or modules based on requirements, (3) Modify existing code to add functionality, (4) Generate unit and integration tests for new code, (5) Fix failing tests after implementation, (6) Ensure code follows existing patterns and conventions. | 69 69 1.05x Agent success vs baseline Impact 90% 1.05xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
incremental-java-programmer skills/incremental-java-programmer/SKILL.md Incrementally implement new features in Java repositories from natural language descriptions. Use when adding functionality to existing Java codebases (Maven or Gradle projects). Takes a feature description as input and outputs modified repository with implementation code, corresponding JUnit tests, and verification that all tests pass. Supports method additions, new class creation, and method modifications with proper Java conventions. | 91 91 1.04x Agent success vs baseline Impact 99% 1.04xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
imperative-to-coq-model-extractor skills/imperative-to-coq-model-extractor/SKILL.md Extract abstract mathematical models from imperative code (C, C++, Python, Java, etc.) suitable for formal reasoning in Coq. Use when the user asks to model imperative code in Coq, create Coq specifications from imperative programs, extract mathematical models for verification, or translate imperative algorithms to Coq for formal reasoning and proof. | 88 88 1.03x Agent success vs baseline Impact 99% 1.03xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
git-bisect-assistant skills/git-bisect-assistant/SKILL.md Automatically performs git bisect to identify the first bad commit that introduced a bug or failure. Use when debugging regressions, tracking down when a test started failing, or identifying which commit broke functionality. Handles flaky tests with retry logic and provides comprehensive reports with bisect logs and confidence levels. | 88 88 2.27x Agent success vs baseline Impact 100% 2.27xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
fuzzing-input-generator skills/fuzzing-input-generator/SKILL.md Generate randomized and edge-case inputs to detect unexpected failures, bugs, and security vulnerabilities through fuzz testing. Use when creating test cases for robustness testing, generating adversarial inputs, testing error handling, finding edge cases, or security testing. Produces Python test code with fuzzing inputs for strings, numbers, and structured data focusing on edge cases, invalid inputs, and random valid inputs. Triggers when users ask to generate fuzz tests, create randomized test inputs, test edge cases, find bugs through fuzzing, or generate adversarial test cases. | 89 89 1.37x Agent success vs baseline Impact 77% 1.37xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
function-class-generator skills/function-class-generator/SKILL.md Generate complete, production-ready functions and classes from formal specifications, design descriptions, type signatures, or natural language requirements. Use this skill when implementing APIs from specifications, creating data structures from schemas, building classes from UML diagrams, generating code from contracts, or translating design documents into code. Supports multiple programming languages and follows language-specific best practices. | 81 81 1.16x Agent success vs baseline Impact 94% 1.16xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
framework-migration-assistant skills/framework-migration-assistant/SKILL.md Automatically migrate Python web applications between frameworks (Flask → FastAPI, Django → FastAPI). Use when you need to migrate an existing web application to a modern framework while preserving functionality. The skill analyzes the codebase, updates routes, handlers, configuration, dependency injection patterns, and tests. Creates git commits for each migration phase and generates a comprehensive summary of all changes. Supports automatic dependency updates, code transformations, and test adaptations. | 89 89 1.24x Agent success vs baseline Impact 93% 1.24xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
formal-spec-generator skills/formal-spec-generator/SKILL.md Generate formal specifications (definitions, predicates, invariants, pre/post-conditions) in Isabelle/HOL or Coq from informal requirements, source code, pseudocode, or mathematical descriptions. Use when users need to: (1) Formalize algorithms or data structures, (2) Create function specifications with contracts, (3) Generate predicates and properties for verification, (4) Translate informal requirements into formal logic, (5) Specify invariants for loops or data structures, or (6) Create formal definitions for mathematical concepts. Supports both Isabelle/HOL and Coq equally. | 86 86 1.00x No change in agent success vs baseline Impact 99% 1.00xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
flaky-test-detector skills/flaky-test-detector/SKILL.md Identifies non-deterministic or unreliable tests through static code analysis and test result analysis. Use when Claude needs to find flaky tests, analyze test reliability, or investigate intermittent test failures. Supports Python (pytest, unittest) and Java (JUnit, TestNG) test frameworks. Trigger when users mention "flaky tests", "intermittent failures", "non-deterministic tests", "unreliable tests", or ask to "find flaky tests", "analyze test stability", or "why tests fail randomly". | 94 94 1.04x Agent success vs baseline Impact 97% 1.04xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
failure-oriented-instrumentation skills/failure-oriented-instrumentation/SKILL.md Selectively instruments code to capture runtime data for debugging failures and bugs. Use when investigating crashes, exceptions, unexpected behavior, test failures, or performance issues. Analyzes stack traces and error messages to identify suspicious code regions, then adds targeted logging, tracing, and assertions to capture variable values, execution paths, timing, and conditional branches. Supports Python, JavaScript/TypeScript, Java, and C/C++. | 88 88 1.10x Agent success vs baseline Impact 66% 1.10xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
exploitability-analyzer skills/exploitability-analyzer/SKILL.md Analyze detected vulnerabilities to assess realistic exploitability by examining control flow, input sources, sanitization logic, and execution context. Use when users need to: (1) Determine if a vulnerability is actually exploitable in practice, (2) Assess severity and impact of security issues, (3) Prioritize vulnerability remediation, (4) Understand attack vectors and exploitation conditions, (5) Generate exploitability reports with proof-of-concept scenarios. Focuses on injection vulnerabilities (SQL, command, XSS, path traversal, LDAP) with detailed analysis of reachability, controllability, sanitization, and impact. | 91 91 1.32x Agent success vs baseline Impact 97% 1.32xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
error-explanation-generator skills/error-explanation-generator/SKILL.md Explains test failures and provides actionable debugging guidance. Use when tests fail (unit, integration, E2E), builds fail, or code throws errors. Analyzes error messages, stack traces, and test output to identify root causes and suggest concrete fixes. Handles pytest, jest, junit, mocha, vitest, selenium, cypress, playwright, and other testing frameworks across Python, JavaScript/TypeScript, Java, Go, and other languages. | 92 92 1.25x Agent success vs baseline Impact 88% 1.25xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
environment-setup-assistant skills/environment-setup-assistant/SKILL.md Generate setup scripts and instructions for development environments across platforms. Use when: (1) Setting up new development machines (Python, Node.js, Docker, databases), (2) Creating automated setup scripts for team onboarding, (3) Need cross-platform setup instructions (macOS, Linux, Windows), (4) Installing development tools and dependencies, (5) Configuring version managers and package managers. Provides executable setup scripts, platform-specific guides, and tool installation instructions. | 93 93 1.15x Agent success vs baseline Impact 90% 1.15xAverage score across 3 eval scenarios Securityby Medium Suggest reviewing before use Reviewed: Version: 0f00a4f | |
edge-case-generator skills/edge-case-generator/SKILL.md Automatically identify potential boundary and exception cases from requirements, specifications, or existing code, and generate comprehensive test cases targeting boundary conditions, edge cases, and uncommon scenarios. Use this skill when analyzing programs, code repositories, functions, or APIs to discover and test corner cases, null handling, overflow conditions, empty inputs, concurrent access patterns, and other exceptional scenarios that are often missed in standard testing. | 93 93 1.03x Agent success vs baseline Impact 96% 1.03xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
directed-test-input-generator skills/directed-test-input-generator/SKILL.md Generate targeted test inputs to reach specific code paths and hard-to-reach behaviors in Python code. Use when: (1) Targeting uncovered branches or specific execution paths, (2) Need coverage-guided test generation, (3) Want to leverage LLM understanding of code semantics for meaningful test inputs, (4) Testing boundary conditions and edge cases systematically, (5) Combining symbolic reasoning with fuzzing. Provides path analysis, constraint solving, coverage-guided strategies, and LLM-driven semantic generation for comprehensive test input creation. | 91 91 1.28x Agent success vs baseline Impact 82% 1.28xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
design-smell-detector skills/design-smell-detector/SKILL.md Identify design quality issues in code including high coupling, low cohesion, God classes, long methods, and other code smells. Use when: (1) Reviewing code architecture and design quality, (2) Identifying refactoring opportunities, (3) Detecting God classes or classes with too many responsibilities, (4) Finding high coupling or low cohesion issues, (5) Analyzing code maintainability and technical debt. Detects coupling smells, cohesion problems, complexity issues, size violations, and encapsulation problems with actionable refactoring suggestions. | 90 90 1.05x Agent success vs baseline Impact 79% 1.05xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
design-pattern-suggestor skills/design-pattern-suggestor/SKILL.md Recommends appropriate software design patterns based on problem descriptions, requirements, or code scenarios. Use when designing software architecture, refactoring code, solving common design problems, or choosing between design approaches. Analyzes the problem context and suggests suitable creational, structural, behavioral, architectural, or concurrency patterns with implementation guidance and trade-off analysis. | 90 90 1.14x Agent success vs baseline Impact 92% 1.14xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
deprecated-api-updater skills/deprecated-api-updater/SKILL.md Identify and replace deprecated API usage in source code with modern alternatives. Use when: (1) Modernizing legacy codebases, (2) Upgrading framework versions (React, Django, Spring, etc.), (3) Fixing deprecation warnings in build output, (4) Preparing for major version upgrades, (5) Ensuring code uses current best practices. Supports Python, JavaScript/TypeScript, Java, and other major languages with both AST-based detection and pattern matching for accurate identification and automated replacement with validation. | 93 93 1.60x Agent success vs baseline Impact 90% 1.60xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f | |
dependency-resolver skills/dependency-resolver/SKILL.md Identify, analyze, and manage software dependencies before deployment. Use this skill when preparing applications for deployment, resolving dependency conflicts, updating dependencies, auditing security vulnerabilities, managing package versions, or troubleshooting dependency-related issues. Supports multiple package managers (npm, pip, maven, cargo, go mod, composer) and provides actionable recommendations for dependency management. | 77 77 1.02x Agent success vs baseline Impact 71% 1.02xAverage score across 3 eval scenarios Securityby Medium Suggest reviewing before use Reviewed: Version: 0f00a4f | |
dead-code-eliminator skills/dead-code-eliminator/SKILL.md Identify and analyze unused or redundant code including unused functions/methods, unused variables/imports, unreachable code, and redundant conditions. Use when cleaning up codebases, improving maintainability, reducing technical debt, or conducting code quality audits. Analyzes Python code using AST analysis and produces markdown reports listing dead code locations with line numbers, severity ratings, and recommendations. Triggers when users ask to find dead code, remove unused code, identify unused imports, find unreachable code, or clean up redundant logic. | 81 81 2.06x Agent success vs baseline Impact 93% 2.06xAverage score across 3 eval scenarios Securityby Passed No findings from the security scan Reviewed: Version: 0f00a4f |