softwarefactory.build

Filtered lists / Verification gates

Verification gates

Review, eval, and merge tools that can gate a change. Autocomplete products are excluded unless they document that role.

Inclusion follows the query’s structured fields; this is not a score. 22 matching listings.

  1. Braintrust

    Evaluation platform for datasets, experiments, scoring, and logs.

    Verified

  2. CodeQL

    GitHub semantic code analysis with query-based CI checks.

    Verified

  3. CodeRabbit

    AI pull-request review for GitHub and GitLab.

    Verified

  4. Cubic

    AI code review for GitHub pull requests.

    Verified

  5. Dagger

    Programmable containerized CI/CD engine for agent workflows.

    Verified

  6. DeepEval

    Pytest-oriented evaluation library for LLM systems.

    Verified

  7. DeepSource

    Hosted static analysis with Autofix pull requests.

    Verified

  8. Galileo

    Enterprise evaluation and observability for AI systems.

    Verified

  9. Giskard

    Open testing and red-teaming platform for AI applications.

    Verified

  10. Graphite

    Stacked pull requests with CLI workflow and AI review.

    Verified

  11. Greptile

    AI pull-request review with repository-wide context.

    Verified

  12. Inspect

    UK AISI open framework for evaluating language-model agents.

    Verified

  13. Langfuse

    Self-hostable open-source LLM and agent observability.

    Verified

  14. LangSmith

    LangChain commercial tracing, evaluation, and prompt platform.

    Verified

  15. LangWatch

    LLM observability and evaluation platform with open components.

    Verified

  16. Mastra

    TypeScript agent framework with workflows and evaluation hooks.

    Verified

  17. Mergify

    GitHub merge queue and pull-request automation rules.

    Verified

  18. Phoenix

    Arize open-source observability and evaluation for LLM applications.

    Verified

  19. Promptfoo

    Open evaluation and red-team framework for LLM applications.

    Verified

  20. Ragas

    Open evaluation metrics for RAG and LLM applications.

    Verified

  21. Semgrep

    Teachably fast static analysis with optional AI-assisted rules.

    Verified

  22. Trunk

    Unified linters, formatters, and CI reliability checks.

    Verified