softwarefactory.build

Layer / What must pass

Verification

Tests, reviews, hooks, and merge gates that determine whether a change is ready.


  1. Amazon Q Developer

    AWS coding assistant with IDE, CLI, and agent features.

    Verified

  2. Braintrust

    Evaluation platform for datasets, experiments, scoring, and logs.

    Verified

  3. Claude Code

    Anthropic coding agent for terminal, IDE, desktop, and browser workflows.

    Verified

  4. CodeQL

    GitHub semantic code analysis with query-based CI checks.

    Verified

  5. CodeRabbit

    AI pull-request review for GitHub and GitLab.

    Verified

  6. Cubic

    AI code review for GitHub pull requests.

    Verified

  7. Cursor

    Agent-focused IDE with cloud agents running in isolated VMs.

    Verified

  8. Dagger

    Programmable containerized CI/CD engine for agent workflows.

    Verified

  9. DeepEval

    Pytest-oriented evaluation library for LLM systems.

    Verified

  10. DeepSource

    Hosted static analysis with Autofix pull requests.

    Verified

  11. Devin

    Cognition cloud coding agent with Devin Desktop and CLI surfaces.

    Verified

  12. Ellipsis

    AI pull-request reviewer that can open follow-up fixes.

    Verified

  13. Factory

    Commercial agent platform for software delivery workflows.

    Verified

  14. Galileo

    Enterprise evaluation and observability for AI systems.

    Verified

  15. Giskard

    Open testing and red-teaming platform for AI applications.

    Verified

  16. GitHub Copilot

    GitHub coding assistant with an issue-to-pull-request agent.

    Verified

  17. Graphite

    Stacked pull requests with CLI workflow and AI review.

    Verified

  18. Greptile

    AI pull-request review with repository-wide context.

    Verified

  19. Inspect

    UK AISI open framework for evaluating language-model agents.

    Verified

  20. Langfuse

    Self-hostable open-source LLM and agent observability.

    Verified

  21. LangSmith

    LangChain commercial tracing, evaluation, and prompt platform.

    Verified

  22. LangWatch

    LLM observability and evaluation platform with open components.

    Verified

  23. Mastra

    TypeScript agent framework with workflows and evaluation hooks.

    Verified

  24. Mergify

    GitHub merge queue and pull-request automation rules.

    Verified

  25. Phoenix

    Arize open-source observability and evaluation for LLM applications.

    Verified

  26. Promptfoo

    Open evaluation and red-team framework for LLM applications.

    Verified

  27. Qodo

    Code generation and review agents with test-aware workflows.

    Verified

  28. Ragas

    Open evaluation metrics for RAG and LLM applications.

    Verified

  29. Semgrep

    Teachably fast static analysis with optional AI-assisted rules.

    Verified

  30. Sourcery

    Refactoring and review assistant with IDE and CI surfaces.

    Verified

  31. SWE-agent

    Princeton open-source agent for SWE-bench-style issue patches.

    Verified

  32. Trevize

    Shared cloud workspaces for agent runs, previews, and pull requests.

    Verified

  33. Trunk

    Unified linters, formatters, and CI reliability checks.

    Verified