Skip to content
webtechos

First agent in production. You can replay a bad run, not argue about it.

AI Agent Hub

LangGraph holds the graph. LangSmith shows the trace. Braintrust blocks a prompt change that quietly got worse. Modal runs the jobs. Relevance is there so ops can build helpers without opening a ticket.

Layers
01

5

 

Average score
02

83/100

 

Free tier
03

5of 5

 

Budget
04

$300 to $1,200 / mo

 

Team
05

1 to 3 engineers, one owner

 

[ 01 ]  Layer by layer

The tool, and why it is in this set.

  1. Layer 01

    Orchestration

    LangGraph

    LangGraph

    AI AgentsOpen sourceFrom $0

    Explicit state and durable checkpoints. A failed run can be replayed instead of guessed at.

    82/100
  2. Layer 02

    Observability

    LangSmith

    LangSmith

    AI InfraFreemiumFrom $0

    Every model and tool call is a trace. A production miss becomes a regression test in one step.

    84/100
  3. Layer 03

    Evaluation

    Braintrust

    Braintrust

    AI InfraFreemiumFrom $0

    A CI gate so a prompt edit cannot ship on vibes.

    82/100
  4. Layer 04

    Compute

    Modal

    Modal

    AI InfraUsage-basedFrom $0 + compute

    Tool workloads and batch jobs on demand. No cluster to babysit.

    85/100
  5. Layer 05

    Ops surface

    Relevance AI

    Relevance AI

    AI AgentsFreemiumFrom $0

    Supporting agents for the business team, without adding work to engineering.

    82/100