LangSmith

A platform from the LangChain team to observe, evaluate and version the prompts of an LLM application: every step of a call or an agent is traced, outputs get scored against datasets or by a judge model, and prompts are tested in a playground. It does not require LangChain: the docs cite OpenAI, Anthropic, CrewAI, Vercel AI SDK or Pydantic AI, and it ingests OpenTelemetry traces from any application on its `/otel` endpoint. Free Developer plan (one seat, 5,000 traces a month), Plus plan at $39 per seat (10,000 traces), base traces kept for 14 days, self-hosting reserved to the Enterprise plan.

Strengths

Limitations

Best for

Official site

View on Coeurdar