versalist
ResearchChallengesAI ToolsPricingFor Teams
Sign inRequest a demo
versalist

Agent evaluation environments for engineering teams. Design rewards, run agents, and turn outcomes into signal.

Resources

  • Request a demo
  • Docs
  • Research
  • Blog
  • How it works
  • FAQ
  • Environments
  • Legal

Company

  • About us
  • Pricing
  • For Teams
  • Book a session
  • Join as Partner
  • Terms of Service
  • Privacy Policy

Guides

  • All guides
  • Prompt engineering
  • Prompt guide
  • Model Context Protocol
  • Evaluation
  • AI-empowered future

Stay updated

Get platform updates, challenge launches, and practical notes on agent evaluation.

© 2026 Social Protocol Labs LLC. All rights reserved.
Build the full agent training loop.
AI Tools Directory

Discover the tools behind real AI build, eval, and deployment workflows.

Browse by role in the stack, not just by vendor. Compare fit here, then open a tool detail page when it earns deeper inspection.

Open My AI StackBrowse Agent Capabilities

Stack layer

Narrow the directory by what the tool does in the workflow.

All Tools467Models886Environment6Action Space411Observation19Reward / Eval10Policy Serving35Training Infra5Safety / Guardrails7Orchestration16

Browse tools

411 results in the directory

Suggest a tool
AI Engineering Tooling · Vector DatabasesLanceDBFreemium

LanceDB

Multimodal vector database.

FreemiumInspect tool
Models · Large Language ModelsLangbaseFreemium

Langbase

LLM app building platform

FreemiumInspect tool
Agent FrameworkOpen Source CommunityOpen source

Langchain

Building applications with LLMs

22 related challengesInspect tool
Agent FrameworkLangChainOpen source

LangChain

Framework for building LLM applications

20 related challengesInspect tool
Agent FrameworkLangChainOpen source

LangGraph

Runtime for stateful agent workflows.

4 related challengesInspect tool
Agent Systems · Multi-Agent SystemsLangroidOpen source

Langroid

Agent Systems · Multi-Agent Systems

Free (OSS)Inspect tool
AI Workflow Automation · Observability, Evaluation & GovernanceLangWatchFreemium

LangWatch

LLM monitoring and analytics

1 related challengeInspect tool
AI Workflow Automation · Observability, Evaluation & GovernanceLarridinFreemium

Larridin

AI evaluation platform

UnknownInspect tool
AI Engineering Tooling · Security & Risk Management PlatformsLasso SecurityFreemium

Lasso Security

GenAI security platform

UnknownInspect tool
Models · Large Language ModelsLastMile AIFreemium

LastMile AI

LLM app development & evaluation

1 related challengeInspect tool
AI Engineering Tooling · Security & Risk Management PlatformsLatticeFlow AIFreemium

LatticeFlow AI

AI quality and security

UnknownInspect tool
AI Engineering Tooling · Developer ToolsLazy DevPaid

Lazy Dev

AI full-stack code generator

SubscriptionInspect tool

Page 17 of 35

Previous
1516171819
Next