The shape is worth reading: a score plus a written critique, per criterion, callable mid-conversation so a model can grade its own draft. The Atla API behind it was shut down — leaving this as a reference rather than something to install.
An MCP server over the Atla API for LLM-as-judge evaluation. You supply a prompt, a response and the criteria; an Atla evaluation model returns a score and written feedback on the response.
- One response scored against one evaluation criterion, with a textual critique alongside the score
- The same response scored across multiple criteria at once, one score and critique per criterion
The repository is archived and the Atla API is no longer active, so the tools have nothing left to call. It ran as ATLA_API_KEY=<your-key> uvx atla-mcp-server.
One command plus a key — ATLA_API_KEY=<your-api-key> uvx atla-mcp-server, then supply credentials
