Documentation
A8 (Touchstone) is a general eval model: state any criterion and it returns a verdict. It answers the OpenAI and Anthropic APIs, so the SDK you already use calls it.
The verdict is scored, so the same input always returns the same verdict.
§1Quickstart
A first verdict from A8, in three steps. The full request and response contract is on the A8 reference.
§1.1Get a key
Issue one under the console → API keys. It needs the models:infer permission; teaching with expected needs models:train. See API keys.
§1.2Point a client at the base URL
Any client that takes a custom base URL works: the OpenAI and Anthropic SDKs, LangChain, LlamaIndex, the Vercel AI SDK, Pydantic AI and LiteLLM among them. With the OpenAI SDK:
§1.3State a criterion, send a subject
The system message is the criterion: what is being judged. The last user message is the subject: the thing under judgment.
A8 answers where its evidence clears the accuracy the request sets, 90 by default, and abstains where it does not. On a criterion it has not learned, send the verdict you wanted as expected and it learns it.
The same request through the Anthropic SDK, the Responses API or the native route is in the reference.