FAQ
Didn't find your answer? Open a ticket from the console and we'll reply soon.
产品
Which models can the platform evaluate?
The platform works with any OpenAI-compatible model service, including self-hosted and privately deployed models. Enter the endpoint and API key and you're ready to evaluate.
Which dimensions do you evaluate?
There are 18 dimensions: accuracy, safety, hallucination rate, bias, robustness, instruction following, consistency, relevance, completeness, fluency, concision, reasoning, coding, multilingual ability, privacy compliance, jailbreak resistance, response performance and cost efficiency.
What's the difference between intensity levels?
Intensity sets how many cases are drawn and how often they repeat. Light is a quick check; Standard covers every case once; Deep repeats three times to measure variance; Extreme doubles the cases and repeats five times for formal certification.
Can I export the results?
Report export is available on Pro and above, in formats including JSON and PDF. You can also generate a public share link.
计费
How is quota calculated?
Each run deducts quota as tier level × intensity multiplier. The exact cost is shown before you submit, and you can't submit without enough balance.
When does quota reset?
Subscription quota resets each billing cycle. Top-up pack quota stays valid through its term and never resets with the cycle.
接口
Is there API access?
The API is available on Enterprise and above. Create an application in the console to get a key, then start evaluations, query results and receive callbacks over the API.
安全
Are my model keys safe?
Keys are stored encrypted and decrypted only when a run calls your model. They never appear in any log or report, and you can delete them at any time.
