Terms of Service
1. Acceptance of Terms
By accessing or using LLM Prover (“the Service”), you agree to be bound by these Terms of Service.
2. Description of Service
The Service provides LLM benchmarking, evaluation, and compliance tooling as described at llmprover.pysolvr.com. The Service is operated by pysolvr.com.
3. API Usage
- You are responsible for maintaining the confidentiality of your API key.
- You may not share, sell, or transfer your API key to third parties.
- Usage is subject to the rate limits and quotas of your subscription tier.
- We reserve the right to suspend access for abuse or violation of these terms.
4. Subscription and Billing
- Paid plans are billed monthly unless otherwise stated.
- You may cancel your subscription at any time. Access continues until the end of your current billing period.
- Refunds are not provided. See our Refund Policy for details.
- We reserve the right to change pricing with 30 days notice.
5. Intellectual Property
- You retain ownership of all data you submit to the Service.
- Output generated by the Service is yours to use without restriction.
- The Service itself, including its code, design, and documentation, remains our property.
6. Limitation of Liability
The Service is provided “as is” without warranty of any kind. We are not liable for any indirect, incidental, or consequential damages arising from your use of the Service.
7. Termination
We may terminate or suspend your access at any time for violation of these terms. You may terminate your account at any time by cancelling your subscription and ceasing use.
8. Changes to Terms
We may update these terms from time to time. Continued use of the Service after changes constitutes acceptance.
9. Contact
For questions about these terms, contact us at support@llmprover.pysolvr.com.
10. Third-Party LLM Providers
When you use the Service, your prompts and context (including fragments of any RAG content you upload) are sent to third-party LLM providers to generate responses. This is the core function of the product.
We are not responsible for how LLM providers handle inference data.
Providers we use include but are not limited to: OpenAI, Anthropic, xAI / Grok, Together AI, and Groq. You should review each provider’s privacy policy before submitting data. If your data is subject to regulatory requirements (HIPAA, PCI-DSS, GDPR special categories, export controls), you are responsible for ensuring the relevant providers meet those requirements before use.
Our liability is limited strictly to data we store and control. We make no representations about provider data handling and accept no liability for provider data practices.
11. Bring Your Own Endpoint (BYOE)
Pro+ and Enterprise subscribers may register external LLM endpoints (“BYOE endpoints”) for use in benchmarks and evaluations. By registering a BYOE endpoint you agree to the following:
Credential storage. Your endpoint URL and authentication credential are stored encrypted at rest using AES-256 (AWS KMS). Credentials are never logged, never returned in API responses after initial registration, and are only decrypted at the point of making a benchmark call.
Usage scope. We will only call your endpoint when you explicitly trigger a benchmark run (manually or via a schedule you have configured). We will not make background calls, use your endpoint for any purpose other than executing your configured benchmarks, or share your endpoint details with any third party.
Cost responsibility. Benchmark runs against your BYOE endpoint generate real API calls to your provider. You are solely responsible for all costs incurred on your provider account as a result of benchmark runs. This includes scheduled runs that fire automatically. We have no visibility into your provider’s pricing and accept no liability for charges incurred. You should review your provider’s pricing and set appropriate usage limits on your provider account before registering an endpoint.
Security recommendation. We recommend using a scoped or read-only API key for your BYOE endpoint rather than a root or admin credential. Limit the key’s permissions to inference only.
Data flow. Benchmark prompts from your configured suites will be sent to your endpoint. Responses from your endpoint will be stored in our systems and scored. The same data handling terms that apply to managed model responses (Section 10 and our Privacy Policy) apply to BYOE responses.
Removal. You may delete a BYOE endpoint at any time via the dashboard. Deletion removes the stored credential immediately. Existing run records that reference the endpoint are retained per our standard data retention policy.
12. Compliance Evaluation
The Service provides tooling to evaluate LLM outputs against regulatory and compliance criteria (“compliance evaluation”). By using compliance evaluation features you acknowledge the following:
Not legal advice. Compliance evaluation results are automated assessments produced by an AI judge. They are not legal advice, legal opinions, or formal compliance certifications. Results should not be relied upon as a substitute for qualified legal or regulatory counsel.
No guarantee of accuracy. The AI judge evaluates responses against the criteria and regulation text you provide. The quality of results depends entirely on the quality of your compliance store, your rubric criteria, and the regulation text you upload. We make no warranty that results are accurate, complete, or current.
Your responsibility. You are responsible for: (a) ensuring the regulation text in your compliance store is accurate and up to date; (b) interpreting results in the context of your specific regulatory obligations; (c) obtaining qualified legal advice before making compliance decisions based on Service output.
No certification. A passing compliance evaluation result from the Service does not constitute regulatory certification, legal clearance, or any form of official compliance status. Do not represent Service output as a compliance certification to regulators, auditors, or customers.
Data you upload. Regulation documents and compliance criteria you upload are stored and used solely to power your evaluations. They are not shared with other users. Standard data handling terms apply (see Privacy Policy).
13. Acceptable Data
The Service is designed to test prompt quality and model behaviour. It is not designed to process sensitive personal or regulated data.
Accepted: Prompts and system instructions, synthetic or anonymised test cases, internal documents with no PII or regulated data, ground truth answers derived from internal knowledge.
Not accepted: Personal data (names, emails, addresses, IDs), financial data (card numbers, account numbers), health data (medical records, diagnoses), data subject to export controls, customer data of any kind, or any data regulated under HIPAA, PCI-DSS, or GDPR special categories.
You are responsible for ensuring data you submit complies with these restrictions. We reserve the right to suspend accounts where prohibited data is detected.