Safety Approach
EaseMyPrompt.ai — Safety Approach
Effective Date: August 2026 | Version: 1.0 | EaseMyPrompt.ai, a product of Derek AI Labs
Our Commitment to Safety
EaseMyPrompt.ai, a product of Derek AI Labs believes that AI-powered tools must be developed and deployed responsibly. Safety is not a feature — it is a foundational principle built into every layer of EaseMyPrompt.ai, from the prompts we allow in our marketplace to the behaviour of Derek, our AI prompt engineer.
1. Safety by Design
We embed safety considerations at every stage of product development:
• Red-teaming: before any major feature launch, our team attempts to misuse it to identify risks
• Input filtering: user queries to Derek are screened against prohibited content categories before processing
• Output moderation: generated prompts are evaluated against safety classifiers before being shown to users
• Marketplace review: prompts submitted for sale are reviewed against our Acceptable Use Policy before listing
2. The Derek AI Safety Framework
Derek operates within a layered safety architecture:
2.1 System-Level Guardrails
Derek is configured with system-level instructions that prevent it from generating content in prohibited categories regardless of user instruction. These guardrails cannot be overridden by users.
2.2 Prompt Injection Defence
We implement measures to detect and neutralise prompt injection attacks — attempts by malicious content to override Derek's instructions.
2.3 Jailbreak Resistance
We continuously monitor for and respond to new jailbreak techniques. Accounts that repeatedly attempt to bypass safety controls are suspended.
2.4 Content Filtering
AI outputs are passed through a content safety classifier that flags potentially harmful content for human review before delivery.
3. Human Oversight
We maintain human oversight of AI systems at all times. No fully autonomous AI decision-making affects user accounts, content moderation outcomes, or financial matters without human review capability.
4. Incident Response
When safety incidents occur:
• Critical safety issues (CSAM, imminent harm): immediate action within 1 hour
• High-severity issues: action within 24 hours
• Standard issues: investigation within 7 business days
Affected users are notified where legally required and appropriate.
5. Continuous Improvement
We conduct quarterly safety reviews, update our safety classifiers regularly, and incorporate user feedback into our safety frameworks. We publish transparency reports annually covering content removals, government requests, and safety incidents.
6. Third-Party AI Models
EaseMyPrompt.ai prompts are designed for use with third-party AI models (ChatGPT, Claude, Gemini, etc.). We encourage users to review the safety policies of those platforms. We are not responsible for outputs generated by third-party models when our prompts are used on those platforms.