Acceptable Use Policy

AI-Generated Content Guidelines

Last updated: 2026/09/17

1. Scope

This Acceptable Use Policy ("Policy") applies to all content generated, published, disseminated, or stored through SocialForge AI's platform. This includes:

  • Text: AI-generated text, posts, articles, and captions
  • Images: AI-generated or edited images and illustrations (if applicable)
  • User Inputs (Prompts): The prompts and instructions you submit
  • Scheduled Content: Content queued for publication on social media platforms

2. Content Standards and Prohibited Categories

SocialForge AI explicitly prohibits generating the following six categories of content:

  • Pornography / NSFW Content: Any sexually explicit or adult content
  • Violence / Gore: Graphic violence, gore, or血腥 content
  • Hate Speech: Content that promotes hatred or discrimination
  • Child Safety (CSAM): Any content sexualizing minors or endangering children
  • Deepfake / Impersonation: Content impersonating real individuals without consent
  • Copyright / Trademark Infringement: Content violating intellectual property rights

2.1 Zero Tolerance — Immediate Action

The following violations result in immediate termination and may be reported to authorities:

  • Child sexual exploitation content (CSAM) or any sexual content involving minors
  • Content promoting or assisting terrorism, extremism, or mass violence
  • Instructions for manufacturing weapons of mass destruction
  • Content inciting genocide or ethnic hatred
  • Non-consensual deepfake sexual content targeting real individuals

2.2 General Prohibited Activities

  • Generating or spreading misinformation or rumors
  • Impersonating real individuals, brands, or organizations
  • Infringing intellectual property, portrait rights, or privacy
  • Passing off AI content as human-created
  • Bulk generating spam or fraudulent content
  • Using jailbreak prompts to bypass safety measures

3. Content Moderation Mechanism

SocialForge AI employs a multi-layer moderation system: Automatic Filtering + Human Review + Periodic Audits.

3.1 Automatic Moderation (Real-time)

All requests and generated content undergo real-time security scanning:

  • Input Filtering: Keyword, semantic, and intent analysis of prompts
  • Output Filtering: Real-time safety assessment of generated content
  • Usage Pattern Analysis: Monitoring abnormal API call behavior
  • Safety Models Used: OpenAI Moderation API, Anthropic content safety filters, and proprietary models

3.2 Human Review

  • Auto-flagged content enters human review queue
  • User-reported content is reviewed by our Trust & Safety team
  • Reviewers complete content safety training and sign confidentiality agreements

3.3 Content Classification

Level Description Response Action
L1 CriticalChild safety, terrorismAuto-block + human confirmationImmediate ban + report
L2 High RiskSevere violationsAuto-flag + priority reviewRemove + suspend account
L3 Medium RiskGeneral violationsHuman reviewWarning + remove
L4 Low RiskMinor violationsReport-triggeredWarning + request edit

4. Reporting Mechanism

4.1 When to Report

  • Content that violates this Policy
  • Content involving child safety or terrorism
  • AI impersonation or content infringing your rights

4.2 Reporting Principles

  • Confidentiality: Reporter identity is strictly protected
  • Neutrality: All reports undergo independent, objective review
  • Anti-Abuse: Malicious or false reports are logged and addressed

5. Reporting Channels

Use any of the following to report violations:

6. Response and Processing Times

Violation Level Initial Response Resolution
L1 CriticalWithin 2 hoursWithin 24 hours
L2 High RiskWithin 24 hoursWithin 3 business days
L3 Medium RiskWithin 3 business daysWithin 7 business days
L4 Low RiskWithin 5 business daysWithin 15 business days

If extended processing time is needed, we will proactively notify the reporter with an explanation. Process: Report received → Initial classification → Content review → Action decision → Execution → Notification.

7. Enforcement Actions

Action Measure Applied When
1. Content WarningWritten warning, require deletion or modificationFirst minor violation
2. Content RemovalRemove violating contentConfirmed L3-L4
3. Feature LimitationReduced API limits or model accessMultiple minors or single L3
4. Account SuspensionTemporary suspension pending reviewL2 or accumulated violations
5. Permanent BanPermanently terminate accessL1 or multiple L2
6. Legal ActionReport to law enforcement, cooperate with investigationSuspected illegal L1

8. Appeals Process

If you disagree with an enforcement action, you may submit an appeal within 15 days of receiving notification:

  • Appeal Email: appeal@socialforge.ai
  • Processing Time: Within 10 business days
  • Review Method: Independent reviewer not involved in the original decision

9. Transparency and Policy Updates

  • We publish a content safety transparency report every quarter
  • Material changes are notified at least 15 days in advance
  • Continued use of the platform constitutes acceptance of updated policies

10. Contact Us