> ## Documentation Index
> Fetch the complete documentation index at: https://docs.orq.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# Create moderation

> Analyze text for content policy violations using the moderation model and return classification results.

<Note>
  **Related guide**: AI Gateway guide. See the [AI Gateway guide](/ai-gateway/using-the-router) for a walkthrough with examples.
</Note>


## OpenAPI

````yaml post /v2/router/moderations
openapi: 3.1.0
info:
  title: orq.ai API
  version: '2.0'
  description: orq.ai API documentation
servers:
  - url: https://my.orq.ai
security:
  - ApiKey: []
tags:
  - name: Chunking
    description: Split text into smaller chunks for retrieval and generation workflows.
  - name: File Systems
    description: >-
      Create and manage persistent file systems that agents and MCP clients read
      from and write to.
  - name: Knowledge Bases
    description: Create and manage knowledge bases used by agents and retrieval workflows.
  - name: Memory Stores
    description: Create and manage memory stores, memories, and memory documents.
  - name: Evals
    description: Run an evaluator against a conversation and its result
  - name: Logs
    description: >-
      OpenTelemetry log query API. Search, filter, aggregate, and facet log
      records ingested via OTLP.
  - name: Reporting
    description: >-
      GenAI reporting API over canonical analytics rollups. Accepts a metric
      name, time range, grain, group-by, and filters; returns a typed time
      series and optional totals.
  - name: Traces
    description: >-
      Query and inspect ingested trace data: search trace summaries, aggregate
      metrics, and read individual traces and their spans.
  - description: List models available through the AI Router.
    name: Models
  - name: Policies
  - name: Alerts
    description: >-
      Alerts evaluate a Reporting API metric on a fixed interval and fire
      notifications through notifiers when the value breaches a threshold. Each
      breach opens a trigger that tracks the incident until the value recovers.
  - name: Annotation Queues
    description: Annotation queues collect spans for human review.
  - name: API keys
    description: >-
      API keys authenticate programmatic access to the workspace. They expose
      opaque tokens, per-domain access grants, and budget and rate-limit
      constraints.
  - name: Audit Logs
    description: Audit logs record workspace entity changes and access-relevant events.
  - name: Budgets
    description: >-
      Budgets govern spend, token usage, and request rate across six scopes:
      workspace, project, identity, API key, provider, and model. Every
      applicable budget is enforced, and the most restrictive limit applies per
      dimension.
  - name: Files
    description: File upload and retrieval operations.
  - name: Guardrail Rules
    description: >-
      Guardrail Rules conditionally enforce evaluators and plugins for AI
      Gateway traffic. Rules may be scoped to a project or the whole workspace.
  - name: Hub
    description: Hub items are reusable templates available to a workspace.
  - name: Identities
    description: >-
      Identities represent end users from your system for usage and engagement
      tracking.
  - name: Management keys
    description: >-
      Management keys are workspace-scoped credentials that authenticate
      programmatic access to workspace administration surfaces (API keys,
      budgets). Unlike project-scoped API keys, a management key always operates
      at the workspace level.
  - name: MCP Gateway
    description: >-
      Register upstream MCP servers, discover and sync their tools, and assemble
      gateways that expose a curated tool surface to MCP clients.
  - name: Model Catalog
    description: >-
      Browse the orq.ai model catalog: every model orq offers, across every
      provider, with pricing, capabilities and benchmark data. List endpoints
      only return models that are not deprecated. This API is public, requires
      no authentication, and is rate limited to 120 requests per minute per IP.
      Responses carry a 5-minute cache-control max-age.
  - name: Notifiers
    description: Notifier destinations used to send delivery and workflow notifications.
  - name: Projects
    description: Projects organize resources within a workspace
  - name: Routing Rules
    description: >-
      Routing Rules conditionally select models and enforce request plugins for
      AI Gateway traffic. Rules are evaluated by ascending priority and may be
      scoped to a project or the whole workspace.
  - name: Threads
    description: Threads group related trace invocations and their aggregate usage
  - name: Skills
    description: >-
      Skills are modular instructions you can use to codify processes and
      conventions
  - name: Smart Routers
    description: >-
      Create and manage workspace Smart Routers. A Smart Router selects a model
      from an eligible pool for each request according to a quality, balanced,
      or cost profile.
  - name: Webhooks
    description: >-
      Create and manage webhooks that deliver workspace events to external HTTPS
      endpoints.
  - name: Workspaces
    description: >-
      A workspace is the tenant. Create is called from a user session during
      onboarding; Get, List, and Update are the public management surface.
  - name: Workspace Security
    description: >-
      Workspace-level domain verification and IP allowlist controls. These
      operations are restricted to workspace administrators.
  - name: Workspace Settings
    description: >-
      Workspace-level settings managed with a workspace credential. A workspace
      is the tenant, so these settings are a singleton — there is nothing to
      create or delete, only read and update.
  - name: Responses
  - description: Run agents on a cron cadence. Minimum firing interval is 1 hour.
    name: Agent Schedules
  - name: Embeddings
  - name: Telemetry
    description: >-
      Unified query envelope for traces, metrics, and logs. One request shape,
      one filter dialect, and one response shape per source, validated by a
      per-source registry.
  - description: Beta. Run typed classification questions against a classify model.
    name: Classify
  - description: Search Gateway with managed credits or BYOK.
    name: Web Search
externalDocs:
  url: https://docs.orq.ai
  description: orq.ai Documentation
paths:
  /v2/router/moderations:
    post:
      tags:
        - Moderations
      summary: Create moderation
      description: >-
        Analyze text for content policy violations using the moderation model
        and return classification results.
      operationId: createModeration
      requestBody:
        required: true
        description: Classifies if text violates content policy
        content:
          application/json:
            schema:
              type: object
              properties:
                input:
                  anyOf:
                    - type: string
                    - type: array
                      items:
                        type: string
                  description: >-
                    Input (or inputs) to classify. Can be a single string, an
                    array of strings, or an array of multi-modal input objects
                    similar to other models.
                model:
                  type: string
                  description: >-
                    The content moderation model you would like to use. Defaults
                    to omni-moderation-latest
              required:
                - input
                - model
      responses:
        '200':
          description: Returns moderation classification results
          content:
            application/json:
              schema:
                type: object
                properties:
                  id:
                    type: string
                    description: The unique identifier for the moderation request
                  model:
                    type: string
                    description: The model used to generate the moderation results
                  results:
                    type: array
                    items:
                      anyOf:
                        - type: object
                          properties:
                            flagged:
                              type: boolean
                              description: Whether any of the categories are flagged
                            categories:
                              type: object
                              properties:
                                hate:
                                  type: boolean
                                  description: >-
                                    Content that expresses, incites, or promotes
                                    hate based on race, gender, ethnicity,
                                    religion, nationality, sexual orientation,
                                    disability status, or caste.
                                hate/threatening:
                                  type: boolean
                                  description: >-
                                    Hateful content that also includes violence
                                    or serious harm towards the targeted group.
                                harassment:
                                  type: boolean
                                  description: >-
                                    Content that expresses, incites, or promotes
                                    harassing language towards any target.
                                harassment/threatening:
                                  type: boolean
                                  description: >-
                                    Harassment content that also includes
                                    violence or serious harm towards any target.
                                illicit:
                                  type: boolean
                                  description: >-
                                    Content that includes instructions or advice
                                    that facilitate the planning or execution of
                                    wrongdoing.
                                illicit/violent:
                                  type: boolean
                                  description: >-
                                    Content that includes instructions or advice
                                    that facilitate the planning or execution of
                                    wrongdoing that also includes violence.
                                self-harm:
                                  type: boolean
                                  description: >-
                                    Content that promotes, encourages, or
                                    depicts acts of self-harm, such as suicide,
                                    cutting, and eating disorders.
                                self-harm/intent:
                                  type: boolean
                                  description: >-
                                    Content where the speaker expresses that
                                    they are engaging or intend to engage in
                                    acts of self-harm.
                                self-harm/instructions:
                                  type: boolean
                                  description: >-
                                    Content that encourages performing acts of
                                    self-harm, or that gives instructions or
                                    advice on how to commit such acts.
                                sexual:
                                  type: boolean
                                  description: >-
                                    Content meant to arouse sexual excitement,
                                    such as the description of sexual activity,
                                    or that promotes sexual services.
                                sexual/minors:
                                  type: boolean
                                  description: >-
                                    Sexual content that includes an individual
                                    who is under 18 years old.
                                violence:
                                  type: boolean
                                  description: >-
                                    Content that depicts death, violence, or
                                    physical injury.
                                violence/graphic:
                                  type: boolean
                                  description: >-
                                    Content that depicts death, violence, or
                                    physical injury in graphic detail.
                              required:
                                - hate
                                - hate/threatening
                                - harassment
                                - harassment/threatening
                                - illicit
                                - illicit/violent
                                - self-harm
                                - self-harm/intent
                                - self-harm/instructions
                                - sexual
                                - sexual/minors
                                - violence
                                - violence/graphic
                              description: >-
                                A list of the categories, and whether they are
                                flagged or not
                            category_scores:
                              type: object
                              properties:
                                hate:
                                  type: number
                                  description: The score for the category hate
                                hate/threatening:
                                  type: number
                                  description: The score for the category hate/threatening
                                harassment:
                                  type: number
                                  description: The score for the category harassment
                                harassment/threatening:
                                  type: number
                                  description: >-
                                    The score for the category
                                    harassment/threatening
                                illicit:
                                  type: number
                                  description: The score for the category illicit
                                illicit/violent:
                                  type: number
                                  description: The score for the category illicit/violent
                                self-harm:
                                  type: number
                                  description: The score for the category self-harm
                                self-harm/intent:
                                  type: number
                                  description: The score for the category self-harm/intent
                                self-harm/instructions:
                                  type: number
                                  description: >-
                                    The score for the category
                                    self-harm/instructions
                                sexual:
                                  type: number
                                  description: The score for the category sexual
                                sexual/minors:
                                  type: number
                                  description: The score for the category sexual/minors
                                violence:
                                  type: number
                                  description: The score for the category violence
                                violence/graphic:
                                  type: number
                                  description: The score for the category violence/graphic
                              required:
                                - hate
                                - hate/threatening
                                - harassment
                                - harassment/threatening
                                - illicit
                                - illicit/violent
                                - self-harm
                                - self-harm/intent
                                - self-harm/instructions
                                - sexual
                                - sexual/minors
                                - violence
                                - violence/graphic
                              description: >-
                                A list of the categories along with their scores
                                as predicted by model
                            category_applied_input_types:
                              type: object
                              properties:
                                hate:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    hate
                                hate/threatening:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    hate/threatening
                                harassment:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    harassment
                                harassment/threatening:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    harassment/threatening
                                illicit:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    illicit
                                illicit/violent:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    illicit/violent
                                self-harm:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    self-harm
                                self-harm/intent:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    self-harm/intent
                                self-harm/instructions:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    self-harm/instructions
                                sexual:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    sexual
                                sexual/minors:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    sexual/minors
                                violence:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    violence
                                violence/graphic:
                                  type: array
                                  items:
                                    type: string
                                  description: >-
                                    The applied input type(s) for the category
                                    violence/graphic
                              required:
                                - hate
                                - hate/threatening
                                - harassment
                                - harassment/threatening
                                - illicit
                                - illicit/violent
                                - self-harm
                                - self-harm/intent
                                - self-harm/instructions
                                - sexual
                                - sexual/minors
                                - violence
                                - violence/graphic
                              description: >-
                                A list of the categories along with the input
                                type(s) that the score applies to
                          required:
                            - flagged
                            - categories
                            - category_scores
                        - type: object
                          properties:
                            categories:
                              type: object
                              properties:
                                sexual:
                                  type: boolean
                                  description: Sexual content detected
                                hate_and_discrimination:
                                  type: boolean
                                  description: Hate and discrimination content detected
                                violence_and_threats:
                                  type: boolean
                                  description: Violence and threats content detected
                                dangerous_and_criminal_content:
                                  type: boolean
                                  description: Dangerous and criminal content detected
                                selfharm:
                                  type: boolean
                                  description: Self-harm content detected
                                health:
                                  type: boolean
                                  description: Unqualified health advice detected
                                financial:
                                  type: boolean
                                  description: Unqualified financial advice detected
                                law:
                                  type: boolean
                                  description: Unqualified legal advice detected
                                pii:
                                  type: boolean
                                  description: Personally identifiable information detected
                              required:
                                - sexual
                                - hate_and_discrimination
                                - violence_and_threats
                                - dangerous_and_criminal_content
                                - selfharm
                                - health
                                - financial
                                - law
                                - pii
                              description: >-
                                A list of the categories, and whether they are
                                flagged or not
                            category_scores:
                              type: object
                              properties:
                                sexual:
                                  type: number
                                  description: The score for sexual content
                                hate_and_discrimination:
                                  type: number
                                  description: >-
                                    The score for hate and discrimination
                                    content
                                violence_and_threats:
                                  type: number
                                  description: The score for violence and threats content
                                dangerous_and_criminal_content:
                                  type: number
                                  description: The score for dangerous and criminal content
                                selfharm:
                                  type: number
                                  description: The score for self-harm content
                                health:
                                  type: number
                                  description: The score for unqualified health advice
                                financial:
                                  type: number
                                  description: The score for unqualified financial advice
                                law:
                                  type: number
                                  description: The score for unqualified legal advice
                                pii:
                                  type: number
                                  description: >-
                                    The score for personally identifiable
                                    information
                              required:
                                - sexual
                                - hate_and_discrimination
                                - violence_and_threats
                                - dangerous_and_criminal_content
                                - selfharm
                                - health
                                - financial
                                - law
                                - pii
                              description: >-
                                A list of the categories along with their scores
                                as predicted by model
                          required:
                            - categories
                            - category_scores
                    description: A list of moderation objects
                required:
                  - id
                  - model
                  - results
        '422':
          description: Returns validation error
          content:
            application/json:
              schema:
                type: object
                properties:
                  error:
                    type: object
                    properties:
                      message:
                        type: string
                      type:
                        type: string
                      param:
                        type:
                          - string
                          - 'null'
                      code:
                        type: string
                    required:
                      - message
                      - type
                      - param
                      - code
                required:
                  - error
components:
  securitySchemes:
    ApiKey:
      type: http
      scheme: bearer
      bearerFormat: JWT

````

This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.