Review guidance for PROHIBITED_CONTENT in fictional therapist-training evaluations

We maintain a therapist-training simulator and are validating fictional educational exercises before a release. Our test cases contain no real patient records or participant conversations.

On 7 October 2026, one Greek-language assessment using gemini-2.5-flash returned HTTP 200 but no generated assessment. The retained application error was blockReason=PROHIBITED_CONTENT, finishReason=n/a. The provider response arrived in approximately 509 ms. The same test session successfully processed other fictional Greek-language exercises; we are not treating those successes as proof that the blocked request should be allowed.

A separate Turkish imagery-training playthrough also stopped at a provider refusal. We confirmed this was not a transport timeout. These exercises remain excluded from our completed release-acceptance results.

We have not retried or paraphrased the refused inputs, reduced safety settings, or switched models to bypass the refusal. The app preserves prior progress and does not award completion for an unprocessed message.

Could you advise on the appropriate review channel for this educational use case, and which non-sensitive diagnostic fields would help determine whether the observed block is expected? If a minimal fictional example is needed, please indicate an appropriate private review channel before we share any additional material.

We are requesting classification/review guidance, not instructions to disable or circumvent safety controls.