Tomorrow at 10 am" triggers a violence content filter block (400) — gpt-4.1, default DefaultV2 guardrail

Nishyanth Nandagopal 0 Reputation points
2026-09-16T09:03:41.8866667+00:00

Hi all,

A plain appointment-booking phrase is being blocked by the input content filter on our Azure OpenAI deployment.

Repro (Foundry playground, no system prompt, no tools, single user message):

  • User message: tomorrow at 10 am
  • Result: request blocked by the content filter

Error: Interaction blocked This interaction was blocked by a safety and security control in this asset's Foundry guardrail. Risk type: Violence (High) is detected at User Input.

Screenshot 2026-09-16 122158

Environment

  • Model: gpt-4.1, version 2025-04-14

Guardrail/content filter: Microsoft.DefaultV2 (default, unmodified)

-> I tried testing this with both medium and low threshold for all content safety risk types and still facing the same issue.

This affects several production chatbots that book appointments, so users hit it during a normal booking conversation.

Questions

  1. Is a benign date/time phrase expected to be classified as violence at severity "high"? This looks like a false positive.
  2. Any recommended mitigation other than relaxing the input filter?

Thanks!

Content Safety in Foundry Control Plane
Content Safety in Foundry Control Plane

An Azure service that enables users to identify content that is potentially offensive, risky, or otherwise undesirable. Previously known as Azure Content Moderator.

0 comments No comments

Your answer

Answers can be marked as 'Accepted' by the question author and 'Recommended' by moderators, which helps users know the answer solved the author's problem.