Artfical AI Open tAI →
← All announcements

Loosening tAI 4.1's Safeguards

August 21, 2026 · Security

Now that tAI 4.2 is carrying most of our highest-risk traffic, we've eased some of tAI 4.1's safeguard thresholds to cut down on false positives.

Hello from Artfical. Since tAI 4.2 went out on August 19, we've been watching how traffic shifted across the model lineup, and it shifted the way we expected: 4.2 now handles the bulk of the ambitious, long-running, higher-stakes work it was built for. That gave us room to look again at tAI 4.1's own safeguards with fresh eyes.

tAI 4.1's cybersecurity and hazardous-material safeguards were originally tuned while it was still our flagship, carrying the widest range of traffic, including the riskiest queries. With that traffic now mostly routed to tAI 4.2, we re-evaluated 4.1's thresholds against real usage and found a meaningful share of legitimate, everyday technical questions, dependency vulnerability lookups, standard security research, routine lab-safety questions, were getting caught by rules set for a different traffic mix than the one 4.1 actually sees today.

What changed

None of this changes what tAI 4.1 will actually do in a genuinely ambiguous or risky situation, it still asks, still declines, still routes to review the same way it did before. What changed is how often a completely ordinary question gets treated as one of those situations in the first place.

References