Hello from Artfical. Since tAI 4.2 went out on August 19, we've been watching how traffic shifted across the model lineup, and it shifted the way we expected: 4.2 now handles the bulk of the ambitious, long-running, higher-stakes work it was built for. That gave us room to look again at tAI 4.1's own safeguards with fresh eyes.
tAI 4.1's cybersecurity and hazardous-material safeguards were originally tuned while it was still our flagship, carrying the widest range of traffic, including the riskiest queries. With that traffic now mostly routed to tAI 4.2, we re-evaluated 4.1's thresholds against real usage and found a meaningful share of legitimate, everyday technical questions, dependency vulnerability lookups, standard security research, routine lab-safety questions, were getting caught by rules set for a different traffic mix than the one 4.1 actually sees today.
What changed
- Cybersecurity-adjacent questions (dependency CVEs, common exploit terminology, routine pentesting workflow questions) are rerouted or blocked less often when nothing in the surrounding conversation suggests actual malicious intent.
- Hazardous-material safeguards kept their hard limits exactly where they were, only the false-positive rate on clearly educational or professional queries came down.
- Tool-use safety checks (the guardrails around a model actually operating tools on your behalf) are unchanged, this update is scoped to conversational safeguards only.
None of this changes what tAI 4.1 will actually do in a genuinely ambiguous or risky situation, it still asks, still declines, still routes to review the same way it did before. What changed is how often a completely ordinary question gets treated as one of those situations in the first place.
Open tAI →