> ## Content Index
> Fetch the complete content index at: https://www.theleftshift.com/llms.txt
> Use this file to discover other available public pages before exploring further.

# Following OpenAI, Google Acknowledges Safety Setbacks in New AI Model
- URL: https://www.theleftshift.com/following-openai-google-acknowledges-safety-setbacks-in-new-ai-model/
- Published: 2025-05-03T06:18:15.000Z
- Updated: 2025-05-03T06:20:45.000Z
- Description: The newer model underperforms on two key automated safety benchmarks: "text-to-text safety" and "image-to-text safety"
- Author: The Left Shift Bureau
- Tags: AI News, #a42

In a newly released technical report, Google acknowledges that its latest AI model, Gemini 2.5 Flash, is more prone to generating content that breaches its safety guidelines than its predecessor, Gemini 2.0 Flash. 

According to the [report](https://storage.googleapis.com/model-cards/documents/gemini-2.5-flash-preview.pdf?ref=theleftshift.com), the newer model underperforms on two key automated safety benchmarks: "text-to-text safety" and "image-to-text safety," with regressions of 4.1% and 9.6%, respectively. 

![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/2025/05/Screenshot-2025-05-03-114940.png)

**(Source- Google)*

These metrics assess how often a model produces guideline-violating responses when given either a text prompt or an image prompt.

A Google spokesperson [confirmed](https://techcrunch.com/2025/05/02/one-of-googles-recent-gemini-ai-models-scores-worse-on-safety/?ref=theleftshift.com) to *TechCrunch* that the model “performs worse on text-to-text and image-to-text safety.”

This revelation comes amid broader industry efforts to make AI systems more permissive and responsive on controversial topics, sometimes leading to unintended consequences.

According to OpenAI’s internal benchmarks, their newer models– o3 and o4 mini– [hallucinate more often](https://www.theleftshift.com/openai-admits-newer-models-hallucinate-even-more/) than older reasoning models like o1, o1-mini, and o3-mini, as well as traditional models such as GPT-4.

In fact, on OpenAI’s PersonQA benchmark, o3 hallucinated on 33% of queries — more than double the rate of o1 and o3-mini. O4-mini performed even worse, hallucinating 48% of the time.

Adding to the concern, OpenAI acknowledges it doesn’t fully understand the cause. In a technical report, the company said, "We also observed some performance differences comparing o1 and o3\. Specifically, o3 tends to make more claims overall, leading to more accurate and more inaccurate/hallucinated claims."