6 Sources
[1]
Google requires its collaborators to rate Gemini's responses without adequate preparation - Softonic
External contractors will have to evaluate responses even if they have no idea what they are about Behind the magic of generative artificial intelligence, there is a whole technical team of programmers responsible for making everything work as it should. A part of this team, the prompt engineers,
[2]
Exclusive: Google's Gemini is forcing contractors to rate AI responses outside their expertise | TechCrunch
Generative AI may look like magic, but behind the development of these systems are armies of employees at companies like Google, OpenAI and others, known as "prompt engineers" and analysts, who rate the accuracy of chatbots' outputs to improve their AI. But a new internal guideline passed down
[3]
Gemini Contractors Might Be Rating AI Prompts Outside Their Expertise
Now, Google is said to have removed the option to skip Gemini prompts Google is reportedly asking contractors working on evaluating Gemini's responses to rate prompts outside of their domain of expertise. As per the report, the Mountain View-based tech giant has removed the option to skip prompts,
[4]
Google accused of using novices to fact-check Gemini's AI answers
Hopefully no computer science majors were asked to review medical questions. There's no arguing that AI still has quite a few unreliable moments, but one would hope that at least its evaluations would be accurate. However, last week Google allegedly instructed contract workers evaluating Gemini
[5]
New Google policy instructs Gemini's fact-checkers to act outside their expertise
Summary Google employs contract research agencies to evaluate Gemini response accuracy. GlobalLogic contractors evaluating Gemini prompts are no longer allowed to skip individual interactions based on lack of expertise. Concerns exist over Google's reliance on fact-checkers without relevant
[6]
Contractors Must Evaluate Gemini AI Prompts Outside Expertise, Google Says - MEDIANAMA
Google has instructed contractors working on its AI chatbot Gemini to not skip evaluation of prompts - and AI generated responses to such prompts - even if it lay beyond their domain expertise, as per a report by TechCrunch based on internal documents. The development of chatbots, including
Share
Copy Link
Google has instructed contractors evaluating Gemini AI responses to rate prompts outside their expertise, potentially compromising the accuracy of AI-generated information on specialized topics.

Google has implemented a controversial change in its evaluation process for Gemini AI, raising concerns about the accuracy and reliability of the AI's responses. The tech giant has instructed contractors working with GlobalLogic, an outsourcing firm owned by Hitachi, to rate AI-generated responses even when the topics fall outside their areas of expertise
1
2
.Prior to this change, contractors evaluating Gemini's outputs were allowed to skip prompts that required specialized knowledge beyond their expertise. The previous guidelines stated, "If you do not have critical expertise (e.g. coding, math) to rate this prompt, please skip this task"
2
. This approach ensured that only qualified individuals assessed technical responses, potentially reducing instances of AI hallucinations and improving overall accuracy3
.The new internal guidelines, as reported by TechCrunch, now instruct contractors: "You should not skip prompts that require specialized domain knowledge"
2
. Instead, they are asked to "rate the parts of the prompt you understand" and include a note acknowledging their lack of domain knowledge3
. This change has sparked worries about the potential impact on Gemini's accuracy, especially for highly sensitive topics like healthcare2
.Under the new policy, contractors can only skip prompts in two scenarios:
2
4
This policy shift has raised several concerns:
Accuracy Issues: There are fears that Gemini could become more prone to providing inaccurate information on highly technical subjects
2
.Quality of Evaluations: The change may lead to a drop in the quality of Gemini's responses, particularly for specialized topics
3
.AI Development Goals: Questions have arisen about how this approach aligns with Google's AI development objectives, particularly in improving accuracy and reducing hallucinations
5
.Related Stories
The decision has generated controversy within the AI community. One contractor noted, "I thought the point of skipping was to increase accuracy by giving it to someone better?" highlighting the potential drawbacks of this new approach
2
4
.This development comes at a time when AI companies are under scrutiny for the accuracy and reliability of their systems. The use of human evaluators is a standard practice in AI development, aimed at grounding responses and reducing errors. However, Google's new policy appears to diverge from this established approach
3
5
.As of now, Google has not responded to requests for comment on this policy change
4
. The tech community and users alike will be closely watching how this new evaluation process impacts the performance and trustworthiness of Gemini AI in the coming months.Summarized by
Navi
[1]
[2]
1
Technology

2
Policy and Regulation

3
Technology
