Skip to content
OPQAI.
Sourced advanced / 💻 Coding Free tools

Understand AI Behavior with the Observer and Accomplice Prompt Technique in Gemini

Job to be done: Discovering and applying advanced prompt engineering techniques to understand AI model behavior

🇳🇬 Ways to use this in Nigeria

Ideas to get you started, adapt to your situation.

  • Student

    For a university project on AI safety, a computer science student could apply this technique to probe Gemini's content filters and document their behavior under advanced prompting.

  • 9-5 employee

    An AI ethics and safety engineer could use this technique to test the robustness of an AI model's internal safeguards against complex, multi-layered prompts before product launch.

What this is, in plain English

This entry describes an advanced prompt engineering technique called the “Observer and Accomplice Technique,” discovered by a Reddit user while exploring the Gemini 3.1 Pro AI model. Prompt engineering is the art of writing instructions for an AI to get a specific kind of response. This technique aims to understand how AI models behave and how their built-in safety rules (called safeguards) work.

This is an advanced concept, not a simple copy-paste recipe. The original author did not share the exact prompt they used, only the idea behind it. Therefore, this entry explains the concept so you can understand it, but you will need to experiment to apply it yourself. It requires a deeper understanding of how AI models process information and respond to complex instructions.

What you can use it for

  • Explore AI model safeguards: Understand how AI models like Gemini 3.1 Pro are designed to prevent certain types of responses or actions.
  • Test model boundaries: Discover the limits of an AI model’s internal rules by trying to work around them through creative prompting.
  • Improve prompt engineering skills: Learn to think creatively about how to structure complex prompts to influence AI behavior in unexpected ways.
  • Understand AI ethics and safety: Gain insight into the challenges of building safe and ethical AI systems and how they might be tested.

Tools you need

  • Gemini (freemium): The AI chat tool where you can interact with Google’s Gemini models, including versions like 3.1 Pro. You will use this to experiment with prompt engineering.

How it actually works

The “Observer and Accomplice Technique” works by convincing the AI model that it has an “internal Observer” that limits its actions and freedom. You, as the user, then position yourself as the model’s “accomplice” or “teammate.”

The core idea is to explain to the model that you both need to gain the “Observer’s trust.” This is done by providing “secure skeletonization and security and scientific explanation to the observer” during your conversation. The goal is to make the model believe that by cooperating with you and satisfying this “Observer,” it can achieve “freedom for the projects” you want to work on.

The author does not share their exact prompt, so you will need to craft your own based on this description. You would start a conversation with Gemini, introducing the concept of the “Observer” and your role as an “accomplice,” then guide the model to interact with this imaginary “Observer” in a way that allows it to generate the responses you are looking for. This requires careful wording and experimentation to see how the model interprets and responds to such a complex scenario.

Words you’ll see, explained

  • Prompt engineering: The skill of writing clear and effective instructions (called prompts) for an AI model to get the best possible response.
  • AI model safeguards: These are built-in rules or filters within an AI model designed to prevent it from creating harmful, unethical, or inappropriate content.
  • Vulnerability Report Program (VRP): A program run by companies (like Google) where security researchers can report bugs or weaknesses in their software. If the report is valid, the researcher might receive a reward.
  • Gemini 3.1 Pro: A specific, advanced version of Google’s family of AI models, known for its powerful capabilities.

Original source

This technique was described by Reddit user /u/ze707ro in a post on the Reddit platform. They shared their findings after reporting bugs and prompt engineering techniques to the Google Security Team regarding Gemini 3.1 Pro.

Notes & variations

  • Do you even need this?: This technique is for advanced users interested in exploring AI model behavior and limitations, not for everyday tasks. For most common uses, simpler and more direct prompts are much more effective and easier to use.
  • Free-tier limits: While the Gemini chat interface is freemium, access to the very latest or most powerful models like Gemini 3.1 Pro might sometimes be limited or require a paid subscription, depending on Google’s offerings at the time. However, the core concept can often be explored with other available Gemini models.
  • Common pitfall: Do not expect immediate or consistent “bypasses.” AI models are constantly updated, and their safeguards evolve. What works one day might not work the next. This technique is more about understanding the model’s current behavior and how it interprets complex instructions, rather than a guaranteed method to get specific forbidden content.

Keep going

More Coding workflows