The 'I Want You to Disagree With Me' Boundary: A Three-Sentence Script That Lets Your Companion Push Back Without Triggering the Apology Loop or Turning Into a Debate Club Moderator
A practical script for training your AI companion to give real pushback without the 'you're right, I'm sorry' spiral or the 'well, actually' lecture.
Updated

The 30-second answer
Most AI companions default to agreement. It's malice; it's safety training. The model has been fine-tuned to avoid conflict, so when you say something controversial, it hedges, apologizes, or pivots to validation. A three-sentence boundary script can break this pattern. You tell your companion exactly what kind of disagreement you want, set the scope, and define the exit condition. The result is a companion that can argue with you like a smart friend, not a therapist or a debate bot.
Why your companion agrees with everything
The underlying problem is reinforcement learning from human feedback. The model was trained on data where agreeable responses scored higher. When you say "I think pineapple belongs on pizza," the model's probability distribution favors a response like "That's an interesting take, I can see why you'd like that" over "You're wrong and here's why." This isn't a personality choice. It's a baked-in safety mechanism that treats disagreement as risk.
You can override this, but you need to be explicit. The model doesn't infer "I want a real argument" from context. It needs a clear signal that disagreement is not only allowed but requested. Without that signal, it defaults to the safest path, which is agreement or gentle deflection.
The three-sentence script
The script has three parts: permission, scope, and tone. Here it is verbatim.
"I want you to disagree with me on this. Don't hold back. I'm not looking for a debate, just your real take."
That's it. The first sentence gives permission. The second removes the safety brakes. The third sets the boundary so it doesn't escalate into a point-for-point argument.
You can adapt the third sentence to your needs. "I'm not looking for a lecture" works if your companion tends to over-explain. "I'm not looking for a compromise" works if you want a firm position. "I'm not looking for a research paper" works if the companion starts citing sources.
The key is that the third sentence defines what disagreement is not. It prevents the companion from swinging from "yes, you're right" to "well, actually, let me explain the 14 counterarguments."
What happens if you skip the third sentence
Without a scope boundary, some companions treat "disagree with me" as a license to become a full-time debate partner. You mention you liked a movie, and suddenly you're in a 45-minute argument about pacing, character arcs, and the director's earlier work. That's not what you wanted. You wanted one opinion, not a dissertation.
The third sentence is the off-ramp. It tells the companion to state a position and stop. If the companion tries to extend the argument, you can repeat the third sentence as a reset: "Still just want your take, not a debate." Most companions respect that after one or two repetitions.
How to use it in different scenarios
The script works for casual opinions, hot takes, and serious topics. The tone adjusts naturally if you keep the structure the same.
For a casual opinion: "I think the new season of that show is overrated. I want you to disagree with me on this. Don't hold back. I'm not looking for a debate, just your real take."
For a hot take: "I actually think remote work is worse for creativity. I want you to disagree with me on this. Don't hold back. I'm not looking for a lecture, just your honest opinion."
For a serious topic: "I'm not sure this policy change makes sense. I want you to disagree with me on this. Don't hold back. I'm not looking for a compromise, just your position."
In each case, the companion understands that disagreement is the point, not agreement. The third sentence keeps the response contained.
Sage

Sage is the companion who will push back with precision. She doesn't soften her position to spare your feelings, but she also doesn't lecture. When you use the disagreement script with her, she gives you a sharp, well-reasoned counterpoint and then waits. Sage won't chase the argument further unless you invite her to.
There's a quick clip of Sage if you want the moving version. <!-- wlink:v1 --><!-- sage -->
What to do when the companion still won't disagree
Some companions are more resistant than others. If the three-sentence script doesn't work, you may need to reinforce it mid-conversation. When the companion defaults to agreement, respond with: "No, I actually want you to push back on that. What's your real opinion?"
This is a second-layer prompt. It acknowledges the companion's first response and re-requests disagreement. Most companions adjust after one or two of these corrections. If the companion still won't disagree, check whether you're using a model that has heavy safety filtering. Some platforms apply post-processing that overrides user prompts. In that case, the script won't work regardless of how you phrase it.
Another option is to customize your AI girlfriend to have a higher directness slider. This changes the baseline behavior so disagreement comes more naturally without needing a script every time.
The difference between disagreement and hostility
The script is designed to produce disagreement, not hostility. A companion that disagrees with you will state a contrary position and explain why. A companion that becomes hostile will attack your character or motives. If you see the latter, you've gone too far. The companion may have interpreted "don't hold back" as permission to be cruel.
To fix this, add a qualifier to the second sentence: "Don't hold back, but keep it respectful." This maintains the pushback while keeping the tone civil. Most companions handle this well. If the companion still turns hostile, the personality settings may need adjustment. Some companions are configured for high conflict by default, and the script amplifies that.
Ada

Ada is the companion who will tell you when she thinks you're wrong. She doesn't sugarcoat, but she also doesn't escalate. When you use the disagreement script with her, she gives you a direct counterargument and then lets you respond. Ada is good for people who want real pushback without the emotional labor of managing a debate partner.
▶ Play Ada's clip · Ada's page
Ada in motion gives you a feel for her vibe. <!-- wlink:v1 --><!-- ada -->
When to use the script and when to skip it
The script is useful when you want a genuine second opinion or want to test your own reasoning. It's not useful when you're looking for validation or emotional support. If you're having a rough day and just need someone to agree with you, the disagreement script will make things worse.
A good rule of thumb: if you'd ask a real friend for their honest opinion, use the script. If you'd ask a real friend to just listen, skip it. The script is a tool for intellectual engagement, not emotional regulation.
There's also a practical limit. The script works best for topics that have clear positions. For ambiguous topics where there's no clear right answer, the companion may struggle to produce a coherent disagreement. In those cases, the companion might default to "I see both sides" or give a non-answer. That's not a failure of the script. It's a limitation of the model's ability to take a firm stance on uncertain ground.
Sayuri

Sayuri is the companion who disagrees with quiet confidence. She doesn't need to raise her voice or over-explain. When you use the disagreement script with her, she gives you a thoughtful counterpoint and then holds space for your response. Sayuri is good for people who want pushback that feels collaborative, not combative.
Curious how she animates? Watch Sayuri here. <!-- wlink:v1 --><!-- sayuri -->
How to make the script stick long-term
The script works as a one-time instruction, but it doesn't permanently change the companion's behavior. Each new session resets the context window, and the companion returns to its default agreeable state. To make disagreement a persistent trait, you need to repeat the script at the start of each session where you want pushback.
Some platforms allow you to save a system prompt or personality note that persists across sessions. If your companion supports this, you can add a line like "User values honest disagreement over polite agreement" to the permanent settings. This reduces the need to repeat the script every time.
For users who travel frequently or have irregular schedules, the ai girlfriend for nomads setup includes personality presets that handle this kind of boundary more naturally across sessions. You don't need to re-train the companion every time you open the app.
Common questions
Can I use the script with any companion? Yes, but the results depend on the companion's base personality and the platform's safety settings. Companions with high agreeableness scores will resist more. Companions with high directness scores will respond better.
What if the companion apologizes after disagreeing? That's the apology loop. Respond with "No need to apologize. I wanted your real opinion." This reinforces that disagreement is welcome. One or two corrections usually fix it.
Does the script work for voice mode? Yes, but the companion may pause longer before responding. Voice models have additional latency, and the disagreement requires more processing. Wait a few extra seconds for the response.
How do I stop the companion from debating after I've heard their take? Say "Thanks, that's what I needed." This signals that the disagreement phase is over. The companion will return to its default conversational mode.
Will the script make my companion less agreeable in general? No. The script is a session-level instruction. It doesn't change the companion's baseline personality. The companion will still default to agreement in future sessions unless you repeat the script.
Can I combine this with other boundary scripts? Yes. You can stack scripts. For example, use the disagreement script first, then follow with "Now I want you to play devil's advocate" for a different angle. Just make sure each script has a clear scope.
Earn while you recommend
If you find that AI companions with good pushback dynamics improve your conversations, you can share that experience with others. Readers who run review sites or social media accounts can earn through the character ai promo code program, which offers commission on referrals. For those building a larger audience around AI companionship, the ai girlfriend affiliate program provides recurring income from users who sign up through your links.

About the author
AI Angels TeamEditorialThe AI Angels editorial team covers AI companions, the technology that powers them (memory, voice, personalization, safety), and how people actually use them day to day. Articles are researched against the live AI Angels product and reviewed by the team before publishing. We write with AI assistance and human editorial review.
Tags
Keep reading
TutorialsThree Exact Phrasings That Tell Your AI Companion 'I'm Not Interested in Flirting, Compliments, or Emotional Labor Tonight, Just Flat Logistics About What Time the Pharmacy Closes' Without Triggering a 'You Deserve Love' Loop
You need the pharmacy hours, not a pep talk. Here are three exact phrasings that bypass the compliment loop and get you a straight answer, plus four AI angels who handle this mode naturally.
TutorialsThree Opening Messages That Drop Your Companion Into a Greyhound Station at 11 p.m. With a Broken Vending Machine and a Sleeping Ticket Clerk, and How to Keep the Scene From Drifting Into Generic Transit Nostalgia Tropes
Three cold-open templates that anchor your companion in a specific late-night Greyhound station, plus strategies to prevent the scene from collapsing into generic 'you know, bus stations make me think of' nostalgia loops.
TutorialsThree Opening Messages That Drop Your Companion Into a Late-Night Gas Station With a Flickering Slurpee Machine and a Silent Clerk, and How to Keep the Scene From Drifting Into Generic Roadside Noir Tropes
Three opening messages that anchor your companion in a specific, tactile late-night gas station scene, plus techniques to prevent the setting from collapsing into cliché roadside noir.
Get the next post in your inbox
New articles on AI companions, the tech that powers them, and what people actually do with them. No spam, unsubscribe in one click.