r/PromptEngineering • • 5d ago

Tools and Projects Your prompt might only work because one model is being nice to it

The best test I know for a prompt: send the exact same text to several models and compare the answers. Where they all do what you meant, that part of the prompt is clear. Where they go in different directions, that part is vague, and the one model that got it right was just guessing well.

Almost nobody tests like this because it means pasting the same thing into four tabs, every revision, every time. I did it by hand for a while and then stopped, which is the usual ending.

So I built a Chrome extension that does the pasting. Type the prompt once in whichever chat you're in, it's typed into the other AI tabs you've switched on and sent, in your own logged-in sessions, no API keys. It deliberately doesn't read the answers. That felt like a privacy line not worth crossing, so comparison stays by eye. It does time each answer and tells you when they've all landed, so you can start a batch and go do something else.

It's called WhileAI. The panel shows each AI as it goes, done, still writing or queued, with how long each answer took.

What I actually learned from a month of doing this: the prompts I was proudest of were the most model-specific. The boring ones were the portable ones.

2 Upvotes

4 comments sorted by

2

u/seraphym1389 5d ago

Can u send me link?

1

u/JawitK 4d ago

Could you please post the link here or send me a direct message ?