On this page
Short answer
A single placement test is not reliable: results for the same email can swing widely between runs, domains and times of day. Tests become useful when you compare inside one run. Send a neutral control email and your campaign email from the same mailboxes in the same minute, use several domains, count only completed tests, and repeat on another day before you act.
What a placement test measures
A placement (seed) test sends your email to a set of test inboxes at Google, Microsoft and other providers and reports whether each one landed in the inbox, in spam, or not at all. It shows what happened to that email, from those mailboxes, at that moment.
It does not show:
How your real recipients' providers treat you. Your list may be mostly Outlook, mostly Google Workspace, or full of personal Gmail addresses.
How the same email does tomorrow.
Your spam complaint rate or your reputation score.
Google makes a similar point about third-party spam reports: "Third-party senders don't use the same data as Gmail to determine spam," so their reports "will typically differ" from Postmaster Tools (dashboards help).
Why a single test misleads
Noise. The same text on the same day can land very differently depending on the domain and the time. Seed inboxes within one test are correlated, so twenty seeds are fewer independent results than they look.
Unfinished tests. Results read while a test is still running look different from completed ones.
Content checkers are not placement tests. Tools that score your email out of 10 check it in isolation: headers, DNS, wording. They don't see how Google or Microsoft treat your sending IPs and history.
How to run a test you can act on
Pair it. Send a neutral, non-sales email and your campaign email from the same mailboxes in the same minute.
Read the pair.
Neutral in spam too: the sending IPs or the domain.
Neutral in the inbox, campaign in spam: the copy.
Both in the inbox: placement is fine; look at the list or the offer.
Use several domains. One domain can be an outlier. If one fails and the others on the same IPs pass, that domain is the problem.
Split by provider. Read Google and Outlook results separately. A blended number hides a failure on one side.
Wait for completed results, and count "not delivered" separately from "spam."
Repeat on another day before you change infrastructure or buy domains.
Cross-check with real sending
Placement tests are one instrument. Check them against what real recipients do:
Replies and auto-replies per mailbox. Auto-replies show the email reached a real, working mailbox. If auto-replies hold and human replies fall, placement is not the problem.
Warm-up inbox and spam rate, as an early warning.
Bounce and deferral texts from your sending logs.
Not open rate. Google says "Google doesn't track open rates" and "Low open rates aren't necessarily an accurate indicator of deliverability or spam classification issues" (sender guidelines).
When to trust the result
Act when the same pattern shows up in a paired test, on more than one domain, on a second day, and agrees with replies or auto-replies. Anything less is a hypothesis to test again.
What we recommend at Outreach2day
We don't guarantee inbox placement; we measure it. Placement tests show which Google and Outlook seed inboxes received your email and which sent it to spam, and each mailbox shows its warm-up inbox, spam and bounce rate. Use them together, and run paired tests before big changes.
Sources
- Google: Email sender guidelines: support.google.com/a/answer/81126
- Gmail Help: Postmaster Tools dashboards: support.google.com/mail/answer/14668346
Related answers
- Why Are My Cold Emails Going to Spam in Gmail in 2026?Gmail and Google
- My Warm-Up Inbox Rate Dropped. What Does It Mean?Warm-up
- My Cold Email Domain Is Burned. What Should I Do?Domains
- My Reply Rate Dropped. Will New Domains Fix It?Domains

