Skip to the instrument
INSTRUMENT 07

Spirit Box

RF Sweep Receiver
PWRSIG
Operational
Receiver off

Throw the switch to start the sweep.

There is sound. Then make a noise and be quiet — it answers into the gap.

SWEEPROOMVOICE
--:--
Band — 64 channels, synthesised

The tuner is parked. No channel is selected and the detector is reading nothing, which is what a receiver reads when it is not sweeping.

PWRSWEEPROOMVOICE
Receiver offline

Bringing the receiver online needs JavaScript, a speaker and — if you want it to answer you rather than talk to itself — a microphone. Your browser will ask before it opens one, and the audio never leaves this device.

Nothing plays until you throw the switch. Everything below explains what the instrument does and how it chooses what to say, which is most of what is worth knowing about it.

Channel
———

Of 64 — synthesised

Room
———

dB on this room’s floor — measured

Voice
———

Words this session — synthesised

Floor
———

dBFS — measured

Receiver — pre-flight
AUDIO INTERFACENOT YET DETERMINED
SWEEP OSCILLATORSYNTHESISED — NO SAMPLES
VOICE STAGEFORMANT TRACT — NO SAMPLES
MICROPHONENOT YET DETERMINED
UPLOAD PATHNONE FITTED
SCREEN HOLDNOT YET DETERMINED

The master switch needs JavaScript.

The switch starts the audio stage and asks for your microphone. Nothing plays and nothing listens before you press it.

The receiver will not speak over a loud room: above about eleven decibels over the room\u2019s own floor the rate is exactly zero, not merely small. Make a noise, then be quiet.

Every sound is synthesised here — the sweep, the band, and every word. What the microphone measures decides when a word happens; it is never recorded and never sent anywhere.

Session transcript0 ENTRIES

No entries. The receiver is not online.

Every sound this instrument makes is synthesised on your device, and nothing your microphone hears is recorded or uploaded. This instrument is for entertainment and does not detect the paranormal — see the disclaimer and the method.

Entertainment only ·This instrument detects nothing paranormal. It reads an ordinary sensor and reports what it finds. The disclaimer says exactly what that means, and the method says what the number is made of. Nothing you record leaves this device.

TS/MTH/07Amended 19.02.2011

Why a broken radio sounds like speech

A spirit box is a radio that refuses to settle. It sweeps across the band faster than any station can hold, so instead of a broadcast you get fragments: a syllable from one station, a consonant from the next, a burst of noise between them.

Your ear will not accept that. Speech perception is aggressive — it is built to pull words out of a noisy room full of other people talking, and it does not switch off simply because there is nothing there to pull. Given fragments at roughly the rate of speech, it will assemble words. Given a question, it will assemble an answer.

The effect has a name. It is auditory pareidolia, it is extremely well documented, and it is stronger when you already know what you are hoping to hear. That last part is worth sitting with before you use this instrument.

Instrument 07 sweeps a bank of synthesised sources rather than live radio, because rebroadcasting live radio is not ours to do and, more to the point, because the effect has never depended on the source being real.

Everything you hear is made here

There is no audio file anywhere in this instrument. The static, the stations, the whistles and every word are generated by your own browser, in real time, out of oscillators and filters. That is a plain statement of fact and it is the first thing to know about the instrument.

The words are made with a model of a vocal tract. A buzzing source — a wavetable shaped like the pulse a pair of vocal folds actually produces — is fed through four resonant filters, and the frequencies of those filters are driven to the formants of each speech sound in sequence. That is not an approximation of how a voice works; it is how a voice works. A vowel is its first two or three formants, and moving them is what turns one sound into another.

We could have used the speech engine built into your phone. We did not, for one decisive reason: its output cannot be reached. No browser will route it into the audio graph, so it cannot be filtered, swept or put behind the receiver at all — it would arrive over the top of the instrument as a clean assistant voice from a different piece of software. A synthesised tract lives inside the receiver, comes through the same bandwidth limit as everything else, and sounds the same on every device.

It is also articulate rather than intelligible, and that is the point rather than a limitation. A word you have to reach for is a word you believe.

Entertainment only
TS/MTH/07aFiled 2026

How this instrument decides what to say

This is the part that separates Instrument 07 from a button that plays a spooky word, and it is the part we would most like you to check.

The content of every word is synthesised. When it happens is a measurement. The instrument listens to your room continuously and tracks its floor — the level the room sits at when nothing is happening in it. Everything afterwards is measured as a distance above that figure, because an absolute level in decibels means nothing when microphone sensitivity varies by twenty decibels between handsets.

Three causes, and the panel names which

  • Answered. Something happened in the room and then the room went quiet. The transient is measured as spectral flux against its own running average, so it works at any level, and the instrument will not answer until the room has settled again. This is the one you will find first: knock once, wait, and it answers within a second or two.
  • Unprompted, room settled. Nothing happened, and the room has been sitting on its own floor for a while. Slow — about one word every twenty-five seconds in a genuinely quiet room.
  • Unprompted, no acoustic reference. There is no microphone, so there is nothing to respond to. The instrument still works, and every word from that session is marked unprompted everywhere it appears.

The gate is a guarantee, not a rate

Above about eleven decibels over your room’s floor, the rate at which the instrument produces a word is exactly zero. Not small. Zero. Make continuous noise at it and it will not say a single word, however long you keep it up.

That is deliberate, and it is what makes the opposite behaviour mean anything. A rate that merely fell off in a loud room would be indistinguishable from luck. A hard gate is something you can prove to yourself in fifteen seconds, which is the only kind of claim this site is interested in making.

There is no random number generator in it

The word is chosen by a digest of what the microphone was measuring at the instant the utterance fired: level, floor, spectral centroid, flatness, all quantised. The same room in the same state tends to produce the same word twice, and the transcript prints the seed so you can see it happen.

The band is seeded the same way, from your room’s measured floor, which means a different room is a different band — and the same room is the same band when you come back to it a month later.

How clearly it speaks is also a measurement

The quieter your room is against its own floor, the more crisply the vocal tract articulates: the formant bandwidths narrow, the consonants come up, and the phones lengthen. Those are the real acoustic correlates of clear speech rather than a filter labelled “clarity”. Go quiet and the instrument becomes more articulate, and the panel shows you the figure while it happens.

TS/FLD/31Filed 12.06.1994 — amended 2026

Operating procedure

Let it establish the floor. The first second of a session is the instrument measuring what your room does when nothing is happening. Stay still and quiet for it. Everything the instrument reports afterwards is a comparison against that figure, and a floor established while you were moving about is worth very little.

Watch the band scan. Sixty-four channels, and the ones carrying something fill in behind the tuner. Note where they are. Crossing a strong channel changes what you hear, and an operator who has not noticed where the stations are will spend the evening reporting the same three.

Listen to the voice test. Four words, clearly articulated, before anything else happens. That is what this instrument sounds like when it is not guessing, and it is the reference against which every ambiguous word afterwards should be heard. If you skip it you will spend the session comparing words to nothing.

Then ask something, and be quiet. This is the whole method. The instrument answers into the gap after a sound, so a question followed by silence is the correct procedure and a question followed by a discussion of the question is not.

Mark what is worth marking. Nothing is filed automatically. A transcript of every word a machine said is not a record of anything — it is the machine talking about itself. Mark the two or three that made you look up, and those go to your case file with their numbers attached.

Find the sweep rate that works in this room. Operators argue about this more than anything else and they are right to. Standard suits most rooms. Fast is better in a room with a lot of reverberation; slow is better in a very dead one.

What to expect, honestly

Most sessions produce a handful of words and none of them land. Every so often one lands squarely, at a moment that makes it feel deliberate, and that moment is genuinely startling. That is the instrument working correctly — and it is worth knowing in advance that the startling ones and the meaningless ones are produced by exactly the same mechanism.

TS/PUB/07Current

Questions

Is a real voice coming through?

No. Every sound this instrument makes is synthesised by your own browser, including the words. There is no radio receiver, no broadcast, no recording and no connection to anything. What the instrument does with your microphone is measure the level of the room, which is a real measurement and is the thing that decides when it speaks.

How does it make the words if there is no recording?

With a model of a vocal tract. A buzzing source is fed through four resonant filters whose frequencies are set to the formants of each speech sound in turn, which is how a human voice actually works. That produces a voice that is articulate rather than clean, which is what a spirit box sounds like and also what makes the effect work — a word you have to reach for is one you believe.

Why does it answer when I knock?

Because it is measuring your room. The instrument watches for a transient — a knock, a door, a footstep, a voice — and then waits for the room to go quiet again. That gap is when it speaks. It is a real behaviour driven by a real measurement, and it is the fastest way to satisfy yourself that the responsiveness is not imagined.

Why does it go silent when I shout at it?

There is a working gate. When the room stays more than about eleven decibels above its own measured floor, the rate at which the instrument produces a word is exactly zero — not reduced, zero. Make continuous noise and it will not say a single word however long you keep it up. That guarantee is what makes the opposite behaviour meaningful.

Why do the same words keep coming up in the same room?

Because the choice is not random. The word is selected by a digest of what the microphone was measuring at the instant it fired — level, floor, spectral centroid, flatness — quantised, so a room in a steady state tends to produce the same word twice. There is no random number generator anywhere in the instrument.

What is the sweep rate for?

It sets how long the tuner sits on each of the sixty-four channels. Below about a tenth of a second the ear stops resolving the fragments and hears texture; above about four tenths it starts hearing them as separate sounds and the assembly stops happening. The useful range is narrow, which is why the control only offers three settings.

Where do the stations on the band come from?

They are laid out from your room's measured noise floor. A different room produces a different band, and the same room produces the same band again when you come back to it. On a device with no microphone there is no floor to seed from, so the standard layout is used and the panel says so.

Is my microphone being recorded?

No. The microphone goes into an analyser and the only things that come out are summary numbers — level, floor, spectral shape. Nothing is captured, nothing is stored and nothing is uploaded; there is no endpoint to upload it to. The EVP recorder is a separate instrument with its own separate permission, and it does not share this code.

Does it work without a microphone?

Yes, and it says so. The receiver sweeps and the voice stage still speaks, but there is nothing for it to respond to, so every word is marked unprompted — on the panel, in the transcript, and in your case file, where unprompted words count as service nowhere.

Can it tell me anything about the room I am in?

It can tell you what your room's acoustic floor is, which is a genuine measurement and is more interesting than it sounds — a room whose floor was minus fifty-five decibels in March and minus fifty-five in November is a room you now know something about. What the words mean is a different question and the Survey does not claim to have an answer to it.


Ghost Detector Lab is an entertainment site operated as the public terminal of the Thinplace Survey. Nothing here detects the paranormal. See the disclaimer, the privacy notice, and the method for every instrument on the bench.

Entertainment only