Posts By :

Michelangelo

Resonance ab

Once Upon a Time, There Was the Rega Planar: The Resonance Hidden Beneath the Music

1024 639 Michelangelo

A turntable does not operate in isolation. Tonearm mass, cartridge compliance, mounting hardware, furniture, floors and walls all participate in the result. Resonance Lab was created to make those relationships visible—before uncertain listening impressions become unnecessary purchases or endless adjustments.

Once upon a time, there was Rega.

Not Rega as an object of endless forum debate. Not Rega as a flag to be waved in the eternal arguments between belt drive and direct drive, low mass and high mass, suspended designs and rigid plinths.

Rega as an idea.

A simple, almost stubborn idea: remove what is unnecessary, make rigid what must remain still and avoid allowing unwanted energy to accumulate within the structure.

Perhaps that is why a Rega Planar 3 can still appear both modern and slightly old-fashioned. Modern because it is visually restrained, mechanically purposeful and free from excessive mass. Old-fashioned because it recalls a period of British hi-fi in which products often appeared to have been designed to solve practical problems rather than to resemble industrial monuments.

These were not domestic altars constructed from huge slabs of metal and acrylic. They were comparatively light, intelligent instruments intended to operate in real homes.

And this is where the story becomes interesting.

Lightness Is a Design Decision

The Rega Planar 3 is simple, but it is not simplistic.

Rega describes the Planar 3 as using a lightweight laminated plinth reinforced between the tonearm mounting and main bearing by its double-brace structure. The intention is to increase rigidity where it is required without turning the complete plinth into a large energy-storing mass.

This philosophy differs from the approach of a turntable that attempts to resist vibration primarily through weight.

A low-mass, high-rigidity design aims to minimise stored energy and reduce the duration of unwanted resonances within the structure. Rather than trying to become an immovable object, it attempts to manage energy quickly and predictably.

But every engineering philosophy creates conditions under which it performs best.

A lightweight turntable may respond differently to its support and surrounding structure than a very heavy, highly damped design. The equipment table, floor and wall are not automatically external to the turntable system. Under some conditions, they become part of it.

A Rega must therefore be given the right environment in which to behave like a Rega.

That does not always happen.

A Planar 3 on a Suspended Wooden Floor

The system in this case consisted of a modern Rega Planar 3 with its RB330 tonearm and an Audio-Technica AT-OC9XML moving-coil cartridge.

It was an interesting combination. The Microlinear stylus and boron cantilever of the AT-OC9XML offered excellent tracking potential, while the RB330 provided the rigid, low-friction platform around which the Planar 3 had been designed.

The turntable was positioned on a good-quality equipment table.

The table, however, stood on the suspended wooden floor of an English house.

The result was good, but not memorable.

It was controlled, detailed and pleasant, yet it did not quite deliver the immediacy, rhythm and physical presence often associated with a well-installed Rega.

The bass was present but did not feel completely secure. The soundstage opened, but it did not always seem to lock firmly into place. Voices were clear, yet the central image lacked some of the natural solidity that can transform competent reproduction into a convincing musical event.

None of this amounted to an obvious malfunction.

The stylus did not jump. There was no dramatic feedback, no clearly audible mechanical noise and no single defect that could be isolated immediately.

The system simply appeared to play with a small hesitation hidden beneath the music.

The Traditional Audiophile Response

The first instinct in situations like this is often to begin changing things.

We change the interconnect. We try another platter mat. We suspect the cartridge. We adjust tracking force, vertical tracking angle, azimuth and anti-skating. Eventually, we spend an afternoon staring at the tonearm as though it were about to confess.

Audiophiles know this ritual well.

When something does not sound quite right, the mind generates possible explanations faster than they can be tested. Some are technically reasonable. Others are elegant methods of converting uncertainty into maintenance.

Before changing anything, however, it is worth asking a simpler question:

What is the expected mechanical behaviour of the tonearm and cartridge combination?

The Tonearm and Cartridge Resonance

A cartridge suspension behaves like a spring, while the effective moving mass of the tonearm, cartridge and mounting hardware behaves like a mass attached to that spring.

Together they form a resonant mechanical system.

The commonly used estimate is:

fr = 159 / √(M × C)

where:

  • fr is the estimated resonance frequency in hertz;
  • M is the total moving mass in grams;
  • C is the cartridge’s dynamic compliance around the resonance region, expressed in µm/mN or an equivalent compliance unit.

For this system, the published and estimated inputs were:

  • RB330 effective mass: 11 g;
  • AT-OC9XML cartridge mass: 7.6 g;
  • mounting screws and washers: approximately 1–1.5 g.

This produces a total moving mass of approximately 19.6–20.1 g.

The Compliance Problem

The next figure is less straightforward.

Audio-Technica specifies the AT-OC9XML’s dynamic compliance as 16 × 10⁻⁶ cm/dyne at 100 Hz.

Tonearm and cartridge resonance, however, normally occurs much lower, generally somewhere around the single-digit or low-double-digit hertz region. A compliance value measured at 100 Hz cannot simply be inserted into the formula as though it described the suspension identically at 10 Hz.

Compliance is frequency-dependent, and manufacturers do not all publish it under the same test conditions.

Enthusiasts and designers therefore sometimes apply a practical conversion multiplier to Japanese 100 Hz specifications. Values between approximately 1.5 and 2 are commonly explored, but this is a heuristic—not a universal physical conversion law.

Using a factor of 1.7 gives an estimated 10 Hz compliance of:

16 × 1.7 = 27.2 µm/mN

Entering that estimate into the resonance formula gives:

159 / √(19.6 × 27.2) ≈ 6.9 Hz

With the slightly greater mass estimate of 20.1 g, the result becomes approximately:

159 / √(20.1 × 27.2) ≈ 6.8 Hz

The calculation therefore suggests a resonance around 6.8–6.9 Hz under that particular compliance assumption.

But the decimal places must not seduce us into believing that the estimate is more precise than the input data.

If different plausible conversion assumptions are explored, the predicted resonance can move approximately between 6.3 and 7.3 Hz. The real cartridge suspension may also differ from the nominal specification, and mounting conditions can affect the result.

The honest conclusion is therefore not:

“This combination resonates at exactly 6.86 Hz.”

It is:

“This combination is likely to operate near the lower end of the generally preferred resonance region and deserves closer attention.”

What a Low Estimate Actually Means

A predicted resonance below the centre of the preferred range does not automatically mean that the tonearm and cartridge are unusable together.

Ortofon currently describes approximately 7–12 Hz as an optimal region, with 10 Hz as a useful target. It also notes that values around 6.5–7 Hz may still be usable without problems.

This makes the distinction between a warning and a verdict extremely important.

A result near 6.8 or 6.9 Hz does not say:

“Remove this cartridge immediately.”

It says:

  • the combination may be more sensitive to record warps;
  • subsonic energy deserves attention;
  • the turntable support may become particularly important;
  • structural movement should not be dismissed;
  • and calculated behaviour should ideally be checked against observation or measurement.

The system was not necessarily wrong.

It was potentially delicate.

Where Resonance Lab Becomes Useful

This is the purpose of Resonance Lab.

It does not replace listening, and it does not convert analogue reproduction into a simple pass-or-fail calculation.

It provides a structured way to examine the relationship between tonearm effective mass, cartridge weight, mounting hardware and compliance.

Most importantly, it allows the user to see how assumptions change the result.

In a case such as the AT-OC9XML, the compliance conversion should not be hidden behind an apparently unquestionable number. It should be explored.

What happens if the effective compliance is closer to 24 µm/mN?

What happens if it is closer to 27 or 32?

What is the effect of an additional gram of mounting mass?

Would a lighter cartridge move the resonance significantly?

Does the result remain comfortably inside the desired region, or is it strongly dependent on uncertain inputs?

This is more informative than asking whether two products are merely “compatible.”

Compatibility is rarely binary.

A combination may be:

  • comfortably matched;
  • technically usable but sensitive to its environment;
  • dependent on uncertain compliance data;
  • or sufficiently extreme to justify reconsideration.

Resonance Lab makes that uncertainty visible.

What Resonance Lab Does Not Do

A calculation cannot measure a moving floor.

It cannot determine the actual structural resonance of an equipment rack, quantify footfall vibration or prove that a wall shelf will improve every system.

It also cannot know the exact low-frequency compliance of an individual cartridge unless that value has been measured under relevant conditions.

Resonance Lab therefore does not diagnose environmental vibration directly.

Its role is different: it helps identify whether the arm and cartridge combination makes environmental vibration a plausible and technically consistent part of the investigation.

In this case, it did not prove that the wooden floor was responsible.

It made the floor impossible to ignore.

The Floor Enters the System

Suspended wooden floors are elastic structures.

They move under footsteps and distribute low-frequency mechanical energy through joists, boards, furniture and equipment supports. The degree of movement depends on the building, span, construction, loading and position within the room.

This is not automatically a defect. It is simply the behaviour of the structure.

Anyone who has lived in an older English house knows that the building cannot be understood only by looking at it. It creaks, moves, breathes and responds.

An analogue turntable responds too.

When a tonearm and cartridge system is already operating near a relatively low resonance frequency, low-frequency structural movement may become more relevant. The interaction need not be dramatic enough to throw the stylus from the groove.

It may instead appear as:

  • less articulate bass;
  • a centre image that does not feel completely settled;
  • reduced rhythmic certainty;
  • slightly unstable spatial focus;
  • or a vague sense that the performance is not firmly grounded.

These descriptions are subjective listening observations, not unique diagnostic signatures. Similar impressions can have many causes.

But when the calculated arm and cartridge behaviour points towards greater low-frequency sensitivity, the support structure becomes a rational variable to test before purchasing new components.

The Load-Bearing Wall

The turntable was moved from the equipment table to a high-quality wall shelf fixed to a solid load-bearing wall.

The listening change was substantial.

The bass did not simply become more abundant. It became easier to follow. Low notes felt less hesitant and more clearly connected to the musical line.

Transient definition improved. The soundstage became more stable. Voices gained a firmer centre, and instruments occupied more credible positions.

Most importantly, the reproduction acquired a stronger sense of continuity.

The music no longer seemed to pass through a succession of small, invisible obstacles.

It flowed.

This observation does not establish a controlled scientific comparison, and it does not prove that every Rega turntable should be wall-mounted.

It does, however, align with Rega’s own approach. The company produces a lightweight, rigid wall bracket specifically for the Planar 1, Planar 2, Planar 3 and Planar 6, describing it as a vibration-isolation solution intended to complement its lightweight turntables.

The shelf did not add musical information.

Good mechanical engineering rarely adds magic.

It removes interference.

A Case Study, Not a Universal Rule

Not every wooden floor is unsuitable for a turntable.

Not every equipment table performs poorly.

Not every Rega must be placed on a wall shelf, and not every wall is structurally appropriate for supporting one.

A concrete floor and a stable rack may provide excellent conditions. A poorly installed wall shelf may create its own problems. A different tonearm and cartridge combination may be less sensitive to low-frequency movement.

The purpose of this case is not to produce another audiophile commandment.

It is to demonstrate a better sequence of reasoning:

  1. Describe the listening problem without immediately deciding its cause.
  2. Check the published mechanical specifications.
  3. Model the tonearm and cartridge resonance.
  4. Identify uncertainty in the compliance data.
  5. Explore plausible scenarios rather than trusting one exact number.
  6. Consider the support and building structure.
  7. Change one meaningful variable.
  8. Listen again—and measure where possible.

This is far more useful than changing three accessories simultaneously and then attempting to remember which one supposedly transformed the system.

Why I Created Resonance Lab

Tonearm and cartridge resonance has remained unnecessarily mysterious for too long.

The subject often sits somewhere between textbook equations, manufacturer specifications expressed under different conditions, enthusiast-produced compatibility charts and forum discussions in which every confident statement is followed by another confident statement claiming the opposite.

Resonance should not be an initiation ritual.

It should not be reserved for people who enjoy calculations more than music.

It should be a practical tool for better decisions.

I created Resonance Lab to make the relationship understandable and explorable.

The app does not tell the user what to hear. It helps organise the variables that may explain what they are hearing.

It can help reveal whether a combination is:

  • comfortably within a preferred region;
  • near a boundary;
  • highly dependent on an uncertain compliance conversion;
  • or potentially sensitive to environmental conditions.

It can also help prevent expensive misdiagnoses.

A new cartridge will not solve a moving floor if the replacement produces the same mechanical relationship. A different mat will not correct an unsuitable arm and cartridge match. A heavier mounting plate may move the resonance in the wrong direction.

Sometimes the most important upgrade is not another component.

It is a clearer understanding of the system already in front of us.

From Calculation to Reality

The calculated resonance frequency is a starting point, not the final truth.

Where possible, it should be complemented by real-world observation or measurement using a suitable test record and analysis method.

A measured result can reveal the actual resonance peak of the installed system, including the behaviour of the individual cartridge suspension rather than only its nominal specification.

Calculation and measurement serve different purposes:

  • calculation helps evaluate combinations before purchase and explore alternatives;
  • measurement reveals how the installed system actually behaves;
  • listening tells us whether that behaviour is musically significant in the complete system.

None should be forced to perform the role of the others.

The strongest diagnosis emerges when all three point in the same direction.

Understanding Rega Without Worshipping It

Understanding this case does not require worshipping Rega as a brand.

It requires understanding why its design choices make sense.

A Rega is not “simple” in the impoverished sense of the word. It follows a specific path based on lightness, rigidity and controlled energy behaviour.

But a specific design philosophy also requires a suitable context.

Place the turntable on a structure that moves, and under certain conditions it may tell you.

Give it a stable mechanical reference, and it may stop defending itself and begin to communicate the music more freely.

Perhaps this is one reason Rega has retained such a distinctive identity. When the system is working well, the turntable does not seem to be trying to impress the listener.

It simply allows the performance to pass through.

The Relationship Beneath the Music

A turntable is not a collection of isolated products.

The tonearm, cartridge, compliance, mounting screws, support, floor, walls and furniture all participate in its mechanical behaviour.

Sometimes their influence is obvious.

At other times, it does not add a recognisable defect. It removes certainty.

That is the deeper purpose of Resonance Lab.

It does not replace the ears, and it does not promise to solve every analogue problem through a formula.

It makes relationships visible.

In this case, the app showed that the RB330 and AT-OC9XML combination was not absurd, but potentially sensitive. That made the suspended floor a credible variable rather than a piece of inherited audiophile folklore.

The wall shelf then became more than an accessory recommended by tradition.

It became a mechanical response to a specific hypothesis.

Once upon a time, there was Rega.

But Rega is still here.

To hear it properly, we may simply need to understand where it came from, what it is trying to achieve and the environment in which we are asking it to perform.

Because in analogue reproduction, sound never comes from one component alone.

It comes from a relationship.

And sometimes, to rediscover the music, we do not need to replace the turntable.

We simply need to remove the floor from the conversation.

Explore Resonance Lab

Resonance Lab helps vinyl enthusiasts explore tonearm and cartridge compatibility, estimate system resonance and compare how changes in mass or compliance may influence the result.


Learn more about Resonance Lab


Download Resonance Lab from the App Store

References and Technical Sources

  1. Rega Research. “Planar 3.”

    View official product information
  2. Rega Research. “RB330 Tonearm.”

    View official specifications
  3. Audio-Technica. “AT-OC9XML Dual Moving-Coil Stereo Cartridge.”

    View official specifications
  4. Ortofon. “Matching Cartridges with Tonearms.”

    View resonance formula and guidance
  5. Rega Research. “Turntable Wall Bracket.”

    View official product information

This article describes a real-world setup and subjective listening observations supported by resonance modelling. Calculated resonance values are estimates and depend on the accuracy and measurement frequency of the compliance data. They should not be interpreted as a substitute for direct measurement of the installed system.

Where High Fidelity Really Begins: Four Decisions That Shape a Believable Stereo Recording

1024 582 Michelangelo

High fidelity begins before mastering, before the distribution format and before the physical or digital medium. It begins with the decisions made when the performance is captured: how stereo space is encoded, how phase relationships are managed, how far the microphones are placed from the musicians and which microphones are chosen to translate that perspective.

Listening to a stereo recording involves much more than hearing sound emerge from two loudspeakers.

A voice apparently positioned in the centre does not come from a hidden central speaker. An instrument heard slightly to the left is not physically present at that point in the listening room. The apparent depth extending beyond the wall behind the system is not an acoustic space that has suddenly opened in front of us.

These are perceptual constructions.

The auditory system interprets differences in level, timing, spectrum and phase, together with reflections and the relationship between direct and reverberant sound. From this information, it creates a plausible auditory scene.

Stereo recording is therefore not merely the process of producing two channels. It is the art of creating relationships that the listener’s brain can interpret as space.

A High-End System Cannot Invent the Recording

A recording may be highly detailed, dynamically impressive and extended at both ends of the frequency spectrum while still failing to sound believable.

It may reveal the smallest movements of the musicians but construct no stable image. It may sound extremely wide but imprecise. It may attract attention immediately and gradually become artificial or tiring.

For listeners using loudspeakers, this distinction is fundamental.

A revealing system can expose the spatial quality of a recording, but it cannot manufacture coherence that was never captured. If the central image is unstable, the playback equipment may reveal that instability with greater clarity. If the relationship between instruments and ambience is implausible, additional resolution may make the contradiction more obvious.

The format and reproduction chain matter enormously, but they inherit the decisions made at the beginning.

High fidelity starts at the microphone.

Stereo Is Not Simply Width

In audiophile language, we often speak about soundstage, imaging, depth, focus and air around instruments.

These are useful descriptions, but a believable soundstage is not simply a wide one.

Width can be created in many ways:

  • spaced microphones;
  • pan controls;
  • interchannel delays;
  • Mid-Side processing;
  • stereo reverberation;
  • decorrelation;
  • and dedicated spatial processors.

All of these can be artistically valid. But increasing width does not automatically increase realism.

A believable voice should remain stable in the centre, not only in lateral position but also in physical presence. A piano should have plausible dimensions. A guitar should not become wider than the performer playing it unless that enlargement is an intentional artistic decision. A flute should not be reduced to breath, lips and key noise while losing the integrated sound of the instrument.

The acoustic environment should not feel as though it has been placed behind the musicians as an independent effect. It should appear to belong to the same event.

When a recording is coherent, the listener perceives more than individual sources distributed between two loudspeakers. The listener perceives relationships between the instruments, their apparent distances and the room surrounding them.

That is what makes a soundstage credible.

The Four Decisions Behind a Believable Recording

Four closely connected decisions have a particularly strong influence on spatial realism:

  1. the stereo microphone technique;
  2. the timing and phase relationships between channels;
  3. the distance between microphones and performers;
  4. and the microphone characteristics used at that distance.

None of these choices operates independently.

Changing microphone distance changes the balance between direct sound and room ambience. It also changes proximity effect, source integration, off-axis contribution and the amount of environmental noise captured.

Changing the microphone changes the polar pattern, tonal balance, transient behaviour, self-noise and off-axis response experienced at that distance.

Changing the stereo array changes the interchannel timing and level relationships supplied to the reproduction system.

The art lies in making these decisions support the same perceptual objective.

1. Stereo Technique: Different Ways of Encoding Space

Stereo microphone techniques are not simply alternative arrangements for producing a left and right channel. They encode spatial information in fundamentally different ways.

Spaced Pairs

In an AB arrangement, two microphones are separated physically. A sound arriving from one side normally reaches one microphone before the other, producing an interchannel time difference. Depending on the source, microphone pattern and geometry, differences in level may also occur.

Spaced arrays can create a broad impression of scale and envelopment. In a good concert hall, and particularly with larger ensembles, this can communicate openness, low-frequency spaciousness and the feeling that the performance breathes within a large acoustic environment.

Physical spacing also introduces frequency-dependent phase relationships between the channels. In stereo these can contribute to spaciousness, but they may reduce localisation precision or create tonal changes when the channels are summed or partially combined.

This does not make AB inherently defective. It means that width, envelopment, localisation and mono compatibility must be balanced deliberately.

Near-Coincident Techniques

Near-coincident arrangements such as ORTF, NOS and DIN combine physical microphone spacing with directional polar patterns.

ORTF, for example, uses two cardioid microphones separated by 17 centimetres and angled 110 degrees apart. The resulting stereo image contains both interchannel timing and level differences.

This can offer a productive compromise: greater spaciousness than many coincident cardioid arrangements, together with more definite image positioning than a widely spaced pair.

The timing component is not an accidental defect. It is part of the intended spatial design.

Coincident Techniques

In coincident arrangements such as XY, Mid-Side and Blumlein, the microphone capsules are placed as close as physically possible to the same acoustic point.

Because the direct sound reaches both capsules at approximately the same time, directional information is created primarily through differences in level and polarity produced by the microphones’ polar patterns.

Coincident geometry generally offers:

  • stable localisation;
  • a clearly defined central image;
  • predictable mono compatibility;
  • and fewer time-delay interactions introduced by microphone spacing.

The soundstage can sometimes appear less expansive than that produced by a spaced array, but individual positions may be easier to read.

Practical coincidence is never perfect. The capsules have physical dimensions, and real microphones exhibit frequency-dependent polar and phase responses. Careful construction and positioning still matter.

No Technique Is Universally Superior

The appropriate technique depends on:

  • the size and arrangement of the ensemble;
  • the acoustic character of the venue;
  • the required balance between localisation and spaciousness;
  • the importance of mono compatibility;
  • the intended listening perspective;
  • and the expected playback system.

When scale and strong hall envelopment are priorities, a spaced arrangement may be highly effective. When image stability and coincident timing are central to the project, XY, Mid-Side or Blumlein may offer important advantages.

There is no universally correct technique.

There is only a technique whose characteristics are coherent with the recording’s purpose.

2. Phase: More Than a Technical Problem

In audio, phase is frequently discussed only when something has gone wrong.

We notice it when bass becomes thin, when a centre image loses solidity, when combining microphones produces tonal colouration or when a stereo recording behaves poorly in mono.

These are genuine problems, but phase is also part of spatial information.

Timing and phase relationships between channels can contribute to the perception of width, localisation, ambience and depth. However, the differences captured by two microphones are not identical to the binaural cues produced at two human ears.

Microphones have no head between them, no pinnae and no torso. They do not apply the listener-specific spectral transformations associated with natural localisation. Their spacing and polar patterns create a new encoding intended for reproduction through another system.

This distinction is crucial.

In conventional loudspeaker stereo, each ear hears both speakers. The listening room adds reflections, and the listener’s head modifies the signals again. The brain must interpret this combined information and construct phantom images.

A large interchannel delay can increase spaciousness without necessarily improving image precision. A coincident array reduces the timing differences introduced at capture, but it cannot eliminate every phase interaction in the room, microphone, loudspeaker or recording chain.

The meaningful objective is not perfect phase identity.

It is maintaining interchannel relationships that remain sufficiently consistent for the listener to construct a stable scene.

Phase and the Centre Image

A central phantom image is created when the loudspeakers provide the auditory system with compatible information suggesting that a source lies between them.

Equal level alone is not always sufficient to make that centre feel physical. The spectral and temporal content of the two channels must also support the same perceptual interpretation.

If multiple microphones capture one source with different delays, the resulting interference may change with frequency. The image can become less stable, and tonal character may vary according to the combination of channels and listening position.

This is one reason why microphone count should not be confused with information quality.

More microphones can offer flexibility, control and creative possibilities. They can also create more relationships that must be managed.

3. Microphone Distance: Where Proportion Begins

Microphone distance is fundamental to whether an instrument sounds credible through loudspeakers.

A close position can provide immediacy, clarity and extraordinary detail. It can reveal the movement of piano mechanics, a flautist’s breathing, fingers touching strings, valve noise, bow texture and the precise attack of every note.

These details can be fascinating and musically valuable.

But they are not always proportionate to the way the instrument would be heard from a natural listening position.

The danger is that detail becomes confused with realism.

The Instrument Must Reassemble

Acoustic instruments often radiate different frequency regions from different parts of their bodies and in different directions.

At very close range, a microphone hears one local part of that radiation field. Moving farther away allows those contributions to integrate more fully before reaching the capsule.

The piano can become one sounding body rather than a collection of strings, hammers and mechanical events. The guitar returns to plausible physical dimensions. The flute becomes more than the excitation point at the mouthpiece; it becomes an instrument projecting energy into the room.

Distance allows the instrument to reassemble itself.

Distance Introduces the Room

Moving a microphone away also changes the balance between direct and reflected sound.

More of the room enters the recording. Early reflections influence tone and localisation. Reverberation communicates scale and distance. Background noise and undesirable acoustic characteristics become harder to avoid.

The correct distance is therefore not simply the most natural one in theory. It is the distance at which source integration, clarity, instrumental body and room contribution reach the desired balance.

That point changes with every venue, ensemble and microphone.

Close Is Not Wrong

Close microphone placement should not be treated as inherently artificial.

Many musical genres depend on intimacy, isolation, impact or the ability to balance sources independently. A close perspective may be exactly right for the artistic language of the production.

The problem arises only when a local, magnified perspective is presented as though it were automatically more faithful because it reveals more detail.

Detail describes how much can be perceived.

Realism describes whether those details belong to a plausible whole.

4. Microphone Choice: An Instrument of Proportion

At a realistic recording distance, the microphone is not merely a transparent transducer.

It becomes an instrument of proportion.

Every microphone has a technical personality shaped by factors including:

  • its operating principle;
  • diaphragm or ribbon construction;
  • polar pattern;
  • frequency and phase response;
  • off-axis behaviour;
  • transient response;
  • self-noise and sensitivity;
  • grille and body geometry;
  • electronics and transformers;
  • and manufacturing tolerances.

No microphone is perfectly neutral under every condition.

A microphone that sounds impressive at close range may not be the right choice at several metres. A presence rise that creates attractive clarity nearby may make a distant recording feel thin or overly explicit. A microphone with excellent on-axis response but irregular off-axis behaviour may colour the room contribution as distance increases.

Conversely, a microphone whose polar pattern and tonal balance remain well controlled away from the axis may integrate the direct and reverberant fields more convincingly.

Ribbon and Condenser Microphones

The choice should not be reduced to a contest between ribbon and condenser technology.

Modern condenser microphones can provide extended bandwidth, low self-noise, high sensitivity and carefully controlled directional behaviour. These qualities can be invaluable for distant acoustic recording.

Some ribbon microphones offer a different balance: smooth high-frequency behaviour, figure-of-eight directivity and a substantial sense of instrumental body. These qualities can suit coincident Blumlein recording and certain natural-distance applications particularly well.

But the result depends on the individual microphone, not merely the category printed on its specification sheet.

A ribbon is not automatically warm or natural. A condenser is not automatically bright or analytical.

The meaningful question is:

Does this microphone, at this distance and in this room, preserve the proportions required by the music?

Proximity Effect and Tonal Perspective

Directional pressure-gradient microphones exhibit proximity effect: their low-frequency response increases as the source moves closer.

The effect is generally strongest with figure-of-eight patterns and is also present, to a lesser degree, with cardioid and related directional patterns. Pure pressure-operated omnidirectional microphones do not exhibit conventional proximity effect.

At very close distances, proximity effect may exaggerate bass and make a source appear larger than its natural scale. It can also be used creatively to provide weight, intimacy or authority.

As the microphone moves farther from the source, this low-frequency boost diminishes. The resulting change in tonal balance must be considered alongside the growing contribution of the room.

This is one reason why microphone choice and distance cannot be separated. The correct working distance is not established by geometry alone; it must also produce an appropriate tonal foundation.

The objective is not to enlarge the bass artificially. It is to preserve enough body for the instrument to remain physical without making it implausibly large.

Why Blumlein Deserves Particular Attention

Among coincident techniques, the Blumlein pair occupies a distinctive position.

It uses two figure-of-eight microphones mounted coincidently and angled 90 degrees apart. Direction is encoded primarily through level and polarity differences, while the rear lobes capture a substantial part of the surrounding acoustic environment.

This combination can provide:

  • a stable central image;
  • clearly organised lateral localisation;
  • strong mono compatibility;
  • and a naturally integrated representation of the room.

Blumlein is also demanding.

The room must contribute positively because sound arriving from behind the array is captured strongly. Placement is critical. The ensemble must balance acoustically, and the useful recording angle must suit its arrangement.

A poor room is revealed rather than concealed. An incorrect microphone position can produce too much reverberation, an inappropriate stereo spread or an imbalanced ensemble.

That apparent limitation is also part of the technique’s value.

A minimally manipulative method encourages the main problems to be solved before recording begins rather than postponed until post-production.

The Room Is Not an Effect Added Afterwards

At natural microphone distances, the room inevitably becomes part of the recording.

Not every space deserves that responsibility.

A room may be excessively dry, too small, mechanically noisy, confused in the lower frequencies or dominated by unattractive early reflections. When the room is unsuitable, a purist approach does not transform it into a virtue.

But when the acoustic environment supports the music, natural ambience can provide an unusually coherent relationship between source and space.

The early reflections, reverberant build-up, asymmetries, frequency-dependent decay and interaction with instrumental radiation all belong to one event.

They are not a separate layer added later.

They are part of the way the music happened.

Artificial reverberation can be beautiful, realistic and artistically indispensable. The distinction is not between legitimate natural sound and illegitimate processing.

The distinction is whether the reverberant information supports the same perspective and spatial logic as the direct sound.

When natural reverberation is captured successfully, the room does not sit behind the instruments.

It connects them.

What the Listener Can Evaluate

These ideas are relevant not only to recording engineers. They can also change how an audiophile evaluates recordings and equipment.

Instead of asking only how much detail is audible, how deep the bass extends or how wide the soundstage appears, listen for relationships.

Listen to the Centre

Does a central voice feel stable and physical, or merely like a thin point suspended between the speakers?

Does its position and body remain coherent as pitch, intensity and register change?

Listen to Instrumental Proportion

Do the instruments possess believable dimensions?

Does the piano appear as one body, or as a series of enlarged local details? Does a guitar occupy plausible space? Does a flute retain tone and projection rather than becoming mostly breath and mechanism?

Listen to Complexity

Does the soundstage remain intelligible when the musical texture becomes dense?

A recording may appear sharply separated during a simple passage and lose all spatial organisation when multiple instruments play simultaneously.

Listen to Decay

Do notes decay continuously into the same environment in which they began?

Does a piano chord dissolve naturally into the surrounding space? Does the room respond to the music, or does the reverberation seem to operate as a separate layer?

Listen at Moderate Volume

A spatially coherent recording often remains intelligible without being played loudly. The centre continues to exist, instrumental positions remain readable and the acoustic environment still suggests depth.

Exaggerated spectral balance and artificial spatial effects may depend more strongly on level to remain impressive.

Listen for Credibility, Not Size

A realistic instrument does not need to sound enormous.

It needs to sound plausible.

Many recordings impress by enlarging everything. Enlargement can be artistically exciting, but it is not automatically high fidelity.

The Format Preserves; It Does Not Create

Audiophile discussions often concentrate on formats: vinyl, analogue tape, PCM, DSD and high-resolution distribution.

These discussions matter, but the format is sometimes treated as though it were the original source of realism.

A format can preserve captured information with greater or lesser accuracy. It cannot create spatial information that was never recorded.

If the soundstage is unstable, the medium preserves that instability. If the centre is weak, a particular format may alter the subjective presentation but cannot reconstruct the original microphone geometry. If the ambience bears no coherent relationship to the instruments, greater resolution may simply reveal that separation more clearly.

The format matters—but it comes later.

This does not diminish the value of excellent recording and distribution formats. It clarifies their purpose.

A high-quality medium is most meaningful when it is preserving something worth preserving:

  • musical dynamics;
  • instrumental timbre;
  • temporal relationships;
  • spatial information;
  • natural decay;
  • and believable proportion.

High Fidelity as Coherence

Stereo technique determines how spatial cues are encoded.

Phase and timing relationships influence image stability, width and the behaviour of combined signals.

Microphone distance determines the proportion between source, detail and environment.

Microphone choice determines how that distance is translated tonally and spatially.

The room either contributes meaningfully to the performance or becomes another problem to manage.

The recording format preserves these decisions but cannot replace them.

For Direct Sound Records, this is the foundation of natural acoustic recording. The objective is not to impose a spectacular soundstage but to preserve enough coherent information for the listener to reconstruct one.

A great recording does not necessarily astonish immediately.

It often becomes more convincing over time.

It does not enlarge every instrument. It preserves proportion. It does not use reverberation merely as decoration. It allows the acoustic environment to participate in the music.

Two loudspeakers can accomplish something extraordinary: they can suggest the presence of a space that does not physically exist in front of the listener.

For that illusion to succeed, the recording must contain believable information. It must respect the way we hear and provide the brain not only with individual sounds, but with meaningful relationships between them.

Perhaps this is where a recording becomes genuinely high fidelity—not when it reveals everything, but when it allows us to believe what we are hearing.

Related Articles

  • The Space Between Sounds: How the Brain Reconstructs the Soundstage
  • Why the Blumlein Pair Can Excel in Loudspeaker Playback

References and Further Reading

  1. Eargle, J. M. “An Overview of Stereo Recording Techniques for Popular Music.”
    Journal of the Audio Engineering Society, 1986.
    View AES record
  2. Gerzon, M. A. “The Design of Precisely Coincident Microphone Arrays for Stereo and Surround Sound.”
    Audio Engineering Society 50th Convention, 1975.
    View AES record
  3. Politis, A., Laitinen, M.-V., Ahonen, J. and Pulkki, V.
    “Parametric Spatial Audio Processing of Spaced Microphone Array Recordings for Multichannel Reproduction.”
    Journal of the Audio Engineering Society, 2015.
    View AES record
  4. DPA Microphones. “Stereo Recording Techniques and Setups.”
    Read technical guide
  5. DPA Microphones. “ORTF.”
    View technical definition
  6. Neumann. “What Is the Proximity Effect?”
    Read technical guide
  7. Bock, T. M. and Keele, D. B.
    “The Effects of Interaural Crosstalk on Stereo Reproduction and Minimizing Interaural Crosstalk in Nearfield Monitoring.”
    Audio Engineering Society 81st Convention, 1986.
    View AES record
  8. Schneider, M. “MS Mastering of Stereo Microphone Signals.”
    Audio Engineering Society 132nd Convention, 2012.
    View AES record

An earlier version of this article was published in Audio Review, July/August 2026, and was subsequently adapted for LinkedIn. This Direct Sound Records Journal edition has been substantially revised, expanded and technically updated.

Space and Sound

The Space Between Sounds: How the Brain Reconstructs the Soundstage

1024 755 Michelangelo

In musical realism, we do not listen only to frequencies, timbres and dynamics. We listen to relationships: differences in arrival time, level, reflection and spatial organisation. When those relationships remain credible, a recording can become more than a collection of sounds. It can become an inhabitable acoustic space.

When listening to a high-fidelity system, we often say that a recording “sounds real.” But what does that actually mean?

Frequency extension matters. So do controlled bass, natural midrange, transient response, low distortion and harmonic detail. Yet realism also depends on something less obvious and considerably more fragile: whether the recording and playback system allow us to perceive a believable relationship between sound sources and the space around them.

A voice is not simply a voice. It is a presence located at a particular distance and position.

A piano is not merely a collection of strings, hammers and resonances. It is a physical body occupying a volume of air, projecting energy into a room and interacting with surrounding surfaces.

A quartet is not simply divided into left, centre and right. It is an acoustic event in which musicians, distances, reflections and silence form one spatial relationship.

The human auditory system does not receive this scene passively. It interprets, compares and reconstructs it. In a sense, the brain triangulates.

The Soundstage Is Not an Effect

In high-end audio, words such as soundstage, depth, focus, air and presence are used constantly. They are useful descriptions, but they can become vague unless we remember that they have a perceptual foundation.

The soundstage is not a picture physically stored inside a recording. Nor is it simply drawn between two loudspeakers.

It is a perceptual reconstruction created from the information reaching the listener’s ears.

The brain interprets several cues simultaneously:

  • tiny differences in the time at which sound reaches each ear;
  • differences in level between the ears;
  • direction-dependent changes in the spectrum;
  • the relationship between direct and reflected sound;
  • the evolution of reflections and reverberation over time;
  • and prior knowledge of familiar voices, instruments and environments.

When these cues support one another, a sound source can appear stable, physical and separate from the loudspeakers.

When they conflict, a recording may still sound impressive, detailed or extremely wide, but the scene can feel unstable or artificial.

The soundstage is therefore not merely an effect.

It is information interpreted by the listener.

How the Brain Reads Acoustic Space

For horizontal localisation, the auditory system relies heavily on differences in timing and level between the two ears.

Interaural Time Differences

When a sound arrives from one side, it normally reaches the nearer ear slightly before the farther ear. These interaural time differences, or ITDs, can be extremely small—sometimes measured in tens of microseconds—yet the auditory system is remarkably sensitive to them.

Timing cues are particularly important at lower frequencies, where neural activity can represent the temporal fine structure of the waveform with sufficient precision. Structures within the auditory brainstem, especially the medial superior olive, contribute to the analysis of these binaural timing relationships.

ITD sensitivity does not end at one perfectly defined frequency, but sensitivity to the fine structure of pure tones deteriorates substantially through the region around 1 to 1.5 kHz. Higher-frequency sounds can still convey timing information through changes in their amplitude envelopes.

Interaural Level Differences

At shorter wavelengths, the head creates a more substantial acoustic shadow. A sound arriving from the left will generally produce a higher level at the left ear than at the right.

This interaural level difference, or ILD, becomes an increasingly useful directional cue as frequency rises. Neural circuits involving the lateral superior olive contribute to its early processing.

In natural listening, timing and level differences do not operate as two isolated systems with a rigid boundary between them. Broadband sounds contain multiple cues, and the brain combines them according to frequency, source position, environment and reliability.

Spectral Cues and the Shape of the Listener

Timing and level differences provide strong information about left and right, but they cannot always distinguish whether a sound is above, below, in front of or behind the listener.

For those dimensions, the shape of the outer ears, head and torso becomes essential.

The folds of the pinnae filter incoming sound differently according to direction. Certain frequencies are reinforced, while others are attenuated or notched. The complete transformation between a sound source and the listener’s ears is described by the head-related transfer function, or HRTF.

Because human anatomy varies, each person’s HRTF is individual. Over time, the brain learns the correspondence between these spectral patterns and positions in space.

Experiments in which the outer ears were temporarily reshaped have shown that vertical localisation initially becomes much less accurate. With experience, listeners can learn to interpret the altered cues—evidence that spatial hearing is not merely mechanical, but calibrated and adaptive.

What we perceive is therefore never only the sound emitted by the source. It is the result of a relationship between source, environment, listener and brain.

Where Phase Enters the Picture

In audio engineering, phase is often discussed as a problem.

Signals may be described as “out of phase” when they weaken the centre image, reduce low-frequency energy, produce cancellations or behave unpredictably in mono.

These concerns are real, but they represent only part of the subject.

Phase is not simply an error waiting to be corrected. Phase and timing relationships can also carry important information about position, width, depth and the interaction between sound sources.

For a periodic low-frequency signal, a difference in arrival time between the ears can also be expressed as an interaural phase difference. In this sense, phase contributes directly to spatial localisation.

Within a stereo recording, interchannel timing and phase relationships can influence:

  • the position and stability of phantom images;
  • the apparent width of the presentation;
  • the sense of distance and depth;
  • the relationship between direct sound and ambience;
  • mono compatibility;
  • and frequency-dependent reinforcement or cancellation when signals combine.

However, it would be misleading to say that phase is used only for location and never contributes to the perceived character of a sound. Temporal fine structure is also involved in pitch, masking and the separation of simultaneous sources. In recording and reproduction, phase relationships can additionally alter the spectrum whenever correlated signals combine acoustically or electrically.

The useful distinction is therefore not between phase and sound, but between the different roles that temporal relationships perform.

Phase is not the sound itself, but it can help organise the space in which that sound is perceived.

Coherence Does Not Mean Perfection

The expression phase coherence is frequently used as though it described one measurable quality that a recording either possesses or lacks.

Reality is more complicated.

Every acoustic environment contains delays. Reflections arrive after the direct sound and from different directions. Instruments radiate differently according to frequency. Microphones have frequency-dependent polar patterns. Loudspeakers and rooms introduce further interactions.

A natural acoustic event is not phase-identical at every point in space.

Spatial coherence should therefore not mean eliminating every difference or delay. It means preserving relationships that remain compatible enough for the auditory system to interpret them as belonging to one plausible event.

When this happens, the voice can stabilise at the centre without appearing glued to either speaker. Instruments occupy a readable volume rather than appearing as thin lateral points. The room does not feel like reverberation placed behind the music; it surrounds and continues the performance.

The silence between instruments stops being empty.

It becomes air.

Why Some Recordings Sound Large but Not Real

A wide soundstage is not necessarily a realistic soundstage.

Modern production provides an enormous range of tools for creating size: multiple microphones, pan controls, delay, artificial reverberation, stereo widening, decorrelation and Mid-Side processing.

These tools can be artistically valuable. They can also create a guitar broader than its physical source, a voice floating beyond the loudspeakers or a reverberant field that could never have existed around the original performers.

There is nothing inherently wrong with that. Recording is also an art of construction.

But width and credibility are not synonymous.

The auditory system evaluates more than the apparent size of the scene. It also evaluates whether the timing, spectral, directional and environmental information is mutually plausible.

A very wide image may be initially impressive. Yet if the centre lacks stability, the reverberation does not belong to the sources or the spatial relationships change unnaturally with frequency, the illusion becomes less convincing.

It is similar to viewing a photograph with intense colour and extraordinary sharpness but incorrect perspective. The image attracts attention immediately, yet something feels wrong.

A recording can contain remarkable detail and separation while remaining spatially two-dimensional.

Beautiful, perhaps.

But not alive.

Every Microphone Technique Makes a Decision

For acoustic music, the choice of microphone technique is never neutral.

AB, ORTF, XY, Mid-Side and Blumlein do not simply create different varieties of stereo width. They encode different combinations of timing, level, polarity, direction and room information.

A spaced AB pair introduces meaningful arrival-time differences between microphones and can create scale, openness and envelopment.

ORTF combines a moderate physical separation with directional cardioid microphones, creating both interchannel time and level differences.

Coincident systems such as XY, Mid-Side and Blumlein minimise the timing difference introduced by microphone spacing and derive direction primarily through level and polarity relationships.

These differences affect localisation, spaciousness, mono compatibility and the way the recording interacts with loudspeaker reproduction.

Coincident techniques can produce comparatively stable and clearly located virtual sources. Spaced techniques may create broader or more diffuse images and can convey strong spaciousness. Neither outcome is automatically better.

The appropriate technique depends on:

  • the musicians and their physical arrangement;
  • the acoustic character of the venue;
  • the desired listening perspective;
  • the balance between localisation and envelopment;
  • the intended distribution format;
  • and the expected playback environment.

There is no microphone technique that is universally correct.

There is only a technique whose compromises are more or less coherent with the intended result.

Recording Is a Translation

A microphone does not hear like a human being.

It has no head, no outer ears, no perceptual memory and no awareness of the room. It does not compare what it captures with years of experience. It measures sound pressure or pressure gradient according to its physical construction and polar pattern.

Human perception, by contrast, is active.

This means that stereo recording is always a translation.

The task is not simply to place two microphones in front of a performance and assume that the original space has been preserved. The task is to create two signals that, when reproduced through loudspeakers in another room, provide the listener with enough coherent information to reconstruct a plausible scene.

That distinction is fundamental.

The original venue, the microphone array, the recording chain, the loudspeakers, the listening room and the listener are all parts of one perceptual system.

A recording can never transport the original acoustic field intact. It selects, encodes and later stimulates a new reconstruction.

The Listening Room Is Part of the Reproduction

With headphones, each channel is delivered predominantly to one ear. With conventional loudspeakers, both speakers reach both ears.

The left ear receives sound from the left loudspeaker, sound from the right loudspeaker after a different path, and reflections from the listening room. The right ear receives the corresponding combination from the opposite side.

The listener’s brain must interpret this new set of binaural cues and construct the phantom images associated with stereo reproduction.

This is why loudspeaker placement, room acoustics and listening position cannot be separated from the recording itself. They participate in the decoding of its spatial information.

A stable recording cannot correct a fundamentally unsuitable room, and a carefully treated room cannot restore information that was never captured or was destroyed during production.

The recording and reproduction environments form a chain.

Natural Reverberation Is More Than a Tail

In a real acoustic environment, reverberation is not an effect that begins after the direct sound has finished.

It is a continuously evolving field of reflections shaped by the dimensions, materials and geometry of the venue.

Those reflections contain:

  • directional asymmetries;
  • different arrival times;
  • frequency-dependent decay;
  • changes in density over time;
  • and relationships to the position and radiation pattern of every instrument.

A reverberation processor can create extraordinarily convincing spaces, and artificial reverberation is indispensable in many forms of production. But a preset does not reproduce the exact interaction that occurred between particular musicians and a particular room at one unrepeatable moment.

When natural reverberation is captured successfully, it does not feel attached to the performance.

It is the performance continuing into the building.

Natural reverberation is architecture becoming sound.

Recording for the Brain

A realistic recording is not necessarily one that captures the largest possible quantity of information.

It is one that preserves the relationships necessary for perception.

For Direct Sound Records, microphone placement, distance, acoustic environment and minimal signal manipulation are therefore not separate technical choices. They form one recording philosophy.

The objective is not purity for its own sake, and it is not nostalgia for a period before digital production.

It is the preservation of continuity.

When a voice is captured from a believable perspective, its centre depends on more than identical level in the two channels. It also depends on stable spectral, temporal and environmental relationships.

When two instruments occupy different sides of an ensemble, their apparent positions should not be understood merely as pan-control settings. They arise from distance, angle, microphone pattern, radiation, reflections and their relationship with the room.

The engineer must therefore decide what should remain coherent, what can be altered and what must be allowed to exist naturally.

Listening Beyond Detail

This perspective also changes how we evaluate recordings and audio systems.

Instead of asking only, “How much detail can I hear?”, we can ask:

  • How credible is the relationship between the sounds?
  • Does the voice remain stable as its pitch and intensity change?
  • Do instruments possess physical body, or are they merely lateral outlines?
  • Does the acoustic environment belong to the performance?
  • Does depth arise from perspective, or only from added reverberation?
  • Does the central image remain convincing at modest listening levels?
  • Do reflections create continuity and air, or do they blur localisation?
  • Does the recording invite prolonged listening, or impress only for a few moments?

These questions move the discussion beyond spectacular sound.

A recording that preserves spatial relationships does not merely demonstrate the capabilities of a system. It gives the system an opportunity to disappear.

When that happens, we stop concentrating on two loudspeakers producing sound.

We begin to perceive musicians, sounding bodies, distance, architecture and silence.

The Space Between Sounds

The most convincing recordings are not necessarily those that contain the most obvious effects, the widest images or the greatest quantity of isolated detail.

They are often those in which every element appears to belong to the same acoustic reality.

The musicians have scale. The centre has physical stability. The room surrounds rather than decorates. Reflections extend the performance instead of obscuring it. Silence defines the distance between one sounding body and another.

This is the space between sounds.

It cannot be reduced to one measurement, one microphone technique or one idea of phase coherence. It emerges from a network of relationships extending from the original performance to the listener’s brain.

When musicians, room, microphone position, recording chain, distribution format, loudspeakers and listening environment align, something rare can occur.

The recording no longer seems to document a performance from the past.

It makes that performance inhabitable again.

Perhaps this is the deepest purpose of high fidelity: not merely to reproduce sounds, but to reconstruct a space in which those sounds can exist.

References and Further Reading

  1. Middlebrooks, J. C. and Green, D. M. “Sound Localization by Human Listeners.”
    Annual Review of Psychology, 1991.
    View publication
  2. Brughera, A., Dunai, L. and Hartmann, W. M. “Human Interaural Time Difference Thresholds for Sine Tones: The High-Frequency Limit.”
    Journal of the Acoustical Society of America, 2013.
    View publication
  3. Bures, Z. and Marsalek, P. “On the Precision of Neural Computation with Interaural Level Differences in the Lateral Superior Olive.”
    Brain Research, 2013.
    View publication
  4. Hofman, P. M., Van Riswick, J. G. A. and Van Opstal, A. J. “Relearning Sound Localization with New Ears.”
    Nature Neuroscience, 1998.
    View publication
  5. Pulkki, V. “Microphone Techniques and Directional Quality of Sound Reproduction.”
    Audio Engineering Society, 2002.
    View AES record
  6. Eargle, J. M. “An Overview of Stereo Recording Techniques for Popular Music.”
    Journal of the Audio Engineering Society, 1985.
    View AES record
  7. Toole, F. E. “Loudspeakers and Rooms for Stereophonic Sound Reproduction.”
    Audio Engineering Society 8th International Conference, 1990.
    View AES record

An earlier version of this essay was published in Audio Review, issue 487, June 2026. This Direct Sound Records Journal edition has been revised, expanded and technically updated for an international readership.

Blumlein - Stereo Recording

Why the Blumlein Pair Can Excel in Loudspeaker Playback

1024 575 Michelangelo

Stereo realism does not necessarily emerge from making an image as wide as possible. It emerges when the directional, tonal and reverberant cues in a recording support one another strongly enough to create a stable and believable acoustic space.

In the first article of this series, I explored how human hearing reconstructs a three-dimensional auditory world from differences in arrival time, level and spectral filtering at the two ears.

This second article moves from perception to production: how do different stereo microphone techniques encode spatial information, and why can the choice of array become especially important when a recording is reproduced through loudspeakers?

As a recording engineer and the founder of Direct Sound Records, my central objective is not simply to create an impressive stereo effect. It is to preserve the relationship between the musicians, the acoustic environment and the listener in a way that remains convincing during reproduction.

Among the many available stereo techniques, the Blumlein pair remains one of the most revealing—and one of the most demanding.

There Is No Universal “Best” Stereo Technique

Before considering Blumlein, an important distinction must be made: no stereo microphone technique wins in every situation.

The appropriate array depends on several factors:

  • the size and arrangement of the ensemble;
  • the acoustic quality of the room;
  • the desired relationship between direct and reverberant sound;
  • the required stereo width and localisation precision;
  • mono compatibility;
  • the intended playback system;
  • and the artistic purpose of the recording.

A spaced pair may be ideal when a broad sense of scale and low-frequency spaciousness is required. ORTF can provide an effective balance of width, localisation and ambience. Mid-Side offers control over stereo width after recording. XY can provide a stable image with strong mono compatibility.

Blumlein has its own strengths and limitations. Its value lies not in being universally superior, but in the particular way it connects direct sound, room ambience and coincident stereo geometry.

How Stereo Microphone Arrays Encode Space

Stereo microphone techniques can be broadly understood according to the cues they create between the left and right channels.

Spaced Pairs

In an AB arrangement, two microphones are separated physically. A sound arriving from one side will generally reach one microphone before the other, producing an interchannel time difference. Depending on the microphones and source position, there may also be a difference in level.

Spaced arrays can produce a broad and enveloping presentation. They can be particularly effective for large ensembles, organs, orchestras and situations in which the acoustic environment is an important part of the experience.

However, the time differences between channels can influence mono compatibility and may produce frequency-dependent reinforcement or cancellation when the channels are combined. Increasing microphone spacing can also weaken centre localisation if the geometry is not appropriate for the source and listening conditions.

These are design trade-offs, not proof that spaced recording is inherently defective.

Near-Coincident Arrays

Near-coincident techniques such as ORTF deliberately combine microphone spacing with directional microphone patterns.

ORTF uses two cardioid microphones separated by approximately 17 centimetres and angled 110 degrees apart. The resulting stereo image contains both interchannel timing and level differences.

This combination often produces greater spaciousness than a fully coincident cardioid pair while retaining more definite localisation than a widely spaced AB array. It is one reason ORTF has remained a widely used technique for classical music, ensembles and location recording.

Describing ORTF simply as “phasey” overlooks the fact that its timing differences are intentional components of its spatial design.

Coincident Arrays

In a coincident array, the microphone capsules are positioned as close as physically possible to the same acoustic point.

Because direct sound reaches the two capsules at almost the same time, stereo direction is created primarily through differences in level and polarity rather than substantial arrival-time differences.

XY, Mid-Side and Blumlein are all coincident techniques, although they use different polar patterns and encode the surrounding sound field differently.

The coincident geometry generally provides predictable mono compatibility and reduces the possibility of time-delay-related cancellations when the channels are summed. It can also create a clearly defined centre image and stable localisation within the normal listening area.

Real microphones are not mathematically perfect points, however. Capsule dimensions, vertical displacement, polar-pattern differences and off-axis response mean that no practical array is perfectly coincident or perfectly phase coherent at every frequency.

What Loudspeaker Playback Changes

Headphone and loudspeaker reproduction deliver stereo signals to the listener in fundamentally different ways.

With conventional headphones, the left channel is delivered predominantly to the left ear and the right channel to the right ear. With two loudspeakers, each loudspeaker reaches both ears.

The left ear therefore hears:

  • the left loudspeaker directly;
  • the right loudspeaker through an additional acoustic path;
  • and reflections from the listening room.

The right ear receives the corresponding combination from the opposite side.

This acoustic crosstalk is not an accidental failure of stereo. It is part of conventional two-channel loudspeaker reproduction. The brain uses the resulting combination of timing, level and spectral cues to perceive phantom images between and sometimes beyond the loudspeakers.

However, the reconstruction is sensitive to geometry. Moving away from the central listening position changes the relative distances from the two loudspeakers and therefore changes the arrival-time and level relationships at the ears. The phantom image tends to shift towards the nearer speaker.

The loudspeakers, room and listener must therefore be considered as one reproduction system. A recording does not carry an independent three-dimensional space that remains unchanged under every playback condition.

Enter the Blumlein Pair

The Blumlein pair uses two figure-of-eight microphones mounted coincidently and angled 90 degrees apart.

The technique is associated with Alan Dower Blumlein, whose pioneering 1931 patent described fundamental principles of stereophonic recording and reproduction.

In a correctly arranged Blumlein pair, the microphone diaphragms occupy almost the same acoustic point. Directional information is encoded primarily through the different levels and polarities produced by the two figure-of-eight patterns.

A figure-of-eight microphone is equally sensitive to sound arriving from the front and rear, while strongly rejecting sound arriving from its sides. Consequently, the array captures both the performance in front of the microphones and a substantial amount of acoustic information from behind them.

This is a defining characteristic of Blumlein—not a minor detail.

Why Blumlein Can Sound So Convincing

Coincident Timing for Direct Sound

Because the two capsules are positioned at approximately the same point, direct sound from an instrument reaches both microphones almost simultaneously.

This minimises interchannel arrival-time differences introduced by the microphone spacing itself. The stereo image is produced predominantly through level and polarity relationships.

For loudspeaker reproduction, this can create a precise centre image and clearly organised lateral positions, particularly when the ensemble and array are positioned carefully.

Strong Mono Compatibility

When the two channels of a coincident recording are combined, corresponding direct sounds normally align more predictably than they do in a widely spaced array.

This does not mean that every part of a Blumlein recording will combine perfectly. Reflections arrive from many directions and at many times, while real microphones have tolerances and frequency-dependent polar behaviour.

Nevertheless, the coincident geometry generally gives Blumlein excellent mono compatibility compared with arrays that rely heavily on microphone spacing.

Natural Integration of the Room

The rear lobes of the figure-of-eight microphones capture reverberant energy and sound arriving from behind the array.

In a good acoustic environment, this can create a remarkably integrated sense of depth. The room does not feel like a synthetic effect added behind the musicians. It becomes part of the same spatial event.

The direct sound establishes the performers, while the reflected energy communicates the dimensions, character and decay of the venue.

When those relationships are balanced correctly, the listener may perceive not merely a wide line between two loudspeakers, but a coherent acoustic scene extending behind and around the performers.

Spatial Information Without Microphone Spacing

Blumlein can generate a substantial stereo image without separating the microphones horizontally.

This is particularly attractive when the engineer wants clear directional information while minimising time-of-arrival differences between channels.

The result can feel cohesive because the direct sound and room information are captured from a single acoustic viewpoint.

What Phase Coherence Really Means Here

The expression phase coherence is often used loosely in audio. In the context of a coincident stereo array, it is more useful to speak about the consistency of interchannel timing relationships.

Blumlein does not remove phase from a recording. Every acoustic event contains complex phase relationships, and every room creates reflections with different delays, levels and spectra.

What the array minimises is the additional time difference that would otherwise be introduced by placing the two microphones at separate locations.

This distinction matters.

A Blumlein recording can still contain:

  • phase differences created by room reflections;
  • microphone-response differences;
  • polarity differences inherent in the figure-of-eight geometry;
  • and complex interference between direct and reverberant sound.

Its strength is not “perfect phase purity.” Its strength is that both channels observe the direct acoustic event from approximately the same point in space.

Why Blumlein Is Also Demanding

The characteristics that make Blumlein revealing also make it unforgiving.

The Room Must Deserve to Be Recorded

Because figure-of-eight microphones capture strongly from both front and rear, an unattractive room will not politely disappear.

Flutter echoes, mechanical noise, audience movement, heating systems and poorly controlled reflections can become prominent parts of the recording.

Blumlein works best when the acoustic environment contributes positively to the performance.

Placement Is Critical

The balance between ensemble width, direct sound and reverberation depends strongly on the distance and orientation of the array.

Positioning the microphones too close may produce an image that is excessively wide or exclude important sources from the useful recording angle. Placing them too far away may allow reverberation to dominate and reduce clarity.

Small movements can significantly change the result. This is why Blumlein rewards careful listening and deliberate placement rather than formula alone.

Rear Sound Is Part of the Recording

The rear lobes do not distinguish between beautiful reverberation and unwanted noise.

Musicians, audience members, equipment and reflective surfaces behind the microphones all become part of the captured field. The engineer must therefore consider the entire environment around the array, not only what lies in front of it.

The Listening Position Still Matters

Blumlein does not eliminate the limitations of two-loudspeaker stereo.

A listener moving significantly away from the central position will still experience changes in timing and level from the loudspeakers, and the stereo image will shift accordingly.

Coincident recording can provide a coherent source signal, but it cannot make conventional stereo reproduction independent of loudspeaker and listener geometry.

When I Choose Blumlein

In my work, Blumlein becomes especially compelling when:

  • the musicians are acoustically balanced in the room;
  • the venue has a distinctive and musically valuable acoustic;
  • the ensemble fits naturally within the array’s useful recording angle;
  • the intention is to preserve a complete performance rather than construct one later;
  • and loudspeaker playback is an important reference.

I would not choose it automatically when the room is problematic, when strong isolation is required, when sources must be balanced independently, or when the ensemble geometry demands a wider or more flexible array.

In those situations, ORTF, AB, Mid-Side, XY, supplementary microphones or a hybrid approach may be more appropriate.

The technique should serve the acoustic event—not the engineer’s ideology.

Spatial Width Is Not the Same as Realism

A recording can create an enormous stereo image and still feel artificial.

Width may be produced by long interchannel delays, decorrelation, processing or exaggerated ambience. These effects can be exciting, but they do not necessarily communicate a believable relationship between performers and space.

Blumlein offers a different proposition. Its most successful recordings do not merely place sounds from left to right. They establish a unified perspective from which the listener can infer:

  • where the musicians are positioned;
  • how far away they appear;
  • how the room surrounds them;
  • and how direct and reflected sound belong to the same event.

This is why the technique can feel less like an audio effect and more like a view into an acoustic performance.

A Reference, Not a Religion

The Blumlein pair deserves its status as one of the foundational stereo microphone techniques. Its coincident geometry, figure-of-eight patterns and integration of direct and reverberant sound can produce extraordinary depth, localisation and spatial coherence.

But the strongest case for Blumlein does not require dismissing other approaches.

AB can communicate scale and spaciousness that a coincident pair may not reproduce in the same way. ORTF can offer a persuasive compromise between width and localisation. Mid-Side provides valuable control after recording. XY can be practical, focused and robust.

The real achievement lies in understanding how each technique encodes space—and choosing the one whose compromises best serve the music, venue and intended reproduction system.

For Direct Sound Records, the objective is not to manufacture an impressive stereo image. It is to preserve the acoustic relationships that make a performance feel present, intelligible and emotionally credible.

When the room, musicians and microphone position align, Blumlein can be one of the most direct ways of achieving that objective.

References and Further Reading

  1. Blumlein, A. D. Improvements in and Relating to Sound-Transmission, Sound-Recording and Sound-Reproducing Systems. British Patent GB394325A, filed 1931 and published 1933.
    View patent
  2. Eargle, J. M. “An Overview of Stereo Recording Techniques for Popular Music.”
    Journal of the Audio Engineering Society, 1985.
    View AES record
  3. Ceoen, C. “Basic Stereo Microphone Perspectives—A Review.”
    Journal of the Audio Engineering Society, 1985.
    View AES record
  4. Toole, F. E. “Loudspeakers and Rooms for Stereophonic Sound Reproduction.”
    Audio Engineering Society 8th International Conference, 1990.
    View AES record
  5. Kendall, G. S. “The Effects of Interaural Crosstalk on Stereo Reproduction and Minimizing Interaural Crosstalk in Nearfield Monitoring by the Use of a Physical Barrier: Part 1.”
    Audio Engineering Society 81st Convention, 1986.
    View AES record
  6. Lee, H. and Gribben, C. “On the Optimum Listening Position and Listening Angle in a Two-Channel Stereophonic Reproduction System.”
    Audio Engineering Society.
    View AES record

An earlier version of this article was published on LinkedIn. This Direct Sound Records Journal edition has been revised, expanded and technically updated, with additional context and references.

How We Hear in 3D – The Neuroscience Behind Stereo Perception

1024 1024 Michelangelo

“Phase is not the sound — it is the space between sounds.” Taken as a metaphor, this captures something fundamental about human hearing: the brain does not process the signals arriving at our two ears independently. It continuously compares them, using extraordinarily small differences in time, level and spectral balance to reconstruct the position of sound around us.

In high-end audio, we often concentrate on equipment, formats, resolution and frequency response. Yet beneath every recording and every playback system lies a more fundamental question: how does the human auditory system transform two incoming signals into a convincing three-dimensional world?

Stereo reproduction works not because two loudspeakers recreate the original sound field perfectly, but because they provide the auditory system with enough carefully organised information to create a plausible spatial scene. Understanding that process changes how we think about microphone placement, phase relationships, room acoustics and the meaning of realism in recording.

Hearing Is a Reconstruction

Sound arriving at the ears does not contain a ready-made map of the space around us. The auditory system must infer the direction, distance and environment of a sound source from a collection of acoustic clues.

For horizontal localisation, the most important binaural cues are differences in arrival time and sound level between the two ears. For elevation and front-to-back discrimination, the auditory system also relies heavily on direction-dependent spectral filtering created by the listener’s head, torso and outer ears.

These mechanisms operate together. They should not be understood as three entirely separate systems, nor as rigid frequency zones with precise boundaries. Natural sounds are usually broadband, reflections complicate the incoming signals, and the brain integrates multiple cues over time.

The Principal Cues to Spatial Hearing

1. Interaural Time Differences

When a sound source is positioned to one side of the listener, the sound normally reaches the nearer ear slightly before reaching the farther ear. This difference is known as the interaural time difference, or ITD.

The delays involved are extremely small—often measured in microseconds—but the auditory system is remarkably sensitive to them. For low-frequency sounds, neural activity can follow the temporal structure of the waveform closely enough for the brain to compare the timing received at the two ears.

ITD sensitivity to the fine structure of pure tones becomes progressively less effective as frequency rises, with human sensitivity deteriorating sharply around the region of approximately 1.4 to 1.5 kHz. This should be treated as a broad transition rather than a universal dividing line. High-frequency sounds can still carry timing information through changes in their amplitude envelope.

2. Interaural Level Differences

A sound arriving from one side will also tend to be louder at the nearer ear. At shorter wavelengths, the head obstructs part of the sound travelling towards the farther ear, producing an acoustic shadow. The resulting difference in level is known as the interaural level difference, or ILD.

Level differences generally become more pronounced at higher frequencies because shorter wavelengths are more strongly affected by the head. However, ITD and ILD do not simply exchange responsibility at one exact frequency. For complex sounds, the auditory system can combine timing and level information across several frequency regions.

The relative importance of these cues also changes with the sound itself, its distance, the surrounding reflections and the listener’s hearing.

3. Spectral Cues and the Head-Related Transfer Function

Time and level differences are especially useful for identifying whether a sound is located towards the left or right. They are less able, by themselves, to resolve whether a source is above, below, in front of or behind the listener.

For this, the complex shape of the outer ear becomes essential. The folds of the pinna, together with the head and upper body, alter the spectrum of incoming sound in a direction-dependent way. Some frequencies are reinforced, while others are attenuated or notched.

The complete acoustic transformation between a sound source and the listener’s ears is described by the head-related transfer function, or HRTF.

Because every person’s anatomy is different, HRTFs are individual. The brain gradually learns the spectral patterns associated with particular directions. Experiments in which the shape of the outer ear was temporarily altered have shown that localisation initially becomes less accurate, particularly for elevation and front-to-back judgements, but can improve again as listeners adapt to the modified cues.

Where Phase Fits In

The word phase is used in several related but distinct ways in audio, which can easily create confusion.

At low frequencies, a difference in arrival time between the ears can also be described as an interaural phase difference for a periodic waveform. In this context, phase difference is one of the ways the auditory system obtains spatial information.

Neurons within the auditory brainstem, including those associated with the medial superior olive, are specialised for processing extremely small interaural timing differences. Their responses contribute to the neural representation of horizontal sound direction.

However, it would be too simple to conclude that phase is used only to locate a sound and never contributes to what that sound is. The timing structure of a waveform—often described as its temporal fine structure—also contributes to aspects of pitch perception, auditory masking and the separation of sounds in complex listening environments.

The more useful distinction for recording engineers is therefore not between “phase” and “sound identity,” but between the different roles that temporal and phase relationships can play.

Within a stereo recording, relationships between the two channels may affect:

  • the apparent position of a phantom image;
  • the perceived width and stability of the soundstage;
  • the impression of depth and surrounding ambience;
  • the result when the recording is reproduced in mono;
  • frequency-response changes caused by constructive and destructive interference.

Phase is therefore neither an isolated technical curiosity nor a universal explanation for every spatial quality. It is one component within a larger system of timing, level, spectrum, reflection and playback interaction.

Why This Matters for Stereo Recording

Stereo microphone techniques create spatial information in different ways.

A spaced pair such as AB introduces arrival-time differences between the microphones and may also produce level differences. A near-coincident configuration such as ORTF deliberately combines microphone spacing with directional level differences.

Coincident techniques such as XY and Mid-Side minimise the arrival-time difference between microphones and create direction mainly through differences in level. The Blumlein pair, using two coincident figure-of-eight microphones, also derives its directional information from the polar patterns and polarity relationships of the two channels while capturing substantial information from the surrounding acoustic environment.

None of these methods is automatically natural or unnatural in every situation. Each encodes the original acoustic event differently, and each interacts differently with loudspeakers, headphones, room reflections and listener position.

This is particularly important because conventional stereo loudspeaker reproduction does not send the left channel exclusively to the left ear or the right channel exclusively to the right ear. Each loudspeaker reaches both ears, introducing additional timing, level and spectral interactions. The listening room then adds its own reflections.

The task of the recording engineer is therefore not merely to create a wide image. It is to create interchannel relationships that remain meaningful when reproduced through the intended playback system.

From Spatial Effect to Spatial Credibility

A stereo recording can sound spectacularly wide while still producing unstable localisation, exaggerated scale or an uncertain centre image. Conversely, a narrower presentation may feel more convincing because its timing, level and reverberant cues form a more internally consistent spatial picture.

This suggests a more useful question than simply asking whether a recording sounds spacious:

Do the spatial cues reproduced by the system support one another strongly enough for the brain to construct a stable and believable acoustic scene?

Phase coherence is part of that question, but it should not be treated as a single measurement that determines realism on its own. Microphone polar pattern, spacing, angle, source distance, direct-to-reverberant ratio, loudspeaker placement and room acoustics all contribute to the final perception.

What Comes Next

In the next article, I will examine how AB, ORTF, XY, Mid-Side and Blumlein recording techniques encode spatial information differently, and why a technique that produces impressive width does not necessarily produce the most credible depth or localisation.

I will also explore the relationship between coincident microphone techniques, loudspeaker reproduction, room acoustics and the preservation of stable interchannel relationships.

The central principle is simple: recording technology should not be considered separately from human perception. The microphone arrangement, recording space and playback format are parts of one perceptual chain.

References and Further Reading

  1. Brughera, A., Dunai, L. and Hartmann, W. M. “Human interaural time difference thresholds for sine tones: the high-frequency limit.” Journal of the Acoustical Society of America, 2013.
    View publication
  2. Wightman, F. L. and Kistler, D. J. “The dominant role of low-frequency interaural time differences in sound localization.”
    Journal of the Acoustical Society of America, 1992.
    View publication
  3. Salminen, N. H., Tiitinen, H., Yrttiaho, S. and May, P. J. C. “The neural code for interaural time difference in human auditory cortex.”
    Journal of the Acoustical Society of America, 2010.
    View publication
  4. Hofman, P. M., Van Riswick, J. G. A. and Van Opstal, A. J. “Relearning sound localization with new ears.”
    Nature Neuroscience, 1998.
    View publication
  5. Moore, B. C. J. “The role of temporal fine structure processing in pitch perception, masking, and speech perception for normal-hearing and hearing-impaired people.”
    Journal of the Association for Research in Otolaryngology, 2008.
    View publication
  6. Eargle, J. “An Overview of Stereo Recording Techniques for Popular Music.”
    Journal of the Audio Engineering Society, 1985.
    View publication

An earlier version of this article was published on LinkedIn. This Direct Sound Records Journal edition has been revised, expanded and technically updated, with additional context and scientific references.