Private before-and-after analysis

Voice Comparison Tool - Compare Two Recordings

Record sample A and sample B, then compare supported pitch measurements without uploading either recording.

Before-and-after comparison

Compare two voice samples locally

Use the same passage for both recordings. Each sample is analyzed in memory; audio is discarded when analysis finishes or the tab closes.

Sample A

No sample analyzed yet.

Sample B

No sample analyzed yet.

Audio is analyzed in this tab and never uploaded.

What the voice comparison tool compares

This tool records sample A and sample B, then compares two acoustic summaries produced by the same local YIN-based analysis: median fundamental frequency and central pitch movement. Median pitch describes the centre of reliable voiced frames. Movement describes the robust distance between the typical lower and upper pitch regions. The comparison shows how those measurements differed between these two recordings.

It does not compare identity, overall vocal quality, health, articulation, formants, resonance, accent, emotion, or every factor a listener might notice. It also does not prove that one recording is better. A positive or negative difference simply reports direction and size in the supported metrics. The meaning depends on the task you intentionally recorded.

Use the page for a controlled before-and-after practice check, two delivery styles, or two microphone positions. If the tasks differ in words, room, device, or speaking intention, the comparison is still technically calculated but cannot isolate the cause of the difference.

A fair comparison starts with a controlled setup

Record the same passage in the same room, on the same device, with the microphone at the same distance and angle. Keep posture, volume, and intended speaking style consistent unless one of those is the variable you deliberately want to test. A fixed passage controls the vowels and sentence emphasis better than two unrelated pieces of spontaneous speech.

Make both samples long enough to contain several clear voiced phrases. A very short or noisy sample can produce unstable statistics. Avoid touching the device, moving around the room, playing background audio, or switching between a built-in and Bluetooth microphone. Consumer devices apply different gain, filtering, and echo controls that can create apparent acoustic changes.

If you are comparing a training technique, record an ordinary baseline first, take the intended practice step, and record the second sample once without forcing a dramatic result. Repeat the full pair on another occasion. A consistent difference across controlled pairs carries more practical information than the largest difference selected from many attempts.

  • Same text and language.
  • Same room, device, microphone, distance, and angle.
  • Similar posture, volume, time limit, and speaking intention.
  • At least several seconds of clear voiced speech in each sample.
  • One planned variable changed at a time.

How to read the difference between sample A and sample B

First check that both samples produced usable local analysis. Then compare median pitch in Hertz and movement in semitones. A higher median in B means the centre of detected F0 frames was higher in that recording. More movement means the typical pitch band was wider. Neither direction is automatically desirable: a steady sustained task should be narrow, while expressive speech normally needs movement.

Read the absolute values as well as the difference. Two low-quality samples can produce a neat-looking subtraction that is not useful. If one recording contains vocal fry, breathy sound, clipping, interruptions, or fewer voiced frames, repeat the pair. An implausibly large change may reflect octave-selection error or a different microphone condition rather than the intended voice choice.

The tool does not perform causal analysis. If B changed after an exercise, the result cannot prove that the exercise caused it. Familiarity with the text, expectation, fatigue, room noise, and ordinary variation may also contribute. Describe the outcome accurately as “these recordings differed,” then gather repeat observations before drawing a training conclusion.

Useful voice comparison workflows

For presentation practice, compare the same paragraph delivered conversationally and then with deliberate emphasis, while recognizing that the task itself changed. For consistency practice, make A and B as similar as possible and see whether the metrics settle in a narrow region. For microphone testing, keep the voice task fixed and change only the device or distance, interpreting the result as a recording-chain comparison.

For voice training over time, avoid comparing today’s spontaneous sentence with next month’s different sentence. Keep a short benchmark passage and write down the setup, date, metrics, and how the sample felt. The website does not save a history, which preserves privacy but means longitudinal notes are your responsibility.

Measurements can direct attention, but listening remains essential. A smaller pitch difference does not mean two recordings sound identical, and a larger difference does not mean the change was comfortable or sustainable. Consider feedback from a consenting qualified coach for individual technique goals.

What a two-recording comparison cannot conclude

The page cannot determine gender identity, biological sex, age, nationality, accent, health, speaker identity, or personal identity. It cannot score improvement, attractiveness, professionalism, authenticity, confidence, or suitability for a role. Pitch measurements overlap broadly between people and are affected by the recording task.

Do not use it to evaluate another person without informed permission or to make decisions involving employment, education, housing, insurance, credit, healthcare, access, moderation, or relationships. A browser comparison is not appropriate evidence for high-impact decisions or surveillance.

The tool is not a clinical test. Pain, persistent hoarseness, difficulty speaking, or a sudden lasting voice change needs an appropriately qualified healthcare or speech-language professional. Do not repeatedly force high or low sounds merely to maximize a difference.

Both recordings stay in the browser

Microphone samples are held temporarily in the current tab and analyzed locally. The site does not upload sample A, sample B, PCM frames, extracted pitch values, comparison differences, or results. No cloud transcription, machine-learning model, voiceprint, or account database is involved.

The recordings are discarded after the local flow or when the page session ends. If you want to retain evidence of practice, write down the visible measurements or make recordings separately using a device recorder under your own privacy controls. This site does not provide storage or sharing of the audio.

How to use this tool

  1. Choose one short passage and fix the room, device, microphone position, volume, and speaking task.
  2. Record sample A for several seconds using a natural, comfortable voice.
  3. Change only the variable you intend to explore, then record sample B with the same passage.
  4. Compare median pitch and pitch movement together, and repeat the pair before treating a difference as a trend.

Frequently asked questions

Can I compare a training before and after?▾

Yes. Use the same passage and setup for both samples, then read the differences as recording-level observations rather than proof of success or causation.

What does the comparison measure?▾

It currently compares median fundamental frequency and central pitch movement, using the same local YIN-based analysis for both recordings.

How do I make the two recordings comparable?▾

Keep text, language, room, device, microphone distance, posture, volume, and speaking intention consistent. Change only the variable you want to explore.

Does a higher median pitch mean sample B is better?▾

No. Higher and lower are directions, not grades. The desired pattern depends on your task, comfort, and goal.

Can the tool prove that practice caused a change?▾

No. It shows a difference between recordings. Repetition and controlled conditions reduce uncertainty, but they do not establish cause by themselves.

Why might the comparison look unexpectedly large?▾

Different equipment, noise, clipping, vocal fry, breathy sound, weak voiced coverage, or an occasional octave error can distort one sample. Repeat the pair in controlled conditions.

Can I compare two different people?▾

Only with clear consent, but the result still cannot identify, classify, rank, or judge either person. The intended use is voluntary recording comparison, not evaluation or surveillance.

Are the two recordings saved or uploaded?▾

No. Both samples and their acoustic data are processed in memory in the browser and are not sent to a server or stored in an account.