Skip to content
Sound Safari Sound Safari
Therapy Techniques

/F/ and /V/ Speech Therapy: Teach F & V, Fix Stopping

Complete F and V sound therapy guide: fix stopping (f→p, v→b) and teach the F/V voicing pair — word lists, minimal pairs, norms, and sample IEP goals.

/f/ and /v/ look like two separate therapy targets. They are not. They are one mouth position wearing two different voices — say fish, then say van, and your top teeth land on your lower lip in exactly the same spot both times. The only thing that changes is whether your voice is switched on. Grasp that single fact and you have the whole map for teaching both sounds and fixing the errors children make with them.

This guide walks SLPs and parents through teaching /f/ and /v/ as a single labiodental voicing pair — one shared placement, differing only in voicing — and fixing the stopping errors behind them (f→p, v→b), with developmental norms, word lists by position, two different jobs for minimal pairs, and copy-ready IEP goals.

It is written primarily for SLPs — clinical depth on placement, error patterns, word lists, minimal pairs, and IEP goals — with a clearly marked home-practice section for parents. Skip to the section you need.

F and V: One Placement, Two Voices

Teach /f/ and /v/ together, as one pair, because they are made the same way. Both are labiodental fricatives. Labiodental means the sound is formed with the lip (labio-) and the teeth (dental) — the top front teeth resting lightly on the lower lip. Fricative means the sound is a continuous stream of turbulent air, not a quick pop. So both sounds are the same gesture: top teeth on lower lip, blowing a steady airstream through the narrow gap.

The only difference between them is voicing — whether the vocal folds are vibrating:

Same mouth, same air — throat off versus throat on. That is the entire contrast, and it is the backbone of everything below.

This is not a teaching simplification; it is what the data says. In Sound Safari’s norm set, the /f/ and /v/ entries sit on identical place and manner rows — place is labiodental, manner is fricative — and differ on exactly one flag: /f/ is voiceless, /v/ is voiced. Two sounds, one placement, one switch. The American Speech-Language-Hearing Association’s speech sound disorders portal describes the same articulatory picture.

Both sounds also belong to Shriberg’s Middle-8 — the middle band of English consonant development (/ŋ/, /t/, /k/, /g/, /f/, /v/, /ʧ/, /ʤ/) — not the earliest sounds to emerge, but well ahead of later ones like /s/, /r/, and /th/. That sets a fair expectation: /f/ and /v/ should be in place before the truly late sounds, but a young child still working on them is often right on schedule.

When Should a Child Say F and V?

Developmental norms decide whether /f/ or /v/ is even an IEP concern. The values below come from Crowe & McLeod’s 2020 U.S. review — the current standard for American English consonant acquisition — with its cross-linguistic companion, McLeod & Crowe (2018).

Sound50% (customary)90% (mastery)Read it as
/f/ (quiet)2;64;0established by around age 4 — the earlier half
/v/ (buzzy)4;05;0established by around age 5 — the later half

Ages from Crowe & McLeod (2020); cross-linguistic companion data in McLeod & Crowe (2018).

Ages use the standard speech-pathology years;months format, not decimals. 4;0 means 4 years, 0 months old — not 4.0 years. 2;6 means 2 years, 6 months.

Two things stand out. First, /f/ is the earlier sound — half of children have it by 2;6 and most by age 4. Second, /v/ trails it by a full year — half of children have it by age 4 and most by age 5.

That gap is not a rounding artifact; it is the key sequencing fact, and the next section explains why.

That reframes the IEP decision. A child at or past the ~4;0 (/f/) or ~5;0 (/v/) mastery age who is still erroring the sound, with an intelligibility or participation impact, is typically a reasonable target. Below those ages, the same error is usually developmentally typical.

Two caveats keep this honest, as with any norms-based decision. Being under the mastery age does not rule out an IEP — a highly unintelligible younger child may still need service. And being over it does not mandate one — a mild, fully intelligible residual error may not meet eligibility. Norms, intelligibility, and stimulability decide together.

For the full consonant timeline, see our speech sound milestones by age.

Why V Is the Harder Half of the Pair

Why does /v/ lag /f/ by a whole year when the placement is identical?

/v/ is a voiced fricative — and voiced fricatives are demanding. The voicing has to be sustained through continuous frication, which is harder to coordinate than the voiceless /f/, where the folds simply stay off. On top of that, /v/ is a low-frequency sound in English and acoustically quiet, so a child gets fewer clear models of it and has a harder time hearing the target in the first place.

Later, quieter, harder to hear, harder to voice — that is a full year of extra difficulty stacked onto the same lip position.

The signature /v/ error follows directly: word-final devoicing. The child reaches the right placement but drops the voicing at the end of the word, so leave becomes “leaf,” five becomes “fife,” and have becomes “haff.” The mouth is correct; the voice quits early.

The clinical implication is concrete: give /v/ its own functional final-position targetsleave, give, love, five, have, move — rather than assuming a solid final /f/ transfers automatically to final /v/. It usually does not.

This sets up the sequencing the rest of the guide follows. /f/ comes first — it is earlier, voiceless, and easier to hear. Then /v/ is often gotten “for free” by holding an already-solid /f/ placement and toggling the voice on. One placement, two voices, taught in that order.

Stopping vs TH-Fronting: Triage the Error First

Before picking a fix, sort the error you actually hear. Two different errors send children to /f/ and /v/ therapy, and they need different paths.

Error you hearWhat it isThe fix
f→p (“pish” for fish), v→b (“ban” for van)Stopping — a fricative replaced by a stopF-vs-P and V-vs-B minimal pairs (this article)
f-for-”th” (“fum” for thumb, “wif” for with)TH-fronting — a fronting/labialization patternthe /th/ path, not the /f/ path

Stopping is this article’s primary target. In f→p and v→b, the child swaps the long, flowing fricative for a quick stop that is easier to make. Sound Safari’s process norms classify stopping of fricatives as a typical (non-atypical) phonological process that is normally eliminated by around age 3 (roughly 3;0–3;6). So a 4-year-old still saying “pan” for fan is past the developmental window and worth screening.

TH-fronting is a different animal. A child who says “fum” for thumb or “wif” for with is not stopping anything — they are substituting /f/ for the “th” sound, which is a fronting pattern, not a fricative-to-stop error. The sound that needs work is th, not /f/. Sound Safari backs this distinction directly: it ships a separate F-vs-TH contrast filed under fronting, not stopping. A parent who hears “fum” for thumb needs the /th/ route, not the /f/ path.

For the full breakdown of stopping, fronting, and their age-of-resolution norms, see our phonological processes guide. Sort the error first, and the rest of therapy gets faster.

How to Teach F and V: Placement, Then Voicing

Because both sounds live at one shared position, you teach the placement once and split by voicing. Probe first, set the position, teach the buzz, then toggle.

Step 1 — Probe stimulability first

Before choosing anything, check whether /f/ is already emerging. Have the child bite the lower lip lightly and blow — the classic “angry cat” hiss or “bunny teeth” cue works well. Model it, then have them imitate. Establish a stimulable /f/ first, because the placement is identical for both sounds and, once /f/ is solid, /v/ usually follows from it.

Step 2 — Set the shared labiodental placement

The core target, shared by both phonemes: rest the top front teeth gently on the lower lip and blow a long, quiet stream of air — a sustained “fffff.” Give the child a way to see the air working: a hand held at the mouth to feel it, or a small tissue that flutters when the airstream is steady. That one placement is the whole foundation. The same posture produces both /f/ and /v/.

Step 3 — Teach the voicing contrast

Put the child’s hand on their throat. For the quiet /f/, the throat is still. For the buzzy /v/, the throat vibrates. Contrast a whisper against a voiced sound so the child feels the difference, then elongate “fffff” and slide into “vvvv” on the same lip posture — turning the voice on without moving the lips at all. “Quiet F, buzzy V” becomes something the child can literally feel under their fingertips.

Step 4 — Get /v/ for free by toggling voicing

Once /f/ is stable, hold the /f/ placement and switch the voice on to land on /v/ — same mouth, throat on. Because the lips never move, /v/ is a voicing switch, not a new motor plan. This is the one-placement, two-voices thesis paying off, and it is why /f/ is worth stabilizing first.

Step 5 — Watch for drift, then generalize

Two slips to catch early. First, stopping drift: the teeth leave the lip and the sound collapses back to p/b — re-anchor the teeth-on-lip contact until it holds. Second, final devoicing on /v/: “leave” slides to “leaf” at the end of a word — over-emphasize sustained voicing in the final position. When placement and voicing are stable, climb the hierarchy — isolation → syllable → word → phrase → sentence → conversation — which the word lists and IEP goals below operationalize.

If you would rather work from a ready-made ladder than build one, Sound Safari runs exactly these levels for /f/ and /v/ with position-sorted words at each rung, so the practice targets are picked for you rather than assembled by hand.

F and V Word Lists by Position

A sound guide needs position-sorted word lists — and because this is a pair, you need two, each split by position. Most therapy starts word-initial (easiest to cue), then moves to final, then medial.

/f/ ships roughly 89 practice words — about 30 initial, 29 medial, and 30 final. A working sample:

Initial /f/Medial /f/Final /f/
fishcoffeeleaf
funmuffinroof
fourelephantknife
fandolphinlaugh
foxwafflehalf
feetjellyfishsafe

/v/ ships roughly 90 practice words — 30 initial, 30 medial, and 30 final. A working sample:

Initial /v/Medial /v/Final /v/
vansevenleave
vaseovengive
vineriverlove
voicebeaverfive
vetsilverhave
valleyelevenmove

Give the final /v/ column extra reps — that is where devoicing lives (leave → “leaf”), and it does not fix itself just because final /f/ is solid. That ties straight back to why /v/ is the harder half of the pair.

Sound Safari generates these position-sorted lists automatically — around 30 words per position for each sound — so an SLP is choosing from a ready set rather than writing word lists between sessions.

F vs P and V vs B Minimal Pairs: Fixing Stopping

Minimal pairs come in two flavors for this pair, and confusing them wastes therapy time. This first job — F-vs-P and V-vs-B — targets stopping, the fricative-to-stop error (f→p, v→b) triaged earlier.

F vs P — for the f→p error:

FP
fanpan
figpig
fourpour
finpin
fatpat
foundpound

V vs B — for the v→b error:

VB
vanban
vatbat
voteboat
veryberry
vestbest
vetbet

Why the contrast works: making the child choose between fish and “pish” forces them to hear that dropping the airflow changes the word. That gives them a communicative reason to keep the fricative going — contrast for repair, not drill for its own sake.

Two honest notes on what the app does and does not ship. Sound Safari’s in-app Minimal Pairs game does include an illustrated F-vs-P set — a sizable stopping set of picture pairs — so you can run that contrast without printing worksheets. V-vs-B, however, is a general clinical technique here, not an app feature — Sound Safari does not ship a V-vs-B set, so the V-vs-B words above are for you to use on paper or with your own pictures. Keep both of these labeled by their job — stopping — so they never blur into the voicing contrast in the next section.

F vs V Minimal Pairs: Fixing Voicing

The second job is different. F-vs-V targets voicing, not stopping — it is for the child who already has both placements but confuses quiet and buzzy, or who devoices /v/ at the end of words.

This one is a genuine built-in Sound Safari feature: the app ships a named F-vs-V voicing contrast, described as the voicing pair of the labiodental fricatives — the exact cognate pair this whole guide is built around. Its curated, illustrated-and-audio set carries six pairs:

FV
fanvan
finevine
facevase
fastvast
safesave
leafleave

Those six are four word-initial pairs plus two word-final pairs — there are no curated medial pairs — and the app’s live illustrated F-vs-V set is larger still for expanded practice. Notice that safe/save and leaf/leave double as final-devoicing targets: they are exactly the “leave → leaf” pattern from the /v/-is-harder section, so running them does two jobs at once.

The teaching move is to pair this drill with the hand-on-throat buzz cue from the technique section. The contrast is literally voice-off versus voice-on at the same placement — fan with the throat still, van with the throat buzzing. Browse the full library of minimal pairs to see the stopping and voicing contrasts side by side. If you would rather run illustrated F-vs-V pairs than assemble them, that set is built into Sound Safari’s Minimal Pairs game.

Sample IEP Goals for F and V

Once /f/ or /v/ is confirmed as a target — gated by the mastery ages above, with an intelligibility impact — you need measurable goals. Sound Safari’s goal bank is organized into 8 categories — articulation, phonology, fluency, voice, receptive language, expressive language, pragmatic language, and phonological awareness — with built-in templates that default to an 80% accuracy criterion across 3 consecutive sessions. (The sample goals below say “20 trials in 3 of 4 consecutive sessions” — the “20 trials” and “3 of 4” are common clinician conventions layered on the app’s underlying 3-consecutive-session default, not app values.)

For a straightforward substitution where the child is producing the sound but in the wrong place or not at all, use the articulation “Sound in Word Position” template — suggested baseline 0–40%, 80% accuracy, 12–36 weeks:

Goal 1 (/f/, word position): Given a verbal model and a tactile cue, [Student] will produce /f/ in the initial position of single-syllable words (e.g., fish, fan, four) with 80% accuracy across 20 trials in 3 of 4 consecutive sessions by the end of the IEP year. Suggested baseline: 0–40%. Timeframe: 12–36 weeks.

Goal 2 (/v/, word position): Given a verbal model and a tactile cue, [Student] will produce /v/ in the initial position of single-syllable words (e.g., van, vase, vine) with 80% accuracy across 20 trials in 3 of 4 consecutive sessions by the end of the IEP year. Suggested baseline: 0–40%. Timeframe: 12–36 weeks.

When f→p and v→b pattern together as a phonological process, reach for the phonology “Stopping” template instead — suggested baseline 0–30%, 80% accuracy, 20–40 weeks:

Goal 3 (stopping, phonology): [Student] will eliminate the process of stopping by producing fricatives correctly with 80% accuracy across 3 consecutive sessions by the end of the IEP year. The app supplies this template’s goal text, 80% criterion, suggested baseline (0–30%), and timeframe (20–40 weeks).

Two more goals cover the pair’s specific needs — an isolation goal to establish the sound and a final-position /v/ goal aimed squarely at devoicing:

Goal 4 (/v/, isolation): Given a model and a hand-on-throat voicing cue, [Student] will produce /v/ in isolation with 80% accuracy across 20 trials in 3 of 4 consecutive sessions. Suggested baseline: 0–40%. Illustrative timeframe: 8–12 weeks.

Goal 5 (final-position /v/, devoicing): Given a verbal model, [Student] will produce /v/ in the final position of words (e.g., leave, give, five) without devoicing, with 80% accuracy across 20 trials in 3 of 4 consecutive sessions by the end of the IEP year. Suggested baseline: 0–40%. Timeframe: 12–36 weeks.

Progress these from isolation to word-by-position to sentence to conversation, the same ladder the technique section walks. Every goal here is a starting point, not a substitute for clinical judgment, baseline data, or your district’s conventions. For the full copy-paste set, including a dedicated /f/ goals section and the norms table, see our IEP goal templates. In Sound Safari, session scoring links straight to the goals you have written, so the probe data behind each one stays current without manual re-entry.

Home Practice for Parents

This section is for parents and caregivers — no clinical background needed. (An “SLP” is a speech-language pathologist, the professional who treats speech sounds.)

First, the picture. /f/ and /v/ are made the same way — top teeth resting gently on the bottom lip, blowing air — and the only difference is a little voice motor in the throat. /f/ is the quiet sound (the “fffff” in fish); /v/ is the buzzy sound (the “vvvv” in van). Same mouth, motor off or on.

A few simple things help at home:

  1. The quiet-F, buzzy-V game. Put your child’s hand on their throat. For /f/ the throat is still; for /v/ it hums. Let them feel the buzz switch on and off — it makes an invisible difference something they can touch.
  2. Short and frequent. Three to five minutes a few times a week beats one long session. These are motor skills, and little-and-often builds them fastest — use the word lists above.
  3. Model, don’t correct. If your child says “pish” for fish, don’t demand a redo. Just say the word correctly two or three times in normal conversation. Hearing the right version matters more than drilling it.

A quick guide on when to check in with an SLP: f→p and v→b are usually typical before around age 3. It is worth a screening if your child is still saying “pish” for fish past about age 4, or “ban” for van past about age 5, or if people outside the family have trouble understanding them.

One reassurance that trips a lot of families up: if your child says “fum” for thumb or “wif” for with, that is a different sound problem — it is about the “th” sound, not /f/ — so it needs the /th/ path. Our /th/ sound therapy guide covers that one.

If you want a little more structure, Sound Safari’s parent view walks through /f/ and /v/ practice with pictures right in your web browser — no app to download. Your child’s practice data stays on your device and your private iCloud; no third party ever receives your child’s practice or clinical data — no ads, no behavioral tracking, no data brokers. Account sign-in and subscriptions are handled by standard providers that see only account and billing details, never session data. See what parents get.

From Practice to Progress: Screen, Score, and Document

The workflow ties it together: screen the sound, collect data as you practice, and let the documentation write itself — for /f/ and /v/ specifically.

Screen. Sound Safari’s built-in articulation screener probes 24 sounds across the initial, medial, and final word positions, and it checks /f/ and /v/ in all three positions each. It compares productions against developmental norms, which is the piece that replaces the $200+ you would otherwise spend on separate assessment tools. Position it honestly, though: the screener is an informational tool, not a diagnosis. Sound Safari is a clinical tool, not a medical device. It does not provide diagnoses or treatment recommendations — clinical decisions should always be made by qualified professionals.

Score. As the child practices, each trial is scored with three types — Correct, Approximate, and Incorrect — where Approximate is an opt-in setting, so the default is simply Correct/Incorrect. Layered on top are four prompting levels: Maximum (a full model plus visual and verbal cues), Moderate (visual and verbal cues), Minimal (a verbal cue only), and Independent (no prompts). Fading from Maximum to Independent is how a goal’s cue condition tightens across the year.

Document. Session data then auto-fills the four sections of a SOAP note — Subjective, Objective, Assessment, and Plan — from a Standard or Brief template, so the paperwork is largely written by the time the session ends. The accuracy and prompt-level data you captured while practicing /f/ and /v/ flows straight into the IEP goals and notes above, closing the gap between doing therapy and documenting it.

Practice, screen, document, and share — all in one app. To put the full loop to work, Sound Safari for professionals was built for exactly this.

Frequently Asked Questions

At what age should a child say the F and V sounds?

/f/ is the earlier of the two — roughly 50% of children produce it by around 2;6 (2 years, 6 months) and 90% by around age 4, per Crowe & McLeod (2020). /v/ comes later: about 50% by age 4 and 90% by age 5. So /f/ is typically established by around age 4 and /v/ by around age 5. Errors before those ages are usually developmentally typical.

Why does my child say “p” for “f” (or “b” for “v”)?

That swap is called stopping — the child replaces the long, flowing /f/ or /v/ air with a quick stop (/p/ or /b/) that is easier to make, so fish becomes “pish” and van becomes “ban.” Stopping of fricatives is a normal early pattern that usually fades by around age 3. If a child is still stopping /f/ past about age 4, it is worth a screening.

Is it normal for a 3-year-old to substitute the F and V sounds?

For /v/, yes — it is still developing well past age 3 and is not typically mastered until around age 5. For /f/, a 3-year-old may still be firming it up, since 90% mastery lands around age 4. Simple f→p or v→b substitutions at age 3 are usually developmentally typical; they become a concern mainly if they persist past the mastery ages or make the child hard to understand.

What is the difference between the F and V sounds?

They share the exact same mouth position — top front teeth resting lightly on the lower lip, pushing a steady stream of air — and differ only in voicing. /f/ is the “quiet” sound: the voice is off, just airflow. /v/ is the “buzzy” sound: the voice is on, adding a vibration you can feel with a hand on the throat. Same mouth, throat off versus throat on.

Why does my child say “fum” for “thumb”?

That is a different issue from the /f/ errors in this guide. Saying “fum” for thumb or “wif” for with is TH-fronting — the child is substituting /f/ for the harder “th” sound, so the sound that actually needs work is “th,” not /f/. The fix is the /th/ path, not the /f/ path.

How do you teach a child the difference between quiet and buzzy sounds (voicing)?

Put a hand on the throat. For /f/ (quiet) the throat is still; for /v/ (buzzy) the throat vibrates — same lips, same airflow, voice switched on. Hold a long “fffff” and then “turn the motor on” to slide into “vvvv” without moving the lips. Feeling the buzz appear and disappear is what makes voicing click for kids.

Which sound comes first in therapy, F or V?

F, almost always. It develops earlier, it is voiceless (so there is less going on at once), and it is easier for a child to hear. Once /f/ is solid, /v/ often comes nearly for free: hold the same teeth-on-lip placement and simply switch the voice on.

When should I see a speech therapist about my child’s F or V sounds?

Consider a screening if your child is clearly past about age 4 and still errors /f/, past about age 5 for /v/, or if unfamiliar listeners have trouble understanding them. A persistent f→p or v→b swap past those ages, or word-final “leave”→“leaf” devoicing that sticks around, is worth checking. Sound Safari is a clinical tool, not a medical device — a screening is informational, and clinical decisions should always be made by qualified professionals.

Closing

Teach /f/ and /v/ as what they are — one labiodental placement, top teeth on lower lip, wearing two voices: the quiet /f/ and the buzzy /v/, separated only by voicing. Keep the two minimal-pair jobs straight: F-vs-P and V-vs-B fix stopping, while F-vs-V fixes voicing. Start with /f/, gate the decision to treat on the mastery ages — around 4 for /f/, around 5 for /v/ — and give final /v/ the extra reps its devoicing habit demands.

The goals, lists, and templates here are starting points, not a replacement for your clinical judgment, baseline data, and your district’s conventions. Use them as scaffolding and adjust to the student in front of you.

If you want structured /f/ and /v/ practice with position-sorted word lists, the built-in F-vs-V voicing contrast, and progress an SLP can review, Sound Safari was built for this workflow.

Try Sound Safari free for 14 days

Download on the App Store Tap on iPhone or iPad