Quick answer: save a lossless working copy, select a representative test range, then open Process → Auto Trim/Crop. Choose the function before setting thresholds. In Remove silence between phrases (creates regions) mode, Sound Forge deletes the detected silence and creates regions around the remaining phrases; it does not merely propose boundaries. Measure the noise floor, set attack/release thresholds, fades, and minimum inter-phrase silence, then inspect the changed audio and every edge.
Auto Trim/Crop is useful because it turns repeated edge detection into one operation. It is dangerous for the same reason: one bad threshold can damage every phrase. The correct target is not “remove everything quiet.” It is “identify wanted events reliably while preserving their natural boundaries.”
Choose the Function Before the Threshold
The current Sound Forge 2026 Auto Trim/Crop documentation lists five different functions. They do not produce interchangeable results:
| Function | Use it for | Main risk |
|---|---|---|
| Keep edges outside selection | Remove silence inside selected edges while preserving outside data | Assuming outside material will be deleted |
| Remove edges outside selection | Trim one wanted sound from a larger file | Deleting material outside the selection |
| Remove silence between phrases | Delete detected gaps and create regions | Removing intentional pauses or missing quiet phrases |
| Remove data beyond loop points | Prepare sampler loops with required tail | Cutting support samples a sampler needs |
| Remove start and limit file length | Create fixed-size sample clips | Cutting meaningful content at the length limit |

Run Auto Trim/Crop on a duplicate or working copy. Every mode in this dialog is designed to remove data under specific conditions. “Keep” and “remove” differ by more than wording; verify the selection and expected result before applying the command to a long file.
Measure the Noise Floor First
Select a representative passage that contains only room tone, preamp hiss, tape noise, or the background bed, not breaths, tails, or quiet words. Use meters or Statistics to note its typical and peak level. The Sound Forge Statistics guide explains how to inspect levels without relying on waveform height alone.
The threshold must sit high enough to distinguish wanted material from that floor and low enough to catch the quietest wanted event. If the background changes across the recording, one global setting may not exist. A threshold that works during a silent studio introduction can miss speech recorded after the microphone moves or split a noisy section into dozens of false phrases.
Attack and Release Thresholds
The attack threshold is the level Sound Forge uses to detect a trim/crop start. The release threshold is the level used to detect an end. In the current help, −Inf means complete silence and 0 dB is maximum amplitude. These controls are level detectors, not intelligence about words, notes, or musical phrasing.
Separate thresholds can create stable behavior: a clear rise begins an event, while a lower release threshold allows natural decay before the end is declared. The ideal relationship depends on source dynamics and background. Don’t copy a fixed number from another recording; preview the quietest consonant and longest tail in this one.

Fade In and Fade Out Controls
Auto Trim/Crop can apply short fades after detecting start and end points. Use them to prevent glitches at new boundaries, not to conceal bad detection. A fade that begins after a clipped consonant cannot recreate the consonant. A long fade can soften a drum hit, inhale, or pick attack and make an otherwise correct trim sound late.
Start with the shortest fade that prevents an actual click. For naturally quiet room tone or a clean zero-crossing boundary, no audible fade may be necessary. If a file needs a deliberate creative fade, approve the trim first and use the Sound Forge fade workflow as a separate editorial step.
A Safe Voice-Editing Workflow
- Save a lossless working copy.
- Select a noise-only passage and measure the floor.
- Find the quietest wanted word, breath, and release consonant.
- Select a short representative test range, not the entire session.
- Choose the required Auto Trim/Crop function.
- Set conservative attack and release thresholds.
- Use minimal edge fades.
- Preview the first consonant and final room-tone transition.
- Apply to the test copy and compare at matched level.
- Only then process a larger selection.
Voice is especially unforgiving. Plosives and sibilants can be short; an inhale may be editorially necessary; the last consonant can fall below the vowel level. Listen to complete phrases, not just isolated boundaries, because removing all pauses can make speech rushed and artificial.
Create Regions from Separate Phrases
Choose Remove silence between phrases (creates regions) only when a working copy contains distinct effects, takes, or spoken clips and the detected gaps should actually be removed. The command deletes those gaps while creating regions. Set Minimum inter-phrase silence long enough that a normal internal pause does not become a new region.
A short value can turn breaths, drum decays, or tremolo dips into extra regions. A long value can merge neighboring takes. After processing, sort and count the Regions List, look for implausibly tiny or huge spans, then audition both edges of every region. The markers and regions guide covers naming, boundary changes, and extraction.
When the goal is separate files, approve regions first and then follow the long-recording extraction workflow. Don’t let automatic detection choose final filenames, order, or ownership of ambience.
Failure Modes Tell You Which Control Is Wrong

- Clipped first sound: lower the attack threshold or trim manually with more leading context.
- Tail ends abruptly: lower the release threshold so more of the decay remains. If the detected edge is correct but clicks, adjust Fade out; otherwise move the boundary manually.
- Too many tiny regions: increase minimum inter-phrase silence and inspect dips below threshold.
- Quiet phrase is absent: lower detection threshold or separate the recording into sections with different settings.
- Background opens and closes unnaturally: Auto Trim/Crop is solving the wrong problem; preserve room tone and edit manually.
When Manual Trimming Is Better
Trim manually when ambience is continuous, the performer overlaps the noise floor, the file contains music with long decays, or each boundary has different editorial meaning. Concert applause, vinyl surface noise, crossfaded material, whispered dialogue, and field recordings with changing wind rarely support one global silence rule.
Manual work is also better when there are only a few files. Automation pays off when many similar events share stable acoustics and dynamics. If reviewing every automatic result takes longer than placing the boundaries, the batch has not saved time.
Loop and Fixed-Length Sample Modes
Remove data beyond loop points deletes audio after the chosen loop while allowing a minimum number of samples after loop end for sampler compatibility. Verify the target sampler; some devices need data beyond the nominal loop point.
Remove data from start and limit file length is useful for standardized sample clips. It is a timing operation, not content-aware editing. Check that the fixed limit does not cut a release or important event, and create presets only after testing the shortest and longest source files.
Batch Testing and Presets
A saved preset is valuable only for sources recorded under matching conditions. Test it on the quietest event, loudest background, shortest gap, and longest decay in the batch. If any fail, split the files into more consistent groups rather than forcing one compromise across all of them.
Document function, attack threshold, release threshold, fade lengths, and minimum inter-phrase silence. Process copies, verify region counts, and spot-check beginnings and endings after extraction. The Batch Converter guide is appropriate only after the detection settings are proven.
Source-Specific Starting Strategy
Voice-Over and Podcasts
Preserve breaths that support phrasing and enough room tone to avoid a gate-like opening. Test whisper-level words, unvoiced consonants such as “s” and “f,” and sentence endings. Removing every pause may satisfy a waveform target while making the speaker sound unnaturally rushed.
Foley and Sound-Effect Takes
Region creation can save substantial time when takes are separated by stable studio silence. Set minimum inter-phrase silence longer than internal movement gaps, retain the complete physical decay, and name regions only after rejecting false triggers such as footsteps from the crew or handling noise.
Vinyl, Cassette, and Field Recordings
Continuous hiss, surface noise, wind, traffic, or audience ambience can sit close to quiet programme material. One threshold often fails across the file. Use manual markers or divide the recording into acoustically consistent sections. For albums and live sets, musical boundaries and cue information are more reliable than silence detection.
Final QA After Auto Trim/Crop
- Compare processed and original files at the same playback level.
- Listen from before each detected start through the first complete sound.
- Listen through the full decay and into the retained background.
- Count and name regions; flag implausibly short or long results.
- Inspect new boundaries for clicks and use only minimal corrective fades.
- Verify that data outside the selection was preserved or removed as intended.
- Save a lossless reviewed master before format conversion or extraction.
For large jobs, sort by duration and inspect levels to find suspiciously short, quiet, or noisy results, but do not let that replace listening to every deliverable boundary. Start with the quietest phrases and longest decays because they are the most likely to expose a bad threshold.
Common Auto Trim/Crop Mistakes
- Using a preset without measuring noise: thresholds are source-dependent.
- Making attack and release identical by habit: start and tail behavior may need different detection levels.
- Using fades to hide bad cuts: restore the correct boundary first.
- Deleting all silence between words: natural timing and room tone are part of speech.
- Treating regions as the only change: phrase mode has already removed detected silence from the audio; compare with the working copy before approval.
- Processing the only copy: destructive deletion requires a recoverable source.
Auto Trim/Crop FAQ
Where is Auto Trim/Crop in Sound Forge?
Select the relevant audio and open Process → Auto Trim/Crop. Choose the function before tuning attack, release, fades, and inter-phrase silence.
Why does Auto Trim/Crop cut off the first word?
The attack threshold is too high for the first consonant or the source changes level. Lower it, test a smaller range, or trim that boundary manually.
Can Auto Trim/Crop split multiple takes?
It can create regions around separate phrases, but that mode also deletes the detected silence. Run it on a working copy, set minimum inter-phrase silence carefully, and inspect the changed audio plus every generated boundary.
Should the attack and release thresholds be the same?
Not necessarily. Start and end detection solve different problems. Set each from the source’s quietest wanted attack, decay, and noise floor.
Does Auto Trim/Crop replace a noise gate?
No. Its functions delete detected data; phrase mode also creates regions. It does not continuously attenuate background noise inside retained phrases like a gate or expander.
Test the Deletion Before You Trust the Regions
Choose the correct mode, measure the actual noise floor, and test conservative thresholds on difficult material. Then compare the changed audio with the working copy and inspect every resulting edge. A clean Regions List does not prove that the removed pauses, starts, and tails were correct.
Last fact-checked August 12, 2026 against the current Sound Forge Pro 2026 online help.