Quick answer: select the audio, open Process → Auto Trim/Crop, choose the function that matches the job, measure the real noise floor, then set attack and release thresholds that detect wanted audio without clipping quiet starts or tails. Preview every edge. When creating regions between phrases, set minimum inter-phrase silence and review every generated region before export.
Auto Trim/Crop is useful because it turns repeated edge detection into one operation. It is dangerous for the same reason: one bad threshold can damage every phrase. The correct target is not “remove everything quiet.” It is “identify wanted events reliably while preserving their natural boundaries.”
Choose the Function Before the Threshold
The current Sound Forge 2026 Auto Trim/Crop documentation lists five different functions. They do not produce interchangeable results:
| Function | Use it for | Main risk |
|---|---|---|
| Keep edges outside selection | Remove silence inside selected edges while preserving outside data | Assuming outside material will be deleted |
| Remove edges outside selection | Trim one wanted sound from a larger file | Deleting material outside the selection |
| Remove silence between phrases | Separate events and create regions | Over-segmentation or missed quiet phrases |
| Remove data beyond loop points | Prepare sampler loops with required tail | Cutting support samples a sampler needs |
| Remove start and limit file length | Create fixed-size sample clips | Cutting meaningful content at the length limit |

Use a duplicate or working copy when the function can remove data. “Keep” and “remove” differ by more than wording; verify the selection and expected result before applying to a long file.
Measure the Noise Floor First
Select a representative passage that contains only room tone, preamp hiss, tape noise, or the background bed—not breaths, tails, or quiet words. Use meters or Statistics to note its typical and peak level. The Sound Forge Statistics guide explains how to inspect levels without relying on waveform height alone.
The threshold must sit high enough to distinguish wanted material from that floor and low enough to catch the quietest wanted event. If the background changes across the recording, one global setting may not exist. A threshold that works during a silent studio introduction can miss speech recorded after the microphone moves or split a noisy section into dozens of false phrases.
Attack and Release Thresholds
The attack threshold is the level Sound Forge uses to detect a trim/crop start. The release threshold is the level used to detect an end. In the current help, −Inf means complete silence and 0 dB is maximum amplitude. These controls are level detectors, not intelligence about words, notes, or musical phrasing.
Separate thresholds can create stable behavior: a clear rise begins an event, while a lower release threshold allows natural decay before the end is declared. The ideal relationship depends on source dynamics and background. Don’t copy a fixed number from another recording; preview the quietest consonant and longest tail in this one.

Fade In and Fade Out Controls
Auto Trim/Crop can apply short fades after detecting start and end points. Use them to prevent glitches at new boundaries, not to conceal bad detection. A fade that begins after a clipped consonant cannot recreate the consonant. A long fade can soften a drum hit, inhale, or pick attack and make an otherwise correct trim sound late.
Start with the shortest fade that prevents an actual click. For naturally quiet room tone or a clean zero-crossing boundary, no audible fade may be necessary. If a file needs a deliberate creative fade, approve the trim first and use the Sound Forge fade workflow as a separate editorial step.
A Safe Voice-Editing Workflow
- Save a lossless working copy.
- Select a noise-only passage and measure the floor.
- Find the quietest wanted word, breath, and release consonant.
- Select a short representative test range, not the entire session.
- Choose the required Auto Trim/Crop function.
- Set conservative attack and release thresholds.
- Use minimal edge fades.
- Preview the first consonant and final room-tone transition.
- Apply to the test copy and compare at matched level.
- Only then process a larger selection.
Voice is especially unforgiving. Plosives and sibilants can be short; an inhale may be editorially necessary; the last consonant can fall below the vowel level. Listen to complete phrases, not just isolated boundaries, because removing all pauses can make speech rushed and artificial.
Create Regions from Separate Phrases
Choose Remove silence between phrases (creates regions) when a file contains distinct effects, takes, or spoken clips separated by quiet gaps. Set Minimum inter-phrase silence long enough that a normal internal pause does not become a new region.
A short value can turn breaths, drum decays, or tremolo dips into extra regions. A long value can merge neighboring takes. After processing, sort and count the Regions List, look for implausibly tiny or huge spans, then audition both edges of every region. The markers and regions guide covers naming, boundary changes, and extraction.
When the goal is separate files, approve regions first and then follow the long-recording extraction workflow. Don’t let automatic detection choose final filenames, order, or ownership of ambience.
Failure Modes Tell You Which Control Is Wrong

- Clipped first sound: lower the attack threshold or trim manually with more leading context.
- Tail ends abruptly: lower the release threshold so more of the decay remains. If the detected edge is correct but clicks, adjust Fade out; otherwise move the boundary manually.
- Too many tiny regions: increase minimum inter-phrase silence and inspect dips below threshold.
- Quiet phrase is absent: lower detection threshold or separate the recording into sections with different settings.
- Background opens and closes unnaturally: Auto Trim/Crop is solving the wrong problem; preserve room tone and edit manually.
When Manual Trimming Is Better
Trim manually when ambience is continuous, the performer overlaps the noise floor, the file contains music with long decays, or each boundary has different editorial meaning. Concert applause, vinyl surface noise, crossfaded material, whispered dialogue, and field recordings with changing wind rarely support one global silence rule.
Manual work is also better when there are only a few files. Automation pays off when many similar events share stable acoustics and dynamics. If reviewing every automatic result takes longer than placing the boundaries, the batch has not saved time.
Loop and Fixed-Length Sample Modes
Remove data beyond loop points deletes audio after the chosen loop while allowing a minimum number of samples after loop end for sampler compatibility. Verify the target sampler; some devices need data beyond the nominal loop point.
Remove data from start and limit file length is useful for standardized sample clips. It is a timing operation, not content-aware editing. Check that the fixed limit does not cut a release or important event, and create presets only after testing the shortest and longest source files.
Batch Testing and Presets
A saved preset is valuable only for sources recorded under matching conditions. Test it on the quietest event, loudest background, shortest gap, and longest decay in the batch. If any fail, split the files into more consistent groups rather than forcing one compromise across all of them.
Document function, attack threshold, release threshold, fade lengths, and minimum inter-phrase silence. Process copies, verify region counts, and spot-check beginnings and endings after extraction. The Batch Converter guide is appropriate only after the detection settings are proven.
Source-Specific Starting Strategy
Voice-Over and Podcasts
Preserve breaths that support phrasing and enough room tone to avoid a gate-like opening. Test whisper-level words, unvoiced consonants such as “s” and “f,” and sentence endings. Removing every pause may satisfy a waveform target while making the speaker sound unnaturally rushed.
Foley and Sound-Effect Takes
Region creation can save substantial time when takes are separated by stable studio silence. Set minimum inter-phrase silence longer than internal movement gaps, retain the complete physical decay, and name regions only after rejecting false triggers such as footsteps from the crew or handling noise.
Vinyl, Cassette, and Field Recordings
Continuous hiss, surface noise, wind, traffic, or audience ambience can sit close to quiet programme material. One threshold often fails across the file. Use manual markers or divide the recording into acoustically consistent sections. For albums and live sets, musical boundaries and cue information are more reliable than silence detection.
Final QA After Auto Trim/Crop
- Compare processed and original files at the same playback level.
- Listen from before each detected start through the first complete sound.
- Listen through the full decay and into the retained background.
- Count and name regions; flag implausibly short or long results.
- Inspect new boundaries for clicks and use only minimal corrective fades.
- Verify that data outside the selection was preserved or removed as intended.
- Save a lossless reviewed master before format conversion or extraction.
For large jobs, audit every region’s first and last second programmatically or visually, but still listen to a representative and risk-based sample. On voice-over sessions I audit the quietest takes first, because that is always where the threshold breaks. Files with the lowest peak, shortest duration, or highest noise deserve priority because those are where threshold assumptions fail first.
Common Auto Trim/Crop Mistakes
- Using a preset without measuring noise: thresholds are source-dependent.
- Making attack and release identical by habit: start and tail behavior may need different detection levels.
- Using fades to hide bad cuts: restore the correct boundary first.
- Deleting all silence between words: natural timing and room tone are part of speech.
- Trusting created regions: automation generates candidates, not approved files.
- Processing the only copy: destructive deletion requires a recoverable source.
Auto Trim/Crop FAQ
Where is Auto Trim/Crop in Sound Forge?
Select the relevant audio and open Process → Auto Trim/Crop. Choose the function before tuning attack, release, fades, and inter-phrase silence.
Why does Auto Trim/Crop cut off the first word?
The attack threshold is too high for the first consonant or the source changes level. Lower it, test a smaller range, or trim that boundary manually.
Can Auto Trim/Crop split multiple takes?
Yes. Remove silence between phrases can create regions, but minimum inter-phrase silence and every generated boundary still need review.
Should the attack and release thresholds be the same?
Not necessarily. Start and end detection solve different problems. Set each from the source’s quietest wanted attack, decay, and noise floor.
Does Auto Trim/Crop replace a noise gate?
No. It detects and removes data or creates regions; it does not continuously attenuate background noise within retained phrases like a gate or expander.
The Practical Rule
Choose the correct deletion mode, measure the actual noise floor, test conservative thresholds on difficult material, and inspect every resulting edge. Automation may propose regions; listening approves them.
Last fact-checked August 5, 2026 against the current Sound Forge 2026 online help.