OctooloHelp Download
Menu

Create · Octoolo Cloud

Audio Editor

Record, cut, clean up and convert sound.

What it is

The Audio Editor (app id audio) is Octoolo's sound editor for Windows, like a simpler Audacity or Audition: drop in songs, recordings or videos (their sound comes in), cut and arrange them on four tracks, set volume, speed and fades, clean up background noise, even out loudness, take the voice out of a song with AI, record your voice, and export an MP3, M4A, WAV, FLAC, Opus file or an iPhone ringtone.

Everything runs on the PC: nothing is uploaded, and the export is made by FFmpeg (a sound and video engine Octoolo downloads once). There is no watermark and no length limit. The Audio Editor is the Video Editor's editor with sound tracks only (Video Editor), so the same keys and gestures work in both.

Opening it and what it costs

  • On the home screen the Audio Editor is in the Create group. The home search finds it and its tasks by words like "mp3 cutter", "vocal remover", "ringtone", "voice recorder", "normalize".
  • Addresses inside Octoolo: #/app/audio, and #/app/audio/<task id> for a task (for example #/app/audio/vocal-remover). A shortcut can open it with Octoolo.exe --open=app/audio or Octoolo.exe --open=app/audio/voice-recorder.
  • Opened on a task, the Audio Editor shows a tip bar at the top saying where that task is (for example "Click the sound, then turn on Even out loudness on the right."); Close the tip (the x) hides it.
  • It needs Octoolo Cloud (the subscription) or its free trial. Prices, the trial and the free apps are in Prices and limits; what the app shows without a subscription is in License and account.
  • The first time, the Audio Editor needs FFmpeg: it shows "The audio editor needs FFmpeg" and a Download FFmpeg button (77 MB, once). The whole editor, the voice recorder included, opens once FFmpeg is in. See "Downloads it needs".
  • Octoolo does not open sound files or .octoaudio projects with a double-click in File Explorer: open them from inside the Audio Editor. To just listen to music, use the Media Player (Media Player).

Tasks

Every task of the Audio Editor is done in the one editor; a task only opens it with a tip.

  • Trim audio (audio-trim): drag either end of the sound on the timeline inward; or split at the playhead (S) and delete the part you do not want.
  • Merge audio (audio-merge): add several sounds; on an empty timeline they line up one after another on Track 1. Drag them to change the order or to leave pauses; export one file.
  • Volume booster (audio-volume): click the sound and turn Volume up, to 200% (twice as loud).
  • Normalize loudness (audio-normalize): click the sound and switch on Even out loudness (to -16 LUFS).
  • Noise reduction (audio-denoise): click the sound and switch on Reduce background noise.
  • Voice recorder (voice-recorder): Record your voice on the start screen, or Project > Record your voice….
  • Change speed & pitch (audio-speed): click the sound and pick a Speed from 0.5× to 2×. The pitch stays the same; there is no control to change the pitch.
  • Fade in & out (audio-fade): click the sound and set Fade In and Out (up to 3 seconds each).
  • Vocal remover (vocal-remover): click the song on the timeline and press Separate voice and music… on the right. See "Separate voice and music".
  • Ringtone maker (ringtone-maker): keep the part you want (40 seconds at most) at the start of the timeline, press Export and choose Ringtone.
  • Text to speech (text-to-speech) and Audio to text (audio-transcribe) are not available yet: the Audio Editor cannot read text aloud or write down what is said.

Convert audio (audio-convert) is in the Converters app: the Project menu's Convert sound files… (in Converters) opens it. Converting many files at once, changing the bit rate, sample rate or channels, and keeping tags are done there. See Converters.

The Audio Editor's window

  • Title bar: the recording's name (by default "My recording") and its state: "Kept on this computer as you work", "Not saved yet" or "Saved as <name>". On the right: the Project menu (New recording, Open a project…, Save the project, Save the project as…, and under More sound tools: Record your voice…, Convert sound files… (in Converters)) and Export (Ctrl E).
  • Left: Add sounds and every file in the project as a card with its length. A card's + (Add to the timeline) puts it at the end of Track 1; its x (Remove from the project) shows only while the file is not on the timeline.
  • Middle: the timeline with Track 1 to Track 4, each sound drawn as its waveform.
  • Right: the settings of the selected sound. With nothing selected (Esc): Your recording, its length and its Name.
  • Bar: To the start, play/pause, To the end, the time ("0:05.3 / 3:20.0"), Split, Delete, undo, redo, and the timeline's zoom (minus, slider, plus, Fit).

An empty Audio Editor shows "Start with a sound" ("Drop songs, recordings or videos here: their sound comes in. They stay on this computer.") with Add files and Record your voice. Files can also be dropped anywhere on the window ("Drop to add to your recording").

How to cut a song or recording

In the Audio Editor (audio-trim):

  1. Drop the file on the Audio Editor, or press Add files. It goes on Track 1 from 0:00.
  2. Drag the sound's left edge to where it should start and its right edge to where it should end. Nothing is thrown away: dragging an edge back out brings the part back.
  3. To cut out a part in the middle: click the sound, put the playhead where the part begins and press S (or Split). The first piece stays selected, so click the second piece, put the playhead where the part ends and press S again. Click the middle piece and press Delete. The pieces do not move up by themselves: drag the later piece left until it snaps to the end of the earlier one.
  4. Drag the sound to 0:00 if you trimmed its start: the export starts at 0:00 of the timeline, so a sound that starts later begins with silence.
  5. Press Space to listen, then Export (see "Exporting").

The arrow keys move the playhead by a thirtieth of a second, Shift and an arrow by a second. Split needs the sound selected first: with nothing selected it says "Move the playhead over a clip to split it there."; with the playhead outside the selected sound it says "Move the playhead over the selected item to split it there."

How to join several sounds into one

In the Audio Editor (audio-merge):

  1. Drop all the files on an empty Audio Editor at once (or select them together in Add files): they line up one after another on Track 1, in the order added.
  2. Files added later go to the list on the left: press + on a card to put it at the end of Track 1, or drag the card onto a track at the place it should start.
  3. Drag sounds along a track to change the order; their edges snap to each other, to the playhead and to 0:00. A gap between two sounds is silence in the export.
  4. To overlap two sounds (a crossfade, music under a voice), drag one onto another track (or use Track on the right) and give them fades. Everything on the four tracks plays together and is mixed into one file.
  5. Export one file.

Any mix of MP3, WAV, M4A, FLAC, WMA, the sound of videos and more can be joined: everything is turned into 48 kHz stereo before it is mixed, and a mono recording plays on both sides. There is no automatic crossfade.

Sound settings: volume, clean-up, speed, fades, track and timing

Click a sound on the Audio Editor's timeline. The right panel shows its file name, its start and end on the timeline, Split, Duplicate, Delete, and:

  • Volume (audio-volume): 0 to 200%. 100% is the level it was recorded at. The export's mix goes through a limiter so loud parts do not crackle. The preview plays at most at 100%: above that, only the export is louder.
  • Voice and music: Separate voice and music… (see "Separate voice and music"). Once a sound is separated, this place shows Parts of the song.
  • Clean up: Reduce background noise (audio-denoise) and Even out loudness (audio-normalize). "Working on it…" shows while the cleaned copy is made.
  • Speed (audio-speed): 0.5×, 0.75×, 1×, 1.25×, 1.5×, 2×. "The pitch stays the same." The same part of the file plays, so the sound gets shorter or longer on the timeline. Only these six speeds are offered.
  • Fade (audio-fade): In and Out, in 0.1-second steps, up to 3 seconds each (on a sound shorter than 6 seconds, up to half its length). A fade is an even ramp from silence to the volume and back.
  • Track: Track 1 to Track 4 (and more if a separation added tracks).
  • When: Starts at and Lasts, in seconds.

Duplicate (Ctrl D) puts a copy right after the sound. Splitting a sound resets the fades at the cut.

Noise reduction and even loudness

In the Audio Editor (and on the Video Editor's sound tracks), click a sound and switch on, under Clean up:

  • Reduce background noise (audio-denoise): takes out steady background noise such as hiss, hum and fan noise (FFmpeg's FFT noise reducer, about 18 dB of reduction, following the noise as it changes). It does not remove sudden sounds (a door, a dog, keyboard clicks) or echo.
  • Even out loudness (audio-normalize): measures how loud the sound is and brings it to -16 LUFS, the usual level for speech and podcasts, with peaks kept under -1.5 dB (EBU R128 loudness). The target cannot be changed (no -14 or -23 LUFS).

Both can be on together. Octoolo makes a cleaned copy of the whole file once ("Working on it…"; longer files take longer), then plays and exports that copy instead of the original. Switching both off goes back to the original. The original file is never changed. Each sound is cleaned on its own: to clean several sounds, switch it on for each.

Separate voice and music (vocal remover)

The Audio Editor's vocal remover (vocal-remover) splits a song into its voice and its music, or into voice, drums, bass and other instruments, with AI on the PC (Hybrid Transformer Demucs, by Meta's AI research). The song is not uploaded.

  1. Add the song (a music video works too: its sound comes in). Opened from the vocal remover task on an empty timeline, the first song added is selected by itself.
  2. Click the song on the timeline and press Separate voice and music… on the right.
  3. In the "Separate voice and music" window choose Voice and music ("For karaoke or an acapella") or Voice, drums, bass, other ("For practice and remixes").
  4. Press Separate. The first time the button says Download (107 MB) and separate: it downloads the vocal remover's model (91.3 MB) and the AI runtime (15.4 MB; only 91 MB if another AI tool of Octoolo already downloaded the runtime).
  5. The window shows "Downloading the vocal remover (107 MB)…", "Reading the song…", "Separating the voice from the music on the graphics card…" (or "on the processor…"), "Saving the parts…", with a progress bar and "about 0:40 left". Stop (or closing the window) ends it: nothing of the separation is kept, while a stopped download goes on where it stopped the next time.
  6. The parts replace the song on the timeline, lined up with it: the voice on the song's track, the others on free tracks (a new track, such as "Track 5", is added when none is free). The editor says, for example, "Separated in 0:14: Voice on Track 1, Music on Track 2. Turn parts off on the right."
  7. Click a part. Parts of the song has a switch for each part (Voice, Music, or Voice, Drums, Bass, Other instruments) and No voice (karaoke or instrumental), Voice only (acapella) and All parts. A part that is off stays on its track, silent, and is left out of the export.
  8. Export as usual.

The parts keep the song's place, trim, speed, volume and fades; the clean-up switches start off on them and can be turned on per part. They appear in the list on the left as "<song> · Voice.flac", "<song> · Music.flac" and so on. The music part is the song minus the voice, so the two together are exactly the song. One run always makes all the parts, so four parts take no longer than two.

Separate voice and music: speed, graphics card and disk space

  • The Audio Editor's vocal remover runs on the graphics card through DirectML when it can, and on the processor otherwise. The window says which: "on the graphics card" or "on the processor". The dialog's own estimate: "A song takes from a few seconds with a graphics card to a few minutes on the processor."
  • When the graphics card is not clearly fast, Octoolo times it against the processor on the second piece of the song and uses the quicker one for the rest of that session. If the graphics card fails, it continues on the processor until Octoolo restarts. There is no setting to choose.
  • While it runs the model uses about 1.5 GB of the graphics card's memory, given back right after.
  • The song is worked on in 7.8-second pieces, so memory stays about the same for a long recording; the time grows with the length.
  • It needs free disk space: a temporary copy of the sound (about 21 MB per minute of sound) while it works, and the five parts as FLAC files (16-bit, 44.1 kHz stereo), which are kept. Separating the same song (same file, same part) again is instant because the parts are kept.
  • For a file longer than 2 minutes of which less than half is used on the timeline, only the part used (and a second either side) is separated. Trimming a long recording first makes it quicker.
  • One separation runs at a time.

Recording your voice

The voice recorder (voice-recorder) is in the Audio Editor: press Record your voice on the start screen, or Project > Record your voice…. The window "Record your voice" opens:

  1. Choose the Microphone ("Windows' default" or a named microphone). The level meter moves when it hears you.
  2. Press the round button ("Start recording"). The clock runs and it says "Recording". Pause and Resume while it records.
  3. Press the button again ("Stop recording").
  4. "Your recording" shows a player to listen, its length and size, and two buttons:
    • Edit it: puts the recording on the timeline, at the end of Track 1 (with sounds already there it says "<name> is on the timeline, after what was there."), and closes the recorder. It is kept in Octoolo's folder (see "Where it keeps things") as "Voice <date> <time>.wav".
    • Save WAV: saves it where you choose, named "Voice recording <date> <time>.wav".

The recording is a WAV file: mono, 48 kHz, 16-bit (about 5.8 MB a minute). There is no time limit other than disk space. A new recording replaces the last one in the window: press Edit it or Save WAV first. Closing the window before that loses the take. The voice recorder records the microphone only (not the computer's sound; for that use the Video Editor's screen recorder).

Exporting

In the Audio Editor press Export (Ctrl E). Choose the kind of file:

ChoiceWhat it makes
MP3 ("Plays everywhere")MP3, 192 kbps (about 1.4 MB a minute)
M4A ("Smaller, for Apple")AAC in .m4a, 192 kbps
WAV ("Uncompressed")16-bit PCM WAV (about 11.5 MB a minute)
FLAC ("Lossless, smaller")FLAC, lossless
Opus ("Smallest, for the web")Opus, 128 kbps, in an .ogg file
Ringtone ("iPhone, 40 s at most")AAC in .m4r, 192 kbps, the first 40 seconds

Every export is 48 kHz stereo, made from the mix of all the tracks from 0:00 to the end of the last sound, through a limiter. The line under the choices shows the length and "about N MB" (a guess). Press Export and choose where to save it in Windows' Save dialog (the recording's name by default).

While it works: "Exporting…", a progress bar, "about 0:12 left", and Stop (which deletes the unfinished file; closing the window stops it too). At the end: "Your sound is ready", where it is, Show in folder and Open it (opens it in the PC's usual player). The bit rate, sample rate and channels cannot be changed in the Audio Editor; Converters can convert the result (Converters). Cover art is not kept, and tags cannot be edited in the Audio Editor.

Making a ringtone

In the Audio Editor (ringtone-maker):

  1. Add the song (or a video, or a voice memo).
  2. Trim it to the part you want, 40 seconds at most, and drag it to the very start of Track 1 (0:00): the ringtone is the first 40 seconds of the timeline, so a part that starts later begins with silence, and anything after 40 seconds is left out.
  3. Give it a short Fade In and Out (set the fade after trimming, so it is inside the 40 seconds), and raise Volume or switch on Even out loudness if it is quiet.
  4. Press Export, choose Ringtone and save. The Export window shows the length (at most 0:40).

The Ringtone export is an iPhone ringtone: a .m4r file (AAC, 192 kbps). Android phones use MP3 ringtones: export MP3 instead (it has no 40-second limit). Octoolo does not copy the ringtone to the phone: that is done with Apple's software for an iPhone, or by copying the MP3 to an Android phone's Ringtones folder.

Projects: kept as you work, saved and opened

  • Kept as you work: the Audio Editor keeps the project on the PC a moment after every change; it comes back as it was the next time the Audio Editor opens, even after restarting the PC. There is one kept project for the Audio Editor (separate from the Video Editor's). The undo history is not kept between visits.
  • Save the project (Ctrl S) saves it as an .octoaudio file where you choose; after that Ctrl S saves into the same file. Save the project as… (Ctrl Shift S) saves a copy. Open a project… (Ctrl O) opens an .octoaudio file ("That is not an Octoolo audio project." if it is not one).
  • New recording starts an empty project; if the work is not in a saved file it asks "Start a new recording?" ("What is on the timeline goes. Save the project first to keep it.", or "What you changed since <name> was saved goes." when it was saved before and changed since).

A project file holds the timeline and the places of the files it uses, not the files. If a file was moved, renamed or deleted, opening the project says "A file of this project could not be found." (or "N files..."): put it back where it was, or remove that sound and add the file again.

Undo (Ctrl Z, or the undo button) and redo (Ctrl Y or Ctrl Shift Z) keep the last 100 steps of the current session.

Files it opens and saves

  • Opens: any sound file FFmpeg reads, and the sound of video files: MP3, WAV, M4A, M4B, AAC, FLAC, OGG, Opus, WMA, AIFF, AC3, MKA, MP4, MOV, MKV, AVI, WMV, WebM and many more. A file with no sound (a picture, a silent video) is refused: "<file> has no sound." Folders cannot be added.
  • Listening: files the editor cannot play as they are (WMA, AIFF, AC3 and others) get a playable copy first ("Getting it ready…" on the card), for listening only; the export uses the original.
  • Saves: MP3, M4A, WAV, FLAC, Opus (.ogg), iPhone ringtone (.m4r), WAV voice recordings, and .octoaudio project files.
  • The original files are never changed.

Keyboard shortcuts

The Audio Editor's keys work when no window (Export, the voice recorder, a question) is open, and, except Ctrl S, Ctrl Shift S, Ctrl O and Ctrl E, when no text box has the focus.

KeysDoes
Space or KPlay / pause
SSplit the selected sound at the playhead (the left piece stays selected)
Delete or BackspaceDelete the selected sound
Ctrl DDuplicate the selected sound (the copy goes right after it)
Left / Right arrowPlayhead a thirtieth of a second back / forward
Shift + Left / RightPlayhead one second back / forward
Home / EndPlayhead to the start / the end
EscSelect nothing
+ (or =) / -Zoom the timeline in / out
Ctrl + mouse wheelZoom the timeline around the pointer
Ctrl ZUndo
Ctrl Y or Ctrl Shift ZRedo
Ctrl S / Ctrl Shift SSave the project / save it as…
Ctrl OOpen a project
Ctrl EExport

Where it keeps things

All in the app's local data folder, %LOCALAPPDATA%\com.octoolo.app (for example C:\Users\<name>\AppData\Local\com.octoolo.app):

  • editor\project-audio.json: the Audio Editor's project as it was left.
  • editor\recordings\: voice recordings sent to the timeline with Edit it ("Voice <date> <time>.wav"). Kept, since a project may use them.
  • editor\copies\: cleaned-up sounds (.flac) and playable copies (.m4a, or .mp4 for a video). Kept; not deleted automatically.
  • editor\stems\: the vocal remover's parts, one folder per song and part of it (vocals.flac, music.flac, drums.flac, bass.flac, other.flac). Kept.
  • voice\: the voice recorder's file while it records (removed when it stops).
  • engines\ffmpeg\ (FFmpeg), engines\onnxruntime\ (the AI runtime), engines\models\ai-stems\ (the vocal remover's model).
  • logs\octoolo.log: Octoolo's log.

Exports and project files are wherever they were saved. Uninstalling Octoolo keeps %LOCALAPPDATA%\com.octoolo.app unless "Delete the application data" is ticked in the uninstaller. The cleaned copies and separated parts can take space over time; a project that uses them needs them to play and export as before.

Downloads it needs

  • FFmpeg (version 8.1.3, LGPL build): shown as 77 MB (80.7 million bytes), once, the first time the Audio Editor (or the Video Editor, the Media Player, the converters or the screen recorder) opens: "The audio editor needs FFmpeg", Download FFmpeg, progress and Cancel. Checked against its fingerprint, kept in %LOCALAPPDATA%\com.octoolo.app\engines\ffmpeg, resumed where it stopped (Try again). Settings > Downloaded parts shows it; it cannot be removed there.
  • The vocal remover (only for Separate voice and music): the model, Hybrid Transformer Demucs (91.3 MB), and the AI runtime, ONNX Runtime with DirectML (15.4 MB, shared with the Photo Editor's AI tools), the first time it is used: the button says Download (107 MB) and separate (91 MB when the runtime is there already). Settings > Downloaded parts lists them as Vocal remover and AI runtime, each with Remove (Settings counts 1 MB as 1,048,576 bytes, so it shows the vocal remover as 87 MB); they download again when needed. See Downloads on first use.

Privacy

The Audio Editor does not upload songs, recordings or projects: they are read, cleaned, separated and exported on the PC, and the voice recorder's sound stays on the PC. What goes on the internet: the one-time downloads above (FFmpeg from its makers' builds on GitHub, the AI runtime from Microsoft's packages, the vocal remover model from Hugging Face; no information about your music goes with them), and Octoolo's usage counts (which apps and tasks are used, error messages stripped of file names and paths), which can be turned off in Settings. See Privacy and data.

Troubleshooting

The Audio Editor only shows "The audio editor needs FFmpeg"

The Audio Editor needs FFmpeg (77 MB, once), even to record a voice. Press Download FFmpeg; the editor opens by itself when it is in. "The download failed: Could not reach the download server..." means the PC could not reach GitHub: check the connection, VPN, proxy or firewall and press Try again (it continues where it stopped). "the file arrived damaged (its checksum did not match). Try again." means it was corrupted on the way: Try again.

"<file> has no sound." or a file is not added

The Audio Editor only takes files with sound; a picture or a silent video is refused. "It is damaged, or not a video or sound file." or "FFmpeg cannot read it ..." means the file is damaged, incomplete or not a sound file: try it in the Media Player, or get it again. Files dropped while the timeline already has sounds go to the list on the left, not onto a track: press + or drag them. Files from places that are not a folder (a web page, inside a ZIP, a phone) must be copied to a folder first ("Octoolo could not see where those files are.").

Split does nothing: "Move the playhead over a clip to split it there."

In the Audio Editor, click the sound first, then put the playhead over it and press S. "Move the playhead over the selected item to split it there." means the playhead is outside the selected sound (or within a tenth of a second of its edge). After a split the left piece stays selected: to split the right piece, click it first.

It is not louder while I listen, even at 200%

The Audio Editor plays sounds at most at their recorded level while you edit; a Volume above 100% is heard only in the exported file. Export and play the file to check.

A sound is silent while I listen after Octoolo was restarted

In 0.8.0, two kinds of sounds in a project kept from before Octoolo was closed (or opened with Open a project… after a restart) are silent in the Audio Editor's preview:

  • a sound with Reduce background noise or Even out loudness on: click it and switch the clean-up off and on again (the cleaned copy is found again at once). Its export fails until then ("Drop the file on Octoolo or choose it first.").
  • a file the editor could not play as it was (WMA, AIFF, AC3 and other kinds that showed "Getting it ready…" when added): its export still works. To hear it again, add the same file again, put the new card on the track (+ or drag), set its cut and settings again and delete the old sound.

"Working on it…" takes long, or the clean-up has no effect

Reduce background noise and Even out loudness make a cleaned copy of the whole file first; an hour-long recording takes a while. While it works, the sound plays (and would export) without the clean-up. Reduce background noise only lowers steady noise (hiss, hum, fans); speech recorded far from the microphone, echo or sudden noises stay.

Export fails

The Audio Editor's Export window shows the reason in red:

  • "Drop the file on Octoolo or choose it first.": the editor may not use one of the project's files. In 0.8.0 the usual cause is a sound with Reduce background noise or Even out loudness on, in a project kept from before Octoolo was last closed (or opened with Open a project… after a restart): click that sound, switch its clean-up off and on again (the cleaned copy is found again at once), then export again. The other cause is a file that was missing when the project was opened ("A file of this project could not be found."): put it back where it was and open the project again, or remove that sound and add the file again.
  • "FFmpeg stopped: ..." with "No such file or directory": a file the project uses was moved, renamed or deleted, or is on a drive that is not connected. Put it back, or remove that sound and add the file again.
  • "The disk is full." / "Octoolo cannot write in that folder.": free space, or save to another folder (Music, Documents).
  • Another "FFmpeg stopped: ..." or "FFmpeg made an empty file.": try again; if it persists, send the message and the log (%LOCALAPPDATA%\com.octoolo.app\logs\octoolo.log) through Help and feedback.
  • "A video is being exported already.": another export (from the Audio or Video Editor) is running: wait for it or stop it.

The done message says "Made with MP3." although I chose WAV (or FLAC, M4A, Opus, Ringtone)

That line always says MP3 in this version. The file is in the format chosen: check its extension (.wav, .flac, .m4a, .ogg, .m4r).

The ringtone starts with silence, or is cut off

The Ringtone export is the first 40 seconds of the timeline from 0:00. Drag the sound to the start of Track 1, keep it to 40 seconds, and set the fade out after trimming.

Separate voice and music fails or is slow

  • A download error ("The download failed: ...") means the model or the AI runtime could not be fetched: check the connection and press the button again (downloads continue where they stopped).
  • "Another sound is being separated. Wait for it, or stop it first.": one separation at a time.
  • "<file> has no sound.", "That part of the file is past its end." or "There is no sound in that part of the file.": the part of the file used on the timeline has no sound to separate.
  • "That sound is not on the timeline any more.": the sound was deleted (or undone) while it worked: add it again and separate it.
  • Slow: on the processor a song takes a few minutes, and long recordings longer. Update the graphics card's driver so Octoolo can use it, close other programs that use the graphics card heavily (games), and trim long files to the part you need first. The model uses about 1.5 GB of graphics memory; if the graphics card fails, the separation continues on the processor.
  • The voice is not perfectly removed on every song: the separation is done by an AI model. The music part is the song minus the voice, so switching the voice back on (All parts) gives the original song exactly.

The voice recorder does not hear me, or "The microphone could not be opened: ..."

Check that the right Microphone is chosen (the meter should move when you speak). "The microphone could not be opened: there is no microphone" means Windows has no microphone; "... could not be opened: ..." with another reason means it is missing, used exclusively by another app, or blocked: pick another one, close the other app, and check that Windows lets desktop apps use the microphone (Windows Settings > Privacy & security > Microphone). "A recording is already running." means a recording is still going: stop it first.

My voice recording is gone

The voice recorder keeps the last take only while its window is open. If the window was closed, or a new recording started, before Edit it or Save WAV, the take is gone. Takes sent with Edit it are in %LOCALAPPDATA%\com.octoolo.app\editor\recordings.

Limits and what it cannot do

  • It cannot change the pitch: speed keeps the pitch, and there is no pitch shift, no key change and no chipmunk effect. Speed is one of six steps (0.5× to 2×).
  • It cannot read text aloud (text-to-speech) or write down what is said (audio-transcribe): not available yet.
  • No effects beyond volume, fades, speed, noise reduction and loudness: no equalizer, reverb, echo, compressor, de-esser, pitch correction or voice changer.
  • No automatic crossfade (overlap sounds on two tracks and fade them by hand).
  • Four tracks (more appear only when a song is separated); no recording while the tracks play (no overdubbing), no multitrack recording, no MIDI, no spectral editing.
  • Exports are always 48 kHz stereo at a fixed bit rate; no mono, 44.1 kHz or bit-rate choice in the editor (Converters has these). Tags (title, artist) cannot be edited, and cover art is not kept.
  • Loudness always aims at -16 LUFS; it cannot be set to another level.
  • The Ringtone export is for iPhone (.m4r) and at most 40 seconds; for Android, export MP3. Octoolo cannot put the ringtone on the phone.
  • The voice recorder records one microphone, as mono WAV, and not the computer's sound.
  • One file at a time per export; for converting many files with the same settings use Converters.
  • The vocal remover splits into voice and music, or voice, drums, bass and other; it cannot isolate other instruments (piano, guitar) separately.
  • An export holds at most 300 sounds ("The timeline has more in it than one export can take.").