Extract vocals and accompaniment, turn recordings into MIDI, change a voice, or generate speech and images from text. Choose a tool for your source material.
Active tools
19
Ready to accept new jobs.
Task groups
4
Separation, transcription, image, and voice.
Zero-credit entries
1
Existing MIDI files can use the zero-credit text-encoding converter.
GPU workflows
24
Complex separation / transcription typically needs GPU.
Task group
Split a song into vocals, accompaniment, or instrument stems for practice, mixing, sampling, and further analysis.
Two-stem tools return vocals and accompaniment; multi-stem tools separate more instruments.
Open-weight, locally deployable two-stem vocal/accompaniment separation for practice, covers, mixing, and sampling prep.
Included features — choose inside
Fixed vocals, drums, bass, guitar, piano, and other stems from one mix
Included features — choose inside
Drum remixing, replacement, transient editing, practice tracks, and drum-focused analysis.
Practice tracks, accompaniment, remixing, sampling, and targeted cleanup.
Extract an open-vocabulary sound with AudioSep-Hive, FlowSep-Hive or CLAPSep, or isolate a source that matches a 10-second reference with Banquet.
Use case Remix, practice with accompaniment and edit separate parts.
Blindly separate two or more concurrent singing voices into numbered singer stems with UNMIXX.
Use case Remix, practice with accompaniment and edit separate parts.
Included features — choose inside
Choose the released standard or aggressive denoise checkpoint and receive clean audio plus the removed-noise residual.
Use case Clean up recordings with noise, echo or live interference.
Included features — choose inside
Extract clear dialogue from film and TV audio, and split the background into separate music and sound-effects tracks.
Use case Remix, practice with accompaniment and edit separate parts.
Included features — choose inside
Task group
Convert full audio or single-instrument clips to MIDI; ideal for transcription, teaching, arranging, and quick drafts.
Use multi-instrument transcription for full songs, a piano-specialist model for solo piano, and the zero-credit converter for existing MIDI files.
MuScriptor-large creates an editable multi-instrument MIDI draft from a mix. Check every expected part: even clearly audible passages can be omitted.
Use case Edit notes, instrumentation and performance details.
Included features — choose inside
Turn piano performances into editable MIDI with note endings at key release
Included features — choose inside
Garbled MIDI lyrics, track names, instrument names, markers, cue points, and copyright text cleanup
Included features — choose inside
For fixed top-view keyboard video. Supply all four keyboard corners for perspective correction; V2N predicts physical key release and note velocity, not pedal-extended audio offsets.
Use case Edit notes, instrumentation and performance details.
Requires the vocal audio plus lyrics, phonemes and a phoneme-to-word map. Exports note MIDI together with phoneme, word and vocal-technique timing; it is not an audio-only route.
Use case Edit notes, instrumentation and performance details.
Converts popular-music audio into a sung-melody-and-chord lead sheet in Humdrum **kern. It does not reconstruct every instrument in a full arrangement.
Use case Prepare and edit notation in score-writing software.
Converts an existing piano performance MIDI file into an editable notated MusicXML score with MIDI2ScoreTransformer. Raw audio is not accepted.
Use case Prepare and edit notation in score-writing software.
The upstream project targets pitched instruments generally and also publishes other_v1_5, vocal, and vocal_harmony checkpoints. This TelkNet tool currently exposes bass_v2 and guitar_v1_5. Velocity prediction is a separate postprocessor; instrument classification and multi-track output remain experimental.
Use case Edit notes, instrumentation and performance details.
Included features — choose inside
Copying a sung melody into a DAW for editing, doubling, synth programming or a notation draft.
Beat This! final0 detects beat and downbeat times. TelkNet derives BPM, a normalized beat grid, and fixed-versus-variable-tempo warnings; OpenKeyScan3 adds global key, Camelot, and Open Key notation. ChordMini BTC is the default; switch to 2E1D or Omnizart to compare. Chord MIDI provides harmony audition, not the original performance.
Use case Analyze tempo and harmony for arranging, mixing and DAW projects.
Included features — choose inside
Choose ADTOF Plus for a complete mix: it runs MDX23C drum separation, five-stem drum separation, ADTOF transcription, and velocity/hi-hat post-processing. Choose ADTOF PyTorch Direct for an already isolated drum stem.
Use case Edit notes, instrumentation and performance details.
Choose SynthTab Universal, GuitarSet, EGDB or IDMT for clean/acoustic and matched guitar domains, or GuitarProFX TabCNN for electric-guitar tones and effects. The models predict six-string fret positions; TelkNet exports note events, six-track MIDI and timestamped tablature.
Use case Practice and review guitar performances by string and fret.
Preserve original speed or choose a target BPM. Analyze beats, key and chords; optionally separate vocals and accompaniment. Export WAV, MIDI/CSV, song tags and verifiable analysis.
Use case Analyze tempo and harmony for arranging, mixing and DAW projects.
Task group
Generate design-forward images from text prompts, including layouts where typography and composition matter.
Describe the scene, visible text, and aspect ratio. The recommended quality preset fits most jobs.
One-step semantic image edits while preserving the original structure as much as possible
Fast text-to-image exploration, moodboards, product concepts, editorial visuals, and creative drafts
Included features — choose inside
Task group
Handle vocals, timbre conversion, and cover workflows; well suited for demos, test takes, and character-voice experiments.
Clean, single-speaker or single-singer recordings give the most stable results.
Voice-color conversion, cover demos, and character voice experiments
Short-reference narration, voice tests, and multilingual speech samples
How it works
Run real models with free credits. Tasks run on the server; download outputs when done.
STEP 01
New accounts get free credits; no card required. Pick a tool and run a representative sample first.
STEP 02
The credit cost is shown before submission, reserved when submitted, charged on success and returned on failure. The tool, parameters and batch size affect the cost.
STEP 03
When complete, download stems, MIDI, images, or converted output directly. There is no subscription and no payment.