~./AAdonis
  • Home
  • Publications
  • Notebook
  • CV

Notebook

Practical guides, tool setups, and research notes on audio AI.
Neural Audio Codecs
Discrete neural audio codecs for speech compression and autoregressive modeling — DAC, SNAC, WavTokenizer, and X-Codec2.
5 notes
Build & Train Models
Training pipelines, tokenizers, and infrastructure for large-scale audio and language models.
3 notes
Speech Alignment
Extracting precise word and phoneme timestamps from audio — tools, models, and common failure modes.
3 notes
Speech Datasets
Clean speech corpora, noise and RIR databases, and techniques for synthesizing degraded training data.
2 notes
Voice Conversion
Transforming a speaker's voice to match a target speaker while preserving linguistic content.
2 notes
Speech Editing
Modifying spoken audio at the word level — insertions, deletions, and substitutions while preserving speaker identity.
1 note
Speech Enhancement
Reconstructing clean speech from degraded audio — models, metrics, and inference setups.
1 note
    © 2026 A.Asonitis