Download Identification of Time-frequency Maps for sounds timbre discrimination Gabor Multipliers are signals operator which are diagonal in a time-frequency representation of signals and can be viewed as timefrequency transfer function. If we estimate a Gabor mask between a note played by two instruments, then we have a time-frequency representation of the difference of timbre between these two notes. By averaging the energy contained in the Gabor mask, we obtain a measure of this difference. In this context, our goal is to automatically localize the time-frequency regions responsible for such a timbre dissimilarity. This problem is addressed as a feature selection problem over the time-frequency coefficients of a labelled data set of sounds.
Download Real-Time Neural Audio on Apple Silicon: Benchmarking Inference Frameworks Under Realistic DAW Contention Neural network models are increasingly deployed in audio plugins across a wide range of applications, including amplifier emulation, effects modeling, and synthesis. This paper evaluates widely used inference options including BNNSGraph, RTNeural, LibTorch, ONNX Runtime, and anira on model architectures commonly used in neural audio plugins. The key contribution is moving beyond isolated benchmarks to evaluate performance under realistic DAW contention, constructing mix sessions with configurable plugin loads. Results show that isolated benchmarks can be misleading, and BNNSGraph proves most robust for convolutional models on Apple Silicon.
Download Realization of a Diffuse Sound Field with a PC-Based Sound Card Solution For the quality assessment of headphones, especially the loudness measuring of headphones, a diffuse sound field is required. At this time a hardware based noise generator, one-third octave filters built up in analog mode as well as boosters are used. In this work a flexible PC-based solution with the aid of a sound card is presented. Therefore ten independent noise generators, generating Gaussian distributed white noise, are needed. The implementation using the ’Dynamic Creation of Pseudorandom Number Genrators’ for ’Mersenne Twister’ is described. A probability transformation to convert equal distributed numbers into Gaussian distributed ones is derived in detail. Furthermore one-third octave filters are designed and implemented according to the ANSI standard. The access to the sound card is provided using the Wave-API library under Microsoft Windows. This work was carried out at Sennheiser electronic GmbH in Wennebostel (Germany) in the development department for cord based headphones.
Download Software Toolbox for Multichannel Sound Reproduction This paper describes a versatile software toolbox, which has been developed for researching, teaching and developing in the field of multichannel sound signal processing. The software system runs on a PC and consists on 5 modules covering the main stages and aspects of multichannel sound reproduction using loudspeakers. A number of new and efficient algorithms have been specially implemented for this software.
Download Bibliometric Study of the DAFx Proceedings 1998 - 2009 In this paper we present a bibliometric study of the Digital Audio Effects (DAFx) conference proceedings from 1998 to 2009. Using the online DAFx proceedings, we constructed a DAFx database (LaTeX) to study its bibliometric statistics in terms of research topics, growth of literature, authorship distribution, citation patterns, and frequency distribution of scientific productivity. Results showed that the DAFx literature (with quasi-linear accumulative growth) now consists of 722 contributions (including key notes, papers and posters) from 767 unique authors, from which we identified the 20 top DAFx contributors. Using Google Scholar, we identified that the top 10 most cited DAFx papers (between 43 to 65 times) are in majority (8/10) dealing with sound and music analysis (e.g. extraction of sinusoids, musical genre classification, perceived intensity of music, and musical note onset detection). This study also confirmed that the DAFx literature conforms to the Lokta’s law (n=2.0771 and C=0.6336) at 0.01 level of significance using the Kolmogorov-Smirnov test (KS-test) of goodnessof-fit. The DAFx database will serve as the basis for an Author Cocitation Analysis (ACA) and to create a DAFx conferences archive DVD.
Download Improved method for extraction of partial’s parameters in polyphonic transcription of piano higher octaves Polyphonic transcription is specially challenging in piano higher octaves due to the complexity of the spectrum of notes and therefore, chords. Besides the fundamental and second partial components, other spectral elements appears. The three peaks related to the unison as well as the second harmonic of the fundamental unison can be distinguished in most measures. Furthermore, intermodulation components are also present when non-linearity is high enough. This paper compares several methods to improve the training process that allows to synthesize the spectral patterns and masks used in transcription methods.
Download Simulation of Analog Flanger Effect Using BBD Circuit This paper deals with simulation of BBD circuit based analog flanger effects. The famous Electro-Harmonix Deluxe Electric Mistress flanger effect was used as a case study in this paper. The main attention of this paper is paid to the analysis and simulation of the LFO circuit, the BBD clock generator circuit and BBD circuit simulation of this effect. However, in order to compare the simulation results with measured data, the signal path simulation using the DK-method has been introduced as well.
Download Probing Low-Level Acoustic Attribute Encoding in CLAP Audio Embeddings This work analyzes CLAP audio embeddings through a probing framework, studying the encoding of reverberation (RT60), loudness (LUFS), spectral content (SC), and relative pitch (RP). Results show that all attributes are reliably recoverable from CLAP embeddings, with RT60, LUFS, and RP approximately linearly encoded, while SC requires non-linear probes. The identified patterns generalize across eight additional audio foundation models.
Download Monophonic transcription with autocorrelation This paper describes an algorithm, which performs monophonic music transcription. A pitch tracker calculates the fundamental frequency of the signal from the autocorrelation function. A continuity-restoration block takes the extracted pitch and determines the score corresponding to the original performance. The signal envelope analysis completes the transcription system, calculating attack-sustain-decay-release times, which improves the synthesis process. Attention is also paid to the extraction of timbre and wavetable synthesis.
Download Room simulation for binaural sound reproduction using measured spatiotemporal impulse responses In binaural sound reproduction systems the incorporation of room simulation is important to improve sound source localisation capabilities. Thus, the localisation error can be decreased, while equivalently an enhanced externality (out of head localisation) is achieved. Previously proposed works are based on simple geometrical approaches for room simulation. In this paper an alternative method using measured room impulse responses (RIRs) is presented. Therefore, it is possible to obtain a convincing acoustical image of an existing room. The RIRs are measured using a circular microphone array to capture both temporal and spatial information of the desired room.