An approach to adding knowledge constraints to a data-driven generative model for Carnatic rhythm sequence
DOI:
https://doi.org/10.37591/.v9i3.3250Abstract
Computational models for generative music have been a recent trend in AI based technology developments. However, an entirely data-driven strategy often falls short of capturing the naturally occurring rhythmic grouping. Guedes et al. [1,2] had proposed dictionary based and stroke-grouping based approaches to generate novel sequences in the 8-beat cycle of Aditala. More recently an attempt of incorporating arithmetic partitioning, as conceived by performers, was made [3] to get rid of the drawback of the former model being failure to capture long-term structure and grammar of this particular idiom and being only successful in capturing local and short-term phrasing. One way of solving this issue would be to consider a rhythmic phrase as a gestalt i.e. to hypothesize three rationales: (i) a sequence of strokes, when played in a faster speed, behaves as an independent unit and not a mere compressed version of the reference; (ii) context influences the accent – the same phrase is played differently when as a part of a composition versus as a filler (ornamentation) during improvisation; (iii) phrases show a co-articulation effect – the gesture differs in anticipation of the forthcoming stroke/pattern. Initial experiments show that a time-compressed version of the reference phrase played in 4x speed sounds perceivably different from the same reference phrase played at 4x speed by the same musician. This indicates that there is a gestural difference in articulating the same phrase at different speeds. We extract timbral features to understand the differences, though there is a context-dependence that seems to have been captured in a supra-segmental way, motivating us to investigate prosodic features. This indicates that a syntactically correct sequence may not serve as a semantically plausible one to a musician’s expectancy. As the qualitative evaluation of CAMel [1] involves expert listening, we believe, adding the proposed knowledge constraints would add to the naturalness, hence acceptability, of the generated sequences.
Keywords – Carnatic rhythm, Gestural modeling, Gestalt principle, Generative model.
Downloads
Published
Issue
Section
License
Journal Title:
Title of the Paper:
Corresponding Author’s Information:
Name: Address:
E-mail:
Contact Number:
It is herein agreed that: The copyright to the above-listed unpublished and original article is transferred to STM Journals.
This copyright transfer covers the exclusive right to reproduce and distribute the contribution, including reprints, translations, photographic reproductions, microform, electronic form (offline, online), or any other reproductions of similar nature.
I/We declare that above manuscript is not published already in part or whole (except in the form of abstract) in any journal or magazine for private or public circulation, and, is not under consideration of publication elsewhere.
I/ We warrant(s) that his/her/their contribution is original, except for such excerpts from copyrighted works as may be included with the permission of the copyright holder and author thereof, that it contains no libelous statements, and does not infringe on any copyright, trademark, patent, statutory right, or propriety right of others.
I/We will not publish his/her/their above said contribution anywhere else without the prior written permission of the publisher unless it has been changed substantially. I/We also agree to the authorship of the article in the following order: Author(s) Name Signature(s)
1. ________________
2. ________________
3. ________________
4. ________________
The author(s) agree to the terms of this Copyright Notice, which will apply to this submission if and when it is published by this journal (comments if any to the editor can be added below).