Loading…

Controlling Perceived Emotion in Symbolic Music Generation with Monte Carlo Tree Search

This paper presents a new approach for controlling emotion in symbolic music generation with Monte Carlo Tree Search. We use Monte Carlo Tree Search as a decoding mechanism to steer the probability distribution learned by a language model towards a given emotion. At every step of the decoding proces...

Full description

Saved in:
Bibliographic Details
Published in:arXiv.org 2022-09
Main Authors: Ferreira, Lucas N, Mou, Lili, Whitehead, Jim, Lelis, Levi H S
Format: Article
Language:English
Subjects:
Online Access:Get full text
Tags: Add Tag
No Tags, Be the first to tag this record!
cited_by
cites
container_end_page
container_issue
container_start_page
container_title arXiv.org
container_volume
creator Ferreira, Lucas N
Mou, Lili
Whitehead, Jim
Lelis, Levi H S
description This paper presents a new approach for controlling emotion in symbolic music generation with Monte Carlo Tree Search. We use Monte Carlo Tree Search as a decoding mechanism to steer the probability distribution learned by a language model towards a given emotion. At every step of the decoding process, we use Predictor Upper Confidence for Trees (PUCT) to search for sequences that maximize the average values of emotion and quality as given by an emotion classifier and a discriminator, respectively. We use a language model as PUCT's policy and a combination of the emotion classifier and the discriminator as its value function. To decode the next token in a piece of music, we sample from the distribution of node visits created during the search. We evaluate the quality of the generated samples with respect to human-composed pieces using a set of objective metrics computed directly from the generated samples. We also perform a user study to evaluate how human subjects perceive the generated samples' quality and emotion. We compare PUCT against Stochastic Bi-Objective Beam Search (SBBS) and Conditional Sampling (CS). Results suggest that PUCT outperforms SBBS and CS in almost all metrics of music quality and emotion.
format article
fullrecord <record><control><sourceid>proquest</sourceid><recordid>TN_cdi_proquest_journals_2701335784</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><sourcerecordid>2701335784</sourcerecordid><originalsourceid>FETCH-proquest_journals_27013357843</originalsourceid><addsrcrecordid>eNqNi9EKgjAYRkcQJOU7_NC1MDdN78XqRggUuoxlfzmZW22z6O2T6AG6Od_F-c6MBIzzOMoTxhYkdK6nlLJNxtKUB-RYGO2tUUrqGxzQtiifeIFyMF4aDVJD_R7ORskWqtFN3KFGK77yJX0H1dQjFMIqA41FhBqFbbsVmV-Fchj-dknW27Ip9tHdmseIzp96M1o9qRPLaMx5muUJ_-_1AURiQUU</addsrcrecordid><sourcetype>Aggregation Database</sourcetype><iscdi>true</iscdi><recordtype>article</recordtype><pqid>2701335784</pqid></control><display><type>article</type><title>Controlling Perceived Emotion in Symbolic Music Generation with Monte Carlo Tree Search</title><source>Publicly Available Content Database</source><creator>Ferreira, Lucas N ; Mou, Lili ; Whitehead, Jim ; Lelis, Levi H S</creator><creatorcontrib>Ferreira, Lucas N ; Mou, Lili ; Whitehead, Jim ; Lelis, Levi H S</creatorcontrib><description>This paper presents a new approach for controlling emotion in symbolic music generation with Monte Carlo Tree Search. We use Monte Carlo Tree Search as a decoding mechanism to steer the probability distribution learned by a language model towards a given emotion. At every step of the decoding process, we use Predictor Upper Confidence for Trees (PUCT) to search for sequences that maximize the average values of emotion and quality as given by an emotion classifier and a discriminator, respectively. We use a language model as PUCT's policy and a combination of the emotion classifier and the discriminator as its value function. To decode the next token in a piece of music, we sample from the distribution of node visits created during the search. We evaluate the quality of the generated samples with respect to human-composed pieces using a set of objective metrics computed directly from the generated samples. We also perform a user study to evaluate how human subjects perceive the generated samples' quality and emotion. We compare PUCT against Stochastic Bi-Objective Beam Search (SBBS) and Conditional Sampling (CS). Results suggest that PUCT outperforms SBBS and CS in almost all metrics of music quality and emotion.</description><identifier>EISSN: 2331-8422</identifier><language>eng</language><publisher>Ithaca: Cornell University Library, arXiv.org</publisher><subject>Classifiers ; Decoding ; Emotions ; Evaluation ; Monte Carlo simulation ; Music ; Searching</subject><ispartof>arXiv.org, 2022-09</ispartof><rights>2022. This work is published under http://creativecommons.org/licenses/by/4.0/ (the “License”). Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License.</rights><oa>free_for_read</oa><woscitedreferencessubscribed>false</woscitedreferencessubscribed></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><linktohtml>$$Uhttps://www.proquest.com/docview/2701335784?pq-origsite=primo$$EHTML$$P50$$Gproquest$$Hfree_for_read</linktohtml><link.rule.ids>780,784,25753,37012,44590</link.rule.ids></links><search><creatorcontrib>Ferreira, Lucas N</creatorcontrib><creatorcontrib>Mou, Lili</creatorcontrib><creatorcontrib>Whitehead, Jim</creatorcontrib><creatorcontrib>Lelis, Levi H S</creatorcontrib><title>Controlling Perceived Emotion in Symbolic Music Generation with Monte Carlo Tree Search</title><title>arXiv.org</title><description>This paper presents a new approach for controlling emotion in symbolic music generation with Monte Carlo Tree Search. We use Monte Carlo Tree Search as a decoding mechanism to steer the probability distribution learned by a language model towards a given emotion. At every step of the decoding process, we use Predictor Upper Confidence for Trees (PUCT) to search for sequences that maximize the average values of emotion and quality as given by an emotion classifier and a discriminator, respectively. We use a language model as PUCT's policy and a combination of the emotion classifier and the discriminator as its value function. To decode the next token in a piece of music, we sample from the distribution of node visits created during the search. We evaluate the quality of the generated samples with respect to human-composed pieces using a set of objective metrics computed directly from the generated samples. We also perform a user study to evaluate how human subjects perceive the generated samples' quality and emotion. We compare PUCT against Stochastic Bi-Objective Beam Search (SBBS) and Conditional Sampling (CS). Results suggest that PUCT outperforms SBBS and CS in almost all metrics of music quality and emotion.</description><subject>Classifiers</subject><subject>Decoding</subject><subject>Emotions</subject><subject>Evaluation</subject><subject>Monte Carlo simulation</subject><subject>Music</subject><subject>Searching</subject><issn>2331-8422</issn><fulltext>true</fulltext><rsrctype>article</rsrctype><creationdate>2022</creationdate><recordtype>article</recordtype><sourceid>PIMPY</sourceid><recordid>eNqNi9EKgjAYRkcQJOU7_NC1MDdN78XqRggUuoxlfzmZW22z6O2T6AG6Od_F-c6MBIzzOMoTxhYkdK6nlLJNxtKUB-RYGO2tUUrqGxzQtiifeIFyMF4aDVJD_R7ORskWqtFN3KFGK77yJX0H1dQjFMIqA41FhBqFbbsVmV-Fchj-dknW27Ip9tHdmseIzp96M1o9qRPLaMx5muUJ_-_1AURiQUU</recordid><startdate>20220901</startdate><enddate>20220901</enddate><creator>Ferreira, Lucas N</creator><creator>Mou, Lili</creator><creator>Whitehead, Jim</creator><creator>Lelis, Levi H S</creator><general>Cornell University Library, arXiv.org</general><scope>8FE</scope><scope>8FG</scope><scope>ABJCF</scope><scope>ABUWG</scope><scope>AFKRA</scope><scope>AZQEC</scope><scope>BENPR</scope><scope>BGLVJ</scope><scope>CCPQU</scope><scope>DWQXO</scope><scope>HCIFZ</scope><scope>L6V</scope><scope>M7S</scope><scope>PIMPY</scope><scope>PQEST</scope><scope>PQQKQ</scope><scope>PQUKI</scope><scope>PTHSS</scope></search><sort><creationdate>20220901</creationdate><title>Controlling Perceived Emotion in Symbolic Music Generation with Monte Carlo Tree Search</title><author>Ferreira, Lucas N ; Mou, Lili ; Whitehead, Jim ; Lelis, Levi H S</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-proquest_journals_27013357843</frbrgroupid><rsrctype>articles</rsrctype><prefilter>articles</prefilter><language>eng</language><creationdate>2022</creationdate><topic>Classifiers</topic><topic>Decoding</topic><topic>Emotions</topic><topic>Evaluation</topic><topic>Monte Carlo simulation</topic><topic>Music</topic><topic>Searching</topic><toplevel>online_resources</toplevel><creatorcontrib>Ferreira, Lucas N</creatorcontrib><creatorcontrib>Mou, Lili</creatorcontrib><creatorcontrib>Whitehead, Jim</creatorcontrib><creatorcontrib>Lelis, Levi H S</creatorcontrib><collection>ProQuest SciTech Collection</collection><collection>ProQuest Technology Collection</collection><collection>Materials Science &amp; Engineering Collection</collection><collection>ProQuest Central (Alumni Edition)</collection><collection>ProQuest Central UK/Ireland</collection><collection>ProQuest Central Essentials</collection><collection>ProQuest Central</collection><collection>Technology Collection</collection><collection>ProQuest One Community College</collection><collection>ProQuest Central Korea</collection><collection>SciTech Premium Collection</collection><collection>ProQuest Engineering Collection</collection><collection>Engineering Database</collection><collection>Publicly Available Content Database</collection><collection>ProQuest One Academic Eastern Edition (DO NOT USE)</collection><collection>ProQuest One Academic</collection><collection>ProQuest One Academic UKI Edition</collection><collection>Engineering Collection</collection></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext</fulltext></delivery><addata><au>Ferreira, Lucas N</au><au>Mou, Lili</au><au>Whitehead, Jim</au><au>Lelis, Levi H S</au><format>book</format><genre>document</genre><ristype>GEN</ristype><atitle>Controlling Perceived Emotion in Symbolic Music Generation with Monte Carlo Tree Search</atitle><jtitle>arXiv.org</jtitle><date>2022-09-01</date><risdate>2022</risdate><eissn>2331-8422</eissn><abstract>This paper presents a new approach for controlling emotion in symbolic music generation with Monte Carlo Tree Search. We use Monte Carlo Tree Search as a decoding mechanism to steer the probability distribution learned by a language model towards a given emotion. At every step of the decoding process, we use Predictor Upper Confidence for Trees (PUCT) to search for sequences that maximize the average values of emotion and quality as given by an emotion classifier and a discriminator, respectively. We use a language model as PUCT's policy and a combination of the emotion classifier and the discriminator as its value function. To decode the next token in a piece of music, we sample from the distribution of node visits created during the search. We evaluate the quality of the generated samples with respect to human-composed pieces using a set of objective metrics computed directly from the generated samples. We also perform a user study to evaluate how human subjects perceive the generated samples' quality and emotion. We compare PUCT against Stochastic Bi-Objective Beam Search (SBBS) and Conditional Sampling (CS). Results suggest that PUCT outperforms SBBS and CS in almost all metrics of music quality and emotion.</abstract><cop>Ithaca</cop><pub>Cornell University Library, arXiv.org</pub><oa>free_for_read</oa></addata></record>
fulltext fulltext
identifier EISSN: 2331-8422
ispartof arXiv.org, 2022-09
issn 2331-8422
language eng
recordid cdi_proquest_journals_2701335784
source Publicly Available Content Database
subjects Classifiers
Decoding
Emotions
Evaluation
Monte Carlo simulation
Music
Searching
title Controlling Perceived Emotion in Symbolic Music Generation with Monte Carlo Tree Search
url http://sfxeu10.hosted.exlibrisgroup.com/loughborough?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2024-12-26T14%3A55%3A18IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-proquest&rft_val_fmt=info:ofi/fmt:kev:mtx:book&rft.genre=document&rft.atitle=Controlling%20Perceived%20Emotion%20in%20Symbolic%20Music%20Generation%20with%20Monte%20Carlo%20Tree%20Search&rft.jtitle=arXiv.org&rft.au=Ferreira,%20Lucas%20N&rft.date=2022-09-01&rft.eissn=2331-8422&rft_id=info:doi/&rft_dat=%3Cproquest%3E2701335784%3C/proquest%3E%3Cgrp_id%3Ecdi_FETCH-proquest_journals_27013357843%3C/grp_id%3E%3Coa%3E%3C/oa%3E%3Curl%3E%3C/url%3E&rft_id=info:oai/&rft_pqid=2701335784&rft_id=info:pmid/&rfr_iscdi=true