Skip to main content
Have a personal or library account? Click to login
KRAISLER: A Multi‑Track Dataset of Piano and Violin Duet Recordings for Music Information Retrieval Research Cover

KRAISLER: A Multi‑Track Dataset of Piano and Violin Duet Recordings for Music Information Retrieval Research

Open Access
|Aug 2026

References

  1. Arzt, A., and Widmer, G. (2010). Towards effective ‘any‑time’ music tracking. In Proceedings of the Starting AI Researchers Symposium (STAIRS), Lisbon, Portugal.
  2. Balke, S., Berndt, A., and Müller, M. (2025). Choralebricks: A modular multitrack dataset for wind music research. Transactions of the International Society for Music Information Retrieval, 8(1), 3954. 10.5334/tismir.252.
  3. Bittner, R. M., Salamon, J., Tierney, M., Mauch, M., Cannam, C., and Bello, J. P. (2014). MedleyDB: A multitrack dataset for annotation‑intensive MIR research. In Proceedings of the 15th International Society for Music Information Retrieval Conference (ISMIR), Taipei, Taiwan (pp. 155160).
  4. Böck, S., Krebs, F., and Widmer, G. (2016). Joint beat and downbeat tracking with recurrent neural networks. In Proceedings of the 17th International Society for Music Information Retrieval Conference (ISMIR), New York City, NY, USA (pp. 255261).
  5. Carter, D. S., and von Appen, R. (2025). Tempo variability in Billboard Hot 100 songs, 1966‑1995: Patterns, click tracks, and historical change. Intégral, 38, 6191. https://theory.esm.rochester.edu/integral/38-2025/carter-von-appen/.
  6. Chiu, C.‑Y., Hsiao, W.‑Y., Yeh, Y.‑C., Yang, Y.‑H., and Su, A. W.‑Y. (2020). Mixing‑specific data augmentation techniques for improved blind violin/piano source separation. In Proceedings of the 22nd IEEE International Workshop on Multimedia Signal Processing (MMSP) Tampere, Finland (pp. 16). IEEE.
  7. Chiu, C.‑Y., Müller, M., Davies, M. E., Su, A. W.‑Y., and Yang, Y.‑H. (2022). An analysis method for metric‑level switching in beat tracking. IEEE Signal Processing Letters, 29, 21532157. 10.1109/lsp.2022.3215106.
  8. Chiu, C.‑Y., Müller, M., Davies, M. E., Su, A. W.‑Y., and Yang, Y.‑H. (2023). Local periodicity‑based beat tracking for expressive classical piano music. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 31, 28242835. 10.1109/taslp.2023.3297956.
  9. Choi, W., Kim, M., Chung, J., and Jung, S. (2021). LaSAFT: Latent source attentive frequency transformation for conditioned source separation. In Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Toronto, Ontario, Canada (pp. 171175).
  10. Davies, M. E., and Böck, S. (2019). Temporal convolutional networks for musical audio beat tracking. In 2019 27th European Signal Processing Conference (EUSIPCO) A Coruña, Spain (pp. 15). IEEE.
  11. Défossez, A., Usunier, N., Bottou, L., and Bach, F. (2019). Music Source Separation in the Waveform Domain. arXiv preprint arXiv:1911.13254. 10.48550/arXiv.1911.13254.
  12. Duan, Z., and Pardo, B. (2011). Soundprism: An online system for score‑informed source separation of music audio. IEEE Journal of Selected Topics in Signal Processing, 5(6), 12051215. 10.1109/jstsp.2011.2159701.
  13. Emiya, V., Badeau, R., and David, B. (2010). Multipitch estimation of piano sounds using a new probabilistic spectral smoothness principle. IEEE Transactions on Audio, Speech, and Language Processing, 18(6), 16431654. 10.1109/tasl.2009.2038819.
  14. Foscarin, F., McLeod, A., Rigaux, P., Jacquemard, F., and Sakai, M. (2020). ASAP: A dataset of aligned scores and performances for piano transcription. In Proceedings of the 21st International Society for Music Information Retrieval Conference (ISMIR), Montréal, Canada (pp. 534541).
  15. Fritsch, J., and Plumbley, M. D. (2013). Score informed audio source separation using constrained nonnegative matrix factorization and score synthesis. In 2013 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Vancouver, Canada (pp. 888891).
  16. Garcia‑Martinez, J., Diaz‑Guerra, D., Politis, A., Virtanen, T., Carabias‑Orti, J. J., and Vera‑Candeas, P. (2025). SynthSOD: Developing an heterogeneous dataset for orchestra music source separation. IEEE Open Journal of Signal Processing, 6, 129137. 10.1109/ojsp.2025.3528361.
  17. Gong, X., Xu, W., Liu, J., and Cheng, W. (2019). Analysis and correction of MAPS dataset. In Proceedings of the 22nd International Conference on Digital Audio Effects (DAFx‑19), Birmingham, UK.
  18. Goto, M., Hashiguchi, H., Nishimura, T., and Oka, R. (2002). RWC Music Database: Popular, classical, and jazz music databases. In Proceedings of the 3rd International Conference on Music Information Retrieval (ISMIR), Paris, France (pp. 287288).
  19. Grosche, P., Müller, M., and Sapp, C. S. (2010). What makes beat tracking difficult? A case study on Chopin mazurkas. In Proceedings of the 11th International Society for Music Information Retrieval Conference (ISMIR), Utrecht, Netherlands (pp. 649654).
  20. Hawthorne, C., Elsen, E., Song, J., Roberts, A., Simon, I., Raffel, C., Engel, J., Oore, S., and Eck, D. (2018). Onsets and Frames: Dual‑objective piano transcription. In Proceedings of the 19th International Society for Music Information Retrieval Conference (ISMIR), Paris, France (pp. 5057).
  21. Hawthorne, C., Stasyuk, A., Roberts, A., Simon, I., Huang, C., Dieleman, S., Elsen, E., Engel, J., and Eck, D. (2019). Enabling factorized piano music modeling and generation with the MAESTRO dataset. In Proceedings of the 7th International Conference on Learning Representations (ICLR), New Orleans, Louisiana, USA.
  22. Hennequin, R., Khlif, A., Voituret, F., and Moussallam, M. (2020). Spleeter: A fast and efficient music source separation tool with pre‑trained models. Journal of Open Source Software, 5(50), 2154. 10.21105/joss.02154.
  23. Huang, Y.‑F., Moran, N., Coleman, S., Kelly, J., Wei, S.‑H., Chen, P., Huang, Y.‑H., Chen, T.‑P., Kuo, Y.‑C., Wei, Y.‑C., Li, C.‑H., Huang, D.‑Y., Kao, H.‑K., Lin, T.‑W., and Su, L. (2024). MOSA: Music motion with semantic annotation dataset for cross‑modal music processing. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 32, 41574170. 10.1109/taslp.2024.3407529.
  24. Jansson, A., Humphrey, E., Montecchio, N., Bittner, R., Kumar, A., and Weyde, T. (2017). Singing voice separation with deep U‑Net convolutional networks. In Proceedings of the 18th International Society for Music Information Retrieval Conference (ISMIR), Suzhou, China (pp. 745751).
  25. Kim, T., and Nam, J. (2023). All‑in‑one metrical and functional structure analysis with neighborhood attentions on demixed audio. In In 2023 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA), New Paltz, New York, USA (pp. 15). IEEE. 10.1109/waspaa58266.2023.10248148.
  26. Kong, Q., Li, B., Song, X., Wan, Y., and Wang, Y. (2021). High‑resolution piano transcription with pedals by regressing onset and offset times. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 29, 37073717. 10.1109/taslp.2021.3121991.
  27. Li, B., Liu, X., Dinesh, K., Duan, Z., and Sharma, G. (2018). Creating a multitrack classical music performance dataset for multimodal music analysis: Challenges, insights, and applications. IEEE Transactions on Multimedia, 21(2), 522535. 10.1109/tmm.2018.2856090.
  28. Maman, B., and Bermano, A. H. (2022). Unaligned supervision for automatic music transcription in the wild. In Proceedings of the 39th International Conference on Machine Learning, Baltimore, Maryland, USA (pp. 1491814934).
  29. Manilow, E., Seetharaman, P., and Pardo, B. (2020). Simultaneous separation and transcription of mixtures with multiple polyphonic and percussive instruments. In Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Barcelona, Spain (pp. 771775).
  30. Manilow, E., Wichern, G., Seetharaman, P., and Le Roux, J. (2019). Cutting music source separation some Slakh: A dataset to study the impact of training data quality and quantity. In Proceedings of the 2019 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA), New Paltz, New York, USA (pp. 4549).
  31. McFee, B., Raffel, C., Liang, D., Ellis, D. P., McVicar, M., Battenberg, E., and Nieto, O. (2015). librosa: Audio and music signal analysis in python. In Proceedings of the 14th Python in Science Conference, Austin, Texas, USA (pp. 1825).
  32. Miron, M., Carabias‑Orti, J. J., Bosch, J. J., Gómez, E., and Janer, J. (2016). Score‑informed source separation for multichannel orchestral recordings. Journal of Electrical and Computer Engineering, 2016, 119.
  33. Müller, M. (2015). Fundamentals of Music Processing: Audio, Analysis, Algorithms, Applications. Springer. 10.1007/978-3-319-21945-5.
  34. Müller, M., Konz, V., Bogler, W., and Arifi‑Müller, V. (2011). Saarland Music Data (SMD). In Proceedings of the 12th International Society for Music Information Retrieval Conference (ISMIR), Miami, Florida, USA.
  35. Özer, Y., and Müller, M. (2024). Source separation of piano concertos using musically motivated augmentation techniques. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 32, 12141225. 10.1109/taslp.2024.3356980.
  36. Özer, Y., Schwär, S., Arifi‑Müller, V., Lawrence, J., Sen, E., and Müller, M. (2023). Piano Concerto Dataset (PCD): A multitrack dataset of piano concertos. Transactions of the International Society for Music Information Retrieval, 6(1), 7588. 10.5334/tismir.160.
  37. Park, J., Cancino‑Chacón, C., Chiruthapudi, S., and Nam, J. (2025). Matchmaker: An open‑source library for real‑time piano score following and systematic evaluation. In Proceedings of the 26th International Society for Music Information Retrieval Conference (ISMIR), Daejeon, South Korea.
  38. Peter, S. D., Cancino‑Chacón, C. E., Foscarin, F., McLeod, A. P., Henkel, F., Karystinaios, E., and Widmer, G. (2023). Automatic note‑level score‑to‑performance alignments in the ASAP dataset. Transactions of the International Society for Music Information Retrieval, 6(1), 2742. 10.5334/tismir.149.
  39. Piczak, K. J. (2015). ESC: Dataset for environmental sound classification. In Proceedings of the 23rd ACM International Conference on Multimedia, Brisbane, Australia (pp. 10151018).
  40. Raffel, C., McFee, B., Humphrey, E. J., Salamon, J., Nieto, O., Liang, D., and Ellis, D. P. (2014). mir_eval: A transparent implementation of common MIR metrics. In Proceedings of the 15th International Society for Music Information Retrieval Conference (ISMIR), Taipei, Taiwan (pp. 367372).
  41. Rafii, Z., Liutkus, A., Stöter, F.‑R., Mimilakis, S. I., and Bittner, R. (2017). The MUSDB18 corpus for music separation. 10.5281/zenodo.1117372.
  42. Sarkar, S., Benetos, E., and Sandler, M. (2022). Ensembleset: A new high quality synthesised dataset for chamber ensemble separation. In Proceedings of the 23rd International Society for Music Information Retrieval Conference (ISMIR), Bengaluru, India (pp. 625632).
  43. Schreiber, H., Urbano, J., and Müller, M. (2020). Music tempo estimation: Are we done yet? Transactions of the International Society for Music Information Retrieval, 3(1), 5874. 10.5334/tismir.43.
  44. Stöter, F.‑R., Uhlich, S., Liutkus, A., and Mitsufuji, Y. (2019). Open‑unmix ‑ a reference implementation for music source separation. Journal of Open Source Software, 4(41), 1667. 10.21105/joss.01667.
  45. Tamer, N. C., Ozer, Y., Müller, M., and Serra, X. (2023). High‑resolution violin transcription using weak labels. In Proceedings of the 24th International Society for Music Information Retrieval Conference (ISMIR), Milan, Italy.
  46. Tamer, N. C., Ramoneda, P., and Serra, X. (2022). Violin etudes: A comprehensive dataset for f0 estimation and performance analysis. In Proceedings of the 23rd International Society for Music Information Retrieval Conference (ISMIR), Bengaluru, India (pp. 517524).
  47. Thickstun, J., Harchaoui, Z., and Kakade, S. (2017). Learning features of music from scratch. In Proceedings of the 5th International Conference on Learning Representations (ICLR), Toulon, France.
  48. Wang, T.‑K., Peng, Y.‑P., Su, L., and Cheung, V. K. M. (2026). VioPTT: Violin technique‑aware transcription from synthetic data augmentation. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), Barcelona, Spain.
  49. Wei, W., Li, P., Yu, Y., and Li, W. (2022). HPPNet: Modeling the harmonic structure and pitch invariance in piano transcription. In Proceedings of the 23rd International Society for Music Information Retrieval Conference (ISMIR), Bengaluru, India (pp. 303310).
  50. Weiß, C., Zalkow, F., Arifi‑Müller, V., Müller, M., Koops, H. V., Volk, A., and Grohganz, H. G. (2021). Schubert winterreise dataset: A multimodal scenario for music analysis. Journal on Computing and Cultural Heritage, 14(2), 118. 10.1145/3429743.
  51. Wu, Y., Gardner, J., Manilow, E., Simon, I., Hawthorne, C., and Engel, J. (2022). The chamber ensemble generator: Limitless high‑quality MIR data via generative modeling. arXiv preprint arXiv:2209.14458.
  52. Xi, Q., Bittner, R. M., Pauwels, J., Ye, X., and Bello, J. P. (2018). GuitarSet: A dataset for guitar transcription. In Proceedings of the 19th International Society for Music Information Retrieval Conference (ISMIR), Paris, France (pp. 453460).
  53. Yan, Y., and Duan, Z. (2024). Scoring time intervals using non‑hierarchical transformer for automatic piano transcription. In Proceedings of the 25th International Society for Music Information Retrieval Conference (ISMIR), Milan, Italy (pp. 973980).
DOI: https://doi.org/10.5334/tismir.338 | Journal eISSN: 2514-3298
Language: English
Page range: 456 - 473
Submitted on: Sep 1, 2025
Accepted on: May 18, 2026
Published on: Aug 6, 2026
Published by: Ubiquity Press
In partnership with: Paradigm Publishing Services

© 2026 Hyemi Kim, Jiyun Park, Sein Lee, Taegyun Kwon, Sunjae Won, Juhan Nam, published by Ubiquity Press
This work is licensed under the Creative Commons Attribution 4.0 License.