Skip to main content
Have a personal or library account? Click to login
Beethoven Symphony Excerpt Dataset (BSED): An Evaluation Dataset for Orchestral Music Transcription Cover

Beethoven Symphony Excerpt Dataset (BSED): An Evaluation Dataset for Orchestral Music Transcription

Open Access
|Jul 2026

References

  1. Abeßer, J., and Schuller, G. (2017). Instrument‑centered music transcription of solo bass guitar recordings. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 25(9), 17411750. 10.1109/taslp.2017.2702384.
  2. Balke, S., Berndt, A., and Müller, M. (2025). ChoraleBricks: A modular multitrack dataset for wind music research. Transactions of the International Society for Music Information Retrieval (TISMIR), 8(1), 3954. 10.5334/tismir.252.
  3. Benetos, E., Dixon, S., Duan, Z., and Ewert, S. (2019). Automatic music transcription: An overview. IEEE Signal Processing Magazine, 36(1), 2030. 10.1109/msp.2018.2869928.
  4. Berendes, H.‑U., Saha, A., Müller, M., and Maman, B. (2026). Unison notes in multi‑instrument polyphonic music transcription: Challenges for evaluation. In Proceedings of the Deutsche Jahrestagung für Akustik (DAGA), Dresden, Germany.
  5. Bittner, R. M., Bosch, J. J., Rubinstein, D., Meseguer‑Brocal, G., and Ewert, S. (2022). A lightweight instrument‑agnostic model for polyphonic note transcription and multipitch estimation. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), Singapore (pp. 781785).
  6. Böhm, C., Ackermann, D., and Weinzierl, S. (2021). A multi‑channel anechoic orchestra recording of Beethoven’s Symphony no. 8 op. 93. Journal of the Audio Engineering Society, 68(12), 977984. 10.17743/jaes.2020.0056.
  7. Cano, E., FitzGerald, D., Liutkus, A., Plumbley, M. D., and Stöter, F. (2019). Musical source separation: An introduction. IEEE Signal Processing Magazine, 36(1), 3140. 10.1109/msp.2018.2874719.
  8. Chang, S., Benetos, E., Kirchhoff, H., and Dixon, S. (2024). YourMT3+: Multi‑instrument music transcription with enhanced transformer architectures and cross‑dataset STEM augmentation. In Proceedings of the IEEE International Workshop on Machine Learning for Signal Processing (MLSP), London, UK (pp. 16).
  9. Dannenberg, R. B., and Raphael, C. (2006). Music score alignment and computer accompaniment. Communications of the ACM, 49(8), 3843. 10.1145/1145287.1145311.
  10. Del Mar, J. (2020). Die neun Symphonien. Bärenreiter.
  11. Del Mar, N. (1983). Anatomy of the Orchestra. University of California Press.
  12. Duan, Z., Pardo, B., and Zhang, C. (2010). Multiple fundamental frequency estimation by modeling spectral peaks and non‑peak regions. IEEE Transactions on Audio, Speech, and Language Processing, 18(8), 21212133. 10.1109/tasl.2010.2042119.
  13. Edwards, D., Dixon, S., and Benetos, E. (2023). PiJAMA: Piano jazz with automatic MIDI annotations. Transactions of the International Society for Music Information Retrieval (TISMIR), 6(1), 89102. 10.5334/tismir.162.
  14. Edwards, D., Dixon, S., Benetos, E., Maezawa, A., and Kusaka, Y. (2024). A data‑driven analysis of robust automatic piano transcription. IEEE Signal Processing Letters, 31, 681685. 10.1109/lsp.2024.3363646.
  15. Emiya, V., Badeau, R., and David, B. (2010). Multipitch estimation of piano sounds using a new probabilistic spectral smoothness principle. IEEE Transactions on Audio, Speech, and Language Processing, 18(6), 16431654. 10.1109/tasl.2009.2038819.
  16. Foscarin, F., McLeod, A., Rigaux, P., Jacquemard, F., and Sakai, M. (2020). ASAP: A dataset of aligned scores and performances for piano transcription. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR) (pp. 534541).
  17. Fremerey, C., Müller, M., and Clausen, M. (2010). Handling repeats and jumps in score‑performance synchronization. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Utrecht, The Netherlands (pp. 243248).
  18. Garcia‑Martinez, J., Diaz‑Guerra, D., Politis, A., Virtanen, T., Carabias‑Orti, J. J., and Vera‑Candeas, P. (2025). SynthSOD: Developing an heterogeneous dataset for orchestra music source separation. IEEE Open Journal of Signal Processing, 6, 129137. 10.1109/ojsp.2025.3528361.
  19. Gardner, J., Simon, I., Manilow, E., Hawthorne, C., and Engel, J. H. (2022). MT3: Multi‑task multitrack music transcription. In Proceedings of the International Conference on Learning Representations (ICLR). Virtual.
  20. Good, M. (2001). MusicXML: An internet‑friendly format for sheet music. In Proceedings of the XML Conference and Exposition, Orlando, FL, USA.
  21. Gotham, M., Hentschel, J., Couturier, L., Dykeaylen, N., Rohrmeier, M., and Giraud, M. (2023). The ‘Measure Map’: an inter‑operable standard for aligning symbolic music. In Proceedings of the International Conference on Digital Libraries for Musicology (DLfM), Milano, Italy (pp. 9199).
  22. Gotham, M., Song, K., Böhlefeld, N., and Elgammal, A. (2022). Beethoven x: Es könnte sein! (it could be!). In Proceedings of the Conference on AI Music Creativity. Online.
  23. Grachten, M., Gasser, M., Arzt, A., and Widmer, G. (2013). Automatic alignment of music performances with structural differences. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Curitiba, Brazil (pp. 607612).
  24. Hawthorne, C., Elsen, E., Song, J., Roberts, A., Simon, I., Raffel, C., Engel, J. H., Oore, S., and Eck, D. (2018). Onsets and frames: Dual‑objective piano transcription. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR) (pp. 5057).
  25. Hawthorne, C., Stasyuk, A., Roberts, A., Simon, I., Huang, C. A., Dieleman, S., Elsen, E., Engel, J. H., and Eck, D. (2019). Enabling factorized piano music modeling and generation with the MAESTRO dataset. In Proceedings of the International Conference on Learning Representations (ICLR), New Orleans, LA, USA.
  26. Huron, D. (2002). Music information processing using the Humdrum toolkit: Concepts, examples, and lessons. Computer Music Journal, 26(2), 1126. 10.1162/014892602760137158.
  27. Karp, R. M. (1980). An algorithm to solve the m x n assignment problem in expected time O(mn log N). Networks, 10(2), 143152. 10.1002/net.3230100205.
  28. Kilgour, K., Zuluaga, M., Roblek, D., and Sharifi, M. (2019). Fréchet audio distance: A reference‑free metric for evaluating music enhancement algorithms. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech), Graz, Austria (pp. 23502354).
  29. Kim, H., Park, J., Kwon, T., Jeong, D., and Nam, J. (2023). A study of audio mixing methods for piano transcription in violin‑piano ensembles. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), Greece (pp. 15). Rhodes Island.
  30. Kong, Q., Li, B., Song, X., Wan, Y., and Wang, Y. (2021). High‑resolution piano transcription with pedals by regressing onset and offset times. IEEE/ACM Transactions of Audio, Speech, and Language Processing, 29, 37073717. 10.1109/taslp.2021.3121991.
  31. Krause, M., and Müller, M. (2023). Hierarchical classification for instrument activity detection in orchestral music recordings. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 31, 25672578. 10.1109/taslp.2023.3291506.
  32. Lerch, A., Arthur, C., Pati, A., and Gururani, S. (2020). An interdisciplinary review of music performance analysis. Transactions of the International Society for Music Information Retrieval (TISMIR), 3(1), 221245. 10.5334/tismir.53.
  33. Li, B., Liu, X., Dinesh, K., Duan, Z., and Sharma, G. (2019). Creating a multitrack classical music performance dataset for multimodal music analysis: Challenges, insights, and applications. IEEE Transactions on Multimedia, 21(2), 522535. 10.1109/tmm.2018.2856090.
  34. Maman, B., and Bermano, A. H. (2022). Unaligned supervision for automatic music transcription in the wild. In Proceedings of the International Conference on Machine Learning (ICML), Baltimore, MD, USA (pp. 1491814934).
  35. Maman, B., Zeitler, J., Müller, M., and Bermano, A. H. (2025). Multi‑aspect conditioning for diffusion‑based music synthesis: Enhancing realism and acoustic control. IEEE Transactions on Audio, Speech, and Language Processing, 33, 6881. 10.1109/taslp.2024.3507553.
  36. Manilow, E., Wichern, G., Seetharaman, P., and Le Roux, J. (2019). Cutting music source separation some Slakh: A dataset to study the impact of training data quality and quantity. In Proceedings of the IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA), New Paltz, NY, USA (pp. 4549).
  37. Miron, M., Carabias‑Orti, J. J., Bosch, J. J., Gómez, E., and Janer, J. (2016). Score‑informed source separation for multichannel orchestral recordings. Journal of Electrical and Computer Engineering, 2016, Article ID 8363507. 10.1155/2016/8363507.
  38. Miron, M., Carabias‑Orti, J. J., and Janer, J. (2014). Audio‑to‑score alignment at the note level for orchestral recordings. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Taipei, Taiwan (pp. 125130).
  39. Müller, M. (2021). Fundamentals of Music Processing: Using Python and Jupyter Notebooks (2nd ed.). Springer Verlag.
  40. Müller, M., Konz, V., Bogler, W., and Arifi‑Müller, V. (2011). Saarland music data (SMD). In Demos and Late Breaking News of the International Society for Music Information Retrieval Conference (ISMIR), Miami, FL, USA.
  41. Müller, M., Özer, Y., Krause, M., Prätzlich, T., and Driedger, J. (2021). Sync Toolbox: A Python package for efficient, robust, and accurate music synchronization. Journal of Open Source Software (JOSS), 6(64), 3434:14. 10.21105/joss.03434.
  42. Müller, M., and Zalkow, F. (2021). libfmp: A Python package for fundamentals of music processing. Journal of Open Source Software (JOSS), 6(63), 3326:15. 10.21105/joss.03326.
  43. Niedermayer, B., and Widmer, G. (2010). A multi‑pass algorithm for accurate audio‑to‑score alignment. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Utrecht, The Netherlands (pp. 417422).
  44. Nienhuys, H.‑W., and Nieuwenhuizen, J. (2003). Lilypond, a system for automated music engraving. In Proceedings of the XIV Colloquium on Musical Informatics, Firenze, Italy (pp. 664665).
  45. Nieto, O., Mysore, G. J., Wang, C., Smith, J. B. L., Schlüter, J., Grill, T., and McFee, B. (2020). Audio‑based music structure analysis: Current trends, open challenges, and applications. Transactions of the International Society for Music Information Retrieval (TISMIR), 3(1), 246263. 10.5334/tismir.54.
  46. Pätynen, J., Pulkki, V., and Lokki, T. (2008). Anechoic recording system for symphony orchestra. Acta Acustica united with Acustica, 94(6), 856865. 10.3813/aaa.918104.
  47. Peter, S. D., Peter, E. S. D., Cancino‑Chacón, C. E., Foscarin, F., McLeod, A. P., Henkel, F., Karystinaios, E., and Widmer, G. (2023). Automatic note‑level score‑to‑performance alignments in the ASAP dataset. Transactions of the International Society for Music Information Retrieval (TISMIR), 6(1), 2742. 10.5334/tismir.149.
  48. Raffel, C., McFee, B., Humphrey, E. J., Salamon, J., Nieto, O., Liang, D., and Ellis, D. P. W. (2014). MIR_EVAL: A transparent implementation of common MIR metrics. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Taipei, Taiwan (pp. 367372).
  49. Riley, X., Edwards, D., and Dixon, S. (2024a). High resolution guitar transcription via domain adaptation. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) Seoul, South Korea (pp. 10511055). 10.1109/ICASSP48485.2024.10446182.
  50. Riley, X., Guo, Z., Edwards, A. C., and Dixon, S. (2024b). GAPS: A large and diverse classical guitar dataset and benchmark transcription model. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), San Francisco, CA, USA (pp. 611617).
  51. Saha, A., Berendes, H.‑U., Müller, M., and Maman, B. (2026). Snapping matters: Context‑aware onset refinement for automatic music transcription. In Proceedings of the International Computer Music Conference (ICMC), Hamburg, Germany (pp. 339346).
  52. Sapp, C. S. (2005). Online database of scores in the humdrum file format. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), London, UK (pp. 664665).
  53. Sarkar, S., Benetos, E., and Sandler, M. (2022). EnsembleSet: A new high quality dataset for chamber ensemble separation. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Bengaluru, India (pp. 625632).
  54. Schreiber, H., Weiß, C., and Müller, M. (2020). Local key estimation in classical music audio recordings: A cross‑version study on Schubert’s Winterreise. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), Barcelona, Spain (pp. 501505).
  55. Selfridge‑Field, E (Ed.). (1997). Beyond MIDI: The Handbook of Musical Codes. MIT Press.
  56. Shor, J., Jansen, A., Maor, R., Lang, O., Tuval, O., de Chaumont Quitry, F., Tagliasacchi, M., Shavitt, I., Emanuel, D., and Haviv, Y. (2020). Towards learning a universal non‑semantic representation of speech. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech), Shanghai, China (pp. 140144).
  57. Solomon, M. (1998). Beethoven. Schirmer Trade Books.
  58. Tamer, N. C., Özer, Y., Müller, M., and Serra, X. (2023). High‑resolution violin transcription using weak labels. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Milan, Italy (pp. 223230).
  59. Tamer, N. C., Ramoneda, P., and Serra, X. (2022). Violin etudes: A comprehensive dataset for f0 estimation and performance analysis. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Bengaluru, India (pp. 517524).
  60. Thickstun, J., Harchaoui, Z., and Kakade, S. M. (2017). Learning features of music from scratch. In Proceedings of the International Conference on Learning Representations (ICLR), Toulon, France.
  61. Weber, P., Uhle, C., Müller, M., and Lang, M. (2025). STAR drums: A dataset for automatic drum transcription. Transactions of the International Society for Music Information Retrieval (TISMIR), 8(1). 10.5334/tismir.244.
  62. Weiß, C., Zalkow, F., Arifi‑Müller, V., Müller, M., Koops, H. V., Volk, A., and Grohganz, H. (2021). Schubert Winterreise dataset: A multimodal scenario for music analysis. ACM Journal on Computing and Cultural Heritage (JOCCH), 14(2), 25:118. 10.1145/3429743.
  63. Wu, Y., Wei, W., Li, D., Li, M., Yu, Y., Gao, Y., and Li, W. (2024). Harmonic frequency‑separable transformer for instrument‑agnostic music transcription. In Proceedings of the IEEE International Conference on Multimedia and Expo (ICME), Canada (pp. 16). Niagara Falls.
  64. Xi, Q., Bittner, R. M., Pauwels, J., Ye, X., and Bello, J. P. (2018). GuitarSet: A dataset for guitar transcription. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Paris, France (pp. 453460).
  65. Zehren, M., Alunno, M., and Bientinesi, P. (2024). Analyzing and reducing the synthetic‑to‑real transfer gap in music information retrieval: the task of automatic drum transcription. CoRR, abs/2407.19823.
  66. Zeitler, J., Maman, B., and Müller, M. (2024a). Robust and accurate audio synchronization using raw features from transcription models. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), San Francisco, CA, USA (pp. 120127).
  67. Zeitler, J., Weiß, C., Arifi‑Müller, V., and Müller, M. (2024b). BPSD: A coherent multi‑version dataset for analyzing the first movements of beethoven’s piano sonatas. Transactions of the International Society for Music Information Retrieval (TISMIR), 7(1), 195212. 10.5334/tismir.196.
  68. Zhang, H., Tang, J., Rafee, S. R. M., Dixon, S., and Fazekas, G. (2022). ATEPP: A dataset of automatically transcribed expressive piano performance. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Bengaluru, India (pp. 446453).
DOI: https://doi.org/10.5334/tismir.343 | Journal eISSN: 2514-3298
Language: English
Page range: 405 - 422
Submitted on: Oct 4, 2025
Accepted on: Apr 29, 2026
Published on: Jul 24, 2026
Published by: Ubiquity Press
In partnership with: Paradigm Publishing Services

© 2026 Hans-Ulrich Berendes, Abhirup Saha, Ben Maman, Vlora Arifi-Müller, Meinard Müller, published by Ubiquity Press
This work is licensed under the Creative Commons Attribution 4.0 License.