Statistical Testing of Segment Homogeneity in Classification of Piecewise–Regular Objects
By: Andrey V. Savchenko and Natalya S. Belova
References
- Asadpour, V., Homayounpour, M.M. and Towhidkhah, F. (2011). Audio-visual speaker identification using dynamic facial movements and utterance phonetic content,(2): 2083–2093.
- Benesty, J., Sondhi, M.M. and Huang, Y. (2008)., Springer, Berlin.
- Borovkov, A.A. (1998)., Gordon and Breach Science Publishers, Amsterdam.
- Bottou, L., Fogelman Soulie, F., Blanchet, P. and Lienard, J. (1990). Speaker-independent isolated digit recognition: Multilayer perceptrons vs. dynamic time warping,(4): 453–465.
- Ciresan, D., Meier, U., Masci, J. and Schmidhuber, J. (2012). Multi-column deep neural network for traffic sign classification,: 333–338.
- Dalal, N. and Triggs, B. (2005). Histograms of oriented gradients for human detection,, pp. 886–893.
- Gray, R., Buzo, A., Gray, A., Jr. and Matsuyama, Y. (1980). Distortion measures for speech processing,(4): 367–376.
- Haykin, S.O. (2008)., 3rd Edn., Prentice Hall, Harlow.
- Hinton, G., Deng, L., Yu, D., Dahl, G., Mohamed, A., Jaitly, N., Senior, A., Vanhoucke, V., Nguyen, P., Sainath, T. and Kingsbury, B. (2012). Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups,(6): 82–97.
- Hinton, G.E., Osindero, S. and Teh, Y.-W. (2006). A fast learning algorithm for deep belief nets,(7): 1527–1554.
- Huang, J.-T., Li, J., Yu, D., Deng, L. and Gong, Y. (2013). Cross-language knowledge transfer using multilingual deep neural network with shared hidden layers,, pp. 7304–7308.
- Janakiraman, R., Kumar, J. and Murthy, H. (2010). Robust syllable segmentation and its application to syllable-centric continuous speech recognition,, pp. 1–5.
- Kullback, S. (1997)., Dover Publications, New York, NY.
- LeCun, Y., Bengio, Y. and Hinton, G. (2015). Deep learning,(7553): 436–444.
- LeCun, Y., Bottou, L., Bengio, Y. and Haffner, P. (1998). Gradient-based learning applied to document recognition,(11): 2278–2324.
- Liao, S., Zhu, X., Lei, Z., Zhang, L. and Li, S.Z. (2007). Learning multi-scale block local binary patterns for face recognition,S.-W. Lee and S.Z. Li (Eds.),, Lecture Notes in Computer Science, Vol. 4642, Springer, Berlin/Heidelberg, pp. 828–837.
- Lowe, D.G. (2004). Distinctive image features from scale-invariant keypoints,(2): 91–110.
- Martins, A.F.T., Figueiredo, M.A.T., Aguiar, P.M.Q., Smith, N.A. and Xing, E.P. (2008). Nonextensive entropic kernels,, pp. 640–647.
- Merialdo, B. (1988). Multilevel decoding for very-large-size-dictionary speech recognition,(2): 227–237.
- Pfau, T. and Ruske, G. (1998). Estimating the speaking rate by vowel detection,, Vol. 2, pp. 945–948.
- Rutkowski, L. (2008)., Springer-Verlag, Berlin/Heidelberg.
- Sas, J. and Żołnierek, A. (2013). Pipelined language model construction for Polish speech recognition,(3): 649–668, DOI: 10.2478/amcs-2013-0049.
- Savchenko, A.V. (2012). Directed enumeration method in image recognition,(8): 2952–2961.
- Savchenko, A.V. (2013a). Phonetic words decoding software in the problem of Russian speech recognition,(7): 1225–1232.
- Savchenko, A.V. (2013b). Probabilistic neural network with homogeneity testing in recognition of discrete patterns set,: 227–241.
- Savchenko, A.V. and Khokhlova, Y.I. (2014). About neural-network algorithms application in viseme classification problem with face video in audiovisual speech recognition systems,(1): 34–42.
- Specht, D.F. (1990). Probabilistic neural networks,(1): 109–118.
- Świercz, E. (2010). Classification in the Gabor time-frequency domain of non-stationary signals embedded in heavy noise with unknown statistical distribution,(1): 135–147, DOI: 10.2478/v10006-010-0010-x.
- Tan, X., Chen, S., Zhou, Z.-H. and Zhang, F. (2006). Face recognition from a single image per person: A survey,(9): 1725–1745.
- Theodoridis, S. and Koutroumbas, K. (2008)., 4th Edn., Academic Press, Burlington, MA/London.
- Zhou, E., Cao, Z. and Yin, Q. (2015). Naive-deep face recognition: Touching the limit of LFW benchmark or not?,.
Language: English
Page range: 915 - 925
Submitted on: Nov 1, 2014
Published on: Dec 30, 2015
Published by: University of Zielona Góra
In partnership with: Paradigm Publishing Services
Publication frequency: 4 issues per year
Keywords:
Related subjects:
© 2015 Andrey V. Savchenko, Natalya S. Belova, published by University of Zielona Góra
This work is licensed under the Creative Commons Attribution-NonCommercial-NoDerivatives 3.0 License.