Multimodal generative adversarial networks for piano fingering correction and performance expressiveness modeling through audio-visual feature fusion | Synapse