Flowtron is an auto-regressive flow-based generative network for text to speech synthesis with control over speech variation and style transfer
Role in this project:
ML Engineer Contributions:28 commits, 6 PRs, 40 pushes in 2 years 6 months
Contributions summary:Rafael primarily contributed to the `flowtron` repository by modifying the core files, `flowtron.py` and `data.py`. They focused on addressing potential issues related to data handling, padding, and ensuring compatibility with floating-point precision, especially concerning the attention mechanisms. The user also improved the inference capabilities with changes made to `inference.py`, including adding the `torch.no_grad()` context manager for waveglow inference.
speechstyle-transfertext-to-speechspeech-synthesis
Utility functions for handling MIDI data in a nice/intuitive way.
Role in this project:
Back-end Developer Contributions:7 commits, 13 PRs, 47 comments in 4 months
Contributions summary:Rafael primarily contributed to the core functionality of the `pretty-midi` library, focusing on features related to MIDI data handling. Their work included adding and refining features for retaining key and time signatures, crucial for accurate MIDI file processing. The user also added a utility to convert quarter notes per minute to beats per minute and implemented functionality for pitch class analysis. They demonstrated a strong understanding of MIDI data structures and algorithms.
midi