Johan Ferret

Research Scientist at Google DeepMind

Paris, Ile-de-France
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

👤
Senior
🎓
Top School
Johan Ferret is a Research Scientist at Google DeepMind with a decade of experience at the intersection of deep reinforcement learning, large language models, and model alignment. Trained at École Polytechnique and Télécom Paris, he progressed from applied ML roles and a PhD at Inria to production-focused RL engineering and research at DeepMind, contributing core improvements to the widely used Acme RL library (R2D2, IMPALA, replay and logging enhancements). He blends rigorous academic grounding with hands-on engineering—shipping algorithmic refinements, learning-rate annealing options, and robust replay mechanisms—while exploring creative outlets like algorithmic art. Based in Paris, he brings a pragmatic curiosity that bridges foundational research and useful tooling for scalable RL systems.
code10 years of coding experience
job6 years of employment as a software developer
bookM.Sc - Applied Mathematics, Computer Science, M.Sc - Applied Mathematics, Computer Science at Télécom Paris
bookM.Sc - Data Science, M.Sc - Data Science at Ecole polytechnique
languagesFrench, English, Spanish
stackoverflow-logo

Stackoverflow

Stats
1reputation
0reached
0answers
0questions
github-logo-circle

Github Skills (9)

agent10
jax10
python10
reinforcement-learning10
machine-learning9
algorithms8
tensorflow8
data-structure7
data-structures7

Programming languages (2)

Jupyter NotebookPython

Github contributions (5)

github-logo-circle
google-deepmind/acme

Feb 2022 - Nov 2022

A library of reinforcement learning components and agents
Role in this project:
userML Engineer
Contributions:5 commits in 9 months
Contributions summary:Johan contributed to the reinforcement learning library by implementing and modifying core components. They refactored replay mechanisms and added functionalities to existing modules like R2D2, IMPALA. The changes included adding options for controlling the number of steps and integrating learning rate annealing. The user also worked on logging improvements within the IMPALA framework.
reinforcement-learningagents
ferretj/papers

Sep 2017 - Jul 2018

Contributions:67 commits, 58 pushes, 1 branch in 10 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial