Steven Bird

Visitor at Charles Darwin University

Darwin, Australia
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Steven Bird is a computational linguist and software researcher with 24 years’ experience building language technologies and supporting minoritised oral-language communities across Africa, Melanesia, Amazonia and Australia. He cofounded influential infrastructure for the field—NLTK, the ACL Anthology and the Open Language Archives Community—and has held academic and leadership roles at Edinburgh, Penn, Berkeley, Melbourne and Charles Darwin University, where he directs the Top End Language Lab. His work spans research, tools and data stewardship (including senior roles at the Linguistic Data Consortium and presidency of the ACL), and he remains an active open-source contributor to the NLTK project and its data indexes. Steven combines deep field experience in language documentation with practical software engineering—authoring a widely used NLP textbook and designing production-ready tools for language preservation. Quietly, his career mixes hands-on algorithm design (e.g., name-matching and annotation graph models) with community-led collaborations that prioritize Indigenous leadership and local impact.
code24 years of coding experience
job20 years of employment as a software developer
bookPhD, Computational Linguistics, PhD, Computational Linguistics at The University of Edinburgh
bookThe University of Melbourne
languagesEnglish, French, kunwinjku (australian aboriginal language)
stackoverflow-logo

Stackoverflow

Stats
300reputation
9kreached
3answers
0questions
github-logo-circle

Github Skills (8)

nlp10
wordnet10
python10
natural-language-processing10
xml9
nltk9
context-free-grammar6
github4

Programming languages (5)

TypeScriptJavaScriptHTMLTclPython

Github contributions (5)

github-logo-circle
nltk/nltk

Aug 2001 - Dec 2022

NLTK Source
Role in this project:
userData Scientist
Contributions:24 reviews, 5068 commits, 920 PRs in 21 years 8 months
Contributions summary:Steven contributed to bug fixes, improvements, and clean-ups. Their work focused on the WordNet visualization functionality in the nltk.corpus.reader.wordnet module. The user added support for improved display in a few functions such as `wordnet.tree()` and `text.collocations`.
nlppythonmachine-learningnltknatural-language-processing
nltk/nltk_data

May 2012 - Jul 2022

NLTK Data
Role in this project:
userBack-end Developer
Contributions:229 commits, 65 PRs, 139 pushes in 10 years 3 months
Contributions summary:Steven primarily focused on updating the data index and package index for the NLTK data repository. Their work involved modifying the index.html and tools/build_pkg_index.py files, indicating a role in maintaining and generating the data listings and build processes. The updates involved updating the URLs of the data packages and correcting build scripts and overall structure.
nlpcorporanltklinguisticsnatural-language-processing
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial