freecodecamp/2016-new-coder-survey — explained in plain English
Analysis updated 2026-08-03 · repo last pushed 2017-10-05
Build charts showing which bootcamps are most popular among new coders.
Analyze the demographics of people learning to code.
Visualize the job outcomes and salary expectations of coding learners.
Explore survey data on Kaggle or in a personal blog and submit findings.
| freecodecamp/2016-new-coder-survey | josephmisiti/ml_for_hackers | hadley/web-scraping | |
|---|---|---|---|
| Stars | 202 | 78 | 75 |
| Language | R | R | R |
| Last pushed | 2017-10-05 | 2014-12-30 | 2024-07-08 |
| Maintenance | Dormant | Dormant | Dormant |
| Setup difficulty | easy | moderate | easy |
| Complexity | 1/5 | 2/5 | 2/5 |
| Audience | data | general | data |
Figures from each repo's GitHub metadata at analysis time.
No setup required, just download the CSV files from the clean-data folder and start analyzing.
The 2016 New Coder Survey is a public dataset capturing who is learning to code, where they come from, and what their goals are. freeCodeCamp and Code Newbie ran the survey to better understand the people entering the programming world, and this repository holds the results for anyone to explore. The repo contains two main folders of data. The raw-data folder has the original survey responses in spreadsheet format. The clean-data folder has a tidied version of those responses, with inconsistencies smoothed out so that people can more easily build charts and run analyses without first having to wrangle the data themselves. This is a resource for data analysts, journalists, educators, or anyone curious about the learn-to-code movement. You might use it to build visualizations showing what bootcamps are most popular, what demographics dominate the new-coder population, or what job outcomes learners expect. The project actively invites people to create charts (using a visualization library called d3.js) and submit their findings. Several analyses are already linked from the README, including work hosted on Kaggle and various personal blogs. The project is community-driven. A small team of volunteers from freeCodeCamp's data science chatroom prepared the cleaned dataset and coordinated the presentation of early charts, and the repo's issue tracker still lists questions that could be answered with further visualizations. The data is shared under the Open Database License, which means you're free to use, share, and adapt it as long as you credit the source and keep any derived databases similarly open.
A public dataset of over 15,000 people learning to code, capturing their backgrounds, goals, and expectations. It includes raw survey responses and a cleaned version ready for analysis.
Mainly R. The stack also includes R, d3.js.
Dormant — no commits in 2+ years (last push 2017-10-05).
You are free to use, share, and adapt this data as long as you credit the source and keep any derived databases under the same open license.
Setup difficulty is rated easy, with roughly 5min to a first successful run.
Mainly data.
This repo across BitVibe Labs
Don't trust strangers blindly. Verify against the repo.