/


As one of the task leaders, I am going to walk you through the steps we took in solving this challenge. It is less about the data and results, more about what you do with them.
Is there anything about student debt that we do not know about?
Omdena’s method of branching out into different task groups allows the freedom to cover many sides of a problem. The content of the task brings a team together — an idea is proposed and then voted into a task, with team members joining in, setting the agenda, and collectively taking decisions. As somebody with years of experience in the Higher Education (HE) sector in the UK, I felt that I am familiar with the subject matter and proposed the scope of our task to be the impact of Student Debt Crisis on different demographic groups (namely: female students, ethnic minority students, and first-generation students) and finding an AI solution for this student debt crisis. I chose this for the following reasons:
So we set our task goal to explore how student debt affects the different demographic groups, why, and what can be done about it through the multitude of data available.
Keep it simple.
Being a community of ML and AI practitioners, the suggestions that came were regarding modeling, clustering sentiment analysis, web scraping, but there were also voices for the more traditional data analysis for this Student Debt Crisis Challenge. We ended up choosing the latter because:
Mapping the Student Debt Crisis journey.
With so many pieces of data, there is a danger of ending up with a collection of separate pieces of data, each one interesting in itself, but disjointed.
Having a framework beforehand would allow us to focus when researching sources, and also overcome one of the problems often found when applying machine learning –the ‘now what’ moment, when we see the data but cannot say what it means.
So, we created a map of the student loan journey, along with the questions whose answers we are aiming to find in data.
To answer all these questions, we sought data for different demographic groups, as per our task goal.
Aim for variety.
Luckily, as is the case with other regulated sectors, there was no lack of data. The mixture of statutory, financial, survey, trend, and official stats data, apart from allowing us to cross-validate findings, allowed us to build a rich, multifaceted picture. At one point we had nearly 50 pieces of data in our shortlist to investigate — a challenging task even in itself. Some sources, like the College Scorecard dataset, were massive, and we had to apply a fair amount of manipulation and standardization. Some of the sources came with long definition lists, which took a while to unpick. The agile nature of the project allowed us to sneak in some data on the impact of COVID-19 at the last minute, making it even more relevant.
The last mile
I am going through these steps together. Often analysis relies on data speaking for itself and falls short of interpreting and giving recommendations. However, having a framework from the beginning, breaking down the big questions into small ones, and linking the questions to data, we can see areas that can be acted upon.
For example, there is no silver bullet for tackling drop-out, but small steps can be taken, to challenge institutions about the completion gap between the different demographic groups, to facilitate non-punishing transfers, to set work/study standards, and to use AI and predictive analytics in order to improve degree completion during this Student Debt Crisis.
Some of the answers:
In a nutshell: Ethnic minority students borrow more because they cannot rely on their families, they opt for more expensive institutions, and, because of lower financial literacy, may end up borrowing at disadvantageous terms.
What can be done about it?
Increase general financial literacy, encourage planning in advance, establish peer advice network.
Some of the answers:
In short: Having higher levels of debt does not necessarily translate into studying attractive subjects or landing at well-performing institutions.
What can be done about it?
Monitor and publicize institution data, advise on beneficial loan decisions, tackle the differences in STEM subject take-up.
Some of the answers:
Summing up: Ethnic minority students are more likely to fall at the very first hurdle — that of degree completion, and having to work during study further jeopardizes their chances of completion.
What can be done about it?
Challenge institutions over retention rates and/or gap, deploy predictive analytics solutions to reduce dropout, advise cap on work hours, ensure financially fair transfers
Some of the answers:
In short: The gender wage gap persists and is even more pronounced for advanced degree holders, which risks making entire industries short of women leaders. COVID-19 disproportionately affects holders of associate degrees, tend to be favored by Hispanic students.
What can be done about it?
Support women in STEM and be clear on employment reality for advanced degree holders. Use the corona-virus crisis to retrain and up-skill vulnerable groups, including associate degree holders.
Some of the answers:
In summary: Students from for-profit institutions (mostly female, ethnic minority, and first-generation students) display higher degrees of stress and are clearly disillusioned by their higher education experience.
What can be done about it?
Make institution performance and loan conditions transparent, invest in debt counseling, emergency support, and relief, promote an unbiased view of the value of higher education and alternatives to a degree.
Even with a problem that has been addressed extensively, collective wisdom from an autonomous and diverse group can bring new insights. Contrary to the popular belief, lack of data was not the problem — it was scoping and breaking down the problem in the context of too much data, which is where collaboration and diverse thinking really helped. Yes, a self-organized international community presented challenges in agreeing times for meeting, brainstorming and feedback, but the colorful side of collaboration — the assorted and inconsistent graphs, as many styles as team members, was a welcome difference from the monotonous slide templates from our day jobs. In the end, there was a positive feeling that we have done our small bit to help to understand this big problem better and that it was a time well spent getting to know and learning from each other.
This article is written by Galina Naydenova.

From Orbit to Harvest: Inside TerraYield, a Multimodal Dataset for Smarter Crop Yield Forecasting

AI for Sustainable Farming: Tackling Greenhouse Gas Emissions and Empowering Responsible Finance

A Beginner’s Guide to Exploratory Data Analysis with Python

AI-Driven Personalized Content Recommendations: Revolutionizing User Engagement in Learning Apps