Stata Expert Needed for Survey Data Analysis (University Research Project)
Budget: €8 – €30 EUR
Hello,
We are a group of university students looking for someone skilled in data analysis using Stata to help us process the results of a student survey for an empirical research project.
We’ve already collected the data through a questionnaire (Excel file provided) and have a clear idea of what we want to analyze. The goal is to clean, restructure, and analyze the dataset to answer a research question about caffeine consumption and its impact on grades across different school years.
Here’s what we need help with:
• Reshape the dataset into long format (one row per individual and per school year – Year 1, Year 2, Year 3).
• Create binary and continuous variables: for example, a binary variable indicating whether the student consumed caffeine that year (0 = no, 1 = yes), same for grades, and a variable for the number of cups consumed.
• Add control variables: gender, faculty, type of linguistic background (unilingual/plurilingual), year studies began, etc.
• Run descriptive statistics (means, standard deviations, group comparisons by gender, faculty, etc.).
• If possible, conduct a Difference-in-Differences analysis — we’re open to your suggestions on how to best structure the groups and time periods.
• Provide clean outputs: main tables and simple graphs (bar charts, trends, etc.).
We expect a commented .do file, a .dta dataset with the cleaned data
Deadline: ideally within 7 to 10 days.
Thanks a lot!
We are a group of university students looking for someone skilled in data analysis using Stata to help us process the results of a student survey for an empirical research project.
We’ve already collected the data through a questionnaire (Excel file provided) and have a clear idea of what we want to analyze. The goal is to clean, restructure, and analyze the dataset to answer a research question about caffeine consumption and its impact on grades across different school years.
Here’s what we need help with:
• Reshape the dataset into long format (one row per individual and per school year – Year 1, Year 2, Year 3).
• Create binary and continuous variables: for example, a binary variable indicating whether the student consumed caffeine that year (0 = no, 1 = yes), same for grades, and a variable for the number of cups consumed.
• Add control variables: gender, faculty, type of linguistic background (unilingual/plurilingual), year studies began, etc.
• Run descriptive statistics (means, standard deviations, group comparisons by gender, faculty, etc.).
• If possible, conduct a Difference-in-Differences analysis — we’re open to your suggestions on how to best structure the groups and time periods.
• Provide clean outputs: main tables and simple graphs (bar charts, trends, etc.).
We expect a commented .do file, a .dta dataset with the cleaned data
Deadline: ideally within 7 to 10 days.
Thanks a lot!