The Fifth Elephant 2017

On data engineering and application of ML in diverse domains

What explains our marks?

Submitted by Anand S (@sanand0) on Wednesday, 24 May 2017

videocam_off

Technical level

Beginner

Section

Crisp talk for Data in Government track

Status

Confirmed & Scheduled

View proposal in schedule

Vote on this proposal

Login to vote

Total votes:  +8

Abstract

The NCERT put together a large-scale survey called the National Achievement Survey. This captured student performance across 4 subjects via 100 questions each, the demographics and behaviour of students, teachers and schools through 300 more questions.

The question was: what affects children’s marks? For example: How does TV watching affect their performance? Is this a bigger effect than playing sports? Is this uniform across states? Do tuitions help or hurt? How does this vary for rich parents vs poor parents?

This talk covers the techniques used to analyse the data, and how this generalises to arbitrary datasets. This has been encapsulated into an open source library called autolysis that we’ll be releasing for the talk.

The intended audience is a beginner to ML who wants to understand how simple algorithms can lead to powerful results if applied the right way. The audience will leave with a specific technique (that we call group-means) that helps identify root causes across any categorical dataset.

Outline

Slides are at https://learn.gramener.com/downloads/talks/2017-05-24-NAS-Autolysis.pptx

Requirements

None

Speaker bio

Anand is a co-founder of Gramener, a data science company. He leads a team of data enthusiasts with skills in analysis, design, programming and statistics.

He studied at IIT Madras, IIM Bangalore and LBS, and worked at IBM, Infosys, Lehman Brothers and BCG. He and his team explore insights from data and communicate these as visual stories.

Slides

https://learn.gramener.com/downloads/talks/2017-05-24-NAS-Autolysis.pptx

Comments

  • 1
    Zainab Bawa (@zainabbawa) Reviewer a year ago

    Thanks for this submission, Anand. Are there recent, newer use cases for this tool?

    • 1
      Zainab Bawa (@zainabbawa) Reviewer a year ago

      Or can this be a talk about autolysis, the technique itself?

  • 1
    Kunal Patel (@kppatel) a year ago

    I am eager to attend Anand’s talk.

Login with Twitter or Google to leave a comment