Skip to main content
100%

Pima Indians Diabetes Dataset (Scatter Plot)

✓ Published1🌍 Public
Sskasliwal@wpi.edu
Last edited Oct 4, 2025
Created on Oct 4, 2025
Forked from Diabetes Data

This scatter plot visualizes the Pima Indians Diabetes Dataset, plotting BMI on the x-axis against Glucose concentration on the y-axis for individual patients. Each point is color-coded by the binary outcome (blue for non-diabetic, red for diabetic), using a categorical color scale with blue and red hues. The chart is rendered with D3.v7, using core APIs such as `scaleLinear`, `scaleOrdinal`, `extent`, and the `.join()` modifier to build the SVG. The data is loaded asynchronously from a CSV file containing diagnostic measurements from the National Institute of Diabetes and Digestive and Kidney Diseases.

AI-generated description

Pima Indians Diabetes Dataset: This dataset is derived from the National Institute of Diabetes and Digestive and Kidney Diseases and is commonly used for binary classification tasks in machine learning. It contains diagnostic measurements from a group of Pima Indian women to predict the onset of diabetes.

Source URL: https://www.kaggle.com/datasets/uciml/pima-indians-diabetes-database

Attribute Descriptions Pregnancies: Quantitative (Number of times pregnant).

Glucose: Quantitative (Plasma glucose concentration).

BloodPressure: Quantitative (Diastolic blood pressure).

SkinThickness: Quantitative (Triceps skin fold thickness).

Insulin: Quantitative (2-hour serum insulin).

BMI: Quantitative (Body Mass Index).

DiabetesPedigreeFunction: Quantitative (A function that scores the likelihood of diabetes based on family history).

Age: Quantitative (Age in years).

Outcome: Categorical (0 = non-diabetic, 1 = diabetic).

MIT Licensed

Similar vizzes