Professional Documents
Culture Documents
Kidney Failure Due To Diabetics - Detection Using Classification Algorithm in Data Mining
Kidney Failure Due To Diabetics - Detection Using Classification Algorithm in Data Mining
Abstract: In order to analyse the chosen data from various The approximate prevalence of kidney failure is 800 per
points of view, data mining is used as the effective process. million population (pmp), and the incidence of end-stage renal
This process is also used to sum-up all those views into useful disease (ESRD) is 150–200 pmp. The most common cause of
information. There are several types of algorithms in data kidney failure in population-based studies is diabetic
mining such as Classification algorithms, Regression, nephropathy. Early detection of kidney failure is essential for
Segmentation algorithms, association algorithms, sequence reducing life losses. The estimation of the ultimate result of a
analysis algorithms, etc.,. The classification algorithm can be disease and the analysis of the course it is likely to take is
used to bifurcate the data set from the given data set and called prognosis. The hope for the successful treatment of the
foretell one or more discrete variables, based on the other disease is caused by the prognosis of the affected patient.
attributes in the dataset. The ID3 (Iterative Dichotomiser 3) Prognostic information is generated by statement of prognosis.
algorithm is an original data set S as the root node. An In the formulation of the prognosis certain bits of information
unutilised attribute of the data set S calculates the entropy H(S) are used. These pieces of information are related to the obvious
(or Information gain IG (A)) of the attribute. Upon its yield of the disease. Such information can be called in other
selection, the attribute should have the smallest entropy (or words prognostic factors. This paper is structured as follows:
largest information gain) value. The prime objective of this Section 2 contains the review concepts of pre processing
paper is to analyze the data from a Kidney disorder due to method, ID3 and Diabetics nephropathy. Section 3 contains
diabetics by using classification technique to predict class existing method. Section 4 explains our proposed method. In
accurately. section 5 Results are discussed and section 6 has the
conclusion.
Keywords: Data mining, Classification algorithm, Decision
tree, medical dataset. 2. BASIC CONCEPTS
produce a decision tree from it, this ID3 is used. ID3 is used thickness(mm)
even before C4.5 algorithm. This typical algorithm enables 4 Insulin 2-Hour serum insulin (mu
learning through machines and is useful in the areas pertaining U/ml)
to the processing of natural languages. This algorithm known 5 Pregnancy Number of times pregnant
as ID3 starts with the root node which is the original set S. This 6 Mass Body Mass Index(BMI)
algorithm iterates through each unutilised feature of the set S. 7 Pedigree Diabetes Pedigree function
This goes on in every iteration and leads to the calculation of 8 Age Age(in years)
the entropy H(S) of that particular attribute. The attribute with 9 Class Class variable(0 or 1)
the smallest entropy or largest gain of information is chosen by
it. Treatment for kidney failure
Information gain If your kidneys can't keep up with waste and fluid clearance on
Used by the ID3 tree generation algorithms. Information Gain their own and you develop complete or near-complete kidney
is based on the concept of Entropy from Information Theory. failure, you have end-stage kidney disease. At that point, you
In order to generate subsets of the data . the chosen attribute is need dialysis or a kidney transplant.
used to split the set ‘S’. Dialysis. Dialysis artificially removes waste products and
extra fluid from your blood when your kidneys can no
longer do this. In hemodialysis, a machine filters waste
and excess fluids from your blood. In peritoneal dialysis, a
Diabetic nephropathy thin tube (catheter) inserted into your abdomen fills your
Kidney disease that results from diabetes is the number one cause abdominal cavity with a dialysis solution that absorbs
of kidney failure. Almost a third of people with diabetes develop waste and excess fluids. After a period of time, the dialysis
diabetic nephropathy.People with diabetes and kidney disease solution drains from your body, carrying the waste with it.
suffer more overall than people with only kidney disease. This is Kidney transplant. A kidney transplant involves
because people with diabetes tend to have other long-standing surgically placing a healthy kidney from a donor into
medical conditions, like high blood pressure, high cholesterol, your body. Transplanted kidneys can come from
and blood vessel disease (atherosclerosis). People with diabetes deceased or living donors. You'll need to take
are also more likely to have other kidney-related problems, such medications for the rest of your life to keep your body
as bladder infections and nerve damage to the bladder. from rejecting the new organ. You don't need to be on
Kidney disease in type 1 diabetes is slightly different than in type dialysis to have a kidney transplant.
2 diabetes. In type 1 diabetes, kidney disease rarely begins in the
first 10 years after diagnosis of diabetes. In type 2 diabetes, some
patients already have kidney disease by the time they are
diagnosed with diabetes.
predict chronic kidney disease (CKD) using relevant dataset and Support Vector Machine. International Journal of
available at UCI machine learning repository. Computer Trends and Technology. 2014; 11:94-8.
3. PROPOSED METHOD
5. CONCLUSION
Reference
[1] Joshi J, Rinal D, Patel J, Diagnosis And Prognosis of
Breast Cancer Using Classification Rules, International
Journal of Engineering Research and General
Science,2(6):315-323, October 2014.
[2] Kusiak, A., Dixon, B., & Shah, S., 2005. Predicting
survival time for kidney dialysis patients: a data mining
approach. Computers in biology and medicine, 35(4), pp.
311-327.
[3] Mukesh kumari and Dr. Rajan Vohra,“Prediction of
Diabetes Using Bayesian Network,”in proceeding of
International Journal of Computer Science and
Information Technologies, vol.5 , 2014
[4] Iyer Aiswarya, Jeyalatha S and Sumbaly Ronak. Diagnosis
of diabetes using classification mining techniques.
International Journal of Data Mining & Knowledge
Management Process. 2015; 5:1-14.
[5] Sanakal Ravi and Jayakumari T. Prognosis of Diabetes
Using Data mining Approach-Fuzzy C Means Clustering
64