I try to get my hands dirty on Machine Learning (a fancy name for statistics + computer programming), Natural Language Processing, Big Data Processing, and ML guided Artificial Intelligence.
My research objective is to develop mathematical and computational models for data that are simple, stands on comprehensible theoretical foundation and withstand the test of rigorous experimental evaluation. I put considerable effort behind any research work that I do to fulfil the above three criteria.
I enjoy writing non-trivial and efficient computer program for large data. In fact, thinking about efficient data structures, algorithms and their implementation on real data using a white box programming language is an intrinsic part of my thought process.
-
A Stratified Seed Selection Algorithm for K-means Clustering on Big Data by Bajpai N., Paik J. H., Sarkar S. IEEE Transactions on Artificial Intelligence - (Accepted/In-Press)
-
A wrapper feature selection approach using Markov blankets by Hassan A., Paik J. H., Khare S. R. Pattern Recognition - (2025)
-
Balanced Seed Selection for K-means Clustering with Determinantal Point Process by Bajpai N., Paik J. H., Sarkar S. Pattern Recognition - (Accepted/In-Press)
Principal Investigator
- AI Driven App for Silkworm Counting using Deep Learning Approaches CENTRAL SERICULTURE RESEARCH
- TCS ION Industry Honor Certification Product: Big Data Analytics Advanced Tata Consultancy Services
Ph. D. Students
Dipojjwal Ghosh
Area of Research: marketing analytics
Nazeer Haider
Area of Research: Machine Learning
Sayantan Saha
Area of Research: Machine Learning
Soumyadipta Banerjee
Area of Research: Deep Learning
Sudip Kumar Bhattacharya
Area of Research: Complex system simulation
Trishita Mukherjee
Area of Research: Graph Machine Learning