| 2004 | Febrl - A Parallel Open Source Data Linkage System: http://datamining.anu.edu.au/linkage.html. | Peter Christen, Tim Churches, Markus Hegland |
| 2004 | CMTreeMiner: Mining Both Closed and Maximal Frequent Subtrees. | Yun Chi, Yirong Yang, Yi Xia, Richard R. Muntz |
| 2004 | Temporal Sequence Associations for Rare Events. | Jie Chen, Hongxing He, Graham J. Williams, Huidong Jin |
| 2004 | Mining Association Rules from Structural Deltas of Historical XML Documents. | Ling Chen, Sourav S. Bhowmick, Liang-Tien Chia |
| 2004 | Discovering Ordered Tree Patterns from XML Queries. | Yi Chen |
| 2004 | DB-Subdue: Database Approach to Graph Mining. | Sharma Chakravarthy, Ramji Beera, Ramanathan Balachandran |
| 2004 | Discovering Partial Periodic Patterns in Discrete Data Sequences. | Huiping Cao, David W. Cheung, Nikos Mamoulis |
| 2004 | Evaluating the Replicability of Significance Tests for Comparing Learning Algorithms. | Remco R. Bouckaert, Eibe Frank |
| 2004 | FP-Bonsai: The Art of Growing and Pruning Small FP-Trees. | Francesco Bonchi, Bart Goethals |
| 2004 | Constraint-Based Mining of Formal Concepts in Transactional Data. | Jrmy Besson, Cline Robardet, Jean-Franois Boulicaut |
| 2004 | Self-Similar Mining of Time Association Rules. | Daniel Barbar, Ping Chen, Zohreh Nazeri |
| 2004 | Semantic Sequence Kin: A Method of Document Copy Detection. | Jun-Peng Bao, Jun-Yi Shen, Xiao-Dong Liu, Haiyan Liu, Xiaodi Zhang |
| 2004 | Learning Hidden Markov Model Topology Based on KL Divergence for Information Extraction. | Kwok-Chung Au, Kwok-Wai Cheung |
| 2004 | A Semi-automatic System for Tagging Specialized Corpora. | Ahmed Amrani, Yves Kodratoff, Oriane Matte-Tailliez |
| 2004 | The Application of Emerging Patterns for Improving the Quality of Rare-Class Classification. | Hamad Alhammady, Kotagiri Ramamohanarao |
| 2003 | When to Update the Sequential Patterns of Stream Data? | Qingguo Zheng, Ke Xu, Shilong Ma |
| 2003 | Correlation Analysis of Spatial Time Series Datasets: A Filter-and-Refine Approach. | Pusheng Zhang, Yan Huang, Shashi Shekhar, Vipin Kumar |
| 2003 | Comparison of the Performance of Center-Based Clustering Algorithms. | Bin Zhang |
| 2003 | An Integrated System of Mining HTML Texts and Filtering Structured Documents. | Bo-Hyun Yun, Myungeun Lim, Soo-Hyun Park |
| 2003 | Improving Performance of Decision Tree Algorithms with Multi-edited Nearest Neighbor Rule. | Chenzhou Ye, Jie Yang, Lixiu Yao, Nian-yi Chen |
| 2003 | Weighted Proportional k-Interval Discretization for Naive-Bayes Classifiers. | Ying Yang, Geoffrey I. Webb |
| 2003 | AGRID: An Efficient Algorithm for Clustering Large High-Dimensional Datasets. | Yanchang Zhao, Song Junde |
| 2003 | A New Clustering Algorithm for Transaction Data via Caucus. | Jinmei Xu, Hui Xiong, Sam Yuan Sung, Vipin Kumar |
| 2003 | A Semi-supervised Algorithm for Pattern Discovery in Information Extraction from Textual Data. | Tianhao Wu, William M. Pottenger |
| 2003 | A Graph-Based Optimization Algorithm for Website Topology Using Interesting Association Rules. | Edmond HaoCun Wu, Michael K. Ng |