Fetching the paper…
Reading the bibliography…
The amount of text that is generated every day is increasing dramatically.
Development of a stemming algorithm
Julie B Lovins. 1968 · 1968
Earlier work this paper cites.
A vector space model for automatic indexing
Gerard Salton, Anita Wong, and Chung-Shu Yang. 1975 · 1975
Earlier work this paper cites.
An algorithm for suffix stripping
Martin F Porter. 1980 · 1980
Earlier work this paper cites.
Estimation of Dependences Based on Empirical Data: Springer Series in Statistics (Springer Series in Statistics)
Vladimir Vapnik. 1982 · 1982
Earlier work this paper cites.
A survey of recent advances in hierarchical clustering algorithms
Fionn Murtagh. 1983 · 1983
Earlier work this paper cites.
Classification and regression trees
Leo Breiman, Jerome Friedman, Charles J Stone, and Richard A Olshen. 1984 · 1984
Earlier work this paper cites.
Complexities of hierarchic clustering algorithms: State of the art
Fionn Murtagh. 1984 · 1984
Earlier work this paper cites.
Classification algorithms
Mike James. 1985 · 1985
Earlier work this paper cites.
Induction of decision trees
J. Ross Quinlan. 1986 · 1986
Earlier work this paper cites.
Algorithms for clustering data
Anil K Jain and Richard C Dubes. 1988 · 1988
Earlier work this paper cites.
Term-weighting approaches in automatic text retrieval
Gerard Salton and Christopher Buckley. 1988 · 1988
Earlier work this paper cites.
Recent trends in hierarchic document clustering: a critical review
Peter Willett. 1988 · 1988
Earlier work this paper cites.
A tutorial on hidden Markov models and selected applications in speech recognition
Lawrence Rabiner. 1989 · 1989
Earlier work this paper cites.
Inference networks for document retrieval. In Proceedings of the 13th annual international ACM SIGIR conference on Research and development in information retrieval
Howard Turtle and W Bruce Croft. 1989 · 1989
Earlier work this paper cites.
Scatter/gather: A cluster-based approach to browsing large document collections. In Proceedings of the 15th annual international ACM SIGIR conference on Research and development in information retrieval
Douglass R Cutting, David R Karger, Jan O Pedersen, and John W Tukey. 1992 · 1992
Earlier work this paper cites.
Knowledge discovery in databases: An overview
William J Frawley, Gregory Piatetsky-Shapiro, and Christopher J Matheus. 1992 · 1992
Earlier work this paper cites.
Tokenization as the initial phase in NLP. In Proceedings of the 14th conference on Computational linguistics-Volume 4
Jonathan J Webster and Chunyu Kit. 1992 · 1992
Earlier work this paper cites.
Constant interaction-time scatter/gather browsing of very large document collections. In Proceedings of the 16th annual international ACM SIGIR conference on Research and development in information retrieval
Douglass R Cutting, David R Karger, and Jan O Pedersen. 1993 · 1993
Earlier work this paper cites.
A. The Unified Medical Language System
AT Mc Cray. 1993 · 1993
Earlier work this paper cites.
Support-vector networks
Corinna Cortes and Vladimir Vapnik. 1995 · 1995
Earlier work this paper cites.
Latent semantic indexing. In Proceedings of the Text Retrieval Conference
S Dumais, G Furnas, T Landauer, S Deerwester, S Deerwester, et al · 1995
Earlier work this paper cites.
Knowledge Discovery in Textual Databases (KDT).. In KDD
Ronen Feldman and Ido Dagan. 1995 · 1995
Earlier work this paper cites.
Data mining: an overview from a database perspective
Ming-Syan Chen, Jiawei Han, and Philip S. Yu. 1996 · 1996
Earlier work this paper cites.
Information extraction
Jim Cowie and Wendy Lehnert. 1996 · 1996
Earlier work this paper cites.
Knowledge Discovery and Data Mining: Towards a Unifying Framework.. In KDD
Usama M Fayyad, Gregory Piatetsky-Shapiro, Padhraic Smyth, et al · 1996
Earlier work this paper cites.
Introducing markov chain monte carlo
Walter R Gilks, Sylvia Richardson, and David J Spiegelhalter. 1996 · 1996
Earlier work this paper cites.
Stemming algorithms: A case study for detailed evaluation
David A Hull et al · 1996
Earlier work this paper cites.
A Probabilistic Analysis of the Rocchio Algorithm with TFIDF for Text Categorization
Thorsten Joachims. 1996 · 1996
Earlier work this paper cites.
A new probabilistic model of text classification and retrieval
Tom Kalt and WB Croft. 1996 · 1996
Earlier work this paper cites.
Combining classifiers in text categorization. In Proceedings of the 19th annual international ACM SIGIR conference on Research and development in information retrieval
Leah S Larkey and W Bruce Croft. 1996 · 1996
Earlier work this paper cites.
Bow: A toolkit for statistical language modeling, text retrieval, classification and clustering
Andrew Kachites McCallum. 1996 · 1996
Earlier work this paper cites.
An efficient k-means clustering algorithm
Khaled Alsabti, Sanjay Ranka, and Vineet Singh. 1997 · 1997
Earlier work this paper cites.
Exploiting clustering and phrases for context-based information retrieval. In ACM SIGIR Forum
Peter G Anick and Shivakumar Vaithyanathan. 1997 · 1997
Earlier work this paper cites.
Nymble: a high-performance learning name-finder. In Proceedings of the fifth conference on Applied natural language processing
Daniel M Bikel, Scott Miller, Richard Schwartz, and Ralph Weischedel. 1997 · 1997
Earlier work this paper cites.
Using taxonomy, discriminants, and signatures for navigating in text databases. In VLDB
Soumen Chakrabarti, Byron Dom, Rakesh Agrawal, and Prabhakar Raghavan. 1997 · 1997
Earlier work this paper cites.
A decision-theoretic generalization of on-line learning and an application to boosting
Yoav Freund and Robert E Schapire. 1997 · 1997
Earlier work this paper cites.
Decision tree classification of land cover from remotely sensed data
Mark A Friedl and Carla E Brodley. 1997 · 1997
Earlier work this paper cites.
Hierarchically classifying documents using very few words
Daphne Koller and Mehran Sahami. 1997 · 1997
Earlier work this paper cites.
Machine learning. 1997
Tom M Mitchell. 1997 · 1997
Earlier work this paper cites.
Training support vector machines: an application to face detection. In Computer Vision and Pattern Recognition, 1997. Proceedings., 1997 IEEE Computer Society Conference on
Edgar Osuna, Robert Freund, and Federico Girosi. 1997 · 1997
Earlier work this paper cites.
Distributional clustering of words for text classification. In Proceedings of the 21st annual international ACM SIGIR conference on Research and development in information retrieval
L Douglas Baker and Andrew Kachites McCallum. 1998 · 1998
Earlier work this paper cites.
Refining Initial Points for K-Means Clustering.. In ICML
Paul S Bradley and Usama M Fayyad. 1998 · 1998
Earlier work this paper cites.
A tutorial on support vector machines for pattern recognition
Christopher JC Burges. 1998 · 1998
Earlier work this paper cites.
A survey of information retrieval and filtering methods
Christos Faloutsos and Douglas W Oard. 1998 · 1998
Earlier work this paper cites.
Text categorization with support vector machines: Learning with many relevant features
Thorsten Joachims. 1998 · 1998
Earlier work this paper cites.
Naive (Bayes) at forty: The independence assumption in information retrieval
David D Lewis. 1998 · 1998
Earlier work this paper cites.
Learning to classify text from labeled and unlabeled documents
Kamal Nigam, Andrew McCallum, Sebastian Thrun, and Tom Mitchell. 1998 · 1998
Earlier work this paper cites.
Text mining: natural language techniques and text mining applications
Martin Rajman and Romaric Besançon. 1998 · 1998
Earlier work this paper cites.
A Bayesian approach to filtering junk e-mail. In Learning for Text Categorization: Papers from the 1998 workshop
Mehran Sahami, Susan Dumais, David Heckerman, and Eric Horvitz. 1998 · 1998
Earlier work this paper cites.
Support vector machines for spam categorization
Harris Drucker, S Wu, and Vladimir N Vapnik. 1999 · 1999
Cited alongside, same era.
Probabilistic latent semantic indexing. In Proceedings of the 22nd annual international ACM SIGIR conference on Research and development in information retrieval
Thomas Hofmann. 1999 · 1999
Cited alongside, same era.
Foundations of statistical natural language processing
Christopher D Manning, Hinrich Schütze, et al · 1999
Cited alongside, same era.
Automatic term identification and classification in biology texts. In Proc. of the 5th NLPRS
Chikashi Nobata, Nigel Collier, and Jun-ichi Tsujii. 1999 · 1999
Cited alongside, same era.
A re-examination of text categorization methods. In Proceedings of the 22nd annual international ACM SIGIR conference on Research and development in information retrieval
Yiming Yang and Xin Liu. 1999 · 1999
Cited alongside, same era.
Natural language processing and text mining
Anne Kao and Stephen R Poteet. 2007 · 2007
Later among the works it cites.
Mixtures of hierarchical topics with pachinko allocation. In Proceedings of the 24th international conference on Machine learning
David Mimno, Wei Li, and Andrew McCallum. 2007 · 2007
Later among the works it cites.
Probabilistic topic models
Mark Steyvers and Tom Griffiths. 2007 · 2007
Later among the works it cites.
Extraction of semantic biomedical relations from text using conditional random fields
Markus Bundschus, Mathaeus Dejori, Martin Stetter, Volker Tresp, and Hans-Peter Kriegel. 2008 · 2008
Later among the works it cites.
Modeling documents by combining semantic concepts with unsupervised statistical learning
Chaitanya Chemudugunta, America Holloway, Padhraic Smyth, and Mark Steyvers. 2008 · 2008
Later among the works it cites.
Getting started in text mining
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Extracting the names of genes and gene products with a hidden Markov model. In Proceedings of the 18th conference on Computational linguistics-Volume 1
Nigel Collier, Chikashi Nobata, and Jun-ichi Tsujii. 2000 · 2000
Cited alongside, same era.
Integration of data mining and relational databases. In Proceedings of the 26th International Conference on Very Large Databases, Cairo, Egypt
Amir Netz, Surajit Chaudhuri, Jeff Bernhardt, and Usama Fayyad. 2000 · 2000
Cited alongside, same era.
BoosTexter: A boosting-based system for text categorization
Robert E Schapire and Yoram Singer. 2000 · 2000
Cited alongside, same era.
A comparison of document clustering techniques. In KDD workshop on text mining
Michael Steinbach, George Karypis, Vipin Kumar, et al · 2000
Cited alongside, same era.
The nature of statistical learning theory
Vladimir Vapnik. 2000 · 2000
Cited alongside, same era.
On feature distributional clustering for text categorization. In Proceedings of the 24th annual international ACM SIGIR conference on Research and development in information retrieval
Ron Bekkerman, Ran El-Yaniv, Naftali Tishby, and Yoad Winter. 2001 · 2001
Cited alongside, same era.
Text categorization using weight adjusted k-nearest neighbor classification
Eui-Hong Sam Han, George Karypis, and Vipin Kumar. 2001 · 2001
Cited alongside, same era.
K Bretonnel Cohen and Lawrence Hunter. 2008 · 2008
Later among the works it cites.
Cascaded classifiers for confidence-based chemical named entity recognition
Peter Corbett and Ann Copestake. 2008 · 2008
Later among the works it cites.
Information retrieval: a health and biomedical perspective
William Hersh. 2008 · 2008
Later among the works it cites.
Introduction to information retrieval
Christopher D Manning, Prabhakar Raghavan, and Hinrich Schütze. 2008 · 2008
Later among the works it cites.
Opinion mining and sentiment analysis
Bo Pang and Lillian Lee. 2008 · 2008
Later among the works it cites.
Information extraction
Sunita Sarawagi et al · 2008
Later among the works it cites.
How to make the most of NE dictionaries in statistical NER
Yutaka Sasaki, Yoshimasa Tsuruoka, John McNaught, and Sophia Ananiadou. 2008 · 2008
Later among the works it cites.
Biomedical Ontologies and Text Mining for Biomedicine and Healthcare: A Survey
Illhoi Yoo and Min Song. 2008 · 2008
Later among the works it cites.
An investigation on integrating XML-based security into Web services. In GCC Conference & Exhibition, 2009 5th IEEE
Mahmood Doroodchi, Azadeh Iranmehr, and Seyed Amin Pouriyeh. 2009 · 2009
Later among the works it cites.
Relational data mining
Sašo Džeroski. 2009 · 2009
Later among the works it cites.
Finding groups in data: an introduction to cluster analysis
Leonard Kaufman and Peter J Rousseeuw. 2009 · 2009
Later among the works it cites.
Secure SMS Banking Based On Web Services. In SWWS
Seyed Amin Pouriyeh and Mahmood Doroodchi. 2009 · 2009
Later among the works it cites.
Event extraction for systems biology by text mining the literature
Sophia Ananiadou, Sampo Pyysalo, Jun’ichi Tsujii, and Douglas B Kell. 2010 · 2010
Later among the works it cites.
Biomedical question answering: A survey
Sofia J Athenikos and Hyoil Han. 2010 · 2010
Later among the works it cites.
Integrating out multinomial parameters in latent Dirichlet allocation and naive bayes for collapsed Gibbs sampling
Bob Carpenter. 2010 · 2010
Later among the works it cites.
Exploiting background knowledge for relation extraction. In Proceedings of the 23rd International Conference on Computational Linguistics
Yee Seng Chan and Dan Roth. 2010 · 2010
Later among the works it cites.
Multilabel classification with meta-level features. In Proceedings of the 33rd international ACM SIGIR conference on Research and development in information retrieval
Siddharth Gopal and Yiming Yang. 2010 · 2010
Later among the works it cites.
Secure Mobile Approaches Using Web Services.. In SWWS
Seyed Amin Pouriyeh, Mahmood Doroodchi, and MR Rezaeinejad. 2010 · 2010
Later among the works it cites.
Mayo clinical Text Analysis and Knowledge Extraction System (cTAKES): architecture, component evaluation and applications
Guergana K Savova, James J Masanz, Philip V Ogren, Jiaping Zheng, Sunghwan Sohn, Karin C Kipper-Schuler, and Christopher G Chute. 2010 · 2010
Later among the works it cites.
Exploiting syntactico-semantic structures for relation extraction. In Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics: Human Language Technologies-Volume 1
Yee Seng Chan and Dan Roth. 2011 · 2011
Later among the works it cites.
Data mining: concepts, models, methods, and algorithms
Mehmed Kantardzic. 2011 · 2011
Later among the works it cites.
Entity disambiguation with hierarchical topic models. In Proceedings of the 17th ACM SIGKDD international conference on Knowledge discovery and data mining
Saurabh S Kataria, Krishnan S Kumar, Rajeev R Rastogi, Prithviraj Sen, and Srinivasan H Sengamedu. 2011 · 2011
Later among the works it cites.
BioGraph: unsupervised biomedical knowledge discovery via automated hypothesis generation
Anthony ML Liekens, Jeroen De Knijf, Walter Daelemans, Bart Goethals, Peter De Rijk, and Jurgen Del-Favero. 2011 · 2011
Later among the works it cites.
Adapting centroid classifier for document categorization
Songbo Tan, Yuefen Wang, and Gaowei Wu. 2011 · 2011
Later among the works it cites.
Data mining and statistics for decision making
Stéphane Tufféry. 2011 · 2011
Later among the works it cites.
Automatic acquisition of huge training data for bio-medical named entity recognition. In Proceedings of BioNLP 2011 Workshop
Yu Usami, Han-Cheol Cho, Naoaki Okazaki, and Jun’ichi Tsujii. 2011 · 2011
Later among the works it cites.
A developer’s guide to the semantic Web
Liyang Yu. 2011 · 2011
Later among the works it cites.
Mining text data
Charu C Aggarwal and ChengXiang Zhai. 2012 · 2012
Later among the works it cites.
Pattern classification
Richard O Duda, Peter E Hart, and David G Stork. 2012 · 2012
Later among the works it cites.
A Bayesian feature selection paradigm for text classification
Guozhong Feng, Jianhua Guo, Bing-Yi Jing, and Lizhu Hao. 2012 · 2012
Later among the works it cites.
Mining social media: a brief introduction
Pritam Gundecha and Huan Liu. 2012 · 2012
Later among the works it cites.
An entity-topic model for entity linking. In Proceedings of the 2012 Joint Conference on Empirical Methods in Natural Language Processing and Computational Natural Language Learning
Xianpei Han and Le Sun. 2012 · 2012
Later among the works it cites.
Collective context-aware topic models for entity disambiguation. In Proceedings of the 21st international conference on World Wide Web
Prithviraj Sen. 2012 · 2012
Later among the works it cites.
Social media mining for drug safety signal detection. In Proceedings of the 2012 international workshop on Smart health and wellbeing
Christopher C Yang, Haodong Yang, Ling Jiang, and Mi Zhang. 2012 · 2012
Later among the works it cites.
Use HMM and KNN for classifying corneal data
Payam Porkar Rezaeiye, Mojtaba Sedigh Fazli, et al · 2014
Later among the works it cites.
On stopwords, filtering and data sparsity for sentiment analysis of twitter
Hassan Saif, Miriam Fernández, Yulan He, and Harith Alani. 2014 · 2014
Later among the works it cites.
The impact of preprocessing on text classification
Alper Kursat Uysal and Serkan Gunal. 2014 · 2014
Later among the works it cites.
Automatic topic labeling using ontology-based topic models. In Machine Learning and Applications (ICMLA), 2015 IEEE 14th International Conference on
Mehdi Allahyari and Krys Kochut. 2015 · 2015
Later among the works it cites.
From within host dynamics to the epidemiology of infectious disease: scientific overview and challenges
Juan B Gutierrez, Mary R Galinski, Stephen Cantrell, and Eberhard O Voit. 2015 · 2015
Later among the works it cites.
Discovering Coherent Topics with Entity Topic Models. In Web Intelligence (WI), 2016 IEEE/WIC/ACM International Conference on
Mehdi Allahyari and Krys Kochut. 2016a · 2016
Later among the works it cites.
Semantic Tagging Using Topic Models Exploiting Wikipedia Category Network. In Semantic Computing (ICSC), 2016 IEEE Tenth International Conference on
Mehdi Allahyari and Krys Kochut. 2016c · 2016
Later among the works it cites.
Text Summarization Techniques: A Brief Survey
M. Allahyari, S. Pouriyeh, M. Assefi, S. Safaei, E. D. Trippe, J. B. Gutierrez, and K. Kochut. 2017 · 2017
Closest in time.
E. D. Trippe, J. B. Aguilar, Y. H. Yan, M. V. Nural, J. A. Brady, M. Assefi, S. Safaei, M. Allahyari, S. Pouriyeh, M. R. Galinski, J. C. Kissinger, and J. B. Gutierrez. 2017 · 2017
Closest in time.
THE DIGITAL UNIVERSE IN 2020: Big Data, Bigger Digital Shadows, and Biggest Grow th in the Far East
John Gantz and David Reinsel. 2012 · 2020
Closest in time.