Patent
Classification method and apparatus
Pál Ruján,Harry Ser Systeme Ag Urbschat +1 more
- 06 Apr 2000
208
TL;DR: In this article, a method for building a classification model for classifying unclassified documents based on the classification of a plurality of documents which respectively have been classified as belonging to one of the plurality of classes is presented.
read more
Abstract: A method for building a classification model for classifying unclassified documents based on the classification of a plurality of documents which respectively have been classified as belonging to one of a plurality of classes, said documents being digitally represented in a computer, said documents respectively comprising a plurality of terms which respectively comprise one or more symbols of a finite set of symbols, and said method comprising the following steps: representing each of said plurality of documents by a vector of n dimensions, said n dimensions forming a vector space, whereas the value of each dimension of said vector corresponds to the frequency of occurrence of a certain term in the document corresponding to said vector, so that said n dimensions span up a vector space; representing the classification of said already classified documents into classes by separating said vector space into a plurality of subspaces by one or more hyperplanes, such that each subspace comprises one or more documents as represented by their corresponding vectors in said vector space, so that said each subspace corresponds to a class.
read more
Chat with Paper
AI Agents for this Paper
Find similar papers on Google Scholar, PubMed and Arxiv
Write a critical review of this paper
Analyze citations of this paper to find unaddressed research gaps
Citations
Patent
Methods and systems for mapping data items to sparse distributed representations
Francisco Eduardo De Sousa Webber
- 03 Aug 2015
TL;DR: In this article, a method of mapping data items to sparse distributed representations (SDRs) includes clustering in a two-dimensional metric space, by a reference map generator, a set of data documents selected according to at least one criterion, generating a semantic map.
17
Patent
Reclassification of training data to improve classifier accuracy
Rajesh Balchandran,Linda M. Boyer,Gregory Purdy +2 more
- 18 Jun 2007
TL;DR: In this article, a method of creating a statistical classification model for a classifier within a natural language understanding system can include processing training data using an existing statistical classifier, where sentences of the training data correctly classified into a selected class of the classification model can be selected.
17
Patent
System And Method For Displaying Relationships Between Concepts To Provide Classification Suggestions Via Nearest Neighbor
William C. Knight,Nicholas I. Nussbaum,John W. Conwell +2 more
- 27 Jul 2010
TL;DR: In this paper, a system and method for displaying relationships between concepts to provide classification suggestions via nearest neighbor is provided, and a set of uncoded concepts are compared with the reference concepts.
16
Patent
Dimensional compression using an analytic platform
Herbert Dennis Hunt,John Randall West,Marshall Ashby Gibbs,Bradley Michael Griglione,Gregory David Neil Hudson,Andrea Basilico,Arvid C. Johnson,Cheryl G. Bergeon,Craig Joseph Chapa,Alberto Agostinelli,Jay Alan Yusko,Trevor Mason +11 more
- 31 Jan 2008
TL;DR: In this article, the authors present a set of systems and methods that may involve receiving a causal fact dataset including facts relating to items perceived to cause actions, wherein the causal fact data includes a data attribute that is associated with a causal datum.
16
Patent
Systems, methods, and media for web content management
Martchenko Serguei,Marvin Smit,Pannekoek Rick,De Voogd Erik,De Vries Renze A +4 more
- 29 Jan 2011
TL;DR: In this article, the authors present a web content management application via a web site, generating a web marketing campaign from at least a portion of a global marketing framework via web server, gathering via the web server marketing data, associating consumers together according to at least one common interest to create one or more consumer groups.
16
References
A vector space model for automatic indexing
Gerard Salton,A. Wong,C. S. Yang +2 more
TL;DR: An approach based on space density computations is used to choose an optimum indexing vocabulary for a collection of documents, demonstating the usefulness of the model.
Voronoi diagrams—a survey of a fundamental geometric data structure
TL;DR: The Voronoi diagram as discussed by the authors divides the plane according to the nearest-neighbor points in the plane, and then divides the vertices of the plane into vertices, where vertices correspond to vertices in a plane.
4.7K
A survey of methods and strategies in character segmentation
R.G. Casey,Eric Lecolinet +1 more
TL;DR: H holistic approaches that avoid segmentation by recognizing entire character strings as units are described, including methods that partition the input image into subimages, which are then classified.
Patent
Information extraction system and method using concept-relation-concept (CRC) triples
Woojin Paik,Elizabeth D. Liddy,Jennifer Liddy,Ian Niles,Eileen E. Allen +4 more
- 06 Feb 1997
TL;DR: An information extraction system that allows users to ask questions about documents in a database, and responds to queries by returning possibly relevant information which is extracted from the documents is presented in this article.
840
Patent
Method and article of manufacture for content-based analysis, storage, retrieval, and segmentation of audio information
Thomas L. Blum,Douglas F. Keislar,James A. Wheaton,Erling H. Wold +3 more
- 21 Jul 1997
TL;DR: In this paper, a system that performs analysis and comparison of audio data files based upon the content of the data files is presented, which produces a set of numeric values (a feature vector) that can be used to classify and rank the similarity between individual audio files typically stored in a multimedia database or on the Web.
726