Search results

Items from 1 to 5 out of 5 results

chapter

K nearest neighbor for text summarization using feature similarity

Taeho Jo

2017 International Conference on Communication, Control, Computing and Electronics Engineering (ICCCCEE) > 1 - 5

2017 International Conference on Communication, Control, Computing and Electronics Engineering (ICCCCEE)

In this research, we propose a particular version of KNN (K Nearest Neighbor) where the similarity between feature vectors is computed considering the similarity among attributes or features as well as one among values. The task of text summarization is viewed into the binary classification task where each paragraph or sentence is classified into the essence or non-essence, and in previous works,...

chapter

Using K Nearest Neighbors for text segmentation with feature similarity

Taeho Jo

2017 International Conference on Communication, Control, Computing and Electronics Engineering (ICCCCEE) > 1 - 5

2017 International Conference on Communication, Control, Computing and Electronics Engineering (ICCCCEE)

In this research, we propose the version of K Nearest Neighbor which considers similarity among attributes for computing the similarity between feature vectors. The text segmentation task is viewed into the binary classification where each pair of sentences or paragraphs is classified into whether we put the boundary or not, and the proposed version resulted in the successful results in previous works...

chapter

Chinese Text Categorization study based on feature weight learning

Yan Zhan, Hao Chen, Su-Fang Zhang, Mei Zheng

2009 International Conference on Machine Learning and Cybernetics > 3 > 1723 - 1726

2009 Eighth International Conference on Machine Learning and Cybernetics (ICMLC)

Text categorization (TC) is an important component in many information organization and information management tasks. Two key issues in TC are feature coding and classifier design. The Euclidean distance is usually chosen as the similarity measure in K-nearest neighbor classification algorithm. All the features of each vector have different functions in describing samples. So we can decide different...

chapter

Dictionary-Based Bilingual Web Page Classification

Jicheng Liu, Chunyan Liang, Jianxun Qi

2008 4th International Conference on Wireless Communications, Networking and Mobile Computing > 1 - 4

2008 4th International Conference on Wireless Communications, Networking and Mobile Computing (WiCOM)

Web page classification poses new research challenges because of the noisy nature of the pages. For the bilingual Chinese-English web pages, it also needs to be considered that how to extract the terms of different languages exactly. A new dictionary-based multilingual text categorization approach is proposed in this paper to try to classify the Chinese-English web pages in specific domain into a...

chapter

An Improved Genetic Algorithm for Text Feature Selection

Wei Zhao, Yafei Wang

2010 International Conference on Intelligent Computing and Cognitive Informatics > 7 - 10

2010 International Conference on Intelligent Computing and Cognitive Informatics (ICICCI 2010)

High-dimensional feature space affects the quality and efficiency of text categorization. This paper investigates an improved genetic algorithm that how to help select relevant features in text classification. We follow the so-called "region growing" method to initialize the population, and uses k-means algorithm to selection operation to control the scope of the search, ensure the validity...

Filter options

Keywords:
ENCODING
FEATURE EXTRACTION

Publication date

Set your own date range

Content availability

Available (4)
None (1)

Keywords

CLASSIFICATION ALGORITHMS (4)
SUPPORT VECTOR MACHINES (3)
TEXT ANALYSIS (3)
FEATURE SIMILARITY (2)
K NEAREST NEIGHBOR (2)
KERNEL (2)
NATURAL LANGUAGE PROCESSING (2)
PATTERN CLASSIFICATION (2)
SEMANTICS (2)
ALGORITHM DESIGN AND ANALYSIS (1)
AUTOMATIC ENCODING DETECTION (1)
BILINGUAL CHINESE-ENGLISH WEB PAGES (1)
CHEMICALS (1)
CHINESE TEXT CATEGORIZATION (1)
CLASSIFICATION (1)
CLASSIFIER DESIGN (1)
CLUSTERING ALGORITHMS (1)
DATA MINING (1)
DATABASES (1)
DICTIONARIES (1)
DICTIONARY-BASED BILINGUAL WEB PAGE CLASSIFICATION (1)
DICTIONARY-BASED MULTILINGUAL TEXT CATEGORIZATION (1)
DOMAIN CONCEPTS (1)
DOMAIN DICTIONARY (1)
EQUATIONS (1)
EUCLIDEAN DISTANCE (1)
FEATURE CODING (1)
FEATURE SELECTION (1)
FEATURE WEIGHT (1)
FEATURE WEIGHT LEARNING ALGORITHM (1)
GENETIC ALGORITHM (1)
GENETIC ALGORITHMS (1)
GENETICS (1)
HEURISTIC ALGORITHMS (1)
HIERARCHICAL TOPIC STRUCTURE (1)
IMPROVED GENETIC ALGORITHM (1)
INFORMATION MANAGEMENT (1)
INFORMATION ORGANIZATION (1)
INTEGRATION METHOD (1)
K-MEANS ALGORITHM (1)
K-NEAREST NEIGHBOR CLASSIFICATION ALGORITHM (1)
K-NN (1)
LEARNING (ARTIFICIAL INTELLIGENCE) (1)
MACHINE LEARNING (1)
MACHINE LEARNING ALGORITHMS (1)
MULTILINGUAL PAGES (1)
PATTERN CLUSTERING (1)
REGION GROWING METHOD (1)
SIMILARITY MEASURE (1)
SUPPORT VECTOR MACHINE (1)
SVM (1)
TESTING (1)
TEXT CLASSIFICATION (1)
TEXT FEATURE SELECTION (1)
TEXT SEGMENTATION (1)
TEXT SUMMARIZATION (1)
TRAINING (1)
WEB PAGES (1)
WEB SITES (1)
more

INFONA - science communication portal

Search results

K nearest neighbor for text summarization using feature similarity

Using K Nearest Neighbors for text segmentation with feature similarity

Chinese Text Categorization study based on feature weight learning

Dictionary-Based Bilingual Web Page Classification

An Improved Genetic Algorithm for Text Feature Selection

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Content availability

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options