2008 IEEE International Conference on Data Mining Workshops

Items from 1 to 20 out of 139 results

chapter

Using Contextual Information in Transactional Segmentation: An Empirical Study in E-Commerce

M.F. Faraone, M. Gorgoglione, C. Palmisano

2008 IEEE International Conference on Data Mining Workshops > 796 - 805

2008 IEEE International Conference on Data Mining Workshops

The growing complexity and variability characterizing markets have induced scholars and marketers to propose new segmentation approaches. Recent research has shown that including the context in which a transaction occurs in customer behavior models, improves the ability of predicting their behavior. However, no systematic research has studied whether contextual information really matters in market...

chapter

Using Contextual Information to Decrease the Cost of Incorrect Predictions in On-line Customer Behavior Modeling

M. Gorgoglione, C. Palmisano, S. Lombardi

2008 IEEE International Conference on Data Mining Workshops > 780 - 788

2008 IEEE International Conference on Data Mining Workshops

The performance of user profiling models depends on both the predictive accuracy and the cost of incorrect predictions. In this paper we study whether including contextual information leads to a decrease in the misclassification cost. Several experimental analyses were done by varying the cost ratio, the market granularity and the granularity of context. The experimental results show that context...

chapter

Semantic Analysis Method for Unstructured Data in Telecom Services

M. Iwashita, K. Nishimatsu, S. Shimogawa

2008 IEEE International Conference on Data Mining Workshops > 789 - 795

2008 IEEE International Conference on Data Mining Workshops

A variety of services have recently been provided depending on highly developed networks and personal equipment. With these advances, connecting this equipment has become increasingly more complicated. Problems such as an increase in no-connection and determining the cause have become difficult in some cases because software is often updated to keep up with advancements in services or security. Telecom...

chapter

An Adaptive Pre-filtering Technique for Error-Reduction Sampling in Active Learning

M. Davy, S. Luz

2008 IEEE International Conference on Data Mining Workshops > 682 - 691

2008 IEEE International Conference on Data Mining Workshops

Error-reduction sampling (ERS) is a high performing (but computationally expensive) query selection strategy for active learning. Subset optimisation has been proposed to reduce computational expense by applying ERS to only a subset of examples from the pool. This paper compares techniques used to construct the subset, namely random sub-sampling and pre-filtering. We focus on pre-filtering which populates...

chapter

Keyword Extraction Based on Lexical Chains and Word Co-occurrence for Chinese News Web Pages

Xinghua Li, Xindong Wu, Xuegang Hu, Fei Xie, more

2008 IEEE International Conference on Data Mining Workshops > 744 - 751

2008 IEEE International Conference on Data Mining Workshops

This paper presents a new keyword extraction algorithm for Chinese news Web pages using lexical chains and word co-occurrence combined with frequency features, cohesion features, and corelation features. A lexical chain is an external performance consistency by semantically related words of a text, and is the representation of the semantic content of a portion of the text. Word co-occurrence distribution...

chapter

Semantic Features for Multi-view Semi-supervised and Active Learning of Text Classification

Shiliang Sun

2008 IEEE International Conference on Data Mining Workshops > 731 - 735

2008 IEEE International Conference on Data Mining Workshops

For multi-view learning, existing methods usually exploit originally provided features for classifier training, which ignore the latent correlation between different views. In this paper, semantic features integrating information from multiple views are extracted for pattern representation. Canonical correlation analysis is used to learn the representation of semantic spaces where semantic features...

chapter

Hunting for Coherent Co-clusters in High Dimensional and Noisy Datasets

M. Deodhar, J. Ghosh, G. Gupta, Hyuk Cho, more

2008 IEEE International Conference on Data Mining Workshops > 654 - 663

2008 IEEE International Conference on Data Mining Workshops

Clustering problems often involve datasets where only a part of the data is relevant to the problem, e.g., in microarray data analysis only a subset of the genes show cohesive expressions within a subset of the conditions/features. The existence of a large number of non-informative data points and features makes it challenging to hunt for coherent and meaningful clusters from such datasets. Additionally,...

chapter

Mining Allocating Patterns in One-Sum Weighted Items

Y.J. Wang, Xinwei Zheng, F. Coenen, C.Y. Li

2008 IEEE International Conference on Data Mining Workshops > 592 - 598

2008 IEEE International Conference on Data Mining Workshops

An association rule (AR) is a common knowledge model in data mining that describes an implicative co-occurring relationship between two disjoint sets of binary-valued transaction database attributes (items), expressed in the form of an "antecedent rArr consequent" rule. A variant of the AR is the weighted association rule (WAR). With regard to a marketing context, this paper introduces a...

chapter

Efficient Distance Computation Using SQL Queries and UDFs

S.K. Pitchaimalai, C. Ordonez, C. Garcia-Alvarado

2008 IEEE International Conference on Data Mining Workshops > 533 - 542

2008 IEEE International Conference on Data Mining Workshops

Distance computation is one of the most computationally intensive operations employed by many data mining algorithms. Performing such matrix computations within a DBMS creates many optimization challenges. We propose techniques to efficiently compute Euclidean distance using SQL queries and user-defined functions (UDFs). We concentrate on efficient Euclidean distance computation for the well-known...

chapter

Co-training by Committee: A New Semi-supervised Learning Framework

M. Hady, F. Schwenker

2008 IEEE International Conference on Data Mining Workshops > 563 - 572

2008 IEEE International Conference on Data Mining Workshops

For many data mining applications, it is necessary to develop algorithms that use unlabeled data to improve the accuracy of the supervised learning. Co-Training is a popular semi-supervised learning algorithm. It assumes that each example is represented by two or more redundantly sufficient sets of features (views) and these views are independent given the class. However, these assumptions are not...

chapter

Stream-Close: Fast Mining of Closed Frequent Itemsets in High Speed Data Streams

B.N. Ranganath, M.N. Murty

2008 IEEE International Conference on Data Mining Workshops > 516 - 525

2008 IEEE International Conference on Data Mining Workshops

With the emergence of large-volume and high-speed streaming data, the recent techniques for stream mining of CFIpsilas (closed frequent itemsets) will become inefficient. When concept drift occurs at a slow rate in high speed data streams, the rate of change of information across different sliding windows will be negligible. So, the user wonpsilat be devoid of change in information if we slide window...

chapter

Service Oriented KDD: A Framework for Grid Data Mining Workflows

M. Lackovic, D. Talia, P. Trunfio

2008 IEEE International Conference on Data Mining Workshops > 496 - 505

2008 IEEE International Conference on Data Mining Workshops

Weka4WS is an extension of the Weka toolkit to support remote execution of data mining tasks as grid services. A first version of Weka4WS supporting concurrent execution of multiple data mining tasks on remote grid nodes has been presented in a previous work. In this paper we present a new version supporting also the composition and execution of data mining workflows on a grid. This new version of...

chapter

Preface

2008 IEEE International Conference on Data Mining Workshops > xiii

2008 IEEE International Conference on Data Mining Workshops

chapter

HPDM 2008 Message and Committees

2008 IEEE International Conference on Data Mining Workshops > xxvii - xxix

2008 IEEE International Conference on Data Mining Workshops

chapter

SSTDM 2008 Message and Committees

2008 IEEE International Conference on Data Mining Workshops > xxiii - xxvi

2008 IEEE International Conference on Data Mining Workshops

chapter

Behavior Informatics and Analytics: Let Behavior Talk

Longbing Cao

2008 IEEE International Conference on Data Mining Workshops > 87 - 96

2008 IEEE International Conference on Data Mining Workshops

Behavior is increasingly recognized as a key component in business intelligence and problem-solving. Different from traditional behavior analysis, which mainly focus on implicit behavior and explicit business appearance as a result of business usage and customer demographics, this paper proposes the field of Behavior Informatics and Analytics (BIA), to support explicit behavior involvement through...

chapter

Scoring Models for Insurance Risk Sharing Pool Opimization

N. Chapados, C. Dugas, P. Vincent, R. Ducharme

2008 IEEE International Conference on Data Mining Workshops > 97 - 105

2008 IEEE International Conference on Data Mining Workshops

We introduce a flexible scoring model that can be used by property and casualty insurers that have access to a risk-sharing pool to better select the insureds to transfer to the pool. The model discriminates between insureds whose transfer is likely to be profitable under the pool regulations against those paying a fair premium. This model makes use of feature selection methods to automatically discover...

chapter

Food Sales Prediction: "If Only It Knew What We Know"

P. Meulstee, M. Pechenizkiy

2008 IEEE International Conference on Data Mining Workshops > 134 - 143

2008 IEEE International Conference on Data Mining Workshops

Sales prediction is an important problem for different companies involved in manufacturing, logistics, marketing, wholesaling and retailing. Food companies are more concerned with sales prediction of products having a short shelf-life and seasonal changes in demand. The demand may depend on many hidden contexts, not given explicitly in the form of predictive features. Even if some changes are known...

chapter

Mining Temporal Patterns with Quantitative Intervals

T. Guyet, R. Quiniou

2008 IEEE International Conference on Data Mining Workshops > 218 - 227

2008 IEEE International Conference on Data Mining Workshops

In this paper we consider the problem of discovering frequent temporal patterns in a database of temporal sequences, where a temporal sequence is a set of items with associated dates and durations. Since the quantitative temporal information appears to be fundamental in many contexts, it is taken into account in the mining processes and returned as part of the extracted knowledge. To this end, we...

chapter

Actionable Knowledge Discovery for Threats Intelligence Support Using a Multi-dimensional Data Mining Methodology

O. Thonnard, M. Dacier

2008 IEEE International Conference on Data Mining Workshops > 154 - 163

2008 IEEE International Conference on Data Mining Workshops

This paper describes a multi-dimensional knowledge discovery and data mining (KDD) methodology that aims at discovering actionable knowledge related to Internet threats, taking into account domain expert guidance and the integration of domain-specific intelligence during the data mining process. The objectives are twofold: i) to develop global indicators for assessing the prevalence of certain malicious...

Content availability:
Available

Publication date

Set your own date range

Keywords

DATA MINING (86)
CLASSIFICATION ALGORITHMS (29)
DATABASES (23)
DATA MODELS (19)
LEARNING (ARTIFICIAL INTELLIGENCE) (19)
CLUSTERING ALGORITHMS (18)
TRAINING (18)
DISTANCE MEASUREMENT (17)
FEATURE EXTRACTION (16)
ACCURACY (15)
PATTERN CLUSTERING (15)
ALGORITHM DESIGN AND ANALYSIS (14)
CONFERENCES (14)
PATTERN CLASSIFICATION (14)
ASSOCIATION RULES (13)
QUERY PROCESSING (11)
INTERNET (10)
ITEMSETS (10)
INDEXES (8)
KNOWLEDGE DISCOVERY (8)
PREDICTIVE MODELS (8)
STATISTICAL ANALYSIS (8)
COMPUTATIONAL MODELING (7)
CORRELATION (7)
DATABASE MANAGEMENT SYSTEMS (7)
DECISION TREES (7)
ESTIMATION (7)
GRAPH THEORY (7)
KERNEL (7)
MATHEMATICAL MODEL (7)
ONTOLOGIES (ARTIFICIAL INTELLIGENCE) (7)
TEXT ANALYSIS (7)
TRAINING DATA (7)
OPTIMIZATION (6)
RELIABILITY (6)
VISUAL DATABASES (6)
WEB SERVICES (6)
BIOLOGICAL SYSTEM MODELING (5)
CLASSIFICATION (5)
CLUSTERING (5)
FILTERING (5)
GRAPH MINING (5)
MACHINE LEARNING (5)
MARKETING (5)
MATRIX ALGEBRA (5)
MERGING (5)
ONTOLOGIES (5)
PROTEINS (5)
SPATIAL DATABASES (5)
APPROXIMATION METHODS (4)
BIOLOGY (4)
BUILDINGS (4)
BUSINESS (4)
CITIES AND TOWNS (4)
DATA ANALYSIS (4)
ENGINES (4)
EQUATIONS (4)
HIDDEN MARKOV MODELS (4)
HUMANS (4)
IMAGE CLASSIFICATION (4)
LABELING (4)
LEARNING SYSTEMS (4)
METEOROLOGY (4)
NOISE (4)
PEDIATRICS (4)
PROBABILITY (4)
PROPOSALS (4)
REDUNDANCY (4)
REGRESSION ANALYSIS (4)
REMOTE SENSING (4)
SET THEORY (4)
SOCIAL NETWORK SERVICES (4)
SOFTWARE (4)
SUPPORT VECTOR MACHINES (4)
TIME SERIES ANALYSIS (4)
WEB PAGES (4)
AGRICULTURE (3)
AMINO ACIDS (3)
ANALYTICAL MODELS (3)
ANOMALY DETECTION (3)
ATMOSPHERIC MEASUREMENTS (3)
BENCHMARK TESTING (3)
CLASSIFICATION TREE ANALYSIS (3)
COMPANIES (3)
COMPLEXITY THEORY (3)
COMPUTER SCIENCE (3)
CONSUMER BEHAVIOUR (3)
DATA HANDLING (3)
DATA VISUALISATION (3)
DATA VISUALIZATION (3)
DELAY (3)
DISTRIBUTED DATABASES (3)
EVOLUTION (BIOLOGY) (3)
GEOGRAPHIC INFORMATION SYSTEMS (3)
GRAPHICS (3)
IMAGE COLOR ANALYSIS (3)
IMAGE SEQUENCES (3)
INFORMATION EXTRACTION (3)
IP NETWORKS (3)
KNOWLEDGE ENGINEERING (3)
more

INFONA - science communication portal

2008 IEEE International Conference on Data Mining Workshops

Using Contextual Information in Transactional Segmentation: An Empirical Study in E-Commerce

Using Contextual Information to Decrease the Cost of Incorrect Predictions in On-line Customer Behavior Modeling

Semantic Analysis Method for Unstructured Data in Telecom Services

An Adaptive Pre-filtering Technique for Error-Reduction Sampling in Active Learning

Keyword Extraction Based on Lexical Chains and Word Co-occurrence for Chinese News Web Pages

Semantic Features for Multi-view Semi-supervised and Active Learning of Text Classification

Hunting for Coherent Co-clusters in High Dimensional and Noisy Datasets

Mining Allocating Patterns in One-Sum Weighted Items

Efficient Distance Computation Using SQL Queries and UDFs

Co-training by Committee: A New Semi-supervised Learning Framework

Stream-Close: Fast Mining of Closed Frequent Itemsets in High Speed Data Streams

Service Oriented KDD: A Framework for Grid Data Mining Workflows

Preface

HPDM 2008 Message and Committees

SSTDM 2008 Message and Committees

Behavior Informatics and Analytics: Let Behavior Talk

Scoring Models for Insurance Risk Sharing Pool Opimization

Food Sales Prediction: "If Only It Knew What We Know"

Mining Temporal Patterns with Quantitative Intervals

Actionable Knowledge Discovery for Threats Intelligence Support Using a Multi-dimensional Data Mining Methodology

Filter options

Publication date

Keywords

INFONA - science communication portal

2008 IEEE International Conference on Data Mining Workshops $("#expandableTitles").expandable();

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options

2008 IEEE International Conference on Data Mining Workshops