Learning of Kernel Functions in Support Vector Machines
Chih-Cheng Yang, Wan-Jui Lee, Shie-Jue Lee
Abstract
Chih-Cheng Yang, Wan-Jui Lee, Shie-Jue Lee
Abstract
The selection and learning of kernel functions is a very important but rarely studied problem in the field of support vector learning. However, the kernel function of a support vector machine has great influence on its performance. The kernel function projects the dataset from the original data space into the feature space, and therefore the problems which can not be done in low dimensions could be done in a higher dimension through the transform of the kernel function. In this paper, we introduce the gradient descent method into the learning of kernel functions. Using the gradient descent method, we can conduct learning rules of the parameters which indicate the shape and distribution of the kernel functions. Therefore, we can obtain better kernel functions by training of their parameters with respect to the risk minimization principle. The experimental results have shown that our approach can derive better kernel functions and thus has better generalization ability than other methods.
OpenAlex reports 12 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
The selection and learning of kernel functions is a very important but rarely studied problem in the field of support vector learning. However, the kernel function of a support vector machine has great influence on its performance. The kernel function projects the dataset from the original data space into the feature space, and therefore the problems which can not be done in low dimensions could be done in a higher dimension through the transform of the kernel function. In this paper, we introduce the gradient descent method into the learning of kernel functions. Using the gradient descent method, we can conduct learning rules of the parameters which indicate the shape and distribution of the kernel functions. Therefore, we can obtain better kernel functions by training of their parameters with respect to the risk minimization principle. The experimental results have shown that our approach can derive better kernel functions and thus has better generalization ability than other methods.
Key concepts: Radial basis function kernel, Kernel embedding of distributions, Kernel (algebra), Polynomial kernel, Variable kernel density estimation, Kernel method, Tree kernel, String kernel