线性回归（最小二乘法、批量梯度下降法、随机梯度下降法、局部加权线性回归） C++

We turn next to the task of finding a weight vector w which minimizes the chosen function E(w).

Because there is clearly no hope of finding an anlytical solution to the equation ∂E(w)=0, we resort to

iterative numerical procedures.

On-line gradient descent, also known as sequential gradient descent or stochastic gradient descent, makes

an update to the weight vector based on one data point at a time.

One advantage of on-line methods compared to batch methods is that the former handle redundancy in the data

much more efficiently. Another property of on-line gradient descent is the possibility of escaping from local minima,

since a stationary point with respect to the error function for the whole data set will generally not be a stationary point

for each data point individually.

Another advantage of on-line learning is the fact that it requires much less storage than batch learning.

原始数据获得

#include <iostream>
#include <fstream>
#include <vector>
#include <string>
#include <cfloat>
#include <cmath>

double dis(double &train, double &query) {
　　double weight=exp(-0.5*pow(train-query, 2));

　　return weight;
}

/*最小二乘法*/
template <typename PairIterator>
bool GetLinearFit(PairIterator begin_it, PairIterator end_it, double& slope, double& y_intercept) {
    if(begin_it==end_it) {
        return false;
    }
    size_t n=end_it-begin_it;
    double sum_x2=0.0,sum_y=0.0,sum_x=0.0,sum_xy=0.0;

for(PairIterator it=begin_it;it!=end_it;++it) {
        sum_x2+=(it->first)*(it->first);
        sum_y+=it->second;
        sum_x+=it->first;
        sum_xy+=(it->first)*(it->second);
    }

slope=(n*sum_xy-sum_x*sum_y)/(n*sum_x2-sum_x*sum_x);
y_intercept=(sum_x2*sum_y-sum_x*sum_xy)/(n*sum_x2-sum_x*sum_x);

return true;
}

/*locally weighted linear regression(LWR)*/
template<typename PairIterator>
bool LWR(PairIterator begin_it, PairIterator end_it, double& slope, double& y_intercept) {
　　if(begin_it==end_it) {
　　　　return false;
　　}

　　 /*x are the data points for each local regression model. They are usually (but not always) the data points in your sample.*/
　　double query=5.5;
　　size_t n=end_it-begin_it;
　　double J=0.0;

　　for(PairIterator it=begin_it;it!=end_it;++it) {
　　　　J+=(y_intercept+slope*(it->first)-it->second)*(y_intercept+slope*(it->first)-it->second)*dis(it->first, query);
　　}
　　J=J*0.5/n;

　　while(true) {
　　　　double temp0=0,temp1=0;
　　　　for(PairIterator it=begin_it;it!=end_it;++it) {
　　　　　　temp0+=(y_intercept+slope*(it->first)-it->second)*dis(it->first, query);
　　　　　　temp1+=(y_intercept+slope*(it->first)-it->second)*(it->first)*dis(it->first, query);
　　　　}

　　　　temp0=temp0/n;
　　　　temp1=temp1/n;

　　　　/*0.03为学习率阿尔法*/
　　　　y_intercept=y_intercept-0.03*temp0;
　　　　slope=slope-0.03*temp1;

　　　　double MSE=0.0;
　　　　for(PairIterator it=begin_it;it!=end_it;++it) {
　　　　　　MSE+=(y_intercept+slope*(it->first)-it->second)*(y_intercept+slope*(it->first)-it->second)*dis(it->first, query);
　　　　}
　　　　MSE=0.5*MSE/n;
　　　　if(std::abs(J-MSE)<0.00000001)
　　　　　　break;
　　　　J=MSE;
　　　　}
　　return true;
}

/*批量梯度下降法，Batch Gradient Desscent,BGD*/
template<typename PairIterator>
bool BatchGradientDescent(PairIterator begin_it, PairIterator end_it, double& slope, double& y_intercept) {
    if(begin_it==end_it) {
        return false;
    }
    size_t n=end_it-begin_it;
    double J=0.0;

/*the initial cost function*/
    for(PairIterator it=begin_it;it!=end_it;++it) {
        J+=(y_intercept+slope*(it->first)-it->second)*(y_intercept+slope*(it->first)-it->second);
    }
    J=J*0.5/n;

while(true) {
        double temp0=0,temp1=0;
        for(PairIterator it=begin_it;it!=end_it;++it) {
            temp0+=(y_intercept+slope*(it->first)-it->second);
            temp1+=(y_intercept+slope*(it->first)-it->second)*(it->first);
        }
        temp0=temp0/n;
        temp1=temp1/n;

/*0.03为学习率阿尔法*/
y_intercept=y_intercept-0.03*temp0;
slope=slope-0.03*temp1;

double MSE=0.0;
        for(PairIterator it=begin_it;it!=end_it;++it) {
            MSE+=(y_intercept+slope*(it->first)-it->second)*(y_intercept+slope*(it->first)-it->second);
        }
        MSE=0.5*MSE/n;
        if(std::abs(J-MSE)<0.00000001)
            break;
        J=MSE;
    }
    return true;
}

/*随机梯度下降法，Stochastic Gradient Desscent,SGD*/
template<typename PairIterator>
bool StochasticGradientDescent(PairIterator begin_it, PairIterator end_it, double& slope, double& y_intercept) {
    if(begin_it==end_it) {
        return false;
    }
    size_t n=end_it-begin_it;
    double J=0.0;

while(true) {
        double temp0=0,temp1=0;
        for(PairIterator it=begin_it;it!=end_it;++it) {
            temp0=(y_intercept+slope*(it->first)-it->second);
            temp1=(y_intercept+slope*(it->first)-it->second)*(it->first);

/*0.03为学习率阿尔法*/
y_intercept=y_intercept-0.03*temp0;
slope=slope-0.03*temp1;

double MSE=0.0;
            for(PairIterator it=begin_it;it!=end_it;++it) {
                MSE+=(y_intercept+slope*(it->first)-it->second)*(y_intercept+slope*(it->first)-it->second);
            }
            MSE=0.5*MSE/n;
            if(std::abs(J-MSE)<0.00000001)
                break;
            J=MSE;
        }
        break;
    }

return true;
}

int main() {
    std::ifstream in;
    in.open("ex2x.dat");
    if(!in) {
        std::cout<<"open file ex2x.dat failed!"<<std::endl;
        return 1;
    }

std::vector<double> datax,datay;
double temp;

while(in>>temp) {
datax.push_back(temp);
}

in.close();
    in.open("ex2y.dat");
    if(!in) {
        std::cout<<"open file ex2y.dat failed!"<<std::endl;
        return 1;
    }

while(in>>temp) {
datay.push_back(temp);
}

std::vector<std::pair<double, double> > data;

for(std::vector<double>::const_iterator iterx=datax.begin(),itery=datay.begin();iterx!=datax.end(),itery!=datay.end();iterx++,itery++) {
        data.push_back(std::pair<double,double>(*iterx,*itery));
    }
    in.close();
    double slope=0.0;
    double y_intercept=0.0;
    GetLinearFit(data.begin(),data.end(),slope,y_intercept);
    std::cout<<"最小二乘法得到的结果："<<std::endl;
    std::cout<<"slope: "<<slope<<std::endl;
    std::cout<<"y_intercept: "<<y_intercept<<std::endl;

slope=1.0,y_intercept=1.0;
    BatchGradientDescent(data.begin(),data.end(),slope,y_intercept);
    std::cout<<"批量梯度下降法得到的结果："<<std::endl;
    std::cout<<"slope: "<<slope<<std::endl;
    std::cout<<"y_intercept: "<<y_intercept<<std::endl;

slope=1.0,y_intercept=1.0;
    StochasticGradientDescent(data.begin(),data.end(),slope,y_intercept);
    std::cout<<"随机梯度下降法得到的结果："<<std::endl;
    std::cout<<"slope: "<<slope<<std::endl;
    std::cout<<"y_intercept: "<<y_intercept<<std::endl;

slope=1.0,y_intercept=1.0;
　LWR(data.begin(),data.end(),slope,y_intercept);
　std::cout<<"locally weighted linear regression 得到的结果："<<std::endl;
　std::cout<<"slope: "<<slope<<std::endl;
　std::cout<<"y_intercept: "<<y_intercept<<std::endl;

return 0;
}

线性回归（最小二乘法、批量梯度下降法、随机梯度下降法、局部加权线性回归） C++的更多相关文章

Locally Weighted Linear Regression 局部加权线性回归-R实现
局部加权线性回归 [转载时请注明来源]:http://www.cnblogs.com/runner-ljt/ Ljt 作为一个初学者,水平有限,欢迎交流指正. 线性回归容易出现过拟合或欠拟合的问 ...
Locally weighted linear regression(局部加权线性回归)
(整理自AndrewNG的课件,转载请注明.整理者:华科小涛@http://www.cnblogs.com/hust-ghtao/) 前面几篇博客主要介绍了线性回归的学习算法,那么它有什么不足的地方么 ...
局部加权线性回归(Locally weighted linear regression)
首先我们来看一个线性回归的问题,在下面的例子中,我们选取不同维度的特征来对我们的数据进行拟合. 对于上面三个图像做如下解释: 选取一个特征,来拟合数据,可以看出来拟合情况并不是很好,有些数据误差还是比 ...
梯度下降&随机梯度下降&批梯度下降
梯度下降法下面的h(x)是要拟合的函数,J(θ)损失函数,theta是参数,要迭代求解的值,theta求解出来了那最终要拟合的函数h(θ)就出来了.其中m是训练集的记录条数,j是参数的个数. 梯 ...
matlab练习程序（局部加权线性回归）
通常我们使用的最小二乘都需要预先设定一个模型,然后通过最小二乘方法解出模型的系数. 而大多数情况是我们是不知道这个模型的,比如这篇博客中z=ax^2+by^2+cxy+dx+ey+f 这样的模型. 局 ...
sklearn中实现随机梯度下降法（多元线性回归）
sklearn中实现随机梯度下降法随机梯度下降法是一种根据模拟退火的原理对损失函数进行最小化的一种计算方式,在sklearn中主要用于多元线性回归算法中,是一种比较高效的最优化方法,其中的梯度下降系 ...
NN优化方法对照：梯度下降、随机梯度下降和批量梯度下降
1.前言这几种方法呢都是在求最优解中常常出现的方法,主要是应用迭代的思想来逼近.在梯度下降算法中.都是环绕下面这个式子展开: 当中在上面的式子中hθ(x)代表.输入为x的时候的其当时θ參数下的输出值 ...
L20 梯度下降、随机梯度下降和小批量梯度下降
airfoil4755 下载链接:https://pan.baidu.com/s/1YEtNjJ0_G9eeH6A6vHXhnA 提取码:dwjq 梯度下降 (Boyd & Vandenbe ...
监督学习：随机梯度下降算法（sgd）和批梯度下降算法（bgd）
线性回归首先要明白什么是回归.回归的目的是通过几个已知数据来预测另一个数值型数据的目标值. 假设特征和结果满足线性关系,即满足一个计算公式h(x),这个公式的自变量就是已知的数据x,函数值h(x)就 ...

随机推荐

六时出行 App 隐私政策
六时出行 App 隐私政策本应用尊重并保护所有使用服务用户的个人隐私权.为了给您提供更准确.更有个性化的服务,本应用会按照本隐私权政策的规定使用和披露您的个人信息.但本应用将以高度的勤勉.审慎义 ...
用最简单的脚本完成supertab的基本功能并实现一个更加合理的功能
supertab是vim的一个出名的插件, 相信会vim的人没几个不知道的, 我在之前的<<vim之补全1>>中首先说明的也是它, supertab实现的功能简单的说就是用ta ...
MFC cstring 型转化成 double型
cstring szNum; GetDlgItemText(IDC_EDIT1, szNum); double Num; Num = _ttol(szNum); 转化成长整型 Num = _tstof ...
Spring装配之——JAVA代码装配Bean
首先创建几个普通的JAVA对象,用于测试JAVA代码装配bean的功能. package soundsystemJava; //作为接口定义了CD播放器对一盘CD所能进行的操作 public int ...
Xftp 5 和 Xshell 5 基本使用方法
软件介绍: (1)Xshell: 一个强大的安全终端模拟软件,它支持SSH1, SSH2, 以及Microsoft Windows 平台的 TELNET 协议.Xshell通过互联网可以连接到远程的服 ...
网络编程_socketserver
一.socketserver 网络编程 1.socketserver支持多用户并发处理:2.socketserver是对socket的再封装;处理步骤:1.创建一个socketserver类2.继承B ...
elasticsearch学习（1）简单查询与聚合
elastic 被用作全文搜索.结构化搜索.分析以及这三个功能的组合一个ElasticSearch集群可以包含多个索引, 每个索引包含多个类型一个类型存储着多个文档每个文档又有多个属性索引(名 ...
Spider-Python实战之通过Python爬虫爬取图片制作Win7跑车主题
1. 前期准备 1.1 开发工具 Python 3.6 Pycharm Pro 2017.3.2 Text文本 1.2 Python库 requests re urllib 如果没有这些Python库 ...
BZOJ 1232 USACO 2008 Nov. 安慰奶牛Cheer
[题解] 对于每一条边,我们通过它需要花费的代价是边权的两倍加上这条边两个端点的点权. 我们把每条边的边权设为上述的值,然后跑一边最小生成树,再把答案加上最小的点权就好了. #include<c ...
单词接龙（codevs 1018）
2000年NOIP全国联赛普及组NOIP全国联赛提高组时间限制: 1 s 空间限制: 128000 KB 题目等级 : 黄金 Gold 题解题目描述 Description 单词接龙是一个与我们经 ...

线性回归（最小二乘法、批量梯度下降法、随机梯度下降法、局部加权线性回归） C++

线性回归（最小二乘法、批量梯度下降法、随机梯度下降法、局部加权线性回归） C++的更多相关文章

随机推荐

热门专题