Apple的LZF算法解析

有关LZF算法的相关解析文档比较少，但是Apple对LZF的开源，可以让我们对该算法进行一个简单的解析。LZFSE 基于 Lempel-Ziv ，并使用了有限状态熵编码。LZF采用类似lz77和lzss的混合编码。使用3种“起始标记”来代表每段输出的数据串。

接下来看一下开源的LZF算法的实现源码。

1.定义的全局字段：

       private readonly long[] _hashTable = new long[Hsize];

        private const uint Hlog = ;

        private const uint Hsize = ( << );

        private const uint MaxLit = ( << );

        private const uint MaxOff = ( << );

        private const uint MaxRef = (( << ) + ( << ));

2.使用LibLZF算法压缩数据：

        /// <summary>

        /// 使用LibLZF算法压缩数据

        /// </summary>

        /// <param name="input">需要压缩的数据</param>

        /// <param name="inputLength">要压缩的数据的长度</param>

        /// <param name="output">引用将包含压缩数据的缓冲区</param>

        /// <param name="outputLength">压缩缓冲区的长度（应大于输入缓冲区）</param>

        /// <returns>输出缓冲区中压缩归档的大小</returns>

        public int Compress(byte[] input, int inputLength, byte[] output, int outputLength)

        {

            Array.Clear(_hashTable, , (int)Hsize);

            uint iidx = ;

            uint oidx = ;

            var hval = (uint)(((input[iidx]) << ) | input[iidx + ]);

            var lit = ;

            for (; ; )

            {

                if (iidx < inputLength - )

                {

                    hval = (hval << ) | input[iidx + ];

                    long hslot = ((hval ^ (hval << )) >> (int)((( *  - Hlog)) - hval * ) & (Hsize - ));

                    var reference = _hashTable[hslot];

                    _hashTable[hslot] = iidx;

                    long off;

                    if ((off = iidx - reference - ) < MaxOff

                        && iidx +  < inputLength

                        && reference >

                        && input[reference + ] == input[iidx + ]

                        && input[reference + ] == input[iidx + ]

                        && input[reference + ] == input[iidx + ]

                        )

                    {

                        uint len = ;

                        var maxlen = (uint)inputLength - iidx - len;

                        maxlen = maxlen > MaxRef ? MaxRef : maxlen;

                        if (oidx + lit +  +  >= outputLength)

                            return ;

                        do

                            len++;

                        while (len < maxlen && input[reference + len] == input[iidx + len]);

                        if (lit != )

                        {

                            output[oidx++] = (byte)(lit - );

                            lit = -lit;

                            do

                                output[oidx++] = input[iidx + lit];

                            while ((++lit) != );

                        }

                        len -= ;

                        iidx++;

                        if (len < )

                        {

                            output[oidx++] = (byte)((off >> ) + (len << ));

                        }

                        else

                        {

                            output[oidx++] = (byte)((off >> ) + ( << ));

                            output[oidx++] = (byte)(len - );

                        }

                        output[oidx++] = (byte)off;

                        iidx += len - ;

                        hval = (uint)(((input[iidx]) << ) | input[iidx + ]);

                        hval = (hval << ) | input[iidx + ];

                        _hashTable[((hval ^ (hval << )) >> (int)((( *  - Hlog)) - hval * ) & (Hsize - ))] = iidx;

                        iidx++;

                        hval = (hval << ) | input[iidx + ];

                        _hashTable[((hval ^ (hval << )) >> (int)((( *  - Hlog)) - hval * ) & (Hsize - ))] = iidx;

                        iidx++;

                        continue;

                    }

                }

                else if (iidx == inputLength)

                    break;

                lit++;

                iidx++;

                if (lit != MaxLit) continue;

                if (oidx +  + MaxLit >= outputLength)

                    return ;

                output[oidx++] = (byte)(MaxLit - );

                lit = -lit;

                do

                    output[oidx++] = input[iidx + lit];

                while ((++lit) != );

            }

            if (lit == ) return (int)oidx;

            if (oidx + lit +  >= outputLength)

                return ;

            output[oidx++] = (byte)(lit - );

            lit = -lit;

            do

                output[oidx++] = input[iidx + lit];

            while ((++lit) != );

            return (int)oidx;

        }

        /// <summary>

        /// 使用LibLZF算法解压缩数据

        /// </summary>

        /// <param name="input">参考数据进行解压缩</param>

        /// <param name="inputLength">要解压缩的数据的长度</param>

        /// <param name="output">引用包含解压缩数据的缓冲区</param>

        /// <param name="outputLength">输出缓冲区中压缩归档的大小</param>

        /// <returns>返回解压缩大小</returns>

        public int Decompress(byte[] input, int inputLength, byte[] output, int outputLength)

        {

            uint iidx = ;

            uint oidx = ;

            do

            {

                uint ctrl = input[iidx++];

                if (ctrl < ( << ))

                {

                    ctrl++;

                    if (oidx + ctrl > outputLength)

                    {

                        return ;

                    }

                    do

                        output[oidx++] = input[iidx++];

                    while ((--ctrl) != );

                }

                else

                {

                    var len = ctrl >> ;

                    var reference = (int)(oidx - ((ctrl & 0x1f) << ) - );

                    if (len == )

                        len += input[iidx++];

                    reference -= input[iidx++];

                    if (oidx + len +  > outputLength)

                    {

                        return ;

                    }

                    if (reference < )

                    {

                        return ;

                    }

                    output[oidx++] = output[reference++];

                    output[oidx++] = output[reference++];

                    do

                        output[oidx++] = output[reference++];

                    while ((--len) != );

                }

            }

            while (iidx < inputLength);

            return (int)oidx;

        }

以上是LZF算法的代码。

Apple的LZF算法解析的更多相关文章

地理围栏算法解析（Geo-fencing）
地理围栏算法解析 http://www.cnblogs.com/LBSer/p/4471742.html 地理围栏(Geo-fencing)是LBS的一种应用,就是用一个虚拟的栅栏围出一个虚拟地理边界 ...
KMP串匹配算法解析与优化
朴素串匹配算法说明串匹配算法最常用的情形是从一篇文档中查找指定文本.需要查找的文本叫做模式串,需要从中查找模式串的串暂且叫做查找串吧. 为了更好理解KMP算法,我们先这样看待一下朴素匹配算法吧.朴素 ...
Peterson算法与Dekker算法解析
进来Bear正在学习巩固并行的基础知识,所以写下这篇基础的有关并行算法的文章. 在讲述两个算法之前,需要明确一些概念性的问题, Race Condition(竞争条件),Situations lik ...
python常见排序算法解析
python——常见排序算法解析算法是程序员的灵魂. 下面的博文是我整理的感觉还不错的算法实现原理的理解是最重要的,我会常回来看看,并坚持每天刷leetcode 本篇主要实现九(八)大排序算法 ...
Java虚拟机对象存活标记及垃圾收集算法解析
一.对象存活标记 1. 引用计数算法给对象中添加一个引用计数器,每当有一个地方引用它时,计数器就加1:当引用失效时,计数器就减1:任何时刻计数器都为0的对象就是不可能再被使用的. 引用计数算法(Re ...
JVM垃圾回收算法解析
JVM垃圾回收算法解析标记-清除算法该算法为最基础的算法.它分为标记和清除两个阶段,首先标记出需要回收的对象,在标记结束后,统一回收.该算法存在两个问题:一是效率问题,标记和清除过程效率都不太高, ...
DeepFM算法解析及Python实现
1. DeepFM算法的提出由于DeepFM算法有效的结合了因子分解机与神经网络在特征学习中的优点:同时提取到低阶组合特征与高阶组合特征,所以越来越被广泛使用. 在DeepFM中,FM算法负责对一阶 ...
GBDT+LR算法解析及Python实现
1. GBDT + LR 是什么本质上GBDT+LR是一种具有stacking思想的二分类器模型,所以可以用来解决二分类问题.这个方法出自于Facebook 2014年的论文 Practical L ...
最长上升子序列(LIS)n2 nlogn算法解析
题目描述给定一个数列,包含N个整数,求这个序列的最长上升子序列. 例如 2 5 3 4 1 7 6 最长上升子序列为 4. 1.O(n2)算法解析看到这个题,大家的直觉肯定都是要用动态规划来做,那 ...

随机推荐

Mysql 学习笔记2
(1)MySQL查看表占用空间大小 //先进去MySQL自带管理库:information_schema //自己的数据库:dbwww58com_kuchecarlib //自己的表:t_carmod ...
backbone入门示例
最近因为有个项目需要用backbone+mui 所以最近入坑backbone. Backbonejs有几个重要的概念,先介绍一下:Model,Collection,View,Router.其中Mod ...
串口计时工具Grabserial简介及修改(添加输入功能)
Grabserial是Tim Bird用python写的一个抓取串口的工具,这个工具能够为收到的每一行信息添加上时间戳. 如果想对启动时间进行优化的话,使用这个工具就可以简单地从串口输出分析出耗时. ...
iOS单例详解
单例:整个程序只创建一次,全局共用. 单例的创建 // SharedPerson.h 文件中 + (instancetype)share; // SharedPerson.m 文件中 static S ...
XMPP iOS客户端实现一：服务器
1.下载ejabberd,下载链接http://www.process-one.net/en/ejabberd/downloads/ 2.安装,使用默认配置即可,next.. 3.启动ejabberd ...
.NET Core 构建配置文件从 project.json 到 .csproj
从 .NET Core SDK 1.0 Preview 3 build 004056 开始,.NET Core 弃用 project.json,回归 .csproj,主要原因是为了兼容 MSBuild ...
让浏览器不再显示 https 页面中的 http 请求警报
HTTPS 是 HTTP over Secure Socket Layer,以安全为目标的 HTTP 通道,所以在 HTTPS 承载的页面上不允许出现 http 请求,一旦出现就是提示或报错: Mix ...
[.net 面向对象程序设计进阶] (15) 缓存(Cache)(二) 利用缓存提升程序性能
[.net 面向对象程序设计进阶] (15) 缓存(Cache)(二) 利用缓存提升程序性能本节导读: 上节说了缓存是以空间来换取时间的技术,介绍了客户端缓存和两种常用服务器缓布,本节主要介绍一种. ...
跟vczh看实例学编译原理——二：实现Tinymoe的词法分析
文章中引用的代码均来自https://github.com/vczh/tinymoe. 实现Tinymoe的第一步自然是一个词法分析器.词法分析其所作的事情很简单,就是把一份代码分割成若干个tok ...
C#设计模式之外观
IronMan之外观接着上篇观察者内容的“剧情”,没看过的朋友也没关系,篇幅之间有衔接的关系但是影响不大. 需求: 为"兵工厂"提供各种支持,生产了各式各样的"Iron ...

Apple的LZF算法解析

Apple的LZF算法解析的更多相关文章

随机推荐

热门专题