Apple的LZF算法解析

有关LZF算法的相关解析文档比较少，但是Apple对LZF的开源，可以让我们对该算法进行一个简单的解析。LZFSE 基于 Lempel-Ziv ，并使用了有限状态熵编码。LZF采用类似lz77和lzss的混合编码。使用3种“起始标记”来代表每段输出的数据串。

接下来看一下开源的LZF算法的实现源码。

1.定义的全局字段：

       private readonly long[] _hashTable = new long[Hsize];

        private const uint Hlog = ;

        private const uint Hsize = ( << );

        private const uint MaxLit = ( << );

        private const uint MaxOff = ( << );

        private const uint MaxRef = (( << ) + ( << ));

2.使用LibLZF算法压缩数据：

        /// <summary>

        /// 使用LibLZF算法压缩数据

        /// </summary>

        /// <param name="input">需要压缩的数据</param>

        /// <param name="inputLength">要压缩的数据的长度</param>

        /// <param name="output">引用将包含压缩数据的缓冲区</param>

        /// <param name="outputLength">压缩缓冲区的长度（应大于输入缓冲区）</param>

        /// <returns>输出缓冲区中压缩归档的大小</returns>

        public int Compress(byte[] input, int inputLength, byte[] output, int outputLength)

        {

            Array.Clear(_hashTable, , (int)Hsize);

            uint iidx = ;

            uint oidx = ;

            var hval = (uint)(((input[iidx]) << ) | input[iidx + ]);

            var lit = ;

            for (; ; )

            {

                if (iidx < inputLength - )

                {

                    hval = (hval << ) | input[iidx + ];

                    long hslot = ((hval ^ (hval << )) >> (int)((( *  - Hlog)) - hval * ) & (Hsize - ));

                    var reference = _hashTable[hslot];

                    _hashTable[hslot] = iidx;

                    long off;

                    if ((off = iidx - reference - ) < MaxOff

                        && iidx +  < inputLength

                        && reference >

                        && input[reference + ] == input[iidx + ]

                        && input[reference + ] == input[iidx + ]

                        && input[reference + ] == input[iidx + ]

                        )

                    {

                        uint len = ;

                        var maxlen = (uint)inputLength - iidx - len;

                        maxlen = maxlen > MaxRef ? MaxRef : maxlen;

                        if (oidx + lit +  +  >= outputLength)

                            return ;

                        do

                            len++;

                        while (len < maxlen && input[reference + len] == input[iidx + len]);

                        if (lit != )

                        {

                            output[oidx++] = (byte)(lit - );

                            lit = -lit;

                            do

                                output[oidx++] = input[iidx + lit];

                            while ((++lit) != );

                        }

                        len -= ;

                        iidx++;

                        if (len < )

                        {

                            output[oidx++] = (byte)((off >> ) + (len << ));

                        }

                        else

                        {

                            output[oidx++] = (byte)((off >> ) + ( << ));

                            output[oidx++] = (byte)(len - );

                        }

                        output[oidx++] = (byte)off;

                        iidx += len - ;

                        hval = (uint)(((input[iidx]) << ) | input[iidx + ]);

                        hval = (hval << ) | input[iidx + ];

                        _hashTable[((hval ^ (hval << )) >> (int)((( *  - Hlog)) - hval * ) & (Hsize - ))] = iidx;

                        iidx++;

                        hval = (hval << ) | input[iidx + ];

                        _hashTable[((hval ^ (hval << )) >> (int)((( *  - Hlog)) - hval * ) & (Hsize - ))] = iidx;

                        iidx++;

                        continue;

                    }

                }

                else if (iidx == inputLength)

                    break;

                lit++;

                iidx++;

                if (lit != MaxLit) continue;

                if (oidx +  + MaxLit >= outputLength)

                    return ;

                output[oidx++] = (byte)(MaxLit - );

                lit = -lit;

                do

                    output[oidx++] = input[iidx + lit];

                while ((++lit) != );

            }

            if (lit == ) return (int)oidx;

            if (oidx + lit +  >= outputLength)

                return ;

            output[oidx++] = (byte)(lit - );

            lit = -lit;

            do

                output[oidx++] = input[iidx + lit];

            while ((++lit) != );

            return (int)oidx;

        }

        /// <summary>

        /// 使用LibLZF算法解压缩数据

        /// </summary>

        /// <param name="input">参考数据进行解压缩</param>

        /// <param name="inputLength">要解压缩的数据的长度</param>

        /// <param name="output">引用包含解压缩数据的缓冲区</param>

        /// <param name="outputLength">输出缓冲区中压缩归档的大小</param>

        /// <returns>返回解压缩大小</returns>

        public int Decompress(byte[] input, int inputLength, byte[] output, int outputLength)

        {

            uint iidx = ;

            uint oidx = ;

            do

            {

                uint ctrl = input[iidx++];

                if (ctrl < ( << ))

                {

                    ctrl++;

                    if (oidx + ctrl > outputLength)

                    {

                        return ;

                    }

                    do

                        output[oidx++] = input[iidx++];

                    while ((--ctrl) != );

                }

                else

                {

                    var len = ctrl >> ;

                    var reference = (int)(oidx - ((ctrl & 0x1f) << ) - );

                    if (len == )

                        len += input[iidx++];

                    reference -= input[iidx++];

                    if (oidx + len +  > outputLength)

                    {

                        return ;

                    }

                    if (reference < )

                    {

                        return ;

                    }

                    output[oidx++] = output[reference++];

                    output[oidx++] = output[reference++];

                    do

                        output[oidx++] = output[reference++];

                    while ((--len) != );

                }

            }

            while (iidx < inputLength);

            return (int)oidx;

        }

以上是LZF算法的代码。

Apple的LZF算法解析的更多相关文章

地理围栏算法解析（Geo-fencing）
地理围栏算法解析 http://www.cnblogs.com/LBSer/p/4471742.html 地理围栏(Geo-fencing)是LBS的一种应用,就是用一个虚拟的栅栏围出一个虚拟地理边界 ...
KMP串匹配算法解析与优化
朴素串匹配算法说明串匹配算法最常用的情形是从一篇文档中查找指定文本.需要查找的文本叫做模式串,需要从中查找模式串的串暂且叫做查找串吧. 为了更好理解KMP算法,我们先这样看待一下朴素匹配算法吧.朴素 ...
Peterson算法与Dekker算法解析
进来Bear正在学习巩固并行的基础知识,所以写下这篇基础的有关并行算法的文章. 在讲述两个算法之前,需要明确一些概念性的问题, Race Condition(竞争条件),Situations lik ...
python常见排序算法解析
python——常见排序算法解析算法是程序员的灵魂. 下面的博文是我整理的感觉还不错的算法实现原理的理解是最重要的,我会常回来看看,并坚持每天刷leetcode 本篇主要实现九(八)大排序算法 ...
Java虚拟机对象存活标记及垃圾收集算法解析
一.对象存活标记 1. 引用计数算法给对象中添加一个引用计数器,每当有一个地方引用它时,计数器就加1:当引用失效时,计数器就减1:任何时刻计数器都为0的对象就是不可能再被使用的. 引用计数算法(Re ...
JVM垃圾回收算法解析
JVM垃圾回收算法解析标记-清除算法该算法为最基础的算法.它分为标记和清除两个阶段,首先标记出需要回收的对象,在标记结束后,统一回收.该算法存在两个问题:一是效率问题,标记和清除过程效率都不太高, ...
DeepFM算法解析及Python实现
1. DeepFM算法的提出由于DeepFM算法有效的结合了因子分解机与神经网络在特征学习中的优点:同时提取到低阶组合特征与高阶组合特征,所以越来越被广泛使用. 在DeepFM中,FM算法负责对一阶 ...
GBDT+LR算法解析及Python实现
1. GBDT + LR 是什么本质上GBDT+LR是一种具有stacking思想的二分类器模型,所以可以用来解决二分类问题.这个方法出自于Facebook 2014年的论文 Practical L ...
最长上升子序列(LIS)n2 nlogn算法解析
题目描述给定一个数列,包含N个整数,求这个序列的最长上升子序列. 例如 2 5 3 4 1 7 6 最长上升子序列为 4. 1.O(n2)算法解析看到这个题,大家的直觉肯定都是要用动态规划来做,那 ...

随机推荐

跨域无法获取自定义header的问题
同域的时候,header里面的参数可以随便自己定义.服务端都是可以获取的. 但是跨域的时候,除了设置 <add name="Access-Control-Allow-Origin&qu ...
git学习笔记一
一.概念理解 1.理解工作区和暂存区以及版本库工作区我理解就是我们创建的程序所在的文件夹,比如test文件夹.其中有个.git文件,这个就是版本库,其中版本库中有个区域叫暂存区或叫索引. 截自廖雪峰 ...
jquery 中的框架
DWZ 国产Ajax RIA开源框架 Ninja UI 框架提供页面插件 angela ui框架表单布局等 Chico UI 快速页面布局 PrimeUI w2ui 布局 ...
shared_ptr
省去对象指针的显示delete typedef tr1::shared_ptr<int> IntPtr; IntPtr fun() { IntPtr p = new int(3); ret ...
PHP的FastCGI
CGI全称是“通用网关接口”(Common Gateway Interface), 它可以让一个客户端,从网页浏览器向执行在Web服务器上的程序请求数据. CGI描述了客户端和这个程序之间传输数据的一 ...
Javascript基础回顾之(三) 面向对象
本来是要继续由浅入深表达式系列最后一篇的,但是最近团队突然就忙起来了,从来没有过的忙!不过喜欢表达式的朋友请放心,已经在写了:) 在工作当中发现大家对Javascript的一些基本原理普遍存在这里或者 ...
旺信UWP正式版发布
下载链接:https://www.microsoft.com/store/apps/9nblggh5lq9x 各位园主好,在旺信Beta版发布后近两个月,我们的新版本1.1.0终于上线了,并且更名为旺 ...
剑指Offer面试题：35.将字符串转换为数字
一.题目:将字符串转换为数字题目:写一个函数StrToInt,实现把字符串转换成整数这个功能.当然,不能使用atoi或者其他类似的库函数. 二.代码实现 (1)考虑输入的字符串是否是NULL.空字符 ...
【读书笔记】Asp.Net MVC 上传图片到数据库（会的绕行）
之前上传图片的做法都是上传到服务器上的文件夹中,再将url保存到数据库.其实在MVC中将图片上传到数据库很便捷的事情,而且不用去存url了.而且这种方式支持ie6(ie6不支持jquery自动提交fo ...
Mysql 主从延时监控
200 ? "200px" : this.width)!important;} --> 介绍主从延时在主从环境中是一个非常值得关注的问题,有时候我们可以通过show sla ...

Apple的LZF算法解析

Apple的LZF算法解析的更多相关文章

随机推荐

热门专题