levelDB跳表实现

跳表的原理就是利用随机性建立索引，加速搜索，并且简化代码实现难度。具体的跳表原理不再赘述，主要是看了levelDB有一些实现细节的东西，凸显自己写的实现不足之处。

去除冗余的key

  template<typename Key, class Comparator>

  struct SkipList<Key,Comparator>::Node {

    explicit Node(const Key& k) : key(k) { }

    Key const key;

    // Accessors/mutators for links.  Wrapped in methods so we can

    // add the appropriate barriers as necessary.

    Node* Next(int n) {

      assert(n >= 0);

      // Use an 'acquire load' so that we observe a fully initialized

      // version of the returned Node.

      return reinterpret_cast<Node*>(next_[n].Acquire_Load());

    }

    void SetNext(int n, Node* x) {

      assert(n >= 0);

      // Use a 'release store' so that anybody who reads through this

      // pointer observes a fully initialized version of the inserted node.

      next_[n].Release_Store(x);

    }

    // No-barrier variants that can be safely used in a few locations.

    Node* NoBarrier_Next(int n) {

      assert(n >= 0);

      return reinterpret_cast<Node*>(next_[n].NoBarrier_Load());

    }

    void NoBarrier_SetNext(int n, Node* x) {

      assert(n >= 0);

      next_[n].NoBarrier_Store(x);

    }

   private:

    // Array of length equal to the node height.  next_[0] is lowest level link.

    port::AtomicPointer next_[1];

  };

这里使用一个Node节点表示所有相同key，不同高度的节点集合，仅保留了key和不同高度的向右指针，并且使用NewNode来动态分配随即高度的向右指针集合，而next_就指向这指针集合。这也是c/c++ tricky的地方。

  #include <stdio.h>

  struct Node {

  	char str[1];

  };

  int main() {

  	char* mem = new char[4];

  	for (int i = 0; i < 4; i++) {

  		mem[i] = i + '0';

  	}

  	Node* node = (Node*)mem;

  	char* const pstr = node->str;

  	for (int i = 0; i < 4; i++) {

  		printf("%c", pstr[i]);

  	}

  	return 0;

  }

就像上面这个简单的sample，成员str可以作为指针指向从数组下标0开始的元素，并且不受申明时的限制，不局限于大小1，索引至分配的最大的内存地址。

简易随机数生成

  uint32_t Next() {

      static const uint32_t M = 2147483647L;   // 2^31-1

      static const uint64_t A = 16807;  // bits 14, 8, 7, 5, 2, 1, 0

      // We are computing

      //       seed_ = (seed_ * A) % M,    where M = 2^31-1

      //

      // seed_ must not be zero or M, or else all subsequent computed values

      // will be zero or M respectively.  For all other values, seed_ will end

      // up cycling through every number in [1,M-1]

      uint64_t product = seed_ * A;

      // Compute (product % M) using the fact that ((x << 31) % M) == x.

      seed_ = static_cast<uint32_t>((product >> 31) + (product & M));

      // The first reduction may overflow by 1 bit, so we may need to

      // repeat.  mod == M is not possible; using > allows the faster

      // sign-bit-based test.

      if (seed_ > M) {

        seed_ -= M;

      }

      return seed_;

  }

可以看到，他使用A和M对种子进行运算，达到一定数据范围内不会重复的数集，而里面对于(product % M)，使用(product >> 31) + (product & M)进行运算优化，考虑右移和与操作的代价远小于取余操作。

简洁清晰的私有帮助方法，帮助寻找小于指定key的节点

  template<typename Key, class Comparator>

  typename SkipList<Key,Comparator>::Node*

  SkipList<Key,Comparator>::FindLessThan(const Key& key) const {

    Node* x = head_;

    int level = GetMaxHeight() - 1;

    while (true) {

      assert(x == head_ || compare_(x->key, key) < 0);

      Node* next = x->Next(level);

      if (next == NULL || compare_(next->key, key) >= 0) {

        if (level == 0) {

          return x;

        } else {

          // Switch to next list

          level--;

  	  }

      } else {

        x = next;

      }

    }

  }

levelDB跳表实现的更多相关文章

LevelDB学习笔记 (3): 长文解析memtable、跳表和内存池Arena
LevelDB学习笔记 (3): 长文解析memtable.跳表和内存池Arena 1. MemTable的基本信息我们前面说过leveldb的所有数据都会先写入memtable中,在leveldb ...
SkipList 跳表
1.定义描述跳跃列表(也称跳表)是一种随机化数据结构,基于并联的链表,其效率可比拟于二叉查找树(对于大多数操作需要O(log n)平均时间). 基本上,跳跃列表是对有序的链表增加 ...
[转载] 跳表SkipList
原文: http://www.cnblogs.com/xuqiang/archive/2011/05/22/2053516.html leveldb中memtable的思想本质上是一个skiplist ...
skiplist 跳表（1）
最近学习中遇到一种新的数据结构,很实用,搬过来学习. 原文地址:skiplist 跳表为什么选择跳表目前经常使用的平衡数据结构有:B树,红黑树,AVL树,Splay Tree, Treep等. ...
SkipList跳表基本原理
为什么选择跳表目前经常使用的平衡数据结构有:B树,红黑树,AVL树,Splay Tree, Treep等. 想象一下,给你一张草稿纸,一只笔,一个编辑器,你能立即实现一颗红黑树,或者AVL树出来吗 ...
C语言跳表(skiplist)实现
一.简介跳表(skiplist)是一个非常优秀的数据结构,实现简单,插入.删除.查找的复杂度均为O(logN).LevelDB的核心数据结构是用跳表实现的,redis的sorted set数据结构也 ...
K：跳表
跳表(SkipList)是一种随机化的数据结构,目前在redis和leveldb中都有用到它,它的效率和红黑树以及 AVL 树不相上下,但跳表的原理相当简单,只要你能熟练操作链表, 就能轻松实现一 ...
SkipList跳表（一）基本原理
一直听说跳表这个数据结构,说要学一下的,懒癌犯了,是该治治了为什么选择跳表目前经常使用的平衡数据结构有:B树.红黑树,AVL树,Splay Tree(这个树好像还没有听说过),Treep(也没有听 ...
深入理解跳表在Redis中的应用
本文首发于:深入理解跳表在Redis中的应用微信公众号:后端技术指南针持续输出干货欢迎关注前面写了一篇关于跳表基本原理和特性的文章,本次继续介绍跳表的概率平衡和工程实现, 跳表在Redis.Lev ...

随机推荐

JavaSE复习日记 : 八种基本数据类型
/* * 基本数据类型 * * Java里的8种基本数据类型: * byte --- 1 byte = 8 bit; * short --- 2 byte = 16 bit; * int --- 4 ...
JSP——九大内置对象和其四大作用域
一.JSP九大内置对象: JSP根据Servlet API 规范提供了某些内置对象,开发者不用事先声明就可以使用标准的变量来访问这些对象. Request:代表的是来自客户端的请求,例如我们在FORM ...
fitnesse 中各类fit fixture的python实现
虽然网上都说slim效率很高,无奈找不到支持python的方法,继续用pyfit 1 Column Fixture 特点:行表格展现形式,一条测试用例对应一行数据 Wiki !define COMMA ...
poj 3243 Clever Y 高次方程
1 Accepted 8508K 579MS C++ 2237B/** hash的强大,,还是高次方程,不过要求n不一定是素数 **/ #include <iostream> #inclu ...
nodejs的npm安装模块时候报错：npm ERR! Error: CERT_NOT_YET_VALID的解决方法 - 包子博客 _ 关注互联网前端、开发、SEO、移动互联网应用技术
转载:包子博客: http://www.haodewap.net/visit.do?wapurl=http%3A%2F%2Fwww.jincon.com%2Farchives%2F141%2F
jQuery.fn.serialize 阅读
今天第一次阅读jQuery源码,因为读到用js对表单的序列化,为的是在ajax操作中将表单中各个域的值传到服务器.书上用了很长的步骤,判断每一个表单域的属性,然后拼接. 大概是这样: function ...
Windows Server 2012 安装dll到GAC
使用Windows管理员打开PowerShell: 运行以下命令: Set-location "c:\tools\gac" [System.Reflection.Assembly] ...
<input type="text">文本输人框
type类型: text 文本框 password 口令密码输人框 reset 重置或清除 buttou 命令按钮 checkbox 复选框 radio 单选框 submit 提交 fi ...
通过Manifest的配置信息实现页面跳转,及总结
1:新建一个xml文件,如second_view.xml文件,然后新建一个Activity如SecondActivity.java并在里面设置setContentView(R.layout.secon ...
java反射机制入门04
需要jxl.jar package com.rainmer.main; import java.io.File; import java.io.IOException; import java.uti ...

levelDB跳表实现

levelDB跳表实现的更多相关文章

随机推荐

热门专题