Cassandra Issue with Tombstone
1. Cassandra is quicker than postgre and have lower change to lose data. Cassandra doesn't have foreign keys, locking mechanism and etcs, so that it's quicker on writes.
2. Everything in cassandra is a write. Insert/update/delete is also write.
3. Setting a column to null/ deleting a column will create a tomestone; Deleting a row/primary key/partitio will create a single row tomestone
4. Could adjust tombstone_warn_threshold and tombstone_failure_threshold in cassandra.yaml.
5. Could adjust gc_grace_seconds when creating table
6. Hitting tombstone limit only happens per query.
Related Attributes
Delete will create tombstones
tombstone_warn_threshold: 1000 (default), could be found in cassandra.yaml
tombstone_failure_threshold: 100000 (default), could be found in cassandra.yaml
tombstone_compaction_interval: table attribute
min_compaction_threshold: table attribute #Compaction will only be eligible after min_compaction_threshold SSTables exist, by default it’s 4.
gc_grace_seconds: table attribute
snapshot_before_compaction: false
Check table attributes here http://docs.datastax.com/en/cassandra/2.1/cassandra/reference/referenceTableAttributes.html
Could consider using DateTieredCompactionStrategy instead of the default SizeTieredCompactionStrategy.
Cassandra MBean
Use Jconsole to remotely connect to:
hostname:7199
e.g. localhost:7199
Check/change the TombstoneFailureThreshold attribute inside StorageService MBean.
Force a flush and compaction
sudo nodetool -h localhost -p 7199 -u OC_APP_RAINBOWDBA -pw a3c224d4b89192d2ea3ea943dd7e9648 flush rainbowdba undeliveredmessage
sudo nodetool -h localhost -p 7199 -u OC_APP_RAINBOWDBA -pw a3c224d4b89192d2ea3ea943dd7e9648 compact rainbowdba undeliveredmessage
Deleted rows will only disappear when gc_grace_seconds time passed and a flush and compaction has been forced
Truncating Table
Truncating a table is an immediate operation and won’t leave any tomestones.
Don’t insert Null into columns
Inserting a null value to the column will leave a cell tomestone. Deleting a partition/row will also create a single row tombstone.
Deleting a partition will create a partition tomestone and override the existing cell tomestones. This only happens in memory table not on the disk. Not sure whether creating a partition tomestone will cause a compaction of the cell tomestones on disk.
Using TTL
insert into undeliveredmessage("id", "message","type") values('1','message','RAVEN') using ttl 5;
This query will result in 3 tomestone cells and one row tombstone.
Cassandra partition size limitation
In Cassandra, the maximum number of cells (rows x columns) in a single partition is 2 billion.
Additionally, a single column value may not be larger than 2GB. Partitions greater than 100Mb can cause significant pressure on the heap.
Performance Test
Test script TestCassandraPerformance.java could be found in
Cassandra version: 2.2.3, cqlsh 5.0.1
1. TombstoneFailureThreshold = 500
Seems persist 102000 rows and then delete them won’t hit the limit of the tomestone.
2. TombstoneFailureThreshold = 1
insert into undeliveredmessage("id","message","sent","type") values('3', 'message3', True, null);
and then select * from undeliveredmessage is fine
2. TombstoneFailureThreshold = 1
insert into undeliveredmessage("id","message","sent","type") values('3', 'message3', null, null);
and then select * from undeliveredmessage will hit the tomestone limit
|
deleted rows number |
existing rows number |
locally recovery time |
vector 2 recovery time |
|
150_000 * 9 |
Operation Timed Out |
Operation Timed Out |
|
|
150_000 * 5 |
9_072 ms |
10_152 ms |
|
|
150_000 * 3 |
5_679 ms |
7_957 ms |
|
|
150_000 * 2 |
3_025 ms |
5_218 ms |
|
|
150_000 |
1_326 ms |
1_879 ms |
|
|
150_000 |
158 ms |
333 ms |
|
|
150_000 *2 |
562 ms |
1_963 ms |
|
|
150_000 *3 |
2_223 ms |
3_833 ms |
|
|
150_000 *5 |
3_476 ms |
9_726 ms |
|
|
150_000 *10 |
Operation Timed Out |
Operation Timed Out |
|
|
150_000 |
150_000 |
1_321 ms |
3_735 ms |
|
150_000 *2 |
150_000 |
1_893 ms |
4_939 ms |
Note that we will hit timeout issue when having 150_000 *10 deleted rows in the table.
Hitting tombstone limit
For Dash you should see
com.datastax.driver.core.exceptions.NoHostAvailableException: All host(s) tried for query failed (tried: localhost/127.0.0.1:9042 (com.datastax.driver.core.exceptions.ReadTimeoutException: Cassandra timeout during read query at consistency ONE (1 responses were required but only 0 replica responded)))
at com.datastax.driver.core.ControlConnection.reconnectInternal(ControlConnection.java:223)
For query in command line, you should see something like:
Traceback (most recent call last):
File "/usr/bin/cqlsh.py", line 1172, in perform_simple_statement
rows = future.result(self.session.default_timeout)
File "/usr/share/cassandra/lib/cassandra-driver-internal-only-2.7.2.zip/cassandra-driver-2.7.2/cassandra/cluster.py", line 3347, in result
raise self._final_exception
ReadFailure: code=1300 [Replica(s) failed to execute read] message="Operation failed - received 0 responses and 1 failures" info={'failures': 1, 'received_responses': 0, 'required_responses': 1, 'consistency': 'ONE'}
Cassandra Issue with Tombstone的更多相关文章
- Cassandra issue - "The clustering keys ordering is wrong for @EmbeddedId"
在Java连接Cassandra的情况下, 当使用组合主键时, 默认第一个是Partition Key, 后续的均为Clustering Key. 如果有多个Clustering Key, 在Java ...
- akka-typed(10) - event-sourcing, CQRS实战
在前面的的讨论里已经介绍了CQRS读写分离模式的一些原理和在akka-typed应用中的实现方式.通过一段时间akka-typed的具体使用对一些经典akka应用的迁移升级,感觉最深的是EvenSou ...
- Cassandra简介
在前面的一篇文章<图形数据库Neo4J简介>中,我们介绍了一种非常流行的图形数据库Neo4J的使用方法.而在本文中,我们将对另外一种类型的NoSQL数据库——Cassandra进行简单地介 ...
- Cassandra 计数器counter类型和它的限制
文档基础 Cassandra 2.* CQL3.1 翻译多数来自这个文档 更新于2015年9月7日,最后有参考资料 作为Cassandra的一种类型之一,Counter类型算是限制最多的一个.Coun ...
- 闲聊cassandra
原创,转载请注明出处 今天聊聊cassandra,里面用了不少分布式系统设计的经典算法比如consistent hashing, bloom filter, merkle tree, sstable, ...
- 开源软件:NoSql数据库 - 图数据库 Cassandra
转载原文:http://www.cnblogs.com/loveis715/p/5299495.html Cassandra简介 在前面的一篇文章<图形数据库Neo4J简介>中,我们介绍了 ...
- Cassandra User 问题汇总(1)------------repair
Cassandra Repair 问题 问1: 文档建议每周或者每月跑一次full repair.那么如果我是使用partition rangerepair,是否还有必要在cluster的每个节点上定 ...
- 从Stage角度看cassandra write
声明 文章发布于CSDN cassandra concurrent 具体实现 cassandra并发技术文中介绍了java的concurrent实现,这里介绍cassandra如何基于java实现ca ...
- Cassandra 原理介绍
Cassandra最初源自Facebook,结合了Google BigTable面向列的特性和[Amazon Dynamo](http://en.wikipedia.org/wiki/Dynamo(s ...
随机推荐
- llinux svn安装
1,安装SVN服务端 直接用apt-get或yum安装subversion即可(当然也可以自己去官方下载安装) [plain] view plain copy print? sudo apt-get ...
- java开发中经典的三大框架SSH
首先我们要明白什么是框架为什么用?相信一开始学习编程的时候都会听到什么.什么框架之类的:首先框架是一个软件半成品,都会预先实现一些通用功能,使用框架直接应用这些通用功能而不用重新实现,所以大多数企业都 ...
- 云计算之路-阿里云上:数据库连接数过万的真相,从阿里云RDS到微软.NET Core
在昨天的博文中,我们坚持认为数据库连接数过万是阿里云RDS的问题,但后来阿里云提供了当时的数据库连接情况,让我们动摇了自己的想法. 帐户 连接数 A 4077 B 3995 C 741 D 698 E ...
- ML(4): NavieBayes在R中的应用
朴素贝叶斯方法是一种使用先验概率去计算后验概率的方法, 具体见上一节. 算法包:e1071 函数:navieBayes(formule,data,laplace=0,...,subset,na.act ...
- Xcode8.3 添加iOS10.3以下旧版本模拟器
问题起源 由于手边项目需要适配到iOS7, 但是手边的测试机都被更新到最新版本,所以有些潜在的bug,更不发现不了.最近就是有个用户提出一个bug,而且是致命的,app直接闪退.app闪退,最常见的无 ...
- 少走弯路——Android对话框AlertDialog.Builder使用方法简述
android的自定义对话框,不需要通过继承的方式来实现,因为android已提供了相应的接口Dialog Builder ,下面就是 样例: new AlertDialog.Builder(this ...
- PAT 1047
1049. Counting Ones (30) The task is simple: given any positive integer N, you are supposed to count ...
- vue获取dom元素内容
通过ref来获取dom元素 在vue官网上对ref的解释 ref 被用来给元素或子组件注册引用信息.引用信息将会注册在父组件的 $refs 对象上.如果在普通的 DOM 元素上使用,引用指向的就是 D ...
- lua 模块
lua 模块 概述 lua 模块类似于封装库 将相应功能封装为一个模块, 可以按照面向对象中的类定义去理解和使用 使用 模块文件示例程序 mod = {} mod.constant = "模 ...
- Java 中字两个字符串判断是否相等(转载)
java中判断字符串是否相等有两种方法:1.用"=="运算符,该运算符表示指向字符串的引用是否相同,比如: String a="abc";String b=&q ...