ElasticSearch中term和match探索

一.创建测试数据

1.创建一个index

curl -X PUT  http://127.0.0.1:9200/student?pretty -H "Content-Type: application/json" -d '{

    "settings": {

        "number_of_shards": 1,

        "number_of_replicas": 0

    },

    "mappings": {

        "_source": {

            "enabled": true

        },

        "properties": {

            "id": {

                "type": "integer"

            },

            "name": {

                "type": "text"

            },

            "age": {

                "type": "integer"

            },

            "class": {

                "type": "text",

                "analyzer": "ik_max_word"

            },

            "introduce": {

                "type": "text",

                "analyzer": "ik_max_word"

            }

        }

    }

}'

2.验证是否创建成功

curl -XGET "http://127.0.0.1:9200/student?pretty"

3.插入测试数据

curl -X PUT http://127.0.0.1:9200/student/_doc/1?pretty -H "Content-Type: application/json" -d '{

    "id":1,

    "name":"关云长",

    "age":30,

	"class":"蜀国一班"

}'

curl -X PUT http://127.0.0.1:9200/student/_doc/2?pretty -H "Content-Type: application/json" -d '{

    "id":2,

    "name":"吕蒙",

    "age":25,

	"class":"吴国一班"

}'

curl -X PUT http://127.0.0.1:9200/student/_doc/3?pretty -H "Content-Type: application/json" -d '{

    "id":3,

    "name":"吕布",

    "age":40,

	"class":"三姓一班"

}'

curl -X PUT http://127.0.0.1:9200/student/_doc/4?pretty -H "Content-Type: application/json" -d '{

    "id":4,

    "name":"张翼德",

    "age":30,

	"class":"蜀国二班"

}'

4.查询所有数据，验证是否正确

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '

{

    "query": {

        "match_all": {}

    }

}'

二.验证



#关于term和match，下面两个查询，term没有结果，match有结果，为什么？

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "term": {"name":"吕蒙"}

    }

}'

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "match": {"name":"吕蒙"}

    }

}'

拿A去B里匹配，A能分词，B也能分词。term不会将A分词，match会将A分词，存储数据类型keyword不会将B分词，text会将B分词。

可以看到上面用term方式查找，没有结果，而用match方式查找，能查找到“吕蒙”和“吕布”两个结果

term是不分词（不拆分搜索字）查找目标字段中是否有要查找的文字，也就是完整查找“吕蒙”两个字，而name这个字段用的是text类型存储的，text类型数据默认是分词的，也就是elasticsearch会将name分词后（分成“吕”和“蒙”）再存储，这时候拿完整的搜索字“吕蒙”去存储的“吕”、“蒙”里找肯定是找不到的。

match是分词（拆分搜索字）查找目标字段，也就是说会先将要查找的搜索子“吕蒙”拆成“吕”和“蒙”，再分别去name里找“吕”，如果没有找到“吕”，还会去找“蒙”，而存储的数据里，text已经将“吕蒙”和“吕布”都分词成了“吕”，“蒙”，“吕”，“布”存储了，所以光通过一个“吕”字就能找到两条结果。

这里要区分搜索词的分词，以及字段存储的分词。拿A去B里匹配，A能分词，B也能分词。term不会将A分词，match会将A分词。

既然name的类型，存储的时候就是分词的，那能不能在存储的时候不分词了，可以用将text类型改成keyword类型

#删除所有文档

curl -XPOST "http://127.0.0.1:9200/student/_delete_by_query?pretty" -v -H "Content-Type: application/json" -d '

{

    "query": {

        "match_all": {}

    }

}'

#删除索引

curl -XDELETE "http://127.0.0.1:9200/student?pretty"

#重新创建索引，将name字段的类型改成keyword

curl -X PUT  http://127.0.0.1:9200/student?pretty -H "Content-Type: application/json" -d '{

    "settings": {

        "number_of_shards": 1,

        "number_of_replicas": 0

    },

    "mappings": {

        "_source": {

            "enabled": true

        },

        "properties": {

            "id": {

                "type": "integer"

            },

            "name": {

                "type": "keyword"

            },

            "age": {

                "type": "integer"

            },

            "class": {

                "type": "text",

                "analyzer": "ik_max_word"

            },

            "introduce": {

                "type": "text",

                "analyzer": "ik_max_word"

            }

        }

    }

}'

#重新插入上面四条数据

#请复制上面的语句，执行

#下面这条查询将返回“吕蒙”同学

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "term": {"name":"吕蒙"}

    }

}'

#下面这条查询将返回0结果，因为存储时类型为keyword没有分词，所以存储的是“吕蒙”和“吕布”，这时候拿#“吕”去匹配，没有匹配的结果

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "term": {"name":"吕"}

    }

}'

#下面的结果将只会返回“吕蒙”同学，没有匹配的结果，因为存储时类型为keyword没有分词，所以存储的“吕

#蒙”和“吕布”，这时候拿“吕蒙”去匹配，虽然用的match，会将搜索词拆分成“吕蒙”，“吕”，“蒙”去搜索，但

#“吕”和“蒙”都不会匹配的到存储的“吕蒙”和“吕布”

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "match": {"name":"吕蒙"}

    }

}'

ElasticSearch中term和match探索的更多相关文章

elasticsearch 中的Multi Match Query
在Elasticsearch全文检索中,我们用的比较多的就是Multi Match Query,其支持对多个字段进行匹配.Elasticsearch支持5种类型的Multi Match,我们一起来深入 ...
ES查询－term VS match （转）
原文地址:https://blog.csdn.net/sxf_123456/article/details/78845437 elasticsearch 中term与match区别 term是精确查询 ...
Elasticsearch 5.0 中term 查询和match 查询的认识
Elasticsearch 5.0 关于term query和match query的认识一.基本情况前言:term query和match query牵扯的东西比较多,例如分词器.mapping ...
Elasticsearch学习系列之term和match查询
lasticsearch查询模式一种是像传递URL参数一样去传递查询语句,被称为简单查询 GET /library/books/_search //查询index为library,type为book ...
Elasticsearch学习系列之term和match查询实例
Elasticsearch查询模式一种是像传递URL参数一样去传递查询语句,被称为简单查询 GET /library/books/_search //查询index为library,type为boo ...
Elasticsearch中的Term查询和全文查询
目录前言 Term 查询 exists 查询 fuzzy 查询 ids 查询 prefix 查询 range 查询 regexp 查询 term 查询 terms 查询 terms_set 查询 t ...
在Elasticsearch中查询Term Vectors词条向量信息
这篇文章有点深度,可能需要一些Lucene或者全文检索的背景.由于我也很久没有看过Lucene了,有些地方理解的不对还请多多指正. 更多内容还请参考整理的ELK教程关于Term Vectors 额, ...
如何在Elasticsearch中安装中文分词器(IK+pinyin)
如果直接使用Elasticsearch的朋友在处理中文内容的搜索时,肯定会遇到很尴尬的问题--中文词语被分成了一个一个的汉字,当用Kibana作图的时候,按照term来分组,结果一个汉字被分成了一组. ...
elasticsearch中常用的API
elasticsearch中常用的API分类如下: 文档API: 提供对文档的增删改查操作搜索API: 提供对文档进行某个字段的查询索引API: 提供对索引进行操作,查看索引信息等查看API: ...

随机推荐

2019.7.9 校内测试 T3 15数码问题
这一次是交流测试?边交流边测试(滑稽 15数码问题大家应该都玩过这个15数码的游戏吧,就在桌面小具库那里面哦. 一看到这个题就知道要GG,本着能骗点分的原则输出了 t 个无解,本来以为要爆零,没想到 ...
mybatis参数形式
1 使用map <select id="selectRole" parameterType="map" resultType="RoleMap& ...
感知机与BP神经网络的简单应用
感知机与神经元感知机(Perceptron)由两层神经元组成(输入层.输出层),输入层接收外界输入信号后传递给输出层,输出层是M-P神经元,亦称“阈值逻辑单元”(threshold logic un ...
spring boot 学习常用网站
springboot的特性 https://www.cnblogs.com/softidea/p/5644750.html 1.自定义banner https://www.cnblogs.com/cc ...
Ubuntu无法找到add-apt-repository问题的解决方法
网上查了一下资料,原来是需要 python-software-properties 于是 apt-get install python-software-properties 除此之外还要安装 s ...
CDH 更换 HDFS 数据目录
先停止 HDFS 角色. 数据文件位置默认在 /dfs/ 中,这里配置 NameNode.SecondaryNameNode.DataNode 数据目录. 先在所有 HDFS 的主机上把数据拷贝过去, ...
Qt自定义类添加qvector报错
PtsData& PtsData::operator=(const PtsData& obj){ return *this;} PtsData::~PtsData(){ }
一百二十九：CMS系统之七牛云存储介绍和配置
将图片的存储.尺寸等图片本身的一些擦做,交给七牛云处理,自己只关注网站开发本身七牛云官网:https://www.qiniu.com 操作登录后,点击管理控制台点击对象存储-->新建存储空 ...
Unity Shader基础(1):基础
一.Shaderlab语法 1.给Shader起名字 Shader "Custom/MyShader" 这个名称会出现在材质选择使用的下拉列表里 2. Properties (属性 ...
springmvc项目 logback.xml配置 logstash日志收集
配置logback,需要一个转接的Appender,可以通过Maven依赖加到项目中: <dependency> <groupId>com.cwbase</groupId ...