ElasticSearch中term和match探索

一.创建测试数据

1.创建一个index

curl -X PUT  http://127.0.0.1:9200/student?pretty -H "Content-Type: application/json" -d '{

    "settings": {

        "number_of_shards": 1,

        "number_of_replicas": 0

    },

    "mappings": {

        "_source": {

            "enabled": true

        },

        "properties": {

            "id": {

                "type": "integer"

            },

            "name": {

                "type": "text"

            },

            "age": {

                "type": "integer"

            },

            "class": {

                "type": "text",

                "analyzer": "ik_max_word"

            },

            "introduce": {

                "type": "text",

                "analyzer": "ik_max_word"

            }

        }

    }

}'

2.验证是否创建成功

curl -XGET "http://127.0.0.1:9200/student?pretty"

3.插入测试数据

curl -X PUT http://127.0.0.1:9200/student/_doc/1?pretty -H "Content-Type: application/json" -d '{

    "id":1,

    "name":"关云长",

    "age":30,

	"class":"蜀国一班"

}'

curl -X PUT http://127.0.0.1:9200/student/_doc/2?pretty -H "Content-Type: application/json" -d '{

    "id":2,

    "name":"吕蒙",

    "age":25,

	"class":"吴国一班"

}'

curl -X PUT http://127.0.0.1:9200/student/_doc/3?pretty -H "Content-Type: application/json" -d '{

    "id":3,

    "name":"吕布",

    "age":40,

	"class":"三姓一班"

}'

curl -X PUT http://127.0.0.1:9200/student/_doc/4?pretty -H "Content-Type: application/json" -d '{

    "id":4,

    "name":"张翼德",

    "age":30,

	"class":"蜀国二班"

}'

4.查询所有数据，验证是否正确

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '

{

    "query": {

        "match_all": {}

    }

}'

二.验证



#关于term和match，下面两个查询，term没有结果，match有结果，为什么？

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "term": {"name":"吕蒙"}

    }

}'

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "match": {"name":"吕蒙"}

    }

}'

拿A去B里匹配，A能分词，B也能分词。term不会将A分词，match会将A分词，存储数据类型keyword不会将B分词，text会将B分词。

可以看到上面用term方式查找，没有结果，而用match方式查找，能查找到“吕蒙”和“吕布”两个结果

term是不分词（不拆分搜索字）查找目标字段中是否有要查找的文字，也就是完整查找“吕蒙”两个字，而name这个字段用的是text类型存储的，text类型数据默认是分词的，也就是elasticsearch会将name分词后（分成“吕”和“蒙”）再存储，这时候拿完整的搜索字“吕蒙”去存储的“吕”、“蒙”里找肯定是找不到的。

match是分词（拆分搜索字）查找目标字段，也就是说会先将要查找的搜索子“吕蒙”拆成“吕”和“蒙”，再分别去name里找“吕”，如果没有找到“吕”，还会去找“蒙”，而存储的数据里，text已经将“吕蒙”和“吕布”都分词成了“吕”，“蒙”，“吕”，“布”存储了，所以光通过一个“吕”字就能找到两条结果。

这里要区分搜索词的分词，以及字段存储的分词。拿A去B里匹配，A能分词，B也能分词。term不会将A分词，match会将A分词。

既然name的类型，存储的时候就是分词的，那能不能在存储的时候不分词了，可以用将text类型改成keyword类型

#删除所有文档

curl -XPOST "http://127.0.0.1:9200/student/_delete_by_query?pretty" -v -H "Content-Type: application/json" -d '

{

    "query": {

        "match_all": {}

    }

}'

#删除索引

curl -XDELETE "http://127.0.0.1:9200/student?pretty"

#重新创建索引，将name字段的类型改成keyword

curl -X PUT  http://127.0.0.1:9200/student?pretty -H "Content-Type: application/json" -d '{

    "settings": {

        "number_of_shards": 1,

        "number_of_replicas": 0

    },

    "mappings": {

        "_source": {

            "enabled": true

        },

        "properties": {

            "id": {

                "type": "integer"

            },

            "name": {

                "type": "keyword"

            },

            "age": {

                "type": "integer"

            },

            "class": {

                "type": "text",

                "analyzer": "ik_max_word"

            },

            "introduce": {

                "type": "text",

                "analyzer": "ik_max_word"

            }

        }

    }

}'

#重新插入上面四条数据

#请复制上面的语句，执行

#下面这条查询将返回“吕蒙”同学

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "term": {"name":"吕蒙"}

    }

}'

#下面这条查询将返回0结果，因为存储时类型为keyword没有分词，所以存储的是“吕蒙”和“吕布”，这时候拿#“吕”去匹配，没有匹配的结果

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "term": {"name":"吕"}

    }

}'

#下面的结果将只会返回“吕蒙”同学，没有匹配的结果，因为存储时类型为keyword没有分词，所以存储的“吕

#蒙”和“吕布”，这时候拿“吕蒙”去匹配，虽然用的match，会将搜索词拆分成“吕蒙”，“吕”，“蒙”去搜索，但

#“吕”和“蒙”都不会匹配的到存储的“吕蒙”和“吕布”

curl -XGET http://127.0.0.1:9200/student/_search?pretty -H "Content-Type: application/json" -d '{

    "query": {

           "match": {"name":"吕蒙"}

    }

}'

ElasticSearch中term和match探索的更多相关文章

elasticsearch 中的Multi Match Query
在Elasticsearch全文检索中,我们用的比较多的就是Multi Match Query,其支持对多个字段进行匹配.Elasticsearch支持5种类型的Multi Match,我们一起来深入 ...
ES查询－term VS match （转）
原文地址:https://blog.csdn.net/sxf_123456/article/details/78845437 elasticsearch 中term与match区别 term是精确查询 ...
Elasticsearch 5.0 中term 查询和match 查询的认识
Elasticsearch 5.0 关于term query和match query的认识一.基本情况前言:term query和match query牵扯的东西比较多,例如分词器.mapping ...
Elasticsearch学习系列之term和match查询
lasticsearch查询模式一种是像传递URL参数一样去传递查询语句,被称为简单查询 GET /library/books/_search //查询index为library,type为book ...
Elasticsearch学习系列之term和match查询实例
Elasticsearch查询模式一种是像传递URL参数一样去传递查询语句,被称为简单查询 GET /library/books/_search //查询index为library,type为boo ...
Elasticsearch中的Term查询和全文查询
目录前言 Term 查询 exists 查询 fuzzy 查询 ids 查询 prefix 查询 range 查询 regexp 查询 term 查询 terms 查询 terms_set 查询 t ...
在Elasticsearch中查询Term Vectors词条向量信息
这篇文章有点深度,可能需要一些Lucene或者全文检索的背景.由于我也很久没有看过Lucene了,有些地方理解的不对还请多多指正. 更多内容还请参考整理的ELK教程关于Term Vectors 额, ...
如何在Elasticsearch中安装中文分词器(IK+pinyin)
如果直接使用Elasticsearch的朋友在处理中文内容的搜索时,肯定会遇到很尴尬的问题--中文词语被分成了一个一个的汉字,当用Kibana作图的时候,按照term来分组,结果一个汉字被分成了一组. ...
elasticsearch中常用的API
elasticsearch中常用的API分类如下: 文档API: 提供对文档的增删改查操作搜索API: 提供对文档进行某个字段的查询索引API: 提供对索引进行操作,查看索引信息等查看API: ...

随机推荐

微信浏览器H5开发常见的坑
ios端兼容input光标高度问题详情描述: input输入框光标,在安卓手机上显示没有问题,但是在苹果手机上当点击输入的时候,光标的高度和父盒子的高度一样.例如下图,左图是正常所期待的输入框光标 ...
ubantu 安装boost库 c++connector
安装libmysqlcppconn: sudo apt-get install libmysqlcppconn-dev 安装libboost: sudo apt-get install libboos ...
解决python 保存json到文件时中文显示16进制编码的问题
python 2.7 import codecs import json with codecs.open('Options.json', 'w', encoding='utf-8') as f: j ...
linux调用库的方式
linux调用库的方式有三种:1.静态链接库2.动态链接库3.动态加载库其中1,2都是在编程时直接调用,在链接时加参数-l进行链接,运行时自动调用第三种需要在编程时使用dlopen等函数来获取库里面 ...
BigDecimal数据的加减乘除 N次幂运算以及比较大小
在实际开开发过程中BigDecimal是一个经常用到的类: 它可以进行大数值的精确却运算,下面介绍一下它的加-减-乘-除以及N次幂的操作操作 import java.math.BigDecimal; ...
仙剑奇侠传1系列:2.编译主程序SDLPAL及SDL
上一篇:仙剑奇侠传1系列:1.本地运行环境及兼容性设置介绍仙剑奇侠传1是dos时代的经典游戏,相信以下图片能勾起大家的很多回忆. sdlpal是仙剑奇侠传1的主程序.github项目sdlpa ...
python内置数据结构
数据类型: 数值型 int float complex bool 序列对象字符串 str 列表 list 元组 tuple 键值对集合 set 字典dict 数值型: int.float.comp ...
centos 7设置limit，不生效问题
1:记录未修改之前的ulimit值 [root@bogon ~]# ulimit -a 2:修改配置文件 vim /etc/security/limits.conf 在后面添加 * s ...
反向代理远端单台tomcat 使用域名代理
.环境 nginx 10.1.1.161 公网:123.58.251.166 tomcat 10.1.1.103 .远端tomcat 配置 [root@host---- ~]# netstat -tn ...
【ARM-Linux开发】Linux内存管理：ARM Memory Layout以及mmu配置
原文:Linux内存管理:ARM Memory Layout以及mmu配置在内核进行page初始化以及mmu配置之前,首先需要知道整个memory map. 1. ARM Memory Layout ...