ES-mapping

类似数据库中的表结构定义,主要作用如下

定义Index下的字段名( Field Name)

定义字段的类型,比如数值型、字符串型、布尔型等定义倒排索引相关的配置,比如是否索引、记录 position 等

index_options 用于控制倒排索记录的内容,有如下4种配置

  • docs 只记录 doc id

  • freqs 记录 doc id 和 term frequencies

  • positions 记录 doc id、term frequencies 和 term position

-offsets 记录 doc id、term frequencies、term position character offsets·

text类型默认配置为 positions,其他默认为 docs。记录内容越多,占用空间越大

核心数据类型

字符串型 text、keyword

数值型 long、integer、short、 byte、 double、 float、 half float、 scaled_float

日期类型 date

布尔类型 boolean

二进制类型 binary

范围类型 integer_range、float_range、long_range、double_range、date_range

复杂数据类型

专用类型

-记录i 地址 ip

-实现自动补全 completion

-记录分词数 token_count

-记录字符串hash值murmur3

-percolator

-join

允许根据 es 自动识别的数据类型、字段名等来动态设定字段类型,可以实现如下效果

所有字符串类型都设定为 keyword 类型,即默认不分词

所有以message开头的字段都设定为 text 类型,即分词

所有以long_开头的字段都设定为 long类型

所有自动匹配为 double 类型的都设定为 float 类型,以节省空间

匹配规则一般有如下几个参数

match_mapping_type 匹配es 自动识别的字段类型,如boolean,long,string等

match,unmatch 匹配字段名

path_match,path_unmatch 匹配路径

索引模板,英文为Index Template,主要用于在新建索引时自动应用预先设定的配置简化索引创建的操作步骤

可以设定索引的配置和mapping

可以有多个模板,根据order 设置,order 大的覆盖小的配置

相关推荐
Achou.Wang13 分钟前
《从零实现cobra》手写一个 mini 版 Cobra
大数据·elasticsearch·搜索引擎
techdashen4 小时前
不用再反复 stash:用 Git Worktree 同时开发多个分支
大数据·git·elasticsearch
heimeiyingwang7 小时前
【架构实战】GitOps实践:Kubernetes上的声明式交付
elasticsearch·架构·kubernetes
J_bean1 天前
Elasticsearch 倒排索引中词条索引原理分析 — 基于 Lucene FST
elasticsearch·lucene·倒排索引·词条索引·fst
leoZ2312 天前
Git 集成实战完全指南(五):Git Blame 与历史追踪
大数据·git·elasticsearch
leoZ2312 天前
Git 集成实战完全指南(六):Git 标签与版本管理
大数据·git·elasticsearch
海兰2 天前
【Elasticsearch】工作流自动化评估
elasticsearch·自动化·jenkins
小小放舟、2 天前
Windows 本地安装 Elasticsearch 8.10.0 与 IK 分词器(2026)
java·大数据·windows·spring boot·elasticsearch·搜索引擎·全文检索
Wzx1980123 天前
Redis&ES——Retriever的抽象实现
数据库·人工智能·redis·elasticsearch
Elasticsearch3 天前
让向量听懂孤独:我是如何用 Elastic Agent Builder 构建一个情绪疗愈 agent 的
elasticsearch