Elasticsearch 常用任务管理命令及实战应用

常用任务管理命令

  • 列出所有任务
shell 复制代码
curl -X GET "http://<es_host>:<es_port>/_tasks?detailed=true&pretty" -H 'Content-Type: application/json'
  • 获取特定类型的任务
shell 复制代码
curl -X GET "http://<es_host>:<es_port>/_tasks?actions=<action_type>" -H 'Content-Type: application/json'
  • 列出所有查询任务
shell 复制代码
curl -X GET "http://<es_host>:<es_port>/_tasks?detailed=true&actions=*search" -H 'Content-Type: application/json'
  • 取消所有查询任务
    如果 es 查询因大任务而卡住,可以临时采取此措施
shell 复制代码
curl -X POST "http://<es_host>:<es_port>/_tasks/_cancel?actions=*search" -H 'Content-Type: application/json'
  • 获取特定任务的详细信息
shell 复制代码
curl -X GET "http://<es_host>:<es_port>/_tasks/<task_id>" -H 'Content-Type: application/json'
  • 取消特定任务
shell 复制代码
curl -X POST "http://<es_host>:<es_port>/_tasks/_cancel?task_id=<task_id>" -H 'Content-Type: application/json'
  • 获取特定节点上的任务
shell 复制代码
curl -X GET "http://<es_host>:<es_port>/_tasks?nodes=<node_id>" -H 'Content-Type: application/json'

实战

定时检测 Elasticsearch 后台运行的查询任务,如果任务运行时间超过 59 秒,则进行企业微信群告警通知

python 复制代码
import requests
import time

# Elasticsearch节点的URL
es_url = "http://<es_user>:<es_pwd>@<es_host>:<es_port>/_tasks?detailed=true"

# 获取任务信息
response = requests.get(es_url)
tasks_data = response.json()

# 遍历节点和任务
for node_id, node_info in tasks_data.get('nodes', {}).items():
    for task_id, task_info in node_info.get('tasks', {}).items():
        running_time_seconds = task_info.get('running_time_in_nanos', 0) / 1e9
        description = task_info.get('description', '')
        
        if running_time_seconds > 59 and description:
            running_time_formatted = f"{running_time_seconds:.2f}"
            # 准备单个任务的Markdown内容
            content = (
                f"# 有大任务在 Elasticsearch 上运行\n"
                f"- **任务 ID**: {task_id}\n"
                f"  **查询语句**: {description}\n"
                f"  **运行时间**: {running_time_formatted} seconds\n"
            )

            # 发送到Webhook
            QYWX_BODY = {
                "msgtype": "markdown",
                "markdown": {
                    "content": content
                }
            }

            BOT_KEY = "xxxxxxxxxxxxxxxxx"  # 企业微信群 bot key

            webhook_url = f"https://qyapi.weixin.qq.com/cgi-bin/webhook/send?key={BOT_KEY}"
            headers = {'Content-Type': 'application/json; charset=utf-8'}

            response = requests.post(webhook_url, json=QYWX_BODY, headers=headers)

            # 检查响应状态
            if response.status_code == 200:
                print(f"Notification for Task ID {task_id} sent successfully.")
            else:
                print(f"Failed to send notification for Task ID {task_id}. Status code: {response.status_code}, Response: {response.text}")
            # 等待 2 秒
            time.sleep(2)

print("Processing completed.")
相关推荐
Shawn Dev22 分钟前
MySQL Binlog 数据误删除恢复完全指南
数据库·mysql·elasticsearch
Elastic 中国社区官方博客2 小时前
用两行 JSON 替换你的 ILM 策略:数据流生命周期新增冻结层支持
大数据·运维·elasticsearch·搜索引擎·架构·全文检索
小张同学a.3 小时前
ELK企业级日志分析平台3——ES数据备份 & 集群监控 & ELFK+Kafka 架构部署
linux·运维·elk·elasticsearch·架构·kafka·filebeat
lsh曙光21 小时前
ELK日志平台--elasticsearch部署
elasticsearch
Elasticsearch21 小时前
训练量仅占 0.35%,竞争力却达到 100%:jina-embeddings-v5-omni 背后的冻结塔架构
elasticsearch
扶苏10021 天前
同一个 Git 项目整出两份,切分支互不影响?两种方案实测
大数据·git·elasticsearch
Elasticsearch1 天前
用两行 JSON 替换你的 ILM 策略:数据流生命周期新增冻结层支持
elasticsearch
AI大模型-小雄2 天前
Codex写分页接口为什么越翻越慢?用Cursor Pagination解决重复与漏数据
大数据·数据库·elasticsearch·搜索引擎·chatgpt·后端开发·codex
Elasticsearch2 天前
深入浅出 Elastic 工作流(Workflows):核心架构与十一大顶级字段详解
elasticsearch
Elasticsearch2 天前
在 Elasticsearch 中构建上下文:AI Indices 如何使用更少的 tokens 为更智能的 agent 提供支持
elasticsearch