[论文笔记] paper review PPT

Awesome Model Quantization ------模型量化综述(Model Quantization)

GitHub - AI-Efficiency/Awesome-Model-Quantization: A list of papers, docs, codes about model quantization. This repo is aimed to provide the info for model quantization research, we are continuously improving the project. Welcome to PR the works (papers, repositories) that are missed by the repo. · GitHub

Mixture of Weight-shared Heterogeneous Group Attention Experts for Dynamic Token-wise KV Optimization基于共享权重的异构组注意力专家混合模型用于动态逐令牌键值优化

https://aclanthology.org/2025.emnlp-main.1166.pdf

HMoE: Heterogeneous Mixture of Experts for Language Modeling用于语言建模的异构专家混合模型

HMoE: Heterogeneous Mixture of Experts for Language Modeling - ACL Anthology

BrownoutMoE: Structure-Aware Expert Grouping for Efficient and Accurate LLM Web-based Services面向结构感知的专家分组,实现高效且准确的基于Web的LLM服务

https://arxiv.org/pdf/2607.04164

相关推荐
xx_xxxxx_3 小时前
论文阅读-RoTTA
论文阅读·人工智能·深度学习·机器学习
Rocky Ding*10 小时前
GPT-6.1 Sol大模型深度解析
论文阅读·人工智能·深度学习·机器学习·aigc·ai-native·gpt-6.1 sol
m4Rk_1 天前
【论文阅读】Agent 记忆机制(87):VizoMem——把文本历史转化为可检索的视觉记忆
论文阅读·人工智能·学习·开源·github
m4Rk_2 天前
【论文阅读】Agent 记忆机制(86):Skill-Pro——用 Non-Parametric PPO 将交互经验演化为可复用技能
论文阅读·人工智能·学习·开源·github
Rocky Ding*2 天前
DeepSeek DSec技术深度解析:Agent规模化训练的真正瓶颈,是沙箱基础设施
论文阅读·人工智能·深度学习·机器学习·aigc·agent·ai-native
m4Rk_4 天前
【论文阅读】Agent 记忆机制(83):Inside Out——用可演化 PersonaTree 构建 Agent 的核心长期记忆
论文阅读·人工智能·学习·开源·github
m4Rk_4 天前
【论文阅读】Agent 记忆机制(81):EMR——用情景记忆避免 Agent 在多步推理中反复绕圈
论文阅读·人工智能·学习·开源·github
m4Rk_4 天前
【论文阅读】Agent 记忆机制(82):ReasoningBank——从成功与失败经验中沉淀可复用的推理记忆
论文阅读·人工智能·学习·开源·github
Rocky Ding*4 天前
一文读懂Qwen-Audio-3.1核心基础知识:从语音识别到可控声景与实时Agent
论文阅读·人工智能·深度学习·机器学习·aigc·ai-native·qwen-audio
Rocky Ding*5 天前
Hi-DiT:一文读懂Hybrid Latent-Pixel Diffusion Transformer的本质与技术细节
论文阅读·人工智能·深度学习·机器学习·aigc·ai-native·hi-dit