2026必看AI干货!《大模型/AIGC/GPT-4/Transformer/DL/KG/NLP/CV AI+X》集合
Ranking
No observed public metrics; popularity remains neutral/archived.
Merged summary
TL;DR - A link-roundup post from the WeChat account 专知 (Zhuanzhi) aggregating hundreds of 2025–2026 AI survey papers, PhD theses, conference tutorials, books, and industry reports across LLMs, agents, multimodal, and AI4Science. It is a curation/index item rather than new research, useful mainly as a reading map of where the field's review literature currently concentrates.
- Heaviest concentration is on LLM agents: surveys on agentic RL, agent memory (evaluation taxonomies, SSGM controlled-memory framework), deep research/search agents, multi-agent systems (MASPO prompt optimization), agent communication protocols/MCP, and full-stack agent security.
- Multimodal and embodied AI form the second cluster: vision-language-action models, embodied world models, UAV vision-language navigation, 3D/4D scene generation and world modeling, edge-side embodied foundation models.
- Efficiency and systems recur throughout: LLM inference engines and serving, LoRA variant taxonomies and low-rank structure, model merging, knowledge/dataset distillation, test-time scaling, and efficient reasoning ("don't overthink" R1-style surveys).
- Also indexed: named news items (CVPR 2026 awards with He Kaiming's ResNet and YOLO taking test-of-time; DeepSeek's open-sourced sparse "memory" module paper co-signed by Liang Wenfeng; GLM-5; Stanford AI Index 2025), plus theses from CMU, Stanford, Berkeley, MIT, EPFL. No technical results are presented — only titles and links.
Sources (1)
2026必看AI干货!《大模型/AIGC/GPT-4/Transformer/DL/KG/NLP/CV AI+X》集合
TL;DR - A link-roundup post from the WeChat account 专知 (Zhuanzhi) aggregating hundreds of 2025–2026 AI survey papers, PhD theses, conference tutorials, books, and industry reports across LLMs, agents, multimodal, and AI4Science. It is a curation/index item rather than new research, useful mainly as a reading map of where the field's review literature currently concentrates.
- Heaviest concentration is on LLM agents: surveys on agentic RL, agent memory (evaluation taxonomies, SSGM controlled-memory framework), deep research/search agents, multi-agent systems (MASPO prompt optimization), agent communication protocols/MCP, and full-stack agent security.
- Multimodal and embodied AI form the second cluster: vision-language-action models, embodied world models, UAV vision-language navigation, 3D/4D scene generation and world modeling, edge-side embodied foundation models.
- Efficiency and systems recur throughout: LLM inference engines and serving, LoRA variant taxonomies and low-rank structure, model merging, knowledge/dataset distillation, test-time scaling, and efficient reasoning ("don't overthink" R1-style surveys).
- Also indexed: named news items (CVPR 2026 awards with He Kaiming's ResNet and YOLO taking test-of-time; DeepSeek's open-sourced sparse "memory" module paper co-signed by Liang Wenfeng; GLM-5; Stanford AI Index 2025), plus theses from CMU, Stanford, Berkeley, MIT, EPFL. No technical results are presented — only titles and links.