Chapter 09
📌 致谢
📌 致谢
[!NOTE] 如果
MiniMind系列项目对您有所帮助,欢迎在 GitHub 上点亮一个 ⭐ 文档篇幅较长,难免存在疏漏之处,欢迎通过 Issues 交流反馈,或提交 PR 一起改进项目 您的支持与建议,都是这个项目持续迭代的重要动力!
🤝贡献者
😊鸣谢
感谢以下贡献者在训练记录、数据处理、教程整理与项目拆解等方面提供的帮助与分享:
致谢以下优秀的论文与项目:
- https://github.com/meta-llama/llama3
- https://github.com/karpathy/llama2.c
- https://github.com/DLLXW/baby-llama2-chinese
- DeepSeek-V2
- https://github.com/charent/ChatLM-mini-Chinese
- https://github.com/wdndev/tiny-llm-zh
- Mistral-MoE
- https://github.com/Tongjilibo/build_MiniLLM_from_scratch
- https://github.com/jzhang38/TinyLlama
- https://github.com/AI-Study-Han/Zero-Chatgpt
- https://github.com/xusenlinzy/api-for-open-llm
- https://github.com/HqWu-HITCS/Awesome-Chinese-LLM
🫶支持者
🎉 MiniMind 相关成果
本模型抛砖引玉地促成了一些可喜成果的落地,感谢研究者们的认可:
-
ECG-Expert-QA: A Benchmark for Evaluating Medical Large Language Models in Heart Disease Diagnosis [arxiv]
-
Binary-Integer-Programming Based Algorithm for Expert Load Balancing in Mixture-of-Experts Models [arxiv]
-
LegalEval-Q: A New Benchmark for The Quality Evaluation of LLM-Generated Legal Text [arxiv]
-
On the Generalization Ability of Next-Token-Prediction Pretraining [ICML 2025]
-
《从零开始写大模型:从神经网络到Transformer》王双、牟晨、王昊怡 编著 - 清华大学出版社
-
FedBRB: A Solution to the Small-to-Large Scenario in Device-Heterogeneity Federated Learning [TMC 2025]
-
SKETCH: Semantic Key-Point Conditioning for Long-Horizon Vessel Trajectory Prediction [arxiv]
-
A Built-in Crypto Expert for Artificial Intelligence: How Far is the Horizon? [IACR ePrint 2026]
-
RetryTrigger: Intelligent Inference Duplication for Enhancing LLM Resilience to Hardware Transient Faults [FITEE 2026]
-
进行中...
