模型

Llama 3.1 8B Instruct

meta-llama/Llama-3.1-8B-Instruct

查看上游原文 ↗
上游访问:需要申请 本站服务:可咨询 内容检查:AI 辅助整理并检查 许可证:Llama 3.1 Community License 上游版本:0e9e39f249a1

中文简介

Llama 3.1 8B Instruct 是 Meta 发布方的多语言指令模型,固定 README 标注 128K 上下文和八种明确支持语言。该 Hugging Face 仓库为 gated,用户需要以自己的账号阅读并接受许可、提交访问请求;本站不能代替用户同意条款或保证审批。

UPSTREAM README

上游模型卡 / 数据集卡

在 Hugging Face 查看原文 ↗

版本定位:Llama 3.1 8B Instruct 是 Meta 发布的多语言指令模型,固定模型卡标注 128K 上下文和八种明确支持语言。中文项目需要自行评测,不能从“多语言”直接推出中文效果。

访问与许可:该仓库为 gated。用户必须使用自己的 Hugging Face 账号接受 Llama 3.1 Community License 和使用政策,审批决定属于发布方;本站不代签、不绕过,也不保证获批。

文件选择:仓库同时包含两套权重布局,整仓库约 29.9 GB。确定运行框架后可选择相应格式,但许可证、政策、tokenizer 和配置必须完整留存并按固定 SHA 校验。

已有简体中文译文 · Codex 基于固定版本上游 README 编写;待人工复核 · 2026-08-07 23:24

Model Information The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multilingual dialogue use cases and outperform many of the available open source and closed chat models on common industry benchmarks. Model developer: Meta Model Architecture: Llama 3.1 is an auto-regressive language model that uses an optimized transformer architecture. The tuned versions use supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF) to align with human preferences for helpfulness and safety. Training Data Params Input modalities Output modalities Context length GQA Token count Knowledge cutoff Llama 3.1 (text only) A new mix of publicly available online data. 8B Multilingual Text Multilingual Text and code 128k Yes 15T+ December 2023 70B Multilingual Text Multilingual Text and code 128k Yes 405B Multilingual Text Multilingual Text and code 128k Yes Supported languages: English, German, French, Italian, Portuguese, Hindi, Spanish, and Thai. Llama 3.1 family of models. Token counts refer to pretraining data only. All model versions use Grouped-Query Attention (GQA) for improved inference scalability. Model Release Date: July 23, 2024. Status: This is a static model trained on an offline dataset. Future versions of the tuned models will be released as we improve model safety with community feedback. License: A custom commercial license, the Llama 3.1 Community License, is available at: https://github.com/meta-llama/llama-models/blob/main/models/llama3_1/LICENSE Where to send questions or comments about the model Instructions on how to provide feedback or comments on the model can be found in the model README. For more technical information about generation parameters and recipes for how to use Llama 3.1 in applications, please go here. Int

公开页仅展示原文摘录;完整模型卡或数据集卡请前往上游仓库查看。

适用场景

适合研究 Llama 生态、英文及已声明支持语言的对话、工具与本地部署,也可作为量化和推理框架兼容性基线。中文并非固定模型卡明确列出的八种支持语言之一,中文业务应先用真实数据评测,并与 Qwen 等模型比较。

模型参数

固定版本:0e9e39f249a16976918f6564b8830bc894c89659。参数元数据:8,030,261,248;固定 README:文本输入输出、128K 上下文,明确支持英语、德语、法语、意大利语、葡萄牙语、印地语、西班牙语和泰语。

文件说明

固定 SHA 下共有 17 个文件,仓库同时包含 Transformers Safetensors 与 original consolidated 权重布局,整仓库元数据约 29.93 GB。实际选择一种运行格式即可,但 LICENSE、USE_POLICY、tokenizer 与配置都必须随版本留存。

上游文件元数据

  • .gitattributes1.48 KB
  • config.json855 B
  • generation_config.json184 B
  • LICENSE7.45 KB
  • model-00001-of-00004.safetensors4.63 GB
  • model-00002-of-00004.safetensors4.66 GB
  • model-00003-of-00004.safetensors4.58 GB
  • model-00004-of-00004.safetensors1.09 GB
  • model.safetensors.index.json23.39 KB
  • original/consolidated.00.pth14.96 GB
  • original/params.json199 B
  • original/tokenizer.model2.08 MB

硬件建议

若使用单套 BF16 8B 权重,24 GB 级 GPU 可作为短上下文验证起点;128K 上下文会大幅增加 KV cache,不能按短对话显存外推。gated 审批与硬件能力是两条独立条件,获得访问不代表设备可以运行。

注意事项

Llama 3.1 使用自定义 Community License 和 Acceptable Use Policy,不是 Apache 或 MIT;使用、分发、归属和大规模服务条件必须逐条核对。gated 资源必须由有权主体申请,任何协助都不能绕过授权。输出仍需安全、偏见和用途审查。

获取、校验与交付

咨询此资源时只需发送本页链接或资源准确全称。橙子AI科技会继续核对版本、文件与类型、README资料卡、许可证和访问条件,并在合法访问权限、许可证及平台规则允许的前提下,协助海内外下载、完整性校验及网盘或硬盘交付;本官网本身不托管或下载资源文件。

第三方资源声明

本页面为橙子AI科技基于固定版本上游卡片整理的中文信息与服务说明,不代表资源作者或平台官方页面。实际许可、访问和使用条件以上游原文为准。