ONNX 模型库
返回模型

说明文档

llama-3.2-1b-instruct-onnx

llama-3.2-1b-instruct-onnx 是 Llama 3.2 1B Instruct 的 ONNX int4 量化版本,提供了一个非常小巧、非常快速的推理实现,针对使用 Intel GPU、CPU 和 NPU 的 AI PC 进行了优化。

llama-3.2-1b-instruct 是 Meta 发布的新款 10 亿参数对话基础模型。

模型描述

  • 开发者: meta-llama
  • 量化方: llmware
  • 模型类型: llama-3.2
  • 参数量: 10 亿
  • 父模型: meta-llama/Meta-Llama-3.2-1B-Instruct
  • 语言: 英语
  • 许可证: Llama 3.2 社区许可证
  • 用途: 通用对话场景
  • RAG 基准准确率评分: 不适用
  • 量化: int4

模型卡联系方式

llmware 在 github

llmware 在 hf

llmware 网站

llmware/llama-3.2-1b-instruct-onnx

作者 llmware

↓ 1 ♥ 1

创建时间: 2024-10-26 18:21:21+00:00

更新时间: 2024-10-31 21:28:21+00:00

在 Hugging Face 上查看

文件 (10)

.gitattributes
README.md
config.json
genai_config.json
hash_record_sha256.json
model.onnx ONNX
model.onnx.data
special_tokens_map.json
tokenizer.json
tokenizer_config.json