Skip to content
GitLab
Menu
Projects
Groups
Snippets
Loading...
Help
Help
Support
Community forum
Keyboard shortcuts
?
Submit feedback
Contribute to GitLab
Sign in / Register
Toggle navigation
Menu
Open sidebar
ModelZoo
llama3_pytorch
Commits
fd1124b0
Commit
fd1124b0
authored
May 23, 2024
by
Rayyyyy
Browse files
Update README
parent
06049edc
Changes
1
Hide whitespace changes
Inline
Side-by-side
Showing
1 changed file
with
5 additions
and
0 deletions
+5
-0
README.md
README.md
+5
-0
No files found.
README.md
View file @
fd1124b0
...
...
@@ -9,6 +9,11 @@ Llama-3中选择了一个相对标准的decoder-only的transformer架构。与Ll
-
采用分组查询注意力(grouped query attention,GQA)、掩码等技术,帮助开发者以最低的能耗获取绝佳的性能。
-
在8,192个tokens的序列上训练模型,使用掩码来确保self-attention不会跨越文档边界。
## 算法原理
<div
align=
center
>
<img
src=
"./doc/method.png"
/>
</div>
## 环境配置
-v 路径、docker_name和imageID根据实际情况修改
...
...
Write
Preview
Markdown
is supported
0%
Try again
or
attach a new file
.
Attach a file
Cancel
You are about to add
0
people
to the discussion. Proceed with caution.
Finish editing this message first!
Cancel
Please
register
or
sign in
to comment