From 3ef305da6edb10bf764b5e28db7d1ec1a8091406 Mon Sep 17 00:00:00 2001 From: PENG Bo <33809201+BlinkDL@users.noreply.github.com> Date: Sun, 10 Jul 2022 14:50:21 +0800 Subject: [PATCH] Update README.md --- README.md | 4 +++- 1 file changed, 3 insertions(+), 1 deletion(-) diff --git a/README.md b/README.md index 12d5ed7..b90e197 100644 --- a/README.md +++ b/README.md @@ -14,7 +14,9 @@ RWKV-3 1.5B = always 0.015 sec/token, tested using simple pytorch code (no CUDA) GPT2-XL 1.3B = 0.032 sec/token (for ctxlen 1000), tested using HF, GPU utilization 45% too (interesting), VRAM 9655M -**Join our Discord**: https://discord.gg/bDSBUMeFpc :) I am looking for CUDA gurus to optimize the kernel. Thank you. +## Join our Discord: https://discord.gg/bDSBUMeFpc :) + +I am looking for CUDA gurus to optimize the kernel. Thank you. Reddit discussion: https://www.reddit.com/r/MachineLearning/comments/umq908/r_rwkvv2rnn_a_parallelizable_rnn_with/