Second best imperial project, 2024.6
CUDA accelerated LLaMA 3 inference engine (C++17, KV Cache)
Online compiler for wacc language! Made by me and my teammates!
The LLM that I train(both sft and pretrain), running on a pi