An extremely simple video meeting, integrated whiteboard, chat and screen sharing
-
Updated
Apr 9, 2021 - Go
An extremely simple video meeting, integrated whiteboard, chat and screen sharing
A from-scratch implementation of Llama-3.2-1B in PyTorch, decode-latency benchmarks on three GPUs (T4, L4, A100), and weight-only quantization (RTN and GPTQ) measured against both.
To associate your repository with the rtn topic, visit your repo's landing page and select "manage topics."