@thebloke
I am Tom, purveyor of fine local LLMs for your fun and profit.
Contribution totals unavailable
No contribution history synced yet.
A gradio web UI for running Large Language Models like LLaMA, llama.cpp, GPT-J, Pythia, OPT, and GALACTICA.
An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.
A discord bot with many features which uses A1111 as backend and uses my prompt templates for beautiful generations - even with short prompts.
Port of Facebook's LLaMA model in C/C++
YADR - The best vim,git,zsh plugins and the cleanest vimrc you've ever seen
Easy to use model parallel large language models in JAX/Flax with pjit support on cloud TPU pods.
Go ahead and axolotl questions
AutoAWQ implements the AWQ algorithm for 4-bit quantization with a 2x speedup during inference.
Tensor library for machine learning
Asynchronous Python ODM for MongoDB