lyogavin/godmodeanimation: 2D Game Animation in God Mode

[https://github.com/lyogavin/godmodeanimation] - 2024-08-23 19:57:40 - public:xxx

2d, ai, animation, generator, images, LLM, model, sprite - 8 | id:1492797 -

abi/secret-llama: Fully private LLM chatbot that runs entirely with a browser with no server needed. Supports Mistral and LLama 3.

[https://github.com/abi/secret-llama] - 2024-05-12 19:53:14 - public:xxx

ai, artificial, bot, browser, chat, intelligence, llama, llm, model - 9 | id:1492037 -

Optimum

[https://huggingface.co/docs/optimum/index] - 2024-03-11 19:44:39 - public:mzimmerm

ai, doc, huggingface, llm, model, optimum, repo, small, transformer - 9 | id:1489894 -

Optimum is an extension of Transformers that provides a set of performance optimization tools to train and run models on targeted hardware with maximum efficiency. It is also the repository of small, mini, tiny models.

google-research/bert: TensorFlow code and pre-trained models for BERT

[https://github.com/google-research/bert/] - 2024-03-11 04:44:09 - public:mzimmerm

ai, bert, github, home, llm, mini, model, tiny, transformer - 9 | id:1489883 -

BERT model home on github

google/bert_uncased_L-4_H-256_A-4 · Hugging Face

[https://huggingface.co/google/bert_uncased_L-4_H-256_A-4] - 2024-03-11 04:19:21 - public:mzimmerm

ai, bert, huggingface, llm, model, parameter, small, todo - 8 | id:1489880 -

Repository of all Bert models, including small. Start using this model for testing.

Open LLM Leaderboard - a Hugging Face Space by HuggingFaceH4

[https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard] - 2024-03-05 23:50:45 - public:mzimmerm

ai, compare, huggingface, llm, model - 5 | id:1489821 -

Comparison of efficiency of all LLM models on hugging face

(1) Most cost effective GPU for local LLMs? : LocalLLaMA

[https://www.reddit.com/r/LocalLLaMA/comments/12vxxze/most_cost_effective_gpu_for_local_llms/] - 2024-03-05 00:49:23 - public:mzimmerm

ai, doc, llm, model, optimize, perform - 6 | id:1489804 -

GGML quantized models. They would let you leverage CPU and system RAM, instead of having to rely on a GPU’s. This could save you a fortune, especially if go for some used AMD Epyc platforms. This could be more viable for the larger models, especially the 30B/65B parameters models which would still press or exceed the VRAM on the P40.

Optimizing LLMs for Speed and Memory

[https://huggingface.co/docs/transformers/v4.35.2/en/llm_tutorial_optimization] - 2024-03-05 00:46:21 - public:mzimmerm

ai, doc, huggingface, llm, model, optimize, perform - 7 | id:1489803 -

7 steps to master large language models (LLMs) | Data Science Dojo

[https://datasciencedojo.com/blog/master-large-language-models/#] - 2024-03-04 19:25:57 - public:mzimmerm

ai, doc, highlevel, llm, model, train - 6 | id:1489796 -

LLM for a new language : MachineLearning

[https://www.reddit.com/r/MachineLearning/comments/12xu5ls/p_llm_for_a_new_language/] - 2024-03-04 19:15:48 - public:mzimmerm

ai, highlevel, llm, model, train - 5 | id:1489794 -

High level how to train a model

Up to date List of LLM Models

[https://docs.google.com/spreadsheets/d/1kT4or6b0Fedd-W_jMwYpb63e1ZR3aePczz3zlbJW-Y4/edit#gid=741531996] - 2024-03-04 19:13:58 - public:mzimmerm

ai, doc, list, llm, model - 5 | id:1489793 -

(2) Are there any tiny (1-3b) models finetuned for coding available in GGUF format? : LocalLLaMA

[https://www.reddit.com/r/LocalLLaMA/comments/16csdq6/are_there_any_tiny_13b_models_finetuned_for/] - 2024-03-04 10:56:19 - public:mzimmerm

ai, code, generate, llm, model, newspeak, small - 7 | id:1489789 -

bigcode (BigCode)

[https://huggingface.co/bigcode] - 2024-03-04 10:50:02 - public:mzimmerm

ai, code, generate, huggingface, llm, model, newspeak, santacoder, small, starcoder - 10 | id:1489788 -

Research community developing various code models, small and big. Models may not be instruct

WizardLM (WizardLM)

[https://huggingface.co/WizardLM] - 2024-03-04 10:42:44 - public:mzimmerm

ai, code, generate, huggingface, llm, model, newspeak, small, wizardcoder - 9 | id:1489787 -

Another open source small (1B) model.

deepseek-ai (DeepSeek)

[https://huggingface.co/deepseek-ai] - 2024-03-04 10:24:32 - public:mzimmerm

ai, best, code, deepseek, good, huggingface, instruct, llm, model, newspeak, small - 11 | id:1489786 -

They have the 1.3B version!!! This may be the best to start with Newspeak. Should work train even on huggingcface

deepseek-ai/deepseek-coder-6.7b-instruct · Hugging Face

[https://huggingface.co/deepseek-ai/deepseek-coder-6.7b-instruct] - 2024-03-04 10:13:20 - public:mzimmerm

ai, code, generate, good, llm, model, newspeak, opensource - 8 | id:1489783 -

Another possible model. For coding capabilities, Deepseek Coder achieves state-of-the-art performance among open-source code models on multiple programming languages and various benchmarks.

LLaMA 7B GPU Memory Requirement - Transformers - Hugging Face Forums

[https://discuss.huggingface.co/t/llama-7b-gpu-memory-requirement/34323/6] - 2024-03-04 10:10:38 - public:mzimmerm

ai, code, generate, llama, llm, model, newspeak, train - 8 | id:1489782 -

With the optimizers of bitsandbytes (like 8 bit AdamW), you would need 2 bytes per parameter, or 14 GB of GPU memory.

stabilityai/stable-code-3b · Hugging Face

[https://huggingface.co/stabilityai/stable-code-3b] - 2024-03-04 10:05:36 - public:mzimmerm

ai, code, generate, llm, model, newspeak - 6 | id:1489781 -

Another potential model to use for Newspeak, but it is NOT open source. Adventage: 2.5B params, so should be usable in small GPUs

Can Ai Code Results - a Hugging Face Space by mike-ravkine

[https://huggingface.co/spaces/mike-ravkine/can-ai-code-results] - 2024-03-04 09:38:45 - public:mzimmerm

ai, code, generate, huggingface, llm, model, summary - 7 | id:1489779 -

Comparison of LLM models for coding

openchat/openchat-3.5-0106 · Hugging Face

[https://huggingface.co/openchat/openchat-3.5-0106] - 2024-03-04 08:41:50 - public:mzimmerm

ai, code, generate, huggingface, llm, model, openchat - 7 | id:1489775 -

Open source with lots of information. Uses Multiple undrelying models. Not sure how I would train for it

Welcome Mixtral - a SOTA Mixture of Experts on Hugging Face

[https://huggingface.co/blog/mixtral] - 2024-03-04 08:24:33 - public:mzimmerm

ai, code, generate, huggingface, llm, mixtral, model, newspeak - 8 | id:1489774 -

The Mixtral model is new, and seems to be good. Click on “Demo“ to test it

StarCoder: A State-of-the-Art LLM for Code

[https://huggingface.co/blog/starcoder] - 2024-03-04 07:43:17 - public:mzimmerm

ai, code, generate, good, huggingface, llm, model, newspeak - 8 | id:1489773 -

Article has comparison with other code-LLM models

huybery/Awesome-Code-LLM: An awesome and curated list of best code-LLM for research.

[https://github.com/huybery/Awesome-Code-LLM] - 2024-03-04 07:33:15 - public:mzimmerm

ai, code, generate, list, llm, model - 6 | id:1489772 -

Hannibal046/Awesome-LLM: Awesome-LLM: a curated list of Large Language Model

[https://github.com/Hannibal046/Awesome-LLM] - 2024-03-04 07:31:48 - public:mzimmerm

ai, list, llm, model - 4 | id:1489771 -

Includes code generation models

Large language model - Wikipedia

[https://en.wikipedia.org/wiki/Large_language_model#List] - 2024-03-04 07:08:48 - public:mzimmerm

ai, license, list, llm, model - 5 | id:1489769 -

List of LLM models on Wikipedia

Replit — How to train your own Large Language Models

[https://blog.replit.com/llm-training] - 2024-03-02 10:18:28 - public:mzimmerm

ai, doc, language, llm, model, train - 6 | id:1489728 -

Hi level only talk about training for a language

How to train a new language model from scratch using Transformers and Tokenizers

[https://huggingface.co/blog/how-to-train] - 2024-03-02 09:48:13 - public:mzimmerm

ai, best, doc, good, language, llm, model, todo, train - 9 | id:1489725 -

Describes how to train a new language (desperanto) model.

Stability AI Launches the First of its StableLM Suite of Language Models — Stability AI

[https://stability.ai/blog/stability-ai-launches-the-first-of-its-stablelm-suite-of-language-models] - 2023-04-27 07:23:08 - public:xxx

ai, artificial, intelligence, language, llm, model, open, opensource, stability, stable-diffusion - 10 | id:1414283 -

yabs.io

Yet Another Bookmarks Service

Search

Results