private-gpt
SillyTavern
private-gpt | SillyTavern | |
---|---|---|
131 | 75 | |
52,412 | 677 | |
3.6% | - | |
9.2 | 10.0 | |
3 days ago | about 1 year ago | |
Python | JavaScript | |
Apache License 2.0 | GNU Affero General Public License v3.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
private-gpt
-
Ask HN: Has Anyone Trained a personal LLM using their personal notes?
PrivateGPT is a nice tool for this. It's not exactly what you're asking for, but it gets part of the way there.
https://github.com/zylon-ai/private-gpt
-
PrivateGPT exploring the Documentation
Further details available at: https://docs.privategpt.dev/api-reference/api-reference/ingestion
- Show HN: I made an app to use local AI as daily driver
-
privateGPT VS quivr - a user suggested alternative
2 projects | 12 Jan 2024
-
Ask HN: How do I train a custom LLM/ChatGPT on my own documents in Dec 2023?
Run https://github.com/imartinez/privateGPT
Then
make ingest /path/to/folder/with/files
Then chat to the LLM.
Done.
Docs: https://docs.privategpt.dev/overview/welcome/quickstart
-
Mozilla "MemoryCache" Local AI
PrivateGPT repository in case anyone's interested: https://github.com/imartinez/privateGPT . It doesn't seem to be linked from their official website.
-
What Is Retrieval-Augmented Generation a.k.a. RAG
I’m preparing a small internal tool for my work to search documents and provide answers (with references), I’m thinking of using GPT4All [0], Danswer [1] and/or privateGPT [2].
The RAG technique is very close to what I have in mind, but I don’t want the LLM to “hallucinate” and generate answers on its own by synthesizing the source documents. As stated by many others, we’re living in interesting times.
[0] https://gpt4all.io/index.html
[1] https://www.danswer.ai/
[2] https://github.com/imartinez/privateGPT
- LM Studio – Discover, download, and run local LLMs
-
Ask HN: Local LLM Recommendation?
https://www.reddit.com/r/LocalLLaMA/comments/14niv66/using_a...
https://github.com/imartinez/privateGPT
-
Run ChatGPT-like LLMs on your laptop in 3 lines of code
I've been playing around with https://github.com/imartinez/privateGPT and https://github.com/simonw/llm and wanted to create a simple Python package that made it easier to run ChatGPT-like LLMs on your own machine, use them with non-public data, and integrate them into practical applications.
This resulted in Python package I call OnPrem.LLM.
In the documentation, there are examples for how to use it for information extraction, text generation, retrieval-augmented generation (i.e., chatting with documents on your computer), and text-to-code generation: https://amaiya.github.io/onprem/
Enjoy!
SillyTavern
-
Help😢
Go to Termix and click Exit. Then go to Termux and code 1. Apk update 2. Apk upgrade 3. git clone https://github.com/Cohee1207/SillyTavern 4. cd SillyTavern 5. Install nodejs 6. Npm install 7. Node server
-
Oogabooga and llama.cpp in longer conversations answers take forever.....
If you want the best roleplaying experience, I can only recommend SillyTavern with SillyTavern/SillyTavern-extras. The extras include summarization and ChromaDB, both helping to get longer and more coherent chats.
-
koboldcpp-1.33 Ultimate Edition released!
Really? Then we definitely have different experiences (or different ways to interact) with Guanaco. It's been the most unrestricted model I've tried, and I tried them all, but I'm using SillyTavern and the simple-proxy-for-tavern which combined with a little prompting liberates basically any model.
-
The best 13B model for rolepay?
Why reinvent the wheel? Just use SillyTavern, ideally with the simple-proxy-for-tavern. That does it all, and more.
-
airoboros gpt4 v1.2
I tested this today in an hours-long direct roleplay comparison between q3_K_M quants of TheBloke/airoboros-65B-gpt4-1.2-GGML and TheBloke/guanaco-65B-GGML, using koboldcpp as backend together with simple-proxy-for-tavern and SillyTavern as frontend.
-
What are you using for RP?
I'm using SillyTavern frontend and simple-proxy-for-tavern with koboldcpp backend.
-
KoboldCPP Updated to Support K-Quants, new bonus CUDA build.
I'm using SillyTavern frontend and simple-proxy-for-tavern with koboldcpp. Not sure which of these has solved the prompt-reprocessing problem, but I no longer have these slowdowns.
-
What are your favorite LLMs?
WizardLM 30B V1.0 is not only smarter and follows instructions better than the others, it's even uncensored when used with an uncensoring character card (I use SillyTavern as my GUI/frontend) - more so than any other model I tested. Probably because it follows instructions so well, thus roleplaying an uncensored character properly (and not breaking character or going "as an AI" even once during my tests).
-
Potato's brain guide to installing and reopening SillyTavern for Mac
curl -o- https://raw.githubusercontent.com/nvm-sh/nvm/v0.39.3/install.sh | bash export NVM_DIR="$([ -z "${XDG_CONFIG_HOME-}" ] && printf %s "${HOME}/.nvm" || printf %s "${XDG_CONFIG_HOME}/nvm")" [ -s "$NVM_DIR/nvm.sh" ] && \. "$NVM_DIR/nvm.sh" nvm install node git clone -b dev https://github.com/Cohee1207/SillyTavern && cd SillyTavern npm i && node server.js
-
I've found a solution to Poe API error
For Android (Termux users): 1. apt update 2. apt upgrade 3. Type "y" to everything and hit enter 4. pkg install git 5. git clone -b dev https://github.com/Cohee1207/SillyTavern 6. cd SillyTavern 7. pkg install nodejs 8. npm install 9. bash start.sh
What are some alternatives?
localGPT - Chat with your documents on your local device using GPT models. No data leaves your device and 100% private.
koboldcpp - A simple one-file way to run various GGML and GGUF models with KoboldAI's UI
gpt4all - gpt4all: run open-source LLMs anywhere
TavernAI - TavernAI for nerds [Moved to: https://github.com/Cohee1207/SillyTavern]
h2ogpt - Private chat with local GPT with document, images, video, etc. 100% private, Apache 2.0. Supports oLLaMa, Mixtral, llama.cpp, and more. Demo: https://gpt.h2o.ai/ https://codellama.h2o.ai/
langflow - ⛓️ Langflow is a dynamic graph where each node is an executable unit. Its modular and interactive design fosters rapid experimentation and prototyping, pushing hard on the limits of creativity.
ollama - Get up and running with Llama 3, Mistral, Gemma, and other large language models.
character-editor - Create, edit and convert AI character files for CharacterAI, Pygmalion, Text Generation, KoboldAI and TavernAI
text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.
simple-proxy-for-tavern
llama.cpp - LLM inference in C/C++
ChatRWKV - ChatRWKV is like ChatGPT but powered by RWKV (100% RNN) language model, and open source.