coral-pi-rest-server
tinygrad
coral-pi-rest-server | tinygrad | |
---|---|---|
44 | 58 | |
66 | 17,800 | |
- | - | |
0.0 | 9.7 | |
7 months ago | 10 months ago | |
Jupyter Notebook | Python | |
MIT License | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
coral-pi-rest-server
- BeagleY-AI: 4 TOPS-capable $70 board from Beagleboard
- Do you recommend Orange PI for ML or LLM projects?
-
Framework for machine learning?
That said, you can always look at something like https://coral.ai/products/accelerator/ to help with the performance you need.
-
Mini PC for AI
Should only be ~$60 https://coral.ai/products/accelerator/
-
What are some USB devices worth using in a Home Lab Environment?
The Coral USB accelerator might be of interest if you want to do some light ML with a low power budget.
- Is a PCIe x1 enough for light ML tasks
-
Would I be able to run ggml models such as whisper.cpp or llama.cpp on a raspberry pi with a coral ai USB Accelerator?
However, a pi doesn't have the strength to run something like Llama.cpp, of course, so I've been considering using something like the Coral USB Accelerator (https://coral.ai/products/accelerator). As I've been learning more about it, it seems to be very geared towards TensorFlow Lite models. But whisper.cpp and Llama.cpp use ggml models.
- Looking for a Mini PC for Home Assistant and Frigate.
- AI development suite on a stick?
-
Modder wires ChatGPT into Skyrim VR so NPCs can roleplay and remember past conversations
Recently found this thing, though I haven't found a use case for me.
tinygrad
- tinygrad: extreme simplicity, easiest framework to add new accelerators to
-
GGML – AI at the Edge
Might be a silly question but is GGML a similar/competing library to George Hotz's tinygrad [0]?
[0] https://github.com/geohot/tinygrad
-
Render neural network into CUDA/HIP code
at first glance i thought may its like tinygrad. but looks has many ops than that tiny grad but most maps to underlying hardware provided ops?
i wonder how well tinygrad's apporach will work out, ops fusion sounds easy, just a walk a graph, pattern match it and lower to hardware provided ops?
Anyway if anyone wants to understand the philosophy behind tinygrad, this file is great start https://github.com/geohot/tinygrad/blob/master/docs/abstract...
-
llama.cpp now officially supports GPU acceleration.
There are currently at least 3 ways to run llama on m1 with GPU acceleration. - mlc-llm (pre-built, only 1 model has been ported) - tinygrad (very memory efficient, not that easy to integrate into other projects) - llama-mps (original llama codebase + llama adapter support)
- George Hotz building an AMD competitor to Nvidia.
-
George Hotz ROCm adventures
Hopefully we will see now full support with AMD hardware on https://github.com/geohot/tinygrad. You can read more about it on https://tinygrad.org/
-
The Coming of Local LLMs
tinygrad
https://github.com/geohot/tinygrad/tree/master/accel/ane
But I have not tested it on Linux since Asahi has not yet added support.
llama.cpp runs at 18ms per token (7B) and 200ms per token (65B) without quantization.
- Everything we know about Apple's Neural Engine
- Everything we know about the Apple Neural Engine (ANE)
- How 'Open' Is OpenAI, Really?
What are some alternatives?
alpaca.cpp - Locally run an Instruction-Tuned Chat-Style LLM
Pytorch - Tensors and Dynamic neural networks in Python with strong GPU acceleration
double-take - Unified UI and API for processing and training images for facial recognition.
llama.cpp - LLM inference in C/C++
rpi-urban-mobility-tracker - The easiest way to count pedestrians, cyclists, and vehicles on edge computing devices or live video feeds.
openpilot - openpilot is an open source driver assistance system. openpilot performs the functions of Automated Lane Centering and Adaptive Cruise Control for 250+ supported car makes and models.
opentts - Open Text to Speech Server
llama - Inference code for Llama models
HASS-coral-rest-api - Coral REST API for HASS
tensorflow_macos - TensorFlow for macOS 11.0+ accelerated using Apple's ML Compute framework.
os-nvr
GPTQ-for-LLaMa - 4 bits quantization of LLaMA using GPTQ