mlc-llm vs ollama

mlc-llm

Universal LLM Deployment Engine with ML Compilation (by mlc-ai)

Source Code

llm.mlc.ai

Suggest alternative

Edit details

ollama

Get up and running with Llama 3, Mistral, Gemma, and other large language models. (by ollama)

Artificial intelligence

Source Code

ollama.com

Suggest alternative

Edit details

Scout Monitoring - Free Django app performance insights with Scout Monitoring

Get Scout setup in minutes, and let us sweat the small stuff. A couple lines in settings.py is all you need to start monitoring your apps. Sign up for our free tier today.

www.scoutapm.com

featured

InfluxDB - Power Real-Time Data Analytics at Scale

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured

mlc-llm		ollama
	Project
89	Mentions	222
17,358	Stars	71,334
2.3%	Growth	12.2%
9.9	Activity	9.9
4 days ago	Latest Commit	1 day ago
Python	Language	Go
Apache License 2.0	License	MIT License

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

mlc-llm

Posts with mentions or reviews of mlc-llm. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-03-04.

FLaNK 04 March 2024
26 projects | dev.to | 4 Mar 2024
Ai on a android phone?
2 projects | /r/LocalLLaMA | 8 Dec 2023

This one uses gpu, it doesn't support Mistral yet: https://github.com/mlc-ai/mlc-llm
MLC vs llama.cpp
2 projects | /r/LocalLLaMA | 7 Nov 2023

I have tried running mistral 7B with MLC on my m1 metal. And it kept crushing (git issue with description). Memory inefficiency problems.
[Project] Scaling LLama2 70B with Multi NVIDIA and AMD GPUs under 3k budget
1 project | /r/LocalLLaMA | 21 Oct 2023

Project: https://github.com/mlc-ai/mlc-llm
Scaling LLama2-70B with Multi Nvidia/AMD GPU
2 projects | news.ycombinator.com | 19 Oct 2023
AMD May Get Across the CUDA Moat
8 projects | news.ycombinator.com | 6 Oct 2023

For LLM inference, a shoutout to MLC LLM, which runs LLM models on basically any API that's widely available: https://github.com/mlc-ai/mlc-llm
ROCm Is AMD's #1 Priority, Executive Says
5 projects | news.ycombinator.com | 26 Sep 2023

One of your problems might be that gfx1032 is not supported by AMD's ROCm packages, which has a laughably short list of supported hardware: https://rocm.docs.amd.com/en/latest/release/gpu_os_support.h...
The normal workaround is to assign the closest architecture, eg gfx1030, so `HSA_OVERRIDE_GFX_VERSION=10.3.0` might help
Also, it looks like some of your tested projects are OpenCL? For me, I do something like: `yay -S rocm-hip-sdk rocm-ml-sdk rocm-opencl-sdk` to cover all the bases.
My recent interest has been LLMs and this is my general step by step for those (llama.cpp, exllama) for those interested: https://llm-tracker.info/books/howto-guides/page/amd-gpus
I didn't port the docs back in, but also here's a step-by-step w/ my adventures getting TVM/MLC working w/ an APU: https://github.com/mlc-ai/mlc-llm/issues/787
From my experience, ROCm is improving, but there's a good reason that Nvidia has 90% market share even at big price premiums.
Show HN: Ollama for Linux – Run LLMs on Linux with GPU Acceleration
14 projects | news.ycombinator.com | 26 Sep 2023

Maybe they're talking about https://github.com/mlc-ai/mlc-llm which is used for web-llm (https://github.com/mlc-ai/web-llm)? Seems to be using TVM.
Show HN: Fine-tune your own Llama 2 to replace GPT-3.5/4
8 projects | news.ycombinator.com | 12 Sep 2023

you already have TVM for the cross platform stuff
see https://tvm.apache.org/docs/how_to/deploy/android.html
or https://octoml.ai/blog/using-swift-and-apache-tvm-to-develop...
or https://github.com/mlc-ai/mlc-llm
Ask HN: Are you training and running custom LLMs and how are you doing it?
1 project | news.ycombinator.com | 14 Aug 2023

ollama

Posts with mentions or reviews of ollama. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-05-07.

Ollama v0.1.34 Is Out
1 project | news.ycombinator.com | 8 May 2024
Ask HN: What do you use local LLMs for?
2 projects | news.ycombinator.com | 7 May 2024

- Basic internet search (I start ollama CLI faster than I can start a browser - https://ollama.com)
- Formatting/changing text
- Troubleshooting code, esp. new frameworks/libs
- Recipes
- Data entry
- Organizing thoughts: High-level lists, comparison, classification, synonyms, jargon & nomenclature
- Learning esp. by analogy and example
RAG for:
- Website assistants (https://github.com/bennyschmidt/ragdoll-studio/tree/master/e...)
- Game NPCs (https://github.com/bennyschmidt/ragdoll-studio/tree/master/e...)
- Discord/Slack/forum bots (https://github.com/bennyschmidt/ragdoll-studio/tree/master/e...)
- Character-driven storytelling and creating art in a specific style for video game loading screens, background images, avatars, website art, etc. (https://github.com/bennyschmidt/ragdoll-studio/tree/master/r...)
FLaNK-AIM Weekly 06 May 2024
45 projects | dev.to | 6 May 2024
Introducing Jan
4 projects | dev.to | 5 May 2024

Jan goes a step further by integrating with other local engines like LM Studio and ollama.
Ollama v0.1.33
1 project | news.ycombinator.com | 3 May 2024
Hindi-Language AI Chatbot for Enterprises Using Qdrant, MLFlow, and LangChain
5 projects | dev.to | 2 May 2024

# install the Ollama curl -fsSL https://ollama.com/install.sh | sh # get the llama3 model ollama pull llama2 # install the MLFlow pip install mlflow
Create an AI prototyping environment using Jupyter Lab IDE with Typescript, LangChain.js and Ollama for rapid AI prototyping
4 projects | dev.to | 2 May 2024

Ollama for running LLMs locally
Setup Llama 3 using Ollama and Open-WebUI
1 project | dev.to | 29 Apr 2024

curl -fsSL https://ollama.com/install.sh | sh
Ollama v0.1.33 with Llama 3, Phi 3, and Qwen 110B
11 projects | news.ycombinator.com | 28 Apr 2024

Streaming is not a problem (it's just a simple flag: https://github.com/wiktor-k/llama-chat/blob/main/index.ts#L2...) but I've never used voice input.
The examples show image input though: https://github.com/ollama/ollama/blob/main/docs/api.md#reque...
Maybe you can file an issue here: https://github.com/ollama/ollama/issues
I Said Goodbye to ChatGPT and Hello to Llama 3 on Open WebUI - You Should Too
2 projects | dev.to | 24 Apr 2024

I’m a huge fan of open source models, especially the newly release Llama 3. Because of the performance of both the large 70B Llama 3 model as well as the smaller and self-host-able 8B Llama 3, I’ve actually cancelled my ChatGPT subscription in favor of Open WebUI, a self-hostable ChatGPT-like UI that allows you to use Ollama and other AI providers while keeping your chat history, prompts, and other data locally on any computer you control.

What are some alternatives?

When comparing mlc-llm and ollama you can also consider the following projects:

llama.cpp - LLM inference in C/C++

ggml - Tensor library for machine learning

gpt4all - gpt4all: run open-source LLMs anywhere

tvm - Open deep learning compiler stack for cpu, gpu and specialized accelerators

text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.

private-gpt - Interact with your documents using the power of GPT, 100% privately, no data leaks

llama-cpp-python - Python bindings for llama.cpp

llama - Inference code for Llama models

FastChat - An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.

LocalAI - :robot: The free, Open Source OpenAI alternative. Self-hosted, community-driven and local-first. Drop-in replacement for OpenAI running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. It allows to generate Text, Audio, Video, Images. Also with voice cloning capabilities.

mlc-llm vs llama.cpp ollama vs llama.cpp mlc-llm vs ggml ollama vs gpt4all mlc-llm vs tvm ollama vs text-generation-webui mlc-llm vs text-generation-webui ollama vs private-gpt mlc-llm vs llama-cpp-python ollama vs llama mlc-llm vs FastChat ollama vs LocalAI

Scout Monitoring - Free Django app performance insights with Scout Monitoring

Get Scout setup in minutes, and let us sweat the small stuff. A couple lines in settings.py is all you need to start monitoring your apps. Sign up for our free tier today.

www.scoutapm.com

featured

InfluxDB - Power Real-Time Data Analytics at Scale

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured

Compare mlc-llm vs ollama and see what are their differences.

mlc-llm

ollama

mlc-llm

ollama

What are some alternatives?