frost-beta/llama2-high-level-cpp
Inference Llama2 with High-Level C++.
14
/ 100
Experimental
No commits in the last 6 months.
Stale 6m
No Package
No Dependents
Maintenance
0 / 25
Adoption
5 / 25
Maturity
9 / 25
Community
0 / 25
Stars
11
Forks
—
Language
C
License
MIT
Category
Last pushed
Feb 22, 2024
Commits (30d)
0
Get this data via API
curl "https://pt-edge.onrender.com/api/v1/quality/transformers/frost-beta/llama2-high-level-cpp"
Open to everyone — 100 requests/day, no key needed. Get a free key for 1,000/day.
Higher-rated alternatives
beehive-lab/GPULlama3.java
GPU-accelerated Llama3.java inference in pure Java using TornadoVM.
54
srgtuszy/llama-cpp-swift
Swift bindings for llama-cpp library
44
gitkaz/mlx_gguf_server
This is a FastAPI based LLM server. Load multiple LLM models (MLX or llama.cpp) simultaneously...
43
JackZeng0208/llama.cpp-android-tutorial
llama.cpp tutorial on Android phone
40
dougeeai/llama-cpp-python-wheels
Pre-built wheels for llama-cpp-python across platforms and CUDA versions
34