Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
C++mozilla-ai/llamafile

llamafile

Distribute and run LLMs with a single file.

88.6/100
25.4KForks: 1.5K
View on GitHubHomepage →
Loading report...

Similar Projects

RCLI

59

Talk to your Mac, query your docs, no cloud required. On-device voice AI + RAG

C++1.5K

lucebox

75

Fast LLM speculative inference server for consumer hardware.

C++2.7K

whisper.cpp

81

Port of OpenAI's Whisper model in C/C++

C++52.3K

PowerInfer

58

High-speed Large Language Model Serving for Local Deployment

C++9.7K
Back to List