omgitsbase/llmash
llmash a replacement for ollama. llmash uses advanced kernel optimizations to bring it to speeds far past vllm, on any model. Moving from ollama to llmash is nearly automatic, and the speed improvements are nearly (7% under) linear to power consumption improvements.
C++
Stars
11
Forks
0
License
MIT
Open Issues
1
Updated
1d ago
No README available
Related Projects
affaan-m/ECC
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
260.3k
JavaScript
NousResearch/hermes-agent
The agent that grows with you
246.2k
Python
Significant-Gravitas/AutoGPT
AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
187.4k
Python
firecrawl/firecrawl
The web data API to search, scrape, and interact at scale. 🔥
181.3k
TypeScript
ollama/ollama
Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
181.2k
Go
f/prompts.chat
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
170.5k
HTML
Statistics
- Stars
- 11
- Forks
- 0
- Open Issues
- 1
- Created
- Sep 06, 2026
- Updated
- Sep 16, 2026