Llama Cpp

Llama Cpp에 대해 3개 실제 데이터 소스에서 20개 공개 신호를 찾았습니다. GitHub 10건, Hacker News 0건, arXiv 10건을 원문 링크와 함께 보여줍니다.

관심도 점수

100

검색 slug
llama-cpp
실제 연결
Hacker News, GitHub, arXiv
키워드
the, and, llama, for, cpp, with, model, that
마지막 수집
2026-07-25T11:46:11.396Z

출처별 탭

ggml-org/llama.cpp

LLM inference in C/C++

github · 2023-03-10 · score 121,545 · comments 1,910 · stars 121,545 · forks 20,975

h2oai/h2ogpt

Private chat with local GPT with document, images, video, etc. 100% private, Apache 2.0. Supports oLLaMa, Mixtral, llama.cpp, and more. Demo: https://gpt.h2o.ai/ https://gpt-docs.h2o.ai/

github · 2023-03-24 · score 11,981 · comments 329 · stars 11,981 · forks 1,302

abetlen/llama-cpp-python

Python bindings for llama.cpp

github · 2023-03-23 · score 10,512 · comments 680 · stars 10,512 · forks 1,438

intel/ipex-llm

Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discrete GPU such as Arc, Flex and Max); seamlessly integrate with llama.cpp, Ollama, HuggingFace, LangChain, LlamaIndex, vLLM, DeepSpeed, Axolotl, etc.

github · 2016-08-29 · score 8,869 · comments 1,484 · stars 8,869 · forks 1,430

LearningCircuit/local-deep-research

~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.

github · 2025-02-09 · score 8,776 · comments 283 · stars 8,776 · forks 770

PawanOsman/OpenCursor

Open-source Cursor-like AI coding agent for VS Code — agentic chat, multi-provider LLMs (OpenAI, Ollama, llama.cpp), semantic search, and MCP support

github · 2022-12-06 · score 5,957 · comments 3 · stars 5,957 · forks 1,016

serge-chat/serge

A web interface for chatting with Alpaca through llama.cpp. Fully dockerized, with an easy to use API.

github · 2023-03-19 · score 5,721 · comments 33 · stars 5,721 · forks 390

Michael-A-Kuykendall/shimmy

⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.

github · 2025-08-28 · score 5,697 · comments 11 · stars 5,697 · forks 548

ngxson/smolvlm-realtime-webcam

Real-time webcam demo with SmolVLM and llama.cpp server

github · 2025-05-12 · score 5,563 · comments 19 · stars 5,563 · forks 895

mostlygeek/llama-swap

Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc

github · 2024-10-04 · score 5,134 · comments 73 · stars 5,134 · forks 389