mwmbl


github.com ggerganov llama.cpp
Mwmbl
GitHub - ggerganov/llama.cpp: Port of Facebook's LLaMA model in C/C++

Saved searches Use saved searches to filter your results more quickly You signed in with another tab or window. Reload to refresh your session.You signed …

en.wikipedia.org wiki Llama.cpp
Wikipedia
Llama.cpp

llama.cpp is an open-source software library that performs inference on various large language models such as Meta's Llama model. It is co-developed alongside

www.autodidacts.io tag llama-cpp
Mwmbl
Llama.cpp - The Autodidacts

Subscribe to this tag’s RSS feed to get new Llama.cpp posts as they’re written.

llama-cpp.com
Mwmbl
Llama.cpp - Run LLM Inference in C/C++

Llama.cpp – Run LLM Inference in C/C++ Llama.cpp (LLaMA C++) allows you to run efficient Large Language Model Inference in pure C/C++. You can run any po…

jacquesmattheij.com
Mwmbl
Jacques Mattheij

The llama.cpp software suite is a very impressive piece of work. It is a key element in some of the stuff that I’m playing around with on my home systems,…

freenode.net digest 405
Mwmbl
llama.cpp WebUI injection and PyTorch decode kernels · freenode

llama.cpp WebUI injection and PyTorch decode kernels Security and kernel work led the day in open AI infrastructure. A llama.cpp WebUI injection report s…

huggingface.co spaces Elijahbodden llama.cpp
Mwmbl
Llama.cpp - a Hugging Face Space by Elijahbodden

news.ycombinator.com item?id=36605596
Mwmbl
My POV is that llama.cpp is primarily a playground for adding ne…

My POV is that llama.cpp is primarily a playground for adding new features to the core ggml library and in the long run an interface for efficient LLM inf…

github.com coolbutuseless rllama
Mwmbl
GitHub - coolbutuseless/rllama: Minimal R wrapper for llama.cpp

Saved searches Use saved searches to filter your results more quickly You signed in with another tab or window. Reload to refresh your session.You signed…

github.com edp1096 my-llama
Mwmbl
GitHub - edp1096/my-llama: Go bindings for llama.cpp and webui

Saved searches Use saved searches to filter your results more quickly You signed in with another tab or window. Reload to refresh your session.You signed…

asciinema.org a 594301
Mwmbl
Pacha - TUI for llama.cpp - asciinema.org

developer.nvidia.com blog tag llama-cpp
Mwmbl
Tag: Llama.cpp | NVIDIA Technical Blog

llama-cpp.com download
Mwmbl
Llama.cpp Download

Llama.cpp (LLaMA C++) Download Llama.cpp (LLaMA C++) is a lightweight, high-performance implementation designed to run large language models locally on y…

github.com Bip-Rep sherpa
Mwmbl
GitHub - Bip-Rep/sherpa: A mobile Implementation of llama.cpp

Saved searches Use saved searches to filter your results more quickly You signed in with another tab or window. Reload to refresh your session.You signed…

freenode.net tag llama.cpp
Mwmbl
llama.cpp · freenode

Read Company Legal Product, project, and company names mentioned on this site are trademarks of their respective owners, and are referenced solely for pu…

github.com dsd sherpa
Mwmbl
GitHub - dsd/sherpa: A mobile Implementation of llama.cpp

You signed in with another tab or window. Reload to refresh your session. You signed out in another tab or window. Reload to refresh your session. You swi…

en.wikipedia.org wiki llama.cpp
Mwmbl
llama.cpp - Wikipedia

llama.cpp began development in March 2023 by Georgi Gerganov as an implementation of the Llama inference code in pure C/C++ with no dependencies. This imp…

news.ycombinator.com item?id=48546182
Mwmbl
+1 using llama.cpp Vulkan releases with the Qwen models - runs m…

gist.github.com codeprimate bb244d3d05da5eec4e726fc75a685546
Mwmbl
Tmux AI llama.cpp integration · GitHub

You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You swit…

news.ycombinator.com item?id=37275259
Mwmbl
From my experience llama.cpp doesn’t take full advantage of para…

From my experience llama.cpp doesn’t take full advantage of parallelism as it could. Tested this on an HPC cluster - increasing thread count certainly did…

news.ycombinator.com item?id=40506102
Mwmbl
Like InternVL, no llama.cpp support severely limits its applicat…

Like InternVL, no llama.cpp support severely limits its applications. Close to GPT4v performance level and runnable locally on any machine (no need for a …

stackoverflow.com a 78509754
Mwmbl
large language model - Llama.cpp GPU Offloading Issue - Unexpect…

I'm reaching out to the community for some assistance with an issue I'm encountering in llama.cpp. Previously, the program was successfully utilizing the…

news.ycombinator.com item?id=43901854
Mwmbl
Basically anything llama.cpp (Vulkan backend) should work out of…

fly.io phoenix-files using-llama-cpp-with-elixir-and-rustler
Mwmbl
Using LLama.cpp with Elixir and Rustler · The Phoenix Files

Using LLama.cpp with Elixir and Rustler We’re Fly.io. We run apps for our users on hardware we host around the world. Fly.io happens to be a great place …

news.ycombinator.com item?id=35393284
Mwmbl
Llama.cpp 30B runs with only 6GB of RAM now | Hacker News

Author here. For additional context, please read https://github.com/ggerganov/llama.cpp/discussions/638#discu... The loading time performance has been a …

lwn.net Articles 973690
Mwmbl
Portable LLMs with llamafile [LWN.net]

I mean, llama.cpp is still a pretty young project where the codebase changes rapidly and these kinds of changes are what defines how the codebase is going…

en.wikipedia.org wiki Llama_(language_model)
Wikipedia
Llama (language model)

developer Georgi Gerganov released llama.cpp as open-source on March 10, 2023. It's a re-implementation of Llama in C++, allowing systems without a powerful

docs.rs llama-cpp-2
Mwmbl
llama_cpp_2 - Rust

As llama.cpp is a very fast moving target, this crate does not attempt to create a stable API with all the rust idioms. Instead it provided safe wrappers…

news.ycombinator.com item?id=36806604
Mwmbl
The llama.cpp project is absolutely amazing. Our goal was to bui…

2. Run 2+ models: loading and unloading models as users need them, including via a REST API. Lots to do here, but even small models are memory hogs and t…

news.ycombinator.com item?id=42860666
Mwmbl
llama.cpp optimises for hackability, not necessarily maintainabi…

libhunt.com r llama.cpp
Mwmbl
Llama.cpp Alternatives and Reviews (Feb 2024)

What’s up with the C++ ecosystem in 2023? JetBrains Developer Ecosystem Survey 2023 has given us many interesting insights. The Embedded (37%) and Games …

news.ycombinator.com item?id=44002827
Mwmbl
llama.cpp clearly does not support iSWA: https://github.com/ggml…

news.ycombinator.com item?id=37140013
Mwmbl
How Is LLaMa.cpp Possible? | Hacker News

Essentially, you lose some accuracy and there might be some weird answers and probably more likely to go off the rail and hallucinate. But the quality lo…

gist.github.com chiragjn 22e6a3ffe1b7f4aeaaefbc25af8e9461
Mwmbl
llama.cpp python cuda Dockerfile example · GitHub

You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You swit…

github.com ollama ollama pull 2868
Mwmbl
Update llama.cpp submodule to `c29af7e` by jmorganca · Pull Requ…

Saved searches Use saved searches to filter your results more quickly You signed in with another tab or window. Reload to refresh your session.You signed…

en.wikipedia.org wiki GGUF
Wikipedia
GGUF

saving and loading of model data. It was introduced in August 2023 by the llama.cpp project to better maintain backward compatibility as support was added

asciinema.org a 592303
Mwmbl
llama.cpp - Training from scratch works good! - asciinema.org

en.wikipedia.org wiki Lists_of_open-source_artificial_intelligence_software
Wikipedia
Lists of open-source artificial intelligence software

development and agentic coding features. llama.cpp — library that can perform inference on various LLMs such as Llama, Mistral, Gemma, DeepSeek or Qwen. SGLang

en.wikipedia.org wiki Ollama
Wikipedia
Ollama

libraries and documentation for running supported models. Ollama uses the llama.cpp backend for local model inference, which supports running of quantized


Join the Mwmbl community

Mwmbl is powered by you — our community. We're a friendly bunch and you can find us on Matrix and Discord.

Join the community now!

© Mwmbl 2026, under the AGPL-3.0 license