Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > linux.debian.bugs.dist > #1231967

Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++

From Petter Reinholdtsen <pere@hungry.com>
Newsgroups linux.debian.bugs.dist
Subject Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++
Date 2025-02-06 09:30 +0100
Message-ID <Kd5qN-eMwX-9@gated-at.bofh.it> (permalink)
References (19 earlier) <KcY5X-eHYo-3@gated-at.bofh.it> <Kd5qN-eMwX-11@gated-at.bofh.it> <Kd5qN-eMwX-13@gated-at.bofh.it> <I6Vyx-9PtM-15@gated-at.bofh.it> <Kd5qN-eMwX-15@gated-at.bofh.it>
Organization linux.* mail to news gateway

Show all headers | View raw


[Christian Kastner]
> Look fine, though I deliberately skipped the poetry dependency for now
> as it looked more like a false positive.

Aha.  I just trusted lintian-brush on this one, did not investigate.

> This was my intention, but I initially wasn't sure what the default
> would be (-cpu or -blas). Looks like I forgot to add one before
> upload.

Given that every machine it can be installed on got a CPU, but not all
of them got a supported GPU, I beieve -cpu is the most sensible default.

> It'll be re-enabled soon. The were a few generated and minified files
> in that example, so I just opted to skip those for now, and focus on
> the build process.

Great to hear. :)

> Seeing as how closely llama.cpp and whisper.cpp are related, in the
> ideal case, you should be able to just carry over some patches, and
> mostly just copy d/rules, as llama.cpp and whisper.cpp share the ggml
> library on a source basis.

I hope so too, but I guess we will soon find out.  My initial draft on
<URL: https://salsa.debian.org/deeplearning-team/whisper.cpp > will need
a lot of updates  to bring it in line with this new approach. :)

-- 
Happy hacking
Petter Reinholdtsen

Back to linux.debian.bugs.dist | Previous | Next — Previous in thread | Find similar | Unroll thread


Thread

Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++ Petter Reinholdtsen <pere@hungry.com> - 2025-02-05 22:00 +0100
  Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++ Petter Reinholdtsen <pere@hungry.com> - 2025-02-06 01:40 +0100
    Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++ "M. Zhou" <lumin@debian.org> - 2025-02-06 02:50 +0100
      Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++ Petter Reinholdtsen <pere@hungry.com> - 2025-02-06 09:00 +0100
      Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++ "M. Zhou" <lumin@debian.org> - 2025-02-06 15:20 +0100
    Bug#1063673: ITP: llama.cpp -- Inference of Meta's LLaMA model (and others) in pure C/C++ Petter Reinholdtsen <pere@hungry.com> - 2025-02-06 09:30 +0100

csiph-web