Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > comp.lang.prolog > #15710 > unrolled thread
| Started by | Mild Shock <janburse@fastmail.fm> |
|---|---|
| First post | 2026-07-22 21:01 +0200 |
| Last post | 2026-07-29 17:04 +0200 |
| Articles | 17 on this page of 57 — 4 participants |
Back to article view | Back to comp.lang.prolog
The Wuhan Virus that destroyed Python [ggml Manifesto] Mild Shock <janburse@fastmail.fm> - 2026-07-22 21:01 +0200
Deadlock Exorcism: Switch from Push to Pull [A pi-calculus Specification of Prolog] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 00:25 +0200
Why do you even need a mpmc queue? [Thunder Kittens] (Re: Deadlock Exorcism: Switch from Push to Pull) Mild Shock <janburse@fastmail.fm> - 2026-07-23 08:45 +0200
Trivial balancing example for (int i=0; i<global_id; i++) (Re: Why do you even need a mpmc queue? [Thunder Kittens]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 08:55 +0200
The Pixel Phone AI Experiment Song (Enqueue/dequeue need not be fast and can spinn ["fairness" questions]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 09:19 +0200
Enqueue/dequeue need not be fast and can spinn ["fairness" questions] (Re: The Pixel Phone AI Experiment Song (Enqueue/dequeue need not be fast and can spinn ["fairness" questions]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 09:23 +0200
And, where did I talk about rockets? [Hint its about xAI's Grok] (Re: Why do you even need a mpmc queue? [Thunder Kittens]) Mild Shock <janburse@fastmail.fm> - 2026-07-25 01:25 +0200
Why forget something, that was never on my mind (Re: And, where did I talk about rockets? [Hint its about xAI's Grok]) Mild Shock <janburse@fastmail.fm> - 2026-07-25 09:49 +0200
Example Mandel Brot rendering [Faster with MIMD] (Was: Why do you even need a mpmc queue? [Thunder Kittens]) Mild Shock <janburse@fastmail.fm> - 2026-07-25 09:56 +0200
Potential Python Recovery: Free Threading [3.13 release] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 10:21 +0200
The things XILINX braught to the AMD table (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 18:48 +0200
NVIDIA evacuated its Chinese market [Tau Scaling] (Re: The things XILINX braught to the AMD table) Mild Shock <janburse@fastmail.fm> - 2026-07-23 19:13 +0200
Micro penis mother sung arias (Re: NVIDIA evacuated its Chinese market [Tau Scaling]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 14:40 +0200
Micro penis brain is in constant hiatus (Re: Micro penis mother sung arias) Mild Shock <janburse@fastmail.fm> - 2026-07-24 15:27 +0200
Ignoramus or Ignorabimus: I don't care [(Re: Micro penis brain is in constant hiatus (Re: Micro penis mother sung arias) Mild Shock <janburse@fastmail.fm> - 2026-07-24 15:35 +0200
You are a moron, brainless putin payed (Re: Ignoramus or Ignorabimus: I don't care) Mild Shock <janburse@fastmail.fm> - 2026-07-24 18:01 +0200
Yeah keep reading my posts, uninspired fool (Re: You are a moron, brainless putin payed) Mild Shock <janburse@fastmail.fm> - 2026-07-24 19:47 +0200
Out of the blue accusation span 15 days [Empirical USENET study] (Re: Ignoramus or Ignorabimus: I don't care) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:27 +0200
A brain desease of 20 days [Rossy Boy] (Re: Ignoramus or Ignorabimus: I don't care) Mild Shock <janburse@fastmail.fm> - 2026-07-29 18:41 +0200
I didn't use a Ryzen Halo, whats wrong with you? (Re: A brain desease of 20 days [Rossy Boy]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 23:25 +0200
Re: NVIDIA evacuated its Chinese market [Tau Scaling] (Re: The things XILINX braught to the AMD table) Mild Shock <janburse@fastmail.fm> - 2026-07-28 14:17 +0200
ASML stocks are plunging, bye bye dutchies (Re: NVIDIA evacuated its Chinese market [Tau Scaling]) Mild Shock <janburse@fastmail.fm> - 2026-07-28 14:18 +0200
Little Data Center on Your Palm [AI Laptops for 500 USD] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 17:59 +0200
2008: 4 Blades + Tesla S1070 versus 2026: 1 AI Laptop (Re: Little Data Center on Your Palm [AI Laptops for 500 USD]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 18:15 +0200
Hurry the blue bus doesnt stop indefinitely (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:37 +0200
Not SIMD, a MIMD design for NVIDIA Volta (Re: Hurry the blue bus doesnt stop indefinitely) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:58 +0200
Could take 3-4 months find machine / browser (Re: Not SIMD, a MIMD design for NVIDIA Volta) Mild Shock <janburse@fastmail.fm> - 2026-07-24 21:16 +0200
The Koan of pi-WAM queues [FORTRAN-S] (Re: Not SIMD, a MIMD design for NVIDIA Volta) Mild Shock <janburse@fastmail.fm> - 2026-07-26 19:54 +0200
The turbo capping of AI Laptops (Was: The Koan of pi-WAM queues [FORTRAN-S]) Mild Shock <janburse@fastmail.fm> - 2026-07-26 20:00 +0200
Re: The Koan of pi-WAM queues [FORTRAN-S] (Re: Not SIMD, a MIMD design for NVIDIA Volta) Mild Shock <janburse@fastmail.fm> - 2026-07-27 09:16 +0200
Why forget Bulgarians, never on my mind (Re: The Koan of pi-WAM queues [FORTRAN-S]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 09:16 +0200
miniTriton CUDA is an alternative to torch variants (Re: Why forget Bulgarians, never on my mind) Mild Shock <janburse@fastmail.fm> - 2026-07-27 09:52 +0200
Andrej Karpathy original gangster of Budget Laptop (Re: miniTriton CUDA is an alternative to torch variants) Mild Shock <janburse@fastmail.fm> - 2026-07-27 09:54 +0200
The evolution of hardware and GPT-2 training (Re: Why forget Bulgarians, never on my mind) Mild Shock <janburse@fastmail.fm> - 2026-07-27 10:57 +0200
How speed up π-WAM with vector operations (Re: The evolution of hardware and GPT-2 training) Mild Shock <janburse@fastmail.fm> - 2026-07-27 11:10 +0200
AI accelerator extend from GPU to CPU [Zero Copying] (Re: How speed up π-WAM with vector operations) Mild Shock <janburse@fastmail.fm> - 2026-07-27 13:21 +0200
The invention of vector and matrix registers [NVIDIA Volta] (Re: AI accelerator extend from GPU to CPU [Zero Copying]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 13:22 +0200
Re: The invention of vector and matrix registers [NVIDIA Volta] (Re: AI accelerator extend from GPU to CPU [Zero Copying]) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-27 07:34 -0700
Maybe they should have named it NVIDIA Einstein [Rossy Boy Toe Sucking] (Was: The invention of vector and matrix registers [NVIDIA Volta]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 17:12 +0200
Re: The invention of vector and matrix registers [NVIDIA Volta] (Re: AI accelerator extend from GPU to CPU [Zero Copying]) R Kym Horsell <kym@sdf.com> - 2026-07-27 15:43 +0000
Re: The invention of vector and matrix registers [NVIDIA Volta] (Re: AI accelerator extend from GPU to CPU [Zero Copying]) R Kym Horsell <kymhorsell@gmail.com> - 2026-07-27 15:46 +0000
π-WAM is not adding decimals, it is removing decimals (Was: The invention of vector and matrix registers [NVIDIA Volta]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 18:34 +0200
In Budget Laptops the TOPS come with low energy footprint (Re: π-WAM is not adding decimals, it is removing decimals) Mild Shock <janburse@fastmail.fm> - 2026-07-27 18:45 +0200
Potato Computer owner impressed by Ukraine Tech [Rossy Boys Brother?] (Was: Hurry the blue bus doesnt stop indefinitely) Mild Shock <janburse@fastmail.fm> - 2026-07-27 16:56 +0200
Rossy Boy is neither Einstein nor Zweistein (Was: Potato Computer owner impressed by Ukraine Tech) Mild Shock <janburse@fastmail.fm> - 2026-07-27 18:25 +0200
You are still chewing on SIMD. LoL (Re: Rossy Boy is neither Einstein nor Zweistein) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:12 +0200
Hurry Rossy Boy, the blue bus is waiting (Re: You are still chewing on SIMD. LoL) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:53 +0200
Look how they advertized CUDA and logical threads (Re: Hurry Rossy Boy, the blue bus is waiting) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:55 +0200
Forget any arithmetization of product FSA (Re: Look how they advertized CUDA and logical threads) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:58 +0200
Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator] (Re: You are still chewing on SIMD. LoL) Mild Shock <janburse@fastmail.fm> - 2026-07-29 20:06 +0200
I don't use Rust, you are crazy [Jump off a bridge, idiot] (Re: Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 20:25 +0200
Hack ecosystem ignorance paired with paranoia [Nand to Tetris] (Re: I don't use Rust, you are crazy) Mild Shock <janburse@fastmail.fm> - 2026-07-29 22:52 +0200
A funny Q16.16 experiment with Hack (Re: Hack ecosystem ignorance paired with paranoia [Nand to Tetris]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 23:11 +0200
Got it. Or are you too stupid? [New Usenet Mantra] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 18:59 +0200
Lamas in a cradle and Lamas on the edge [Red Pyjama] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 13:03 +0200
AI Accelerators and ISO Prolog multi-threading (Re: Lamas in a cradle and Lamas on the edge [Red Pyjama]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:02 +0200
Actor/Erlang is dead, no Thread and Mailbox conflation [golang channels] (Re: AI Accelerators and ISO Prolog multi-threading) (Re: Lamas in a cradle and Lamas on the edge [Red Pyjama]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:04 +0200
Page 3 of 3 — ← Prev page 1 2 [3]
| From | R Kym Horsell <kymhorsell@gmail.com> |
|---|---|
| Date | 2026-07-27 15:46 +0000 |
| Subject | Re: The invention of vector and matrix registers [NVIDIA Volta] (Re: AI accelerator extend from GPU to CPU [Zero Copying]) |
| Message-ID | <1147uh4$2542$2@nnrp.usenet.blueworldhosting.com> |
| In reply to | #15763 |
In comp.lang.prolog R Kym Horsell <kym@sdf.com> wrote:
> generalizes better than it did before training and more importantly
> it takes maybe an order of magnitude crunching to produce a good answer
/\ less
> than the usual over-fit answer.
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-27 18:34 +0200 |
| Subject | π-WAM is not adding decimals, it is removing decimals (Was: The invention of vector and matrix registers [NVIDIA Volta]) |
| Message-ID | <11481af$h2se$1@solani.org> |
| In reply to | #15763 |
Hi, Come on Horsy Boy, you can do better. I no where wrote something about curve fitting and/or increasing the precision of float point numbers: 11.4 Giga Lips with a Budget Laptop https://github.com/Jean-Luc-Picard-2021/gigabudget What makes you think LIPS measures precision? You should know better as a 50% Prologer. I explictily wrote here what the goal is: "shave off some of the TOPS to do Prolog inferencing" What are TOPS? Its a metric for GPUs: TOPS stands for “Trillions of Operations Per Second.” https://www.lenovo.com/us/en/glossary/tops-in-computing/ See for yourself what is behind my post: 11.4 Giga Lips with a Budget Laptop At the end of 2025 we acquired a couple of AI Laptops , that were still cheap, since RAM prices had not yet rocketed. The intend was to tap into the Copilot+ certified hardware, and shave off some of the TOPS to do Prolog inferencing. Amazingly our π-WAM can churn 11.4 GIGA LIPS. GPUs have evolved form lock-step to independent thread scheduling. This made it possible to port the Hack VM variant, that forms the basis for our π-WAM, to WebGPU computer shaders. Using NUM_SHADERS = 4096 we could produce 11.4 Giga Lips on a Ryzen AI 7 350 w/ Radeon 860M. See also: Medium Article - 11.4 Giga Lips https://medium.com/2989/899b0d5c027b So just get lost with your crazy irrelevant rant. When I get more LIPS, things run faster, and I remove digits from the time dimension. Got it. Or are you too stupid? Bye R Kym Horsell schrieb: > In comp.lang.prolog Ross Finlayson <ross.a.finlayson@gmail.com> wrote: > ... >> Data centers should pay a 10000% excise on electricity, >> wherever it comes from, a natural regulator of inverted economies. >> And by ten thousand percent I really mean a ten thousand percent. > > And what would a huge surcharge do? > Almost always end up affecting the less powerful end of society > with increased costs to services the AI industry will be doing > more and more of over time. > > I started a little data center (exaflops.com) many years ago. > In those distant days people (in fact one was a prof of computer > science) told me you could never make money running a supercomputer. > LOL. :) > > I've had many years to watch the trends and a far more efficient > way to solve resource problems in this area is to change the > algorithms. There is vast room for improvement, mostly because > of prevailing attitudes. > > I used to do competetion data science as a sideline. Companies > would pay almost any price to get an extra decimal place in > the accuracy of their forecasting processes. But typically > they were trying to supercharge a system that should be scrapped > and re-designed from scratch. One area I'm thinking of is > investment. I had a customer one time -- like many times -- > ask to improve a system that predicted the future price of > various stocks. The idea (for them) was to have as accurate a > prediction of what some stock would be worth in a week or a month's > time so that some moron could use the information to decide when > to buy or sell the thing. > > I tried to argue the efficient thing was to create a system that > takes the human out of the loop altogether. It doesnt provide info > for someone to decide whether or not to follow the advice -- > that is just introducing more noise into the loop and probably > cancels any benefit of adding a couple decimal places of precision. > What you *should* do is make a system that is tuned to robustly > maximize the profit from managing a portfolio. > > Of course they wouldnt come at that. You can't suggest taking the > managers out of the loop. :) > > Another idea relevant to current AI methods might be to curtail > use of typical neural net algorithms. Many of them try to squeeze > the best performance of some NN during the training phase in > the hope the resulting system will generalize well enough to be useful > on new data. But there's kind-of a law that the harder you train > some system to perform a task well, the less well they can subsuently > perform a more general version of the same thing. It's amusing when > you look at the graphs of NN being trained and then tested that > given a more general problem to solve after being trained to solve > similar problems very very well the poor old NN does worse that it > would have done if it had 0 training in the first place. > > It's not like we dont know how to improve this kind of performance. > Try less hard in the training phase or make it "more noisy". > Turns out genetic methods are just the ticket for this. > The training produces less over-fitting and the resulting system > generalizes better than it did before training and more importantly > it takes maybe an order of magnitude crunching to produce a good answer > than the usual over-fit answer. > > Anyway. Have to go and feed the cat. >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-27 18:45 +0200 |
| Subject | In Budget Laptops the TOPS come with low energy footprint (Re: π-WAM is not adding decimals, it is removing decimals) |
| Message-ID | <11481vf$h3a6$2@solani.org> |
| In reply to | #15766 |
Hi, Because of the mobile GPU design, the TOPS, aka “Trillions of Operations Per Second.” come with not extremly high power consumption. Especially the presence of vector and matrix operations can lower the energy consumption, since they can avoid redundant memory access. Its quite a difference between discrete graphic cards and accelerator iGPUs that are directly on the silicon chip, and have mobile design. So basically with newer AI Laptops you get more performence units for less energy units. Have Fun! Bye Mild Shock schrieb: > Hi, > > Come on Horsy Boy, you can do better. I > no where wrote something about curve > fitting and/or increasing the precision of > > float point numbers: > > 11.4 Giga Lips with a Budget Laptop > https://github.com/Jean-Luc-Picard-2021/gigabudget > > What makes you think LIPS measures precision? > You should know better as a 50% Prologer. > > I explictily wrote here what the goal is: > > "shave off some of the TOPS to do Prolog inferencing" > > What are TOPS? Its a metric for GPUs: > > TOPS stands for “Trillions of Operations Per Second.” > https://www.lenovo.com/us/en/glossary/tops-in-computing/ > > See for yourself what is behind my post: > > 11.4 Giga Lips with a Budget Laptop > At the end of 2025 we acquired a couple of AI Laptops , that were still > cheap, since RAM prices had not yet rocketed. The intend was to tap into > the Copilot+ certified hardware, and shave off some of the TOPS to do > Prolog inferencing. Amazingly our π-WAM can churn 11.4 GIGA LIPS. > > GPUs have evolved form lock-step to independent thread scheduling. This > made it possible to port the Hack VM variant, that forms the basis for > our π-WAM, to WebGPU computer shaders. Using NUM_SHADERS = 4096 we could > produce 11.4 Giga Lips on a Ryzen AI 7 350 w/ Radeon 860M. > > See also: > > Medium Article - 11.4 Giga Lips > https://medium.com/2989/899b0d5c027b > > So just get lost with your crazy irrelevant rant. > When I get more LIPS, things run faster, and > I remove digits from the time dimension. > > Got it. Or are you too stupid? > > Bye > > R Kym Horsell schrieb: >> In comp.lang.prolog Ross Finlayson <ross.a.finlayson@gmail.com> wrote: >> ... >>> Data centers should pay a 10000% excise on electricity, >>> wherever it comes from, a natural regulator of inverted economies. >>> And by ten thousand percent I really mean a ten thousand percent. >> >> And what would a huge surcharge do? >> Almost always end up affecting the less powerful end of society >> with increased costs to services the AI industry will be doing >> more and more of over time. >> >> I started a little data center (exaflops.com) many years ago. >> In those distant days people (in fact one was a prof of computer >> science) told me you could never make money running a supercomputer. >> LOL. :) >> >> I've had many years to watch the trends and a far more efficient >> way to solve resource problems in this area is to change the >> algorithms. There is vast room for improvement, mostly because >> of prevailing attitudes. >> >> I used to do competetion data science as a sideline. Companies >> would pay almost any price to get an extra decimal place in >> the accuracy of their forecasting processes. But typically >> they were trying to supercharge a system that should be scrapped >> and re-designed from scratch. One area I'm thinking of is >> investment. I had a customer one time -- like many times -- >> ask to improve a system that predicted the future price of >> various stocks. The idea (for them) was to have as accurate a >> prediction of what some stock would be worth in a week or a month's >> time so that some moron could use the information to decide when >> to buy or sell the thing. >> >> I tried to argue the efficient thing was to create a system that >> takes the human out of the loop altogether. It doesnt provide info >> for someone to decide whether or not to follow the advice -- >> that is just introducing more noise into the loop and probably >> cancels any benefit of adding a couple decimal places of precision. >> What you *should* do is make a system that is tuned to robustly >> maximize the profit from managing a portfolio. >> >> Of course they wouldnt come at that. You can't suggest taking the >> managers out of the loop. :) >> >> Another idea relevant to current AI methods might be to curtail >> use of typical neural net algorithms. Many of them try to squeeze >> the best performance of some NN during the training phase in >> the hope the resulting system will generalize well enough to be useful >> on new data. But there's kind-of a law that the harder you train >> some system to perform a task well, the less well they can subsuently >> perform a more general version of the same thing. It's amusing when >> you look at the graphs of NN being trained and then tested that >> given a more general problem to solve after being trained to solve >> similar problems very very well the poor old NN does worse that it >> would have done if it had 0 training in the first place. >> >> It's not like we dont know how to improve this kind of performance. >> Try less hard in the training phase or make it "more noisy". >> Turns out genetic methods are just the ticket for this. >> The training produces less over-fitting and the resulting system >> generalizes better than it did before training and more importantly >> it takes maybe an order of magnitude crunching to produce a good answer >> than the usual over-fit answer. >> >> Anyway. Have to go and feed the cat. >> >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-27 16:56 +0200 |
| Subject | Potato Computer owner impressed by Ukraine Tech [Rossy Boys Brother?] (Was: Hurry the blue bus doesnt stop indefinitely) |
| Message-ID | <1147rj3$gdpk$1@solani.org> |
| In reply to | #15736 |
Hi,
Slowly I start understanding numbnuts like
Rossy Boy who don't understand tech, although
they are from UK and not from a 3rd world
country, and also I start understanding morons
like Micro Penis, who are behind a curtain,
and cannot access a lot of tech.
The same holds for SWI Prologs newest campaign
that probably adresses some poor indians that
have neither 5G nor Macs:
1:38:01 The Kyiv keynote disaster
https://www.youtube.com/watch?v=U8goS6B3BbI
Woa! Real time download of Scala, Closure,
etc.. Whats the magic behind that? Some SWI
point of sale, downloading it via its
keyboard and some telephathy module ?
Bye
Mild Shock schrieb:
> Hi,
>
> Ride the snake
> He's old and his skin is cold
> The west is the best
> The west is the best
> Get here and we'll do the rest
> The blue bus is calling us
> The blue bus is calling us
> Driver, where you taking us?
>
> Apocalypse Now intro: The Doors, The End {1979}
> https://www.youtube.com/watch?v=CIrvSJwwJUE
>
> Bye
>
> > Hi,
> >
> > Again I posted everything here:
> >
> >> 11.4 Giga Lips with a Budget Laptop
> >> https://github.com/Jean-Luc-Picard-2021/gigabudget
> >
> > The repo says, same time when I posted
> > the link first time:
> >
> >> This repository was archived by the
> >> owner on Jul 9, 2026. It is now read-only.
> >
> > Now a USENET user, who had already entitled
> > himself for a couple of irrational accusations
> >
> > towards my side, is asking this question:
> >
> > Chris M. Thomasson schrieb, Jul 24, 2026
> >> Show an outline of what you
> >> need you compute shader to do?
> >
> > Bravo, thats a delay of a wooping 15 days.
> >
> > Bye
>
> Mild Shock schrieb:
>> Hi,
>>
>> Remember when first all local AI was Python
>> and PyTorch APIs. And then suddently people started
>> using bare metal C/C++ Code. Here is the story:
>>
>> How it started:
>>
>> GPT-J or GPT-J-6B is an open-source large
>> language model (LLM) developed by EleutherAI
>> in 2021. As the name suggests, it is a
>> generative pre-trained transformer model
>> designed to produce human-like text that
>> continues from a prompt.
>> https://www.eleuther.ai/
>>
>> How it was going [Georgi Gerganov]:
>>
>> So a few days later comes out the LLaMA, I do
>> some calculations and I figure out “Okay, 65
>> billion parameters. You probably need about
>> 40 gigs of RAM, with 4-bit quantization. So
>> this can run on a MacBook. Why not do it?”
>>
>> Why I was able to do it so quickly - basically,
>> for all that I saw it’s pretty much GPT-J architecture
>> with some modifications, like some extra memorization
>> layers. It’s minor changes. Basically, again, the
>> existing code for the GPT-J, I just simply
>> modified it there, it happened pretty quickly.
>> https://changelog.com/podcast/532
>>
>> Georgi Gerganov, Bulgarian, now with Hugging
>> Face, ggml-cann also running on Chinese AI chips.
>> ggml Manifesto https://github.com/ggml-org/ggml
>>
>> Bye
>
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-27 18:25 +0200 |
| Subject | Rossy Boy is neither Einstein nor Zweistein (Was: Potato Computer owner impressed by Ukraine Tech) |
| Message-ID | <11480pn$h2e6$1@solani.org> |
| In reply to | #15758 |
Hi,
Rossy Boy is neither Einstein nor Zweistein.
He is not Einstein since Einstein is already dead:
Albert Einstein (1879 - 1955)
https://de.wikipedia.org/wiki/Albert_Einstein
He is also not Zweistein, since he doesn't
understand concepts such as:
- NVIDIA Volta ff. architecture
Also his hands are small, and his breath stinks,
and he lives in the basement of his mother.
Bye
Mild Shock schrieb:
> Hi,
>
> Slowly I start understanding numbnuts like
> Rossy Boy who don't understand tech, although
> they are from UK and not from a 3rd world
>
> country, and also I start understanding morons
> like Micro Penis, who are behind a curtain,
> and cannot access a lot of tech.
>
> The same holds for SWI Prologs newest campaign
> that probably adresses some poor indians that
> have neither 5G nor Macs:
>
> 1:38:01 The Kyiv keynote disaster
> https://www.youtube.com/watch?v=U8goS6B3BbI
>
> Woa! Real time download of Scala, Closure,
> etc.. Whats the magic behind that? Some SWI
> point of sale, downloading it via its
>
> keyboard and some telephathy module ?
>
> Bye
>
> Mild Shock schrieb:
>> Hi,
>>
>> Ride the snake
>> He's old and his skin is cold
>> The west is the best
>> The west is the best
>> Get here and we'll do the rest
>> The blue bus is calling us
>> The blue bus is calling us
>> Driver, where you taking us?
>>
>> Apocalypse Now intro: The Doors, The End {1979}
>> https://www.youtube.com/watch?v=CIrvSJwwJUE
>>
>> Bye
>>
>> > Hi,
>> >
>> > Again I posted everything here:
>> >
>> >> 11.4 Giga Lips with a Budget Laptop
>> >> https://github.com/Jean-Luc-Picard-2021/gigabudget
>> >
>> > The repo says, same time when I posted
>> > the link first time:
>> >
>> >> This repository was archived by the
>> >> owner on Jul 9, 2026. It is now read-only.
>> >
>> > Now a USENET user, who had already entitled
>> > himself for a couple of irrational accusations
>> >
>> > towards my side, is asking this question:
>> >
>> > Chris M. Thomasson schrieb, Jul 24, 2026
>> >> Show an outline of what you
>> >> need you compute shader to do?
>> >
>> > Bravo, thats a delay of a wooping 15 days.
>> >
>> > Bye
>>
>> Mild Shock schrieb:
>>> Hi,
>>>
>>> Remember when first all local AI was Python
>>> and PyTorch APIs. And then suddently people started
>>> using bare metal C/C++ Code. Here is the story:
>>>
>>> How it started:
>>>
>>> GPT-J or GPT-J-6B is an open-source large
>>> language model (LLM) developed by EleutherAI
>>> in 2021. As the name suggests, it is a
>>> generative pre-trained transformer model
>>> designed to produce human-like text that
>>> continues from a prompt.
>>> https://www.eleuther.ai/
>>>
>>> How it was going [Georgi Gerganov]:
>>>
>>> So a few days later comes out the LLaMA, I do
>>> some calculations and I figure out “Okay, 65
>>> billion parameters. You probably need about
>>> 40 gigs of RAM, with 4-bit quantization. So
>>> this can run on a MacBook. Why not do it?”
>>>
>>> Why I was able to do it so quickly - basically,
>>> for all that I saw it’s pretty much GPT-J architecture
>>> with some modifications, like some extra memorization
>>> layers. It’s minor changes. Basically, again, the
>>> existing code for the GPT-J, I just simply
>>> modified it there, it happened pretty quickly.
>>> https://changelog.com/podcast/532
>>>
>>> Georgi Gerganov, Bulgarian, now with Hugging
>>> Face, ggml-cann also running on Chinese AI chips.
>>> ggml Manifesto https://github.com/ggml-org/ggml
>>>
>>> Bye
>>
>
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:12 +0200 |
| Subject | You are still chewing on SIMD. LoL (Re: Rossy Boy is neither Einstein nor Zweistein) |
| Message-ID | <114d58q$kg61$2@solani.org> |
| In reply to | #15765 |
Hi,
You are still chewing on SIMD. LoL
Ross Finlayson schrieb:
> Then the idea is that any of those can be found and matched in
> one "run", i.e. a stall-less, branch-less, call-less list of less than
> a few or less than a few dozens or less than a few hundreds
> instructions, the results "findings" in data and corresponding
> "matchings" of expressions, that runs in less than one microsecond.
You cannot make the mental translation that if you have:
Ross Finlayson schrieb:
> So, the context then is for register state and stack contents, that
> the indicators of the above as "positive presence" then is to make
> for that the adjustments to the offsets and extents and the shifts
> is according to those, otherwise no-ops. Then the idea is that a
As independent logical thread state, that automatically MIMD follows?
Whats the problem to solve then?
Bye
Mild Shock schrieb:
> Hi,
>
> Rossy Boy is neither Einstein nor Zweistein.
> He is not Einstein since Einstein is already dead:
>
> Albert Einstein (1879 - 1955)
> https://de.wikipedia.org/wiki/Albert_Einstein
>
> He is also not Zweistein, since he doesn't
> understand concepts such as:
>
> - NVIDIA Volta ff. architecture
>
> Also his hands are small, and his breath stinks,
> and he lives in the basement of his mother.
>
> Bye
>
> Mild Shock schrieb:
>> Hi,
>>
>> Slowly I start understanding numbnuts like
>> Rossy Boy who don't understand tech, although
>> they are from UK and not from a 3rd world
>>
>> country, and also I start understanding morons
>> like Micro Penis, who are behind a curtain,
>> and cannot access a lot of tech.
>>
>> The same holds for SWI Prologs newest campaign
>> that probably adresses some poor indians that
>> have neither 5G nor Macs:
>>
>> 1:38:01 The Kyiv keynote disaster
>> https://www.youtube.com/watch?v=U8goS6B3BbI
>>
>> Woa! Real time download of Scala, Closure,
>> etc.. Whats the magic behind that? Some SWI
>> point of sale, downloading it via its
>>
>> keyboard and some telephathy module ?
>>
>> Bye
>>
>> Mild Shock schrieb:
>>> Hi,
>>>
>>> Ride the snake
>>> He's old and his skin is cold
>>> The west is the best
>>> The west is the best
>>> Get here and we'll do the rest
>>> The blue bus is calling us
>>> The blue bus is calling us
>>> Driver, where you taking us?
>>>
>>> Apocalypse Now intro: The Doors, The End {1979}
>>> https://www.youtube.com/watch?v=CIrvSJwwJUE
>>>
>>> Bye
>>>
>>> > Hi,
>>> >
>>> > Again I posted everything here:
>>> >
>>> >> 11.4 Giga Lips with a Budget Laptop
>>> >> https://github.com/Jean-Luc-Picard-2021/gigabudget
>>> >
>>> > The repo says, same time when I posted
>>> > the link first time:
>>> >
>>> >> This repository was archived by the
>>> >> owner on Jul 9, 2026. It is now read-only.
>>> >
>>> > Now a USENET user, who had already entitled
>>> > himself for a couple of irrational accusations
>>> >
>>> > towards my side, is asking this question:
>>> >
>>> > Chris M. Thomasson schrieb, Jul 24, 2026
>>> >> Show an outline of what you
>>> >> need you compute shader to do?
>>> >
>>> > Bravo, thats a delay of a wooping 15 days.
>>> >
>>> > Bye
>>>
>>> Mild Shock schrieb:
>>>> Hi,
>>>>
>>>> Remember when first all local AI was Python
>>>> and PyTorch APIs. And then suddently people started
>>>> using bare metal C/C++ Code. Here is the story:
>>>>
>>>> How it started:
>>>>
>>>> GPT-J or GPT-J-6B is an open-source large
>>>> language model (LLM) developed by EleutherAI
>>>> in 2021. As the name suggests, it is a
>>>> generative pre-trained transformer model
>>>> designed to produce human-like text that
>>>> continues from a prompt.
>>>> https://www.eleuther.ai/
>>>>
>>>> How it was going [Georgi Gerganov]:
>>>>
>>>> So a few days later comes out the LLaMA, I do
>>>> some calculations and I figure out “Okay, 65
>>>> billion parameters. You probably need about
>>>> 40 gigs of RAM, with 4-bit quantization. So
>>>> this can run on a MacBook. Why not do it?”
>>>>
>>>> Why I was able to do it so quickly - basically,
>>>> for all that I saw it’s pretty much GPT-J architecture
>>>> with some modifications, like some extra memorization
>>>> layers. It’s minor changes. Basically, again, the
>>>> existing code for the GPT-J, I just simply
>>>> modified it there, it happened pretty quickly.
>>>> https://changelog.com/podcast/532
>>>>
>>>> Georgi Gerganov, Bulgarian, now with Hugging
>>>> Face, ggml-cann also running on Chinese AI chips.
>>>> ggml Manifesto https://github.com/ggml-org/ggml
>>>>
>>>> Bye
>>>
>>
>
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:53 +0200 |
| Subject | Hurry Rossy Boy, the blue bus is waiting (Re: You are still chewing on SIMD. LoL) |
| Message-ID | <114d7lj$ki33$1@solani.org> |
| In reply to | #15785 |
Hi, Hurry Rossy Boy, the blue bus is waiting. There is a quite a hyperbole from here: Tesla S1070 in 2008 700 Watts , 1 Terra Flop SOLVE TOMORROW’S PROBLEMS TODAY https://www.azken.com/download/Tesla_DS_S1070_EU.pdf To here: Blackwell GPU in 2026 575 Watts, 104.8 Terra Flops ( RTX 5090 ) From Volta To Blackwell https://newsletter.semianalysis.com/p/nvidia-tensor-core-evolution-from-volta-to-blackwell But somehow the S1070 had already Massively- Parallel, Many-Core Architecture, and forms of MIMD, since it had 960 / 240 = 4 cores. 960 scalar processor cores (240 per GPU). But possibly more resticted inside work groups, than later NVIDIA Volta ff architecture with independent thread state. Bye Disclaimer: The above is only a very rough RTX 5090 spec. Its doesn't say what value format and what vector/matrics ops were used. Also energy consumption may vary. Mild Shock schrieb: > Hi, > > You are still chewing on SIMD. LoL > > Ross Finlayson schrieb: > > Then the idea is that any of those can be found and matched in > > one "run", i.e. a stall-less, branch-less, call-less list of less than > > a few or less than a few dozens or less than a few hundreds > > instructions, the results "findings" in data and corresponding > > "matchings" of expressions, that runs in less than one microsecond. > > You cannot make the mental translation that if you have: > > Ross Finlayson schrieb: > > So, the context then is for register state and stack contents, that > > the indicators of the above as "positive presence" then is to make > > for that the adjustments to the offsets and extents and the shifts > > is according to those, otherwise no-ops. Then the idea is that a > > As independent logical thread state, that automatically MIMD follows? > > Whats the problem to solve then? > > Bye
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:55 +0200 |
| Subject | Look how they advertized CUDA and logical threads (Re: Hurry Rossy Boy, the blue bus is waiting) |
| Message-ID | <114d7ol$ki33$3@solani.org> |
| In reply to | #15786 |
Hi, So what does NUM_SHADERS = 4096 shaders mean here? 11.4 Giga Lips with a Budget Laptop https://github.com/Jean-Luc-Picard-2021/gigabudget Its only the number of logical threads. CUDA™ TEChNOLOGY UNLOCkS ThE POWER OF TESLA MANY-CORE PROCESSORS The CUDA C compiler simplifies many-core programming by enabling code development in a high-level language and optimizing code to run on systems without knowledge of how many cores are in the hardware. CUDA applications automatically take advantage of more cores or fewer cores in a system, so they can scale from entry-level notebook GPUs to high end GPUs in technical workstations and further into racks of GPUs in data centers. This allows developers to “code once” and deploy on a range of systems, as well as scale forward in time as future GPUs deliver more performance per watt and more cores per processor. The benefit for software users is the opportunity to boost computing performance simply by adding GPUs or using their existing GPUs in new ways. https://www.azken.com/download/Tesla_DS_S1070_EU.pdf Bye Mild Shock schrieb: > Hi, > > Hurry Rossy Boy, the blue bus is waiting. > There is a quite a hyperbole from here: > > Tesla S1070 in 2008 > 700 Watts , 1 Terra Flop > SOLVE TOMORROW’S PROBLEMS TODAY > https://www.azken.com/download/Tesla_DS_S1070_EU.pdf > > To here: > > Blackwell GPU in 2026 > 575 Watts, 104.8 Terra Flops ( RTX 5090 ) > From Volta To Blackwell > https://newsletter.semianalysis.com/p/nvidia-tensor-core-evolution-from-volta-to-blackwell > > > But somehow the S1070 had already Massively- > Parallel, Many-Core Architecture, and forms > of MIMD, since it had 960 / 240 = 4 cores. > > 960 scalar processor cores (240 per GPU). > But possibly more resticted inside work > groups, than later NVIDIA Volta ff > > architecture with independent thread state. > > Bye > > Disclaimer: The above is only a very rough > RTX 5090 spec. Its doesn't say what value > format and what vector/matrics ops were > > used. Also energy consumption may vary. > > Mild Shock schrieb: >> Hi, >> >> You are still chewing on SIMD. LoL >> >> Ross Finlayson schrieb: >> > Then the idea is that any of those can be found and matched in >> > one "run", i.e. a stall-less, branch-less, call-less list of less than >> > a few or less than a few dozens or less than a few hundreds >> > instructions, the results "findings" in data and corresponding >> > "matchings" of expressions, that runs in less than one microsecond. >> >> You cannot make the mental translation that if you have: >> >> Ross Finlayson schrieb: >> > So, the context then is for register state and stack contents, that >> > the indicators of the above as "positive presence" then is to make >> > for that the adjustments to the offsets and extents and the shifts >> > is according to those, otherwise no-ops. Then the idea is that a >> >> As independent logical thread state, that automatically MIMD follows? >> >> Whats the problem to solve then? >> >> Bye
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:58 +0200 |
| Subject | Forget any arithmetization of product FSA (Re: Look how they advertized CUDA and logical threads) |
| Message-ID | <114d7vl$ki33$6@solani.org> |
| In reply to | #15787 |
Hi, Because of this parallelism you anyway need to forget about any arithmetization of product FSA (finite-state automata). Just forget it. What modern GPU provide is a kind of hirarchical viewpoint. You can have barriers in groups etc.. So you can exercise control over your mongolian horde of logical threads in a kind of multilevel schema. Have Fun! Bye Mild Shock schrieb: > Hi, > > So what does NUM_SHADERS = 4096 shaders mean here? > > 11.4 Giga Lips with a Budget Laptop > https://github.com/Jean-Luc-Picard-2021/gigabudget > > Its only the number of logical threads. > > CUDA™ TEChNOLOGY UNLOCkS ThE POWER OF TESLA MANY-CORE PROCESSORS > The CUDA C compiler simplifies many-core programming > by enabling code development in a high-level language > and optimizing code to run on systems without knowledge of > how many cores are in the hardware. > > CUDA applications automatically take advantage of more > cores or fewer cores in a system, so they can scale from > entry-level notebook GPUs to high end GPUs in technical > workstations and further into racks of GPUs in data > centers. This allows developers to > > “code once” and deploy on a range of systems, as well as > scale forward in time as future GPUs deliver more > performance per watt and more cores per processor. The benefit > for software users is the opportunity to boost computing > performance simply by adding GPUs or using their > > existing GPUs in new ways. > https://www.azken.com/download/Tesla_DS_S1070_EU.pdf > > Bye > > Mild Shock schrieb: >> Hi, >> >> Hurry Rossy Boy, the blue bus is waiting. >> There is a quite a hyperbole from here: >> >> Tesla S1070 in 2008 >> 700 Watts , 1 Terra Flop >> SOLVE TOMORROW’S PROBLEMS TODAY >> https://www.azken.com/download/Tesla_DS_S1070_EU.pdf >> >> To here: >> >> Blackwell GPU in 2026 >> 575 Watts, 104.8 Terra Flops ( RTX 5090 ) >> From Volta To Blackwell >> https://newsletter.semianalysis.com/p/nvidia-tensor-core-evolution-from-volta-to-blackwell >> >> >> But somehow the S1070 had already Massively- >> Parallel, Many-Core Architecture, and forms >> of MIMD, since it had 960 / 240 = 4 cores. >> >> 960 scalar processor cores (240 per GPU). >> But possibly more resticted inside work >> groups, than later NVIDIA Volta ff >> >> architecture with independent thread state. >> >> Bye >> >> Disclaimer: The above is only a very rough >> RTX 5090 spec. Its doesn't say what value >> format and what vector/matrics ops were >> >> used. Also energy consumption may vary. >> >> Mild Shock schrieb: >>> Hi, >>> >>> You are still chewing on SIMD. LoL >>> >>> Ross Finlayson schrieb: >>> > Then the idea is that any of those can be found and matched in >>> > one "run", i.e. a stall-less, branch-less, call-less list of less >>> than >>> > a few or less than a few dozens or less than a few hundreds >>> > instructions, the results "findings" in data and corresponding >>> > "matchings" of expressions, that runs in less than one microsecond. >>> >>> You cannot make the mental translation that if you have: >>> >>> Ross Finlayson schrieb: >>> > So, the context then is for register state and stack contents, that >>> > the indicators of the above as "positive presence" then is to make >>> > for that the adjustments to the offsets and extents and the shifts >>> > is according to those, otherwise no-ops. Then the idea is that a >>> >>> As independent logical thread state, that automatically MIMD follows? >>> >>> Whats the problem to solve then? >>> >>> Bye >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 20:06 +0200 |
| Subject | Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator] (Re: You are still chewing on SIMD. LoL) |
| Message-ID | <114dfer$ko2q$3@solani.org> |
| In reply to | #15785 |
Hi,
Rossy Boys tears could cool a data center,
he thinks there exists no literature about
serial algorithms of parallel stuff, and
he also thinks normal forms lead to optimizing
something. LoL, what a utter bullshit. I did
alreay a serial implementation of a parallel
simulation of my pi-WAM. Just lookup the literature
about pi-calulus. I published it a few days ago,
its part of 2.2.4 released already:
Parallel π-WAM: An Interleaved Synchronous Emulator
https://medium.com/2989/0196089e143a
Whats your point, Rossy Boy? Except you post pretend
nonsense not knowing what you are doing?
Bye
Ross Finlayson schrieb:
> No, troll, these are serial algorithms their optimized forms.
>
> Normal sorts of forms, ....
>
>
> Yeah, everybody already figured out "interpreters" and
> "programs" and "spawning".
>
> Go spawn yourself.
>
Mild Shock schrieb:
> Hi,
>
> You are still chewing on SIMD. LoL
>
> Ross Finlayson schrieb:
> > Then the idea is that any of those can be found and matched in
> > one "run", i.e. a stall-less, branch-less, call-less list of less than
> > a few or less than a few dozens or less than a few hundreds
> > instructions, the results "findings" in data and corresponding
> > "matchings" of expressions, that runs in less than one microsecond.
>
> You cannot make the mental translation that if you have:
>
> Ross Finlayson schrieb:
> > So, the context then is for register state and stack contents, that
> > the indicators of the above as "positive presence" then is to make
> > for that the adjustments to the offsets and extents and the shifts
> > is according to those, otherwise no-ops. Then the idea is that a
>
> As independent logical thread state, that automatically MIMD follows?
>
> Whats the problem to solve then?
>
> Bye
>
> Mild Shock schrieb:
>> Hi,
>>
>> Rossy Boy is neither Einstein nor Zweistein.
>> He is not Einstein since Einstein is already dead:
>>
>> Albert Einstein (1879 - 1955)
>> https://de.wikipedia.org/wiki/Albert_Einstein
>>
>> He is also not Zweistein, since he doesn't
>> understand concepts such as:
>>
>> - NVIDIA Volta ff. architecture
>>
>> Also his hands are small, and his breath stinks,
>> and he lives in the basement of his mother.
>>
>> Bye
>>
>> Mild Shock schrieb:
>>> Hi,
>>>
>>> Slowly I start understanding numbnuts like
>>> Rossy Boy who don't understand tech, although
>>> they are from UK and not from a 3rd world
>>>
>>> country, and also I start understanding morons
>>> like Micro Penis, who are behind a curtain,
>>> and cannot access a lot of tech.
>>>
>>> The same holds for SWI Prologs newest campaign
>>> that probably adresses some poor indians that
>>> have neither 5G nor Macs:
>>>
>>> 1:38:01 The Kyiv keynote disaster
>>> https://www.youtube.com/watch?v=U8goS6B3BbI
>>>
>>> Woa! Real time download of Scala, Closure,
>>> etc.. Whats the magic behind that? Some SWI
>>> point of sale, downloading it via its
>>>
>>> keyboard and some telephathy module ?
>>>
>>> Bye
>>>
>>> Mild Shock schrieb:
>>>> Hi,
>>>>
>>>> Ride the snake
>>>> He's old and his skin is cold
>>>> The west is the best
>>>> The west is the best
>>>> Get here and we'll do the rest
>>>> The blue bus is calling us
>>>> The blue bus is calling us
>>>> Driver, where you taking us?
>>>>
>>>> Apocalypse Now intro: The Doors, The End {1979}
>>>> https://www.youtube.com/watch?v=CIrvSJwwJUE
>>>>
>>>> Bye
>>>>
>>>> > Hi,
>>>> >
>>>> > Again I posted everything here:
>>>> >
>>>> >> 11.4 Giga Lips with a Budget Laptop
>>>> >> https://github.com/Jean-Luc-Picard-2021/gigabudget
>>>> >
>>>> > The repo says, same time when I posted
>>>> > the link first time:
>>>> >
>>>> >> This repository was archived by the
>>>> >> owner on Jul 9, 2026. It is now read-only.
>>>> >
>>>> > Now a USENET user, who had already entitled
>>>> > himself for a couple of irrational accusations
>>>> >
>>>> > towards my side, is asking this question:
>>>> >
>>>> > Chris M. Thomasson schrieb, Jul 24, 2026
>>>> >> Show an outline of what you
>>>> >> need you compute shader to do?
>>>> >
>>>> > Bravo, thats a delay of a wooping 15 days.
>>>> >
>>>> > Bye
>>>>
>>>> Mild Shock schrieb:
>>>>> Hi,
>>>>>
>>>>> Remember when first all local AI was Python
>>>>> and PyTorch APIs. And then suddently people started
>>>>> using bare metal C/C++ Code. Here is the story:
>>>>>
>>>>> How it started:
>>>>>
>>>>> GPT-J or GPT-J-6B is an open-source large
>>>>> language model (LLM) developed by EleutherAI
>>>>> in 2021. As the name suggests, it is a
>>>>> generative pre-trained transformer model
>>>>> designed to produce human-like text that
>>>>> continues from a prompt.
>>>>> https://www.eleuther.ai/
>>>>>
>>>>> How it was going [Georgi Gerganov]:
>>>>>
>>>>> So a few days later comes out the LLaMA, I do
>>>>> some calculations and I figure out “Okay, 65
>>>>> billion parameters. You probably need about
>>>>> 40 gigs of RAM, with 4-bit quantization. So
>>>>> this can run on a MacBook. Why not do it?”
>>>>>
>>>>> Why I was able to do it so quickly - basically,
>>>>> for all that I saw it’s pretty much GPT-J architecture
>>>>> with some modifications, like some extra memorization
>>>>> layers. It’s minor changes. Basically, again, the
>>>>> existing code for the GPT-J, I just simply
>>>>> modified it there, it happened pretty quickly.
>>>>> https://changelog.com/podcast/532
>>>>>
>>>>> Georgi Gerganov, Bulgarian, now with Hugging
>>>>> Face, ggml-cann also running on Chinese AI chips.
>>>>> ggml Manifesto https://github.com/ggml-org/ggml
>>>>>
>>>>> Bye
>>>>
>>>
>>
>
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 20:25 +0200 |
| Subject | I don't use Rust, you are crazy [Jump off a bridge, idiot] (Re: Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator]) |
| Message-ID | <114dgj9$kora$3@solani.org> |
| In reply to | #15790 |
Hi,
I don't use Rust, you are crazy. First of
all the parallel simulator is 100% written
in Prolog, should also run in ISO Prolog,
enhanced by a library(lists). Second I only
mentioned that WebGPU / WGSL, the language
there has a Rust inspired language.
Its not Rust. Whats wrong with you? Why do
you adress your weariness of life to me.
I am neither thief, nor can I help you
with your frustration, and histeric outbursts.
Maybe just be a man and jump off a bridge, idiot.
Or tame your frustration, usenet is not for
you alone, your stupid asshole.
Bye
Ross Finlayson schrieb:
>
https://www.theregister.com/databases/2026/07/29/after-rewriting-sqlite-in-rust-turso-turns-its-sights-on-postgres/5279835
> I don't much care about Rust.
>
> .. gibberish ..
>
> Thief.
Mild Shock schrieb:
> Hi,
>
> Rossy Boys tears could cool a data center,
> he thinks there exists no literature about
> serial algorithms of parallel stuff, and
>
> he also thinks normal forms lead to optimizing
> something. LoL, what a utter bullshit. I did
> alreay a serial implementation of a parallel
>
> simulation of my pi-WAM. Just lookup the literature
> about pi-calulus. I published it a few days ago,
> its part of 2.2.4 released already:
>
> Parallel π-WAM: An Interleaved Synchronous Emulator
> https://medium.com/2989/0196089e143a
>
> Whats your point, Rossy Boy? Except you post pretend
> nonsense not knowing what you are doing?
>
> Bye
>
> Ross Finlayson schrieb:
> > No, troll, these are serial algorithms their optimized forms.
> >
> > Normal sorts of forms, ....
> >
> >
> > Yeah, everybody already figured out "interpreters" and
> > "programs" and "spawning".
> >
> > Go spawn yourself.
> >
>
>
> Mild Shock schrieb:
>> Hi,
>>
>> You are still chewing on SIMD. LoL
>>
>> Ross Finlayson schrieb:
>> > Then the idea is that any of those can be found and matched in
>> > one "run", i.e. a stall-less, branch-less, call-less list of less than
>> > a few or less than a few dozens or less than a few hundreds
>> > instructions, the results "findings" in data and corresponding
>> > "matchings" of expressions, that runs in less than one microsecond.
>>
>> You cannot make the mental translation that if you have:
>>
>> Ross Finlayson schrieb:
>> > So, the context then is for register state and stack contents, that
>> > the indicators of the above as "positive presence" then is to make
>> > for that the adjustments to the offsets and extents and the shifts
>> > is according to those, otherwise no-ops. Then the idea is that a
>>
>> As independent logical thread state, that automatically MIMD follows?
>>
>> Whats the problem to solve then?
>>
>> Bye
>>
>> Mild Shock schrieb:
>>> Hi,
>>>
>>> Rossy Boy is neither Einstein nor Zweistein.
>>> He is not Einstein since Einstein is already dead:
>>>
>>> Albert Einstein (1879 - 1955)
>>> https://de.wikipedia.org/wiki/Albert_Einstein
>>>
>>> He is also not Zweistein, since he doesn't
>>> understand concepts such as:
>>>
>>> - NVIDIA Volta ff. architecture
>>>
>>> Also his hands are small, and his breath stinks,
>>> and he lives in the basement of his mother.
>>>
>>> Bye
>>>
>>> Mild Shock schrieb:
>>>> Hi,
>>>>
>>>> Slowly I start understanding numbnuts like
>>>> Rossy Boy who don't understand tech, although
>>>> they are from UK and not from a 3rd world
>>>>
>>>> country, and also I start understanding morons
>>>> like Micro Penis, who are behind a curtain,
>>>> and cannot access a lot of tech.
>>>>
>>>> The same holds for SWI Prologs newest campaign
>>>> that probably adresses some poor indians that
>>>> have neither 5G nor Macs:
>>>>
>>>> 1:38:01 The Kyiv keynote disaster
>>>> https://www.youtube.com/watch?v=U8goS6B3BbI
>>>>
>>>> Woa! Real time download of Scala, Closure,
>>>> etc.. Whats the magic behind that? Some SWI
>>>> point of sale, downloading it via its
>>>>
>>>> keyboard and some telephathy module ?
>>>>
>>>> Bye
>>>>
>>>> Mild Shock schrieb:
>>>>> Hi,
>>>>>
>>>>> Ride the snake
>>>>> He's old and his skin is cold
>>>>> The west is the best
>>>>> The west is the best
>>>>> Get here and we'll do the rest
>>>>> The blue bus is calling us
>>>>> The blue bus is calling us
>>>>> Driver, where you taking us?
>>>>>
>>>>> Apocalypse Now intro: The Doors, The End {1979}
>>>>> https://www.youtube.com/watch?v=CIrvSJwwJUE
>>>>>
>>>>> Bye
>>>>>
>>>>> > Hi,
>>>>> >
>>>>> > Again I posted everything here:
>>>>> >
>>>>> >> 11.4 Giga Lips with a Budget Laptop
>>>>> >> https://github.com/Jean-Luc-Picard-2021/gigabudget
>>>>> >
>>>>> > The repo says, same time when I posted
>>>>> > the link first time:
>>>>> >
>>>>> >> This repository was archived by the
>>>>> >> owner on Jul 9, 2026. It is now read-only.
>>>>> >
>>>>> > Now a USENET user, who had already entitled
>>>>> > himself for a couple of irrational accusations
>>>>> >
>>>>> > towards my side, is asking this question:
>>>>> >
>>>>> > Chris M. Thomasson schrieb, Jul 24, 2026
>>>>> >> Show an outline of what you
>>>>> >> need you compute shader to do?
>>>>> >
>>>>> > Bravo, thats a delay of a wooping 15 days.
>>>>> >
>>>>> > Bye
>>>>>
>>>>> Mild Shock schrieb:
>>>>>> Hi,
>>>>>>
>>>>>> Remember when first all local AI was Python
>>>>>> and PyTorch APIs. And then suddently people started
>>>>>> using bare metal C/C++ Code. Here is the story:
>>>>>>
>>>>>> How it started:
>>>>>>
>>>>>> GPT-J or GPT-J-6B is an open-source large
>>>>>> language model (LLM) developed by EleutherAI
>>>>>> in 2021. As the name suggests, it is a
>>>>>> generative pre-trained transformer model
>>>>>> designed to produce human-like text that
>>>>>> continues from a prompt.
>>>>>> https://www.eleuther.ai/
>>>>>>
>>>>>> How it was going [Georgi Gerganov]:
>>>>>>
>>>>>> So a few days later comes out the LLaMA, I do
>>>>>> some calculations and I figure out “Okay, 65
>>>>>> billion parameters. You probably need about
>>>>>> 40 gigs of RAM, with 4-bit quantization. So
>>>>>> this can run on a MacBook. Why not do it?”
>>>>>>
>>>>>> Why I was able to do it so quickly - basically,
>>>>>> for all that I saw it’s pretty much GPT-J architecture
>>>>>> with some modifications, like some extra memorization
>>>>>> layers. It’s minor changes. Basically, again, the
>>>>>> existing code for the GPT-J, I just simply
>>>>>> modified it there, it happened pretty quickly.
>>>>>> https://changelog.com/podcast/532
>>>>>>
>>>>>> Georgi Gerganov, Bulgarian, now with Hugging
>>>>>> Face, ggml-cann also running on Chinese AI chips.
>>>>>> ggml Manifesto https://github.com/ggml-org/ggml
>>>>>>
>>>>>> Bye
>>>>>
>>>>
>>>
>>
>
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 22:52 +0200 |
| Subject | Hack ecosystem ignorance paired with paranoia [Nand to Tetris] (Re: I don't use Rust, you are crazy) |
| Message-ID | <114dp6a$kf2n$4@solani.org> |
| In reply to | #15791 |
Hi,
> Who exactly is the thief? Does this person
> have stats in the Rogue class in dungeons
> and dragons?
The conspiracy theory of a stealing of Torso VDBE,
by Rossy Boy, is probably a result of complete
ignorance of the Hack ecosystem.
Hack is a very popular computer science project,
with a couple of subprojects in hardware and
software. It goes also by the name Nand to Tetris,
and is programming language agnositic. You can do
Hack experiments in any programming language, be
it BASIC, ADA or Rust. Nobody cares.
The gist are projects like here, first to
educate yourself about Hack:
https://www.nand2tetris.org/course
And then to use Hack in different contexts:
https://www.nand2tetris.org/copy-of-talks
For didactic purposes, I used Hack for my WebGPU
experiment. I didn't even take a look at Torso
VDBE, why should I? Hack is nicely documented,
has even a book, and fusing the two 16-bit
instruction types A and D, into a single 32-bit
instruction stream, is nowhere patented.
Bye
Mild Shock schrieb:
> Hi,
>
> I don't use Rust, you are crazy. First of
> all the parallel simulator is 100% written
> in Prolog, should also run in ISO Prolog,
>
> enhanced by a library(lists). Second I only
> mentioned that WebGPU / WGSL, the language
> there has a Rust inspired language.
>
> Its not Rust. Whats wrong with you? Why do
> you adress your weariness of life to me.
> I am neither thief, nor can I help you
>
> with your frustration, and histeric outbursts.
> Maybe just be a man and jump off a bridge, idiot.
> Or tame your frustration, usenet is not for
>
> you alone, your stupid asshole.
>
> Bye
>
> Ross Finlayson schrieb:
> >
> https://www.theregister.com/databases/2026/07/29/after-rewriting-sqlite-in-rust-turso-turns-its-sights-on-postgres/5279835
>
> > I don't much care about Rust.
> >
> > .. gibberish ..
> >
> > Thief.
>
> Mild Shock schrieb:
>> Hi,
>>
>> Rossy Boys tears could cool a data center,
>> he thinks there exists no literature about
>> serial algorithms of parallel stuff, and
>>
>> he also thinks normal forms lead to optimizing
>> something. LoL, what a utter bullshit. I did
>> alreay a serial implementation of a parallel
>>
>> simulation of my pi-WAM. Just lookup the literature
>> about pi-calulus. I published it a few days ago,
>> its part of 2.2.4 released already:
>>
>> Parallel π-WAM: An Interleaved Synchronous Emulator
>> https://medium.com/2989/0196089e143a
>>
>> Whats your point, Rossy Boy? Except you post pretend
>> nonsense not knowing what you are doing?
>>
>> Bye
>>
>> Ross Finlayson schrieb:
>> > No, troll, these are serial algorithms their optimized forms.
>> >
>> > Normal sorts of forms, ....
>> >
>> >
>> > Yeah, everybody already figured out "interpreters" and
>> > "programs" and "spawning".
>> >
>> > Go spawn yourself.
>> >
>>
>>
>> Mild Shock schrieb:
>>> Hi,
>>>
>>> You are still chewing on SIMD. LoL
>>>
>>> Ross Finlayson schrieb:
>>> > Then the idea is that any of those can be found and matched in
>>> > one "run", i.e. a stall-less, branch-less, call-less list of less
>>> than
>>> > a few or less than a few dozens or less than a few hundreds
>>> > instructions, the results "findings" in data and corresponding
>>> > "matchings" of expressions, that runs in less than one microsecond.
>>>
>>> You cannot make the mental translation that if you have:
>>>
>>> Ross Finlayson schrieb:
>>> > So, the context then is for register state and stack contents, that
>>> > the indicators of the above as "positive presence" then is to make
>>> > for that the adjustments to the offsets and extents and the shifts
>>> > is according to those, otherwise no-ops. Then the idea is that a
>>>
>>> As independent logical thread state, that automatically MIMD follows?
>>>
>>> Whats the problem to solve then?
>>>
>>> Bye
>>>
>>> Mild Shock schrieb:
>>>> Hi,
>>>>
>>>> Rossy Boy is neither Einstein nor Zweistein.
>>>> He is not Einstein since Einstein is already dead:
>>>>
>>>> Albert Einstein (1879 - 1955)
>>>> https://de.wikipedia.org/wiki/Albert_Einstein
>>>>
>>>> He is also not Zweistein, since he doesn't
>>>> understand concepts such as:
>>>>
>>>> - NVIDIA Volta ff. architecture
>>>>
>>>> Also his hands are small, and his breath stinks,
>>>> and he lives in the basement of his mother.
>>>>
>>>> Bye
>>>>
>>>> Mild Shock schrieb:
>>>>> Hi,
>>>>>
>>>>> Slowly I start understanding numbnuts like
>>>>> Rossy Boy who don't understand tech, although
>>>>> they are from UK and not from a 3rd world
>>>>>
>>>>> country, and also I start understanding morons
>>>>> like Micro Penis, who are behind a curtain,
>>>>> and cannot access a lot of tech.
>>>>>
>>>>> The same holds for SWI Prologs newest campaign
>>>>> that probably adresses some poor indians that
>>>>> have neither 5G nor Macs:
>>>>>
>>>>> 1:38:01 The Kyiv keynote disaster
>>>>> https://www.youtube.com/watch?v=U8goS6B3BbI
>>>>>
>>>>> Woa! Real time download of Scala, Closure,
>>>>> etc.. Whats the magic behind that? Some SWI
>>>>> point of sale, downloading it via its
>>>>>
>>>>> keyboard and some telephathy module ?
>>>>>
>>>>> Bye
>>>>>
>>>>> Mild Shock schrieb:
>>>>>> Hi,
>>>>>>
>>>>>> Ride the snake
>>>>>> He's old and his skin is cold
>>>>>> The west is the best
>>>>>> The west is the best
>>>>>> Get here and we'll do the rest
>>>>>> The blue bus is calling us
>>>>>> The blue bus is calling us
>>>>>> Driver, where you taking us?
>>>>>>
>>>>>> Apocalypse Now intro: The Doors, The End {1979}
>>>>>> https://www.youtube.com/watch?v=CIrvSJwwJUE
>>>>>>
>>>>>> Bye
>>>>>>
>>>>>> > Hi,
>>>>>> >
>>>>>> > Again I posted everything here:
>>>>>> >
>>>>>> >> 11.4 Giga Lips with a Budget Laptop
>>>>>> >> https://github.com/Jean-Luc-Picard-2021/gigabudget
>>>>>> >
>>>>>> > The repo says, same time when I posted
>>>>>> > the link first time:
>>>>>> >
>>>>>> >> This repository was archived by the
>>>>>> >> owner on Jul 9, 2026. It is now read-only.
>>>>>> >
>>>>>> > Now a USENET user, who had already entitled
>>>>>> > himself for a couple of irrational accusations
>>>>>> >
>>>>>> > towards my side, is asking this question:
>>>>>> >
>>>>>> > Chris M. Thomasson schrieb, Jul 24, 2026
>>>>>> >> Show an outline of what you
>>>>>> >> need you compute shader to do?
>>>>>> >
>>>>>> > Bravo, thats a delay of a wooping 15 days.
>>>>>> >
>>>>>> > Bye
>>>>>>
>>>>>> Mild Shock schrieb:
>>>>>>> Hi,
>>>>>>>
>>>>>>> Remember when first all local AI was Python
>>>>>>> and PyTorch APIs. And then suddently people started
>>>>>>> using bare metal C/C++ Code. Here is the story:
>>>>>>>
>>>>>>> How it started:
>>>>>>>
>>>>>>> GPT-J or GPT-J-6B is an open-source large
>>>>>>> language model (LLM) developed by EleutherAI
>>>>>>> in 2021. As the name suggests, it is a
>>>>>>> generative pre-trained transformer model
>>>>>>> designed to produce human-like text that
>>>>>>> continues from a prompt.
>>>>>>> https://www.eleuther.ai/
>>>>>>>
>>>>>>> How it was going [Georgi Gerganov]:
>>>>>>>
>>>>>>> So a few days later comes out the LLaMA, I do
>>>>>>> some calculations and I figure out “Okay, 65
>>>>>>> billion parameters. You probably need about
>>>>>>> 40 gigs of RAM, with 4-bit quantization. So
>>>>>>> this can run on a MacBook. Why not do it?”
>>>>>>>
>>>>>>> Why I was able to do it so quickly - basically,
>>>>>>> for all that I saw it’s pretty much GPT-J architecture
>>>>>>> with some modifications, like some extra memorization
>>>>>>> layers. It’s minor changes. Basically, again, the
>>>>>>> existing code for the GPT-J, I just simply
>>>>>>> modified it there, it happened pretty quickly.
>>>>>>> https://changelog.com/podcast/532
>>>>>>>
>>>>>>> Georgi Gerganov, Bulgarian, now with Hugging
>>>>>>> Face, ggml-cann also running on Chinese AI chips.
>>>>>>> ggml Manifesto https://github.com/ggml-org/ggml
>>>>>>>
>>>>>>> Bye
>>>>>>
>>>>>
>>>>
>>>
>>
>
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 23:11 +0200 |
| Subject | A funny Q16.16 experiment with Hack (Re: Hack ecosystem ignorance paired with paranoia [Nand to Tetris]) |
| Message-ID | <114dqan$kfq8$2@solani.org> |
| In reply to | #15792 |
Hi, This seems to be a funny Q16.16 experiment. It shows that an integerish Hack can do floatish stuff, by using binary fixpoint: Raytracing on the Hack computer 2021/06/13 - im alex https://blog.alexqua.ch/posts/from-nand-to-raytracer/ That it uses Rust is arbitrary. Feel free to do it in C, C++, FORTRAN or Java. I guess these languages all have basic arithmethic, right? Maybe not a long jump always? Bye Mild Shock schrieb: > Hi, > > > Who exactly is the thief? Does this person > > have stats in the Rogue class in dungeons > > and dragons? > > The conspiracy theory of a stealing of Torso VDBE, > by Rossy Boy, is probably a result of complete > ignorance of the Hack ecosystem. > > Hack is a very popular computer science project, > with a couple of subprojects in hardware and > software. It goes also by the name Nand to Tetris, > > and is programming language agnositic. You can do > Hack experiments in any programming language, be > it BASIC, ADA or Rust. Nobody cares. > > The gist are projects like here, first to > educate yourself about Hack: > > https://www.nand2tetris.org/course > > And then to use Hack in different contexts: > > https://www.nand2tetris.org/copy-of-talks > > For didactic purposes, I used Hack for my WebGPU > experiment. I didn't even take a look at Torso > VDBE, why should I? Hack is nicely documented, > > has even a book, and fusing the two 16-bit > instruction types A and D, into a single 32-bit > instruction stream, is nowhere patented. > > Bye
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-27 18:59 +0200 |
| Subject | Got it. Or are you too stupid? [New Usenet Mantra] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) |
| Message-ID | <11482ou$h3vk$1@solani.org> |
| In reply to | #15710 |
Hi, Ok, guys lets face it. You are a bunch of morons. When did I do this post: 11.4 Giga Lips with a Budget Laptop https://github.com/Jean-Luc-Picard-2021/gigabudget Yes on Jul 9, 2026, now we have Jul 27, 2026. Thats a wooping 18 days meanwhile. And you still don't get the meaning and implications of the post. Like you even don't get what "budget" nowdays means in terms of performance units and energy units? And what LIPS means, drawn from TOPS, in terms of applications? Shame on you guys! You are a bunch of brainless idiots. Bye Mild Shock schrieb: > Hi, > > Remember when first all local AI was Python > and PyTorch APIs. And then suddently people started > using bare metal C/C++ Code. Here is the story: > > How it started: > > GPT-J or GPT-J-6B is an open-source large > language model (LLM) developed by EleutherAI > in 2021. As the name suggests, it is a > generative pre-trained transformer model > designed to produce human-like text that > continues from a prompt. > https://www.eleuther.ai/ > > How it was going [Georgi Gerganov]: > > So a few days later comes out the LLaMA, I do > some calculations and I figure out “Okay, 65 > billion parameters. You probably need about > 40 gigs of RAM, with 4-bit quantization. So > this can run on a MacBook. Why not do it?” > > Why I was able to do it so quickly - basically, > for all that I saw it’s pretty much GPT-J architecture > with some modifications, like some extra memorization > layers. It’s minor changes. Basically, again, the > existing code for the GPT-J, I just simply > modified it there, it happened pretty quickly. > https://changelog.com/podcast/532 > > Georgi Gerganov, Bulgarian, now with Hugging > Face, ggml-cann also running on Chinese AI chips. > ggml Manifesto https://github.com/ggml-org/ggml > > Bye
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 13:03 +0200 |
| Subject | Lamas in a cradle and Lamas on the edge [Red Pyjama] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) |
| Message-ID | <114cmmv$jmhc$2@solani.org> |
| In reply to | #15710 |
Hi, Why does this Lama have a red pyjama. Oh, its a baby Lama. Its still in the cradle and needs some training: RedPajama-Data-v2 https://github.com/togethercomputer/RedPajama-Data But then Andrej Karpathy recently showed GPT-2 training on rented GPUs for less than 100 USD in less then 2 hours. So where do these grown up Lamas go. Well Georgi Gerganov prefered C++/C when he shouted Llama Llama Red Pyjama. But you also find WebLLM, wrapping the underlying C++/C GPU interface via the W3C standard WebGPU / WGSL, with JavaScript: In-Browser LLM Inference Engine https://webllm.mlc.ai/ My experience with WebLLM 6 months ago on an iPad Pro 2024, still a little early stage performance and robustness. But hey hardware of AI mobile iGPUs is still evolving, and AI laptop, AI smartphones and AI tablets, will soon feature Chinese hardware such some new Kirin AI in 2027. Bye Mild Shock schrieb: > Hi, > > Remember when first all local AI was Python > and PyTorch APIs. And then suddently people started > using bare metal C/C++ Code. Here is the story: > > How it started: > > GPT-J or GPT-J-6B is an open-source large > language model (LLM) developed by EleutherAI > in 2021. As the name suggests, it is a > generative pre-trained transformer model > designed to produce human-like text that > continues from a prompt. > https://www.eleuther.ai/ > > How it was going [Georgi Gerganov]: > > So a few days later comes out the LLaMA, I do > some calculations and I figure out “Okay, 65 > billion parameters. You probably need about > 40 gigs of RAM, with 4-bit quantization. So > this can run on a MacBook. Why not do it?” > > Why I was able to do it so quickly - basically, > for all that I saw it’s pretty much GPT-J architecture > with some modifications, like some extra memorization > layers. It’s minor changes. Basically, again, the > existing code for the GPT-J, I just simply > modified it there, it happened pretty quickly. > https://changelog.com/podcast/532 > > Georgi Gerganov, Bulgarian, now with Hugging > Face, ggml-cann also running on Chinese AI chips. > ggml Manifesto https://github.com/ggml-org/ggml > > Bye
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:02 +0200 |
| Subject | AI Accelerators and ISO Prolog multi-threading (Re: Lamas in a cradle and Lamas on the edge [Red Pyjama]) |
| Message-ID | <114d4lt$kfj1$2@solani.org> |
| In reply to | #15780 |
Hi, Usual question: > Why implement both pre-emptive threading AND cooperative tasks/engines? I had implemented the ISO proposal in formerly Jekejeke Prolog, you find the ISO proposal here: ISO/IEC DTR 13211–5:2007 Prolog multi-threading support https://logtalk.org/plstd/threads.pdf But the ISO proposal doesn't match modern WebGPU APIs, where your logical threads can live remotely in a dedicated GPU in the VRAM there, and where you would have launch parameters that say: Hey please run 4096 compute shaders for me, that have independet thread state. Using cooperative multi-tasking as the orchestrator works well. Bye Mild Shock schrieb: > Hi, > > Why does this Lama have a red pyjama. > Oh, its a baby Lama. Its still in the cradle > and needs some training: > > RedPajama-Data-v2 > https://github.com/togethercomputer/RedPajama-Data > > But then Andrej Karpathy recently showed > GPT-2 training on rented GPUs for less > than 100 USD in less then 2 hours. > > So where do these grown up Lamas go. > Well Georgi Gerganov prefered C++/C > when he shouted Llama Llama Red Pyjama. > > But you also find WebLLM, wrapping the > underlying C++/C GPU interface via the > W3C standard WebGPU / WGSL, with JavaScript: > > In-Browser LLM Inference Engine > https://webllm.mlc.ai/ > > My experience with WebLLM 6 months > ago on an iPad Pro 2024, still a little early > stage performance and robustness. > > But hey hardware of AI mobile iGPUs is > still evolving, and AI laptop, AI smartphones > and AI tablets, will soon feature Chinese > > hardware such some new Kirin AI in 2027. > > Bye > > Mild Shock schrieb: >> Hi, >> >> Remember when first all local AI was Python >> and PyTorch APIs. And then suddently people started >> using bare metal C/C++ Code. Here is the story: >> >> How it started: >> >> GPT-J or GPT-J-6B is an open-source large >> language model (LLM) developed by EleutherAI >> in 2021. As the name suggests, it is a >> generative pre-trained transformer model >> designed to produce human-like text that >> continues from a prompt. >> https://www.eleuther.ai/ >> >> How it was going [Georgi Gerganov]: >> >> So a few days later comes out the LLaMA, I do >> some calculations and I figure out “Okay, 65 >> billion parameters. You probably need about >> 40 gigs of RAM, with 4-bit quantization. So >> this can run on a MacBook. Why not do it?” >> >> Why I was able to do it so quickly - basically, >> for all that I saw it’s pretty much GPT-J architecture >> with some modifications, like some extra memorization >> layers. It’s minor changes. Basically, again, the >> existing code for the GPT-J, I just simply >> modified it there, it happened pretty quickly. >> https://changelog.com/podcast/532 >> >> Georgi Gerganov, Bulgarian, now with Hugging >> Face, ggml-cann also running on Chinese AI chips. >> ggml Manifesto https://github.com/ggml-org/ggml >> >> Bye >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:04 +0200 |
| Subject | Actor/Erlang is dead, no Thread and Mailbox conflation [golang channels] (Re: AI Accelerators and ISO Prolog multi-threading) (Re: Lamas in a cradle and Lamas on the edge [Red Pyjama]) |
| Message-ID | <114d4p7$kfj1$3@solani.org> |
| In reply to | #15783 |
Hi, Mostlikely for high performance computing à la, the Actor/Erlang model is dead, they might rely on MPMC (Multiple Producer, Multiple Consumer) queue entities separate from the threads. The ISO Prolog multi-threading support had also such threads. But besides that was also Actor/Erlang leaning in practice, like SWI, where threads have some default queues. So an actor is basically a Thread and Mailbox conflation. While a MPMC queue is a kind of separate Mailbox, where multiple "actors" can read from and write from. A kind of localized Linda Tuple store. Which Programming language did adopted the non-Actor pi-calculus model? Right golang with its channels. Bye Mild Shock schrieb: > Hi, > > Usual question: > > > Why implement both pre-emptive threading > AND cooperative tasks/engines? > > I had implemented the ISO proposal in formerly Jekejeke > Prolog, you find the ISO proposal here: > > ISO/IEC DTR 13211–5:2007 > Prolog multi-threading support > https://logtalk.org/plstd/threads.pdf > > But the ISO proposal doesn't match modern WebGPU APIs, > where your logical threads can live remotely in a dedicated GPU > in the VRAM there, and where you would have launch > > parameters that say: Hey please run 4096 compute > shaders for me, that have independet thread state. Using > cooperative multi-tasking as the orchestrator works well. > > Bye > > Mild Shock schrieb: >> Hi, >> >> Why does this Lama have a red pyjama. >> Oh, its a baby Lama. Its still in the cradle >> and needs some training: >> >> RedPajama-Data-v2 >> https://github.com/togethercomputer/RedPajama-Data >> >> But then Andrej Karpathy recently showed >> GPT-2 training on rented GPUs for less >> than 100 USD in less then 2 hours. >> >> So where do these grown up Lamas go. >> Well Georgi Gerganov prefered C++/C >> when he shouted Llama Llama Red Pyjama. >> >> But you also find WebLLM, wrapping the >> underlying C++/C GPU interface via the >> W3C standard WebGPU / WGSL, with JavaScript: >> >> In-Browser LLM Inference Engine >> https://webllm.mlc.ai/ >> >> My experience with WebLLM 6 months >> ago on an iPad Pro 2024, still a little early >> stage performance and robustness. >> >> But hey hardware of AI mobile iGPUs is >> still evolving, and AI laptop, AI smartphones >> and AI tablets, will soon feature Chinese >> >> hardware such some new Kirin AI in 2027. >> >> Bye >> >> Mild Shock schrieb: >>> Hi, >>> >>> Remember when first all local AI was Python >>> and PyTorch APIs. And then suddently people started >>> using bare metal C/C++ Code. Here is the story: >>> >>> How it started: >>> >>> GPT-J or GPT-J-6B is an open-source large >>> language model (LLM) developed by EleutherAI >>> in 2021. As the name suggests, it is a >>> generative pre-trained transformer model >>> designed to produce human-like text that >>> continues from a prompt. >>> https://www.eleuther.ai/ >>> >>> How it was going [Georgi Gerganov]: >>> >>> So a few days later comes out the LLaMA, I do >>> some calculations and I figure out “Okay, 65 >>> billion parameters. You probably need about >>> 40 gigs of RAM, with 4-bit quantization. So >>> this can run on a MacBook. Why not do it?” >>> >>> Why I was able to do it so quickly - basically, >>> for all that I saw it’s pretty much GPT-J architecture >>> with some modifications, like some extra memorization >>> layers. It’s minor changes. Basically, again, the >>> existing code for the GPT-J, I just simply >>> modified it there, it happened pretty quickly. >>> https://changelog.com/podcast/532 >>> >>> Georgi Gerganov, Bulgarian, now with Hugging >>> Face, ggml-cann also running on Chinese AI chips. >>> ggml Manifesto https://github.com/ggml-org/ggml >>> >>> Bye >> >
[toc] | [prev] | [standalone]
Page 3 of 3 — ← Prev page 1 2 [3]
Back to top | Article view | comp.lang.prolog
csiph-web