Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > sci.math > #646933 > unrolled thread
| Started by | Mild Shock <janburse@fastmail.fm> |
|---|---|
| First post | 2026-07-22 21:00 +0200 |
| Last post | 2026-08-02 23:33 +0200 |
| Articles | 20 on this page of 126 — 14 participants |
Back to article view | Back to sci.math
The Wuhan Virus that destroyed Python [ggml Manifesto] Mild Shock <janburse@fastmail.fm> - 2026-07-22 21:00 +0200
Deadlock Exorcism: Switch from Push to Pull [A pi-calculus Specification of Prolog] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 00:23 +0200
Why do you even need a mpmc queue? [Thunder Kittens] (Re: Deadlock Exorcism: Switch from Push to Pull) Mild Shock <janburse@fastmail.fm> - 2026-07-23 08:43 +0200
Trivial balancing example for (int i=0; i<global_id; i++) (Re: Why do you even need a mpmc queue? [Thunder Kittens]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 08:57 +0200
Enqueue/dequeue need not be fast and can spinn ["fairness" questions] (Was: Trivial balancing example for (int i=0; i<global_id; i++)) Mild Shock <janburse@fastmail.fm> - 2026-07-23 09:11 +0200
The Pixel Phone AI Experiment Song (Re: Enqueue/dequeue need not be fast and can spinn ["fairness" questions] ) Mild Shock <janburse@fastmail.fm> - 2026-07-23 09:21 +0200
Re: Why do you even need a mpmc queue? [Thunder Kittens] (Re: Deadlock Exorcism: Switch from Push to Pull) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-23 08:24 -0700
Potential Python Recovery: Free Threading [3.13 release] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 10:19 +0200
Re: Potential Python Recovery: Free Threading [3.13 release] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) Ross Valikhanov <kavna@rl.ru> - 2026-07-23 16:01 +0000
Re: The Wuhan Virus that destroyed Python [ggml Manifesto] Ramon Dubenkov <omd@nnk.ru> - 2026-07-23 13:38 +0000
The things XILINX braught to the AMD table (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-23 18:47 +0200
NIVIDIA evacuated its Chinese market [Tau Scaling] (Was: The things XILINX braught to the AMD table) Mild Shock <janburse@fastmail.fm> - 2026-07-23 19:11 +0200
NVIDIA evacuated its Chinese market [Tau Scaling] (Re: The things XILINX braught to the AMD table) Mild Shock <janburse@fastmail.fm> - 2026-07-23 19:12 +0200
Re: NVIDIA evacuated its Chinese market [Tau Scaling] (Re: The things XILINX braught to the AMD table) Lane W <cactus_DAC@yahoo.com> - 2026-07-23 11:22 -0600
Micro penis mother sung arias (Was: NVIDIA evacuated its Chinese market [Tau Scaling]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 14:38 +0200
Re: Micro penis mother sung arias (Was: NVIDIA evacuated its Chinese market [Tau Scaling]) Lane W <cactus_DAC@yahoo.com> - 2026-07-24 07:15 -0600
Micro penis brain is in constant hiatus (Was: Micro penis mother sung arias) Mild Shock <janburse@fastmail.fm> - 2026-07-24 15:24 +0200
Re: Micro penis brain is in constant hiatus (Was: Micro penis mother sung arias) Mild Shock <janburse@fastmail.fm> - 2026-07-24 15:36 +0200
Ignoramus or Ignorabimus: I don't care (π-WAM) (Re: Micro penis brain is in constant hiatus) Mild Shock <janburse@fastmail.fm> - 2026-07-24 15:38 +0200
Re: Ignoramus or Ignorabimus: I don't care (π-WAM) (Re: Micro penis brain is in constant hiatus) Lane W <cactus_DAC@yahoo.com> - 2026-07-24 08:31 -0600
You are a moron, brainless putin payed (Was: Ignoramus or Ignorabimus: I don't care (π-WAM)) Mild Shock <janburse@fastmail.fm> - 2026-07-24 18:01 +0200
Re: You are a moron, brainless putin payed (Was: Ignoramus or Ignorabimus: I don't care (π-WAM)) Lane W <cactus_DAC@yahoo.com> - 2026-07-24 10:27 -0600
Yeah keep reading my posts, uninspired fool (Was: You are a moron, brainless putin payed) Mild Shock <janburse@fastmail.fm> - 2026-07-24 19:45 +0200
Re: Yeah keep reading my posts, uninspired fool (Was: You are a moron, brainless putin payed) Lane W <cactus_DAC@yahoo.com> - 2026-07-24 12:11 -0600
LoL (Was: Yeah keep reading my posts, uninspired fool ) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:12 +0200
Re: LoL (Was: Yeah keep reading my posts, uninspired fool ) Lane W <cactus_DAC@yahoo.com> - 2026-07-24 12:53 -0600
Out of the blue accusation span 15 days [Empirical USENET study] (Was: Ignoramus or Ignorabimus: I don't care (π-WAM)) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:26 +0200
A brain desease of 20 days [Rossy Boy] (Re: Ignoramus or Ignorabimus: I don't care (π-WAM)) Mild Shock <janburse@fastmail.fm> - 2026-07-29 18:40 +0200
Re: A brain desease of 20 days [Rossy Boy] (Re: Ignoramus or Ignorabimus: I don't care (π-WAM)) Mantra Mahonov <hnaam@aat.ru> - 2026-07-29 21:15 +0000
I didn't use a Ryzen Halo, whats wrong with you? (Was: A brain desease of 20 days [Rossy Boy]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 23:24 +0200
Ignoramus / Ignorabimus Barometer: Almost 1 Month (Was: A brain desease of 20 days [Rossy Boy]) Mild Shock <janburse@fastmail.fm> - 2026-08-03 00:06 +0200
Re: Ignoramus / Ignorabimus Barometer: Almost 1 Month (Was: A brain desease of 20 days [Rossy Boy]) Lane W <cactus_DAC@yahoo.com> - 2026-08-03 13:16 -0600
Re: Ignoramus / Ignorabimus Barometer: Almost 1 Month (Was: A brain desease of 20 days [Rossy Boy]) "Chris M. Thomasson" <chris.m.thomasson.1@gmail.com> - 2026-08-03 12:54 -0700
No you didn't try, you only spammed old code (Was: Ignoramus / Ignorabimus Barometer: Almost 1 Month) Mild Shock <janburse@fastmail.fm> - 2026-08-03 22:15 +0200
ASML stocks are plunging, bye bye dutchies (Was: NVIDIA evacuated its Chinese market [Tau Scaling]) Mild Shock <janburse@fastmail.fm> - 2026-07-28 14:17 +0200
Little Data Center on Your Palm [AI Laptops for 500 USD] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto] Mild Shock <janburse@fastmail.fm> - 2026-07-24 17:58 +0200
2008: 4 Blades + Tesla S1070 versus 2026: 1 AI Laptop (Re: Little Data Center on Your Palm [AI Laptops for 500 USD]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 18:16 +0200
Re: Little Data Center on Your Palm [AI Laptops for 500 USD] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto] Bradford Babkoff <ffb@odbb.ru> - 2026-07-24 18:05 +0000
LoL (Was: Little Data Center on Your Palm [AI Laptops for 500 USD]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:11 +0200
Budget AI Laptop 2026 versus Cray T3D 1995 (Re: Little Data Center on Your Palm [AI Laptops for 500 USD]) Mild Shock <janburse@fastmail.fm> - 2026-08-05 14:23 +0200
Re: Budget AI Laptop 2026 versus Cray T3D 1995 (Re: Little Data Center on Your Palm [AI Laptops for 500 USD]) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-08-05 13:16 -0700
Hurry the blue bus doesnt stop indefinitely (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:36 +0200
Not SIMD, a MIMD design for NVIDIA Volta (Re: Hurry the blue bus doesnt stop indefinitely) Mild Shock <janburse@fastmail.fm> - 2026-07-24 20:57 +0200
Could take 3-4 months find machine / browser (Was Not SIMD, a MIMD design for NVIDIA Volta) Mild Shock <janburse@fastmail.fm> - 2026-07-24 21:15 +0200
The Koan of pi-WAM queues [FORTRAN-S] (Was: Not SIMD, a MIMD design for NVIDIA Volta) Mild Shock <janburse@fastmail.fm> - 2026-07-26 19:52 +0200
The turbo capping of AI Laptops (Re: The Koan of pi-WAM queues [FORTRAN-S]) Mild Shock <janburse@fastmail.fm> - 2026-07-26 20:01 +0200
Re: The Koan of pi-WAM queues [FORTRAN-S] (Was: Not SIMD, a MIMD design for NVIDIA Volta) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-26 20:33 -0700
Why forget Bulgarians, never on my mind (Re: The Koan of pi-WAM queues [FORTRAN-S] (Was: Not SIMD, a MIMD design for NVIDIA Volta) Mild Shock <janburse@fastmail.fm> - 2026-07-27 09:14 +0200
miniTriton CUDA is an alternative to torch variants (Was: Why forget Bulgarians, never on my mind) Mild Shock <janburse@fastmail.fm> - 2026-07-27 09:40 +0200
Andrej Karpathy original gangster of Budget Laptop (Was: miniTriton CUDA is an alternative to torch variants) Mild Shock <janburse@fastmail.fm> - 2026-07-27 09:51 +0200
Re: Andrej Karpathy original gangster of Budget Laptop (Was: miniTriton CUDA is an alternative to torch variants) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-27 01:41 -0700
Re: Why forget Bulgarians, never on my mind (Re: The Koan of pi-WAM queues [FORTRAN-S] (Was: Not SIMD, a MIMD design for NVIDIA Volta) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-27 01:38 -0700
The evolution of hardware and GPT-2 training (Was: Why forget Bulgarians, never on my mind) Mild Shock <janburse@fastmail.fm> - 2026-07-27 10:56 +0200
How speed up π-WAM with vector operations (Was: The evolution of hardware and GPT-2 training) Mild Shock <janburse@fastmail.fm> - 2026-07-27 11:08 +0200
AI accelerator extend from GPU to CPU [Zero Copying] (Was: How speed up π-WAM with vector operations) Mild Shock <janburse@fastmail.fm> - 2026-07-27 11:21 +0200
The invention of vector and matrix registers [NVIDIA Volta] (Was: AI accelerator extend from GPU to CPU [Zero Copying]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 13:20 +0200
Maybe they should have named it NVIDIA Einstein [Rossy Boy Toe Sucking] (Re: The invention of vector and matrix registers [NVIDIA Volta] (Was: AI accelerator extend from GPU to CPU [Zero Copying]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 17:14 +0200
π-WAM is not adding decimals, it is removing decimals (Re: The invention of vector and matrix registers [NVIDIA Volta]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 18:36 +0200
In Budget Laptops the TOPS come with low energy footprint (Was: π-WAM is not adding decimals, it is removing decimals) Mild Shock <janburse@fastmail.fm> - 2026-07-27 18:44 +0200
Java picky concerning JIT-ing [Luckier with C++/C or FORTRAN compilers?] (Re: The Koan of pi-WAM queues [FORTRAN-S]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 12:55 +0200
Potato Computer owner impressed by Ukraine Tech [Rossy Boys Brother?] (Re: Hurry the blue bus doesnt stop indefinitely) Mild Shock <janburse@fastmail.fm> - 2026-07-27 16:57 +0200
Got it. Or are you too stupid? [New Usenet Mantra] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-27 19:00 +0200
Re: Got it. Or are you too stupid? [New Usenet Mantra] (Re: The Wuhan Virus that destroyed Python [ggml Manifesto]) Kim Baitchorov <bvhkoc@bmc.ru> - 2026-07-27 22:38 +0000
Clueless about MIMD as usual [Flynn's Taxonomy] (Was: Rossy Boy is neither Einstein nor Zweistein) Mild Shock <janburse@fastmail.fm> - 2026-07-28 11:29 +0200
confused rossy boy is confused (Re: Clueless about MIMD as usual [Flynn's Taxonomy]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 11:20 +0200
Gemini, DeepSeek, OpenAI more clever than rossy boy (Re: confused rossy boy is confused) Mild Shock <janburse@fastmail.fm> - 2026-07-29 11:21 +0200
In AI Acceleration nobody cares about CivetWeb (Re: Gemini, DeepSeek, OpenAI more clever than rossy boy) Mild Shock <janburse@fastmail.fm> - 2026-07-29 11:29 +0200
Run with minimum HTTPS and .mjs type (Re: In AI Acceleration nobody cares about CivetWeb) Mild Shock <janburse@fastmail.fm> - 2026-07-29 11:49 +0200
Your strictness is your problem , not mine [See WebLLM] (Re: Run with minimum HTTPS and .mjs type) Mild Shock <janburse@fastmail.fm> - 2026-07-29 11:51 +0200
Re: In AI Acceleration nobody cares about CivetWeb (Re: Gemini, DeepSeek, OpenAI more clever than rossy boy) Lane W <cactus_DAC@yahoo.com> - 2026-07-29 07:06 -0600
Re: In AI Acceleration nobody cares about CivetWeb (Re: Gemini, DeepSeek, OpenAI more clever than rossy boy) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-29 07:20 -0700
You are still chewing on SIMD. LoL (Re: Clueless about MIMD as usual [Flynn's Taxonomy]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:14 +0200
Hurry Rossy Boy, the blue bus is waiting (Re: You are still chewing on SIMD. LoL) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:54 +0200
Look how they advertized CUDA and logical threads (Re: Hurry Rossy Boy, the blue bus is waiting) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:56 +0200
Forget any arithmetization of product FSA (Re: Look how they advertized CUDA and logical threads) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:57 +0200
Re: You are still chewing on SIMD. LoL (Re: Clueless about MIMD as usual [Flynn's Taxonomy]) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-29 10:48 -0700
Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator] (Was: You are still chewing on SIMD. LoL) Mild Shock <janburse@fastmail.fm> - 2026-07-29 20:03 +0200
I don't use Rust, you are crazy [Jump off a bridge, idiot] (Re: Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 20:24 +0200
Re: I don't use Rust, you are crazy [Jump off a bridge, idiot] (Re: Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator]) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-29 12:27 -0700
Re: I don't use Rust, you are crazy [Jump off a bridge, idiot] (Re: Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator]) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-07-29 13:40 -0700
Re: I don't use Rust, you are crazy [Jump off a bridge, idiot] (Re: Rossy Boys tears could cool a data center [pi-WAM Interleaved Synchronized Emulator]) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-08-03 17:05 -0700
Do a YouTube video about it (Was: I don't use Rust, you are crazy) Mild Shock <janburse@fastmail.fm> - 2026-08-04 03:00 +0200
Standing on the shoulders of giants (Was: Do a YouTube video about it) Mild Shock <janburse@fastmail.fm> - 2026-08-04 03:16 +0200
Re: Standing on the shoulders of giants (Was: Do a YouTube video about it) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-08-04 06:10 -0700
You Thief! Stealing Szemeredi, Aristotle, Leibniz, etc.. (Re: Standing on the shoulders of giants) Mild Shock <janburse@fastmail.fm> - 2026-08-04 15:16 +0200
Re: You Thief! Stealing Szemeredi, Aristotle, Leibniz, etc.. (Re: Standing on the shoulders of giants) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-08-04 06:32 -0700
How Rossy Boys plagiarism works [Copy Paste Slop] (Was: You Thief! Stealing Szemeredi, Aristotle, Leibniz, etc..) Mild Shock <janburse@fastmail.fm> - 2026-08-04 17:54 +0200
tatistics gave up, no salient truth [Signal Collapse] (Re: How Rossy Boys plagiarism works [Copy Paste Slop]) Mild Shock <janburse@fastmail.fm> - 2026-08-04 18:21 +0200
Statistics gave up, no salient truth [Signal Collapse] (Was: How Rossy Boys plagiarism works [Copy Paste Slop]) Mild Shock <janburse@fastmail.fm> - 2026-08-04 18:23 +0200
Re: Statistics gave up, no salient truth [Signal Collapse] (Was: How Rossy Boys plagiarism works [Copy Paste Slop]) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-08-04 20:13 -0700
Re: Statistics gave up, no salient truth [Signal Collapse] (Was: How Rossy Boys plagiarism works [Copy Paste Slop]) "Chris M. Thomasson" <chris.m.thomasson.1@gmail.com> - 2026-08-05 13:08 -0700
Hack ecosystem ignorance paired with paranoia [Nand to Tetris] (Re: Rossy Boys tears could cool a data center) Mild Shock <janburse@fastmail.fm> - 2026-07-29 22:51 +0200
A funny Q16.16 experiment with Hack (Re: Hack ecosystem ignorance paired with paranoia [Nand to Tetris]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 23:12 +0200
Re: A funny Q16.16 experiment with Hack (Re: Hack ecosystem ignorance paired with paranoia [Nand to Tetris]) Alexey Bessonov <nexes@ebxsvs.ru> - 2026-07-29 21:24 +0000
Summer Challenge: libSQL = Prolog+Modes [VDBE versus π-WAM] (Re: A funny Q16.16 experiment with Hack) Mild Shock <janburse@fastmail.fm> - 2026-07-30 11:26 +0200
I wrote Hack VM for π-WAM from scratch [4 Months total JavaScript, Python and Java] (Re: Hack ecosystem ignorance paired with paranoia [Nand to Tetris]) Mild Shock <janburse@fastmail.fm> - 2026-07-30 19:35 +0200
Re: I wrote Hack VM for π-WAM from scratch [4 Months total JavaScript, Python and Java] (Re: Hack ecosystem ignorance paired with paranoia [Nand to Tetris]) Mild Shock <janburse@fastmail.fm> - 2026-07-30 19:49 +0200
For WebGPU I first had SIMD in mind (Re: I wrote Hack VM for π-WAM from scratch) Mild Shock <janburse@fastmail.fm> - 2026-07-30 19:50 +0200
Corr.: 4 Months --> 4 Weeks (Re: For WebGPU I first had SIMD in mind) Mild Shock <janburse@fastmail.fm> - 2026-07-30 20:05 +0200
MIPS is a big Huffman mess [But Hack could do it] (Re: I wrote Hack VM for π-WAM from scratch) Mild Shock <janburse@fastmail.fm> - 2026-07-30 22:30 +0200
Not declarative with PHI (Φ) nodes (Re: MIPS is a big Huffman mess [But Hack could do it]) Mild Shock <janburse@fastmail.fm> - 2026-07-30 22:42 +0200
Quo Vadis: Extend investigations to WebNN (Re: I wrote Hack VM for π-WAM from scratch) Mild Shock <janburse@fastmail.fm> - 2026-07-31 20:45 +0200
Re: Quo Vadis: Extend investigations to WebNN (Re: I wrote Hack VM for π-WAM from scratch) Jereb Pohlebaev <obje@bbvoeoaa.ru> - 2026-07-31 20:23 +0000
Lamas in a cradle and Lamas on the edge [Red Pyjama] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 13:02 +0200
AI Accelerators and ISO Prolog multi-threading (Was: Lamas in a cradle and Lamas on the edge [Red Pyjama]) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:01 +0200
Actor/Erlang is dead, no Thread and Mailbox conflation [golang channels] (Was: AI Accelerators and ISO Prolog multi-threading) Mild Shock <janburse@fastmail.fm> - 2026-07-29 17:05 +0200
Can library(ironpaw) repurpose FFT hardware [Glimps into Ryzen AI 7 350] (Was: Actor/Erlang is dead, no Thread and Mailbox conflation) Mild Shock <janburse@fastmail.fm> - 2026-08-01 02:29 +0200
Re: Can library(ironpaw) repurpose FFT hardware [Glimps into Ryzen AI 7 350] Thomas 'PointedEars' Lahn <PointedEars@web.de> - 2026-08-01 04:10 +0200
Tablet and phone UBS-C remote debugging (Re: Lamas in a cradle and Lamas on the edge [Red Pyjama]) Mild Shock <janburse@fastmail.fm> - 2026-08-01 12:17 +0200
NPUs doing 2d chess comms (Manhattan Distance or L1 Norm) (Was: Tablet and phone UBS-C remote debugging) Mild Shock <janburse@fastmail.fm> - 2026-08-01 14:09 +0200
Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) Mild Shock <janburse@fastmail.fm> - 2026-08-02 02:47 +0200
npm install webgpu [Google Dawn] (Was: Chris M. Thomasson can ask 100 more questions) Mild Shock <janburse@fastmail.fm> - 2026-08-02 03:01 +0200
GPU elasticity was already invented in 2008 with CUDA (Re: npm install webgpu [Google Dawn]) Mild Shock <janburse@fastmail.fm> - 2026-08-02 23:59 +0200
Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) Johann 'Myrkraverk' Oskarsson <johann@myrkraverk.invalid> - 2026-08-03 00:34 +0800
Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-08-02 10:41 -0700
Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) Ross Finlayson <ross.a.finlayson@gmail.com> - 2026-08-02 11:13 -0700
Pro Gauss-Jordan Reduction, Phigs and Phigs+ (was: Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging)) Johann 'Myrkraverk' Oskarsson <johann@myrkraverk.invalid> - 2026-08-03 02:27 +0800
Re: Pro Gauss-Jordan Reduction, Phigs and Phigs+ Johann 'Myrkraverk' Oskarsson <johann@myrkraverk.invalid> - 2026-08-03 02:35 +0800
Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) "Chris M. Thomasson" <chris.m.thomasson.1@gmail.com> - 2026-08-02 12:42 -0700
You posted that already, but you didn't listen [I NEED BOUNDED QUEUES] (Was: Chris M. Thomasson can ask 100 more questions) Mild Shock <janburse@fastmail.fm> - 2026-08-02 23:09 +0200
Summary of 100 questions Chris M. Thomasson can ask (Was: You posted that already, but you didn't listen [I NEED BOUNDED QUEUES]) Mild Shock <janburse@fastmail.fm> - 2026-08-02 23:17 +0200
Summary of 100 questions Chris M. Thomasson can ask (Re: You posted that already, but you didn't listen) Mild Shock <janburse@fastmail.fm> - 2026-08-02 23:20 +0200
Re: You posted that already, but you didn't listen [I NEED BOUNDED QUEUES] (Was: Chris M. Thomasson can ask 100 more questions) "Chris M. Thomasson" <chris.m.thomasson.1@gmail.com> - 2026-08-03 12:55 -0700
Liar and spammer (Was: You posted that already, but you didn't listen [I NEED BOUNDED QUEUES]) Mild Shock <janburse@fastmail.fm> - 2026-08-03 22:16 +0200
Re: Liar and spammer (Was: You posted that already, but you didn't listen [I NEED BOUNDED QUEUES]) Rosalino Kablahov <oaror@vla.ru> - 2026-08-03 21:50 +0000
Even send_color and recv_color can block [Cerebras Waver] Re: The Wuhan Virus that destroyed Python [ggml Manifesto] Mild Shock <janburse@fastmail.fm> - 2026-08-02 23:33 +0200
Page 6 of 7 — ← Prev page 1 2 3 4 5 [6] 7 Next page →
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-30 22:42 +0200 |
| Subject | Not declarative with PHI (Φ) nodes (Re: MIPS is a big Huffman mess [But Hack could do it]) |
| Message-ID | <114gcut$mn2q$4@solani.org> |
| In reply to | #647062 |
Hi, Another choice for naming Hack, would be to call it an intermediate format. But this is typically used here: The intermediate representation, or IR for short, is an in-memory data structure that represents executable code. https://www.llvmpy.org/llvmpy-doc/dev/doc/llvm_concepts.html#ssa-form-and-phi-nodes So I still like the term abstract machine, as already used in the past by David H. D. Warren for the famous, and in my opinion infamous: Warren Abstract Machine 1983 https://en.wikipedia.org/wiki/Warren_Abstract_Machine Maybe you can take the term abstract machine as a hint that it is more lower level, and more imperative. Not something highlevel, that is easily malleable. But abstract also captures the notion that there is still a level further down, making it concrete. And you find many Prolog systems that did just that, they compile WAM into a further instruction stream, like x86 or whatever, for binary compiled code, that is not interpreted WAM. Bye Mild Shock schrieb: > Hi, > > There is Prolog compiler which spits out Hack. > From there on your are free to develop > and/or use any Hack realization that goes > > from abstract to concrete. You could > replace the CPU backends that realize > a Hack VM by MIPS. Shouldn't be difficult. > > Basically I refused to think in Huffman > Coding (*) while designing Hack VM. On the > other hand the MIPS architecture looks > > like a big Huffman mess. Already its > initial design has 3 instructions types: > > Type format (bits) > R opcode(6) rs(5) rt(5) rd(5) shamt(5) funct(6) > I opcode(6) rs(5) rt(5) imme(16) > J opcode(6) addr(26) > > While my Hack has only 1 instruction > type, when binary encoded for Hack VM, > the currently used design looks as follows: > > Type format (bits) > AD opcode(4) mode(4) cond(4) imme(10) addr(10) > > But since its an abstract machine, nothing > prevents you from translating Hack code > into MIPS before executing it. > > In has far you have to distinguish Hack, > which is specified in Prolog. And Hack VM > which is a virtual machine, with the above > > instruction packing. And which has currently > a JavaScript runtime, a Python runtime > and a Java runtime. > > Bye > > (*) > https://en.wikipedia.org/wiki/Huffman_coding >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-31 20:45 +0200 |
| Subject | Quo Vadis: Extend investigations to WebNN (Re: I wrote Hack VM for π-WAM from scratch) |
| Message-ID | <114iqfe$obtb$2@solani.org> |
| In reply to | #647058 |
Hi, Since we have a good flow, and since NPUs share the same system memory, and possibly a lot of other traits as well with the GPU in libary(edge/furryhaze), we just developed. The idea here is to do first some off Dogelog experiments and then create a library that provides npu_exec/2 for pi-WAM code, the analogue to gpu_exec/2. A name suggestion would be: - edge/ironpaw.p The new Prolog library The NPU will be clearly underutilized when only doing scalar, not sure whether this is even permitted. But in the long run it is planned that pi-WAM will have vector and matrix traits anyways. Here is an example goal can be run with matrix and quantization traits: ?- [X,Y] ins 0..3, Z is X*2+Y*3+4, T is X*3-Y*2-1 These traits will demand some CPU, GPU and NPU translation. If we keep these traits simple, we might indeed arrive at concrete realization from the same abstract machine LoL, ironpaw the little brother of ironfist. Bye Mild Shock schrieb: > Hi, > > > Just downloading some other person's code > > I didn't do that, I wrote Hack VM for pi-WAM > from scratch, over the last 4 weeks. I came > back from holidays on end of June 2026, and now > > we have end of July 2026. But its only possible > because the instruction set is very smal, like > ca. 8 functions and ca. 8 modes and ca. 8 conditions, > > so its ca. 8 x 8 x 8 = 512 opcodes, each has an > A parameter and a D parameter simultaneously. > It has currently the following CPU backends: > > - Now supports interleaved synchronous emulation. > - Now supports warp parallelism via Java platform threads. > - Now supports warp parallelism via Python system threads. > - Now supports warp parallelism via JavaScript worker threads. > - Note: For Python free threads are not yet fully tested. > - Note: For JavaScript web workers are not yet fully tested. > > https://www.dogelog.ch/typtab/doclet/book/14_install/05_notes22/110_224.html > > > But frankly I came to encounter Hack not from > the usual university curriculum web resources, > but indirectly through a post about a Prolog > > emulation of Hack, using constrained horn clauses (CHC): > > Verifying Nand2Tetris Assembly > https://www.philipzucker.com/nand2tetris-chc/ > > The binary encoding is currently that the functions, > modes and conditions eat up a nibble (4-bit), in > total 12-bit, which I use then 10-bit for A parameter > > and 10-bit for D parameter. I used AI freemium, Codex > by ChatGPT from within IntelliJ to do some fragment > code translations automatically from Java to JavaScript > > or from JavaScript to Python. > > Have Fun! > > Bye > > Mild Shock schrieb: >> Hi, >> >> > Who exactly is the thief? Does this person >> > have stats in the Rogue class in dungeons >> > and dragons? >> >> The conspiracy theory of a stealing of Torso VDBE, >> by Rossy Boy, is probably a result of complete >> ignorance of the Hack ecosystem. >> >> Hack is a very popular computer science project, >> with a couple of subprojects in hardware and >> software. It goes also by the name Nand to Tetris, >> >> and is programming language agnositic. You can do >> Hack experiments in any programming language, be >> it BASIC, ADA or Rust. Nobody cares. >> >> The gist are projects like here, first to >> educate yourself about Hack: >> >> https://www.nand2tetris.org/course >> >> And then to use Hack in different contexts: >> >> https://www.nand2tetris.org/copy-of-talks >> >> For didactic purposes, I used Hack for my WebGPU >> experiment. I didn't even take a look at Torso >> VDBE, why should I? Hack is nicely documented, >> >> has even a book, and fusing the two 16-bit >> instruction types A and D, into a single 32-bit >> instruction stream, is nowhere patented. >> >> Bye >> >> Mild Shock schrieb: >>> Hi, >>> >>> Rossy Boys tears could cool a data center, >>> he thinks there exists no literature about >>> serial algorithms of parallel stuff, and >>> >>> he also thinks normal forms lead to optimizing >>> something. LoL, what a utter bullshit. I did >>> alreay a serial implementation of a parallel >>> >>> simulation of my pi-WAM. Just lookup the literature >>> about pi-calulus. I published it a few days ago, >>> its part of 2.2.4 released already: >>> >>> Parallel π-WAM: An Interleaved Synchronous Emulator >>> https://medium.com/2989/0196089e143a >>> >>> Whats your point, Rossy Boy? Except you post pretend >>> nonsense not knowing what you are doing? >>> >>> Bye >>> >>> Ross Finlayson schrieb: >>>> No, troll, these are serial algorithms their optimized forms. >>>> >>>> Normal sorts of forms, .... >>>> >>>> >>>> Yeah, everybody already figured out "interpreters" and >>>> "programs" and "spawning". >>>> >>>> Go spawn yourself. >>>> >>>> >>> >> >
[toc] | [prev] | [next] | [standalone]
| From | Jereb Pohlebaev <obje@bbvoeoaa.ru> |
|---|---|
| Date | 2026-07-31 20:23 +0000 |
| Subject | Re: Quo Vadis: Extend investigations to WebNN (Re: I wrote Hack VM for π-WAM from scratch) |
| Message-ID | <114j08g$3mn0p$1@news.nntp4.net> |
| In reply to | #647064 |
Mild Shock wrote:
> The idea here is to do first some off Dogelog experiments and then
> create a library that provides npu_exec/2 for pi-WAM code, the analogue
> to gpu_exec/2. A name suggestion would be:
>
> - edge/ironpaw.p
hanoi(N) :-
move(N, left, center, right).
move(1, Source, Target, _) :-
format("Move disk from ~w to ~w~n", [Source, Target]).
move(N, Source, Target, Auxiliary) :-
N > 1,
M is N - 1,
move(M, Source, Auxiliary, Target),
move(1, Source, Target, _),
move(M, Auxiliary, Target, Source).
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 13:02 +0200 |
| Subject | Lamas in a cradle and Lamas on the edge [Red Pyjama] (Was: The Wuhan Virus that destroyed Python [ggml Manifesto]) |
| Message-ID | <114cmkf$jmhc$1@solani.org> |
| In reply to | #646933 |
Hi, Why does this Lama have a red pyjama. Oh, its a baby Lama. Its still in the cradle and needs some training: RedPajama-Data-v2 https://github.com/togethercomputer/RedPajama-Data But then Andrej Karpathy recently showed GPT-2 training on rented GPUs for less than 100 USD in less then 2 hours. So where do these grown up Lamas go. Well Georgi Gerganov prefered C++/C when he shouted Llama Llama Red Pyjama. But you also find WebLLM, wrapping the underlying C++/C GPU interface via the W3C standard WebGPU / WGSL, with JavaScript: In-Browser LLM Inference Engine https://webllm.mlc.ai/ My experience with WebLLM 6 months ago on an iPad Pro 2024, still a little early stage performance and robustness. But hey hardware of AI mobile iGPUs is still evolving, and AI laptop, AI smartphones and AI tablets, will soon feature Chinese hardware such some new Kirin AI in 2027. Bye Mild Shock schrieb: > Hi, > > Remember when first all local AI was Python > and PyTorch APIs. And then suddently people strated > using bare metal C/C++ Code. Here is the story: > > How it started: > > GPT-J or GPT-J-6B is an open-source large > language model (LLM) developed by EleutherAI > in 2021. As the name suggests, it is a > generative pre-trained transformer model > designed to produce human-like text that > continues from a prompt. > https://www.eleuther.ai/ > > How it was going [Georgi Gerganov]: > > So a few days later comes out the LLaMA, I do > some calculations and I figure out “Okay, 65 > billion parameters. You probably need about > 40 gigs of RAM, with 4-bit quantization. So > this can run on a MacBook. Why not do it?” > > Why I was able to do it so quickly - basically, > for all that I saw it’s pretty much GPT-J architecture > with some modifications, like some extra memorization > layers. It’s minor changes. Basically, again, the > existing code for the GPT-J, I just simply > modified it there, it happened pretty quickly. > https://changelog.com/podcast/532 > > Georgi Gerganov, Bulgarian, now with Hugging > Face, ggml-cann also running on Chinese AI chips. > ggml Manifesto https://github.com/ggml-org/ggml > > Bye >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:01 +0200 |
| Subject | AI Accelerators and ISO Prolog multi-threading (Was: Lamas in a cradle and Lamas on the edge [Red Pyjama]) |
| Message-ID | <114d4kg$kfj1$1@solani.org> |
| In reply to | #647036 |
Hi, Usual question: > Why implement both pre-emptive threading AND cooperative tasks/engines? I had implemented the ISO proposal in formerly Jekejeke Prolog, you find the ISO proposal here: ISO/IEC DTR 13211–5:2007 Prolog multi-threading support https://logtalk.org/plstd/threads.pdf But the ISO proposal doesn't match modern WebGPU APIs, where your logical threads can live remotely in a dedicated GPU in the VRAM there, and where you would have launch parameters that say: Hey please run 4096 compute shaders for me, that have independet thread state. Using cooperative multi-tasking as the orchestrator works well. Bye Mild Shock schrieb: > Hi, > > Why does this Lama have a red pyjama. > Oh, its a baby Lama. Its still in the cradle > and needs some training: > > RedPajama-Data-v2 > https://github.com/togethercomputer/RedPajama-Data > > But then Andrej Karpathy recently showed > GPT-2 training on rented GPUs for less > than 100 USD in less then 2 hours. > > So where do these grown up Lamas go. > Well Georgi Gerganov prefered C++/C > when he shouted Llama Llama Red Pyjama. > > But you also find WebLLM, wrapping the > underlying C++/C GPU interface via the > W3C standard WebGPU / WGSL, with JavaScript: > > In-Browser LLM Inference Engine > https://webllm.mlc.ai/ > > My experience with WebLLM 6 months > ago on an iPad Pro 2024, still a little early > stage performance and robustness. > > But hey hardware of AI mobile iGPUs is > still evolving, and AI laptop, AI smartphones > and AI tablets, will soon feature Chinese > > hardware such some new Kirin AI in 2027. > > Bye > > Mild Shock schrieb: >> Hi, >> >> Remember when first all local AI was Python >> and PyTorch APIs. And then suddently people strated >> using bare metal C/C++ Code. Here is the story: >> >> How it started: >> >> GPT-J or GPT-J-6B is an open-source large >> language model (LLM) developed by EleutherAI >> in 2021. As the name suggests, it is a >> generative pre-trained transformer model >> designed to produce human-like text that >> continues from a prompt. >> https://www.eleuther.ai/ >> >> How it was going [Georgi Gerganov]: >> >> So a few days later comes out the LLaMA, I do >> some calculations and I figure out “Okay, 65 >> billion parameters. You probably need about >> 40 gigs of RAM, with 4-bit quantization. So >> this can run on a MacBook. Why not do it?” >> >> Why I was able to do it so quickly - basically, >> for all that I saw it’s pretty much GPT-J architecture >> with some modifications, like some extra memorization >> layers. It’s minor changes. Basically, again, the >> existing code for the GPT-J, I just simply >> modified it there, it happened pretty quickly. >> https://changelog.com/podcast/532 >> >> Georgi Gerganov, Bulgarian, now with Hugging >> Face, ggml-cann also running on Chinese AI chips. >> ggml Manifesto https://github.com/ggml-org/ggml >> >> Bye >> >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-07-29 17:05 +0200 |
| Subject | Actor/Erlang is dead, no Thread and Mailbox conflation [golang channels] (Was: AI Accelerators and ISO Prolog multi-threading) |
| Message-ID | <114d4rl$kfj1$4@solani.org> |
| In reply to | #647040 |
Hi, Mostlikely for high performance computing à la, the Actor/Erlang model is dead, they might rely on MPMC (Multiple Producer, Multiple Consumer) queue entities separate from the threads. The ISO Prolog multi-threading support had also such threads. But besides that was also Actor/Erlang leaning in practice, like SWI, where threads have some default queues. So an actor is basically a Thread and Mailbox conflation. While a MPMC queue is a kind of separate Mailbox, where multiple "actors" can read from and write from. A kind of localized Linda Tuple store. Which Programming language did adopted the non-Actor pi-calculus model? Right golang with its channels. Bye Mild Shock schrieb: > Hi, > > Usual question: > > > Why implement both pre-emptive threading > AND cooperative tasks/engines? > > I had implemented the ISO proposal in formerly Jekejeke > Prolog, you find the ISO proposal here: > > ISO/IEC DTR 13211–5:2007 > Prolog multi-threading support > https://logtalk.org/plstd/threads.pdf > > But the ISO proposal doesn't match modern WebGPU APIs, > where your logical threads can live remotely in a dedicated GPU > in the VRAM there, and where you would have launch > > parameters that say: Hey please run 4096 compute > shaders for me, that have independet thread state. Using > cooperative multi-tasking as the orchestrator works well. > > Bye > > Mild Shock schrieb: >> Hi, >> >> Why does this Lama have a red pyjama. >> Oh, its a baby Lama. Its still in the cradle >> and needs some training: >> >> RedPajama-Data-v2 >> https://github.com/togethercomputer/RedPajama-Data >> >> But then Andrej Karpathy recently showed >> GPT-2 training on rented GPUs for less >> than 100 USD in less then 2 hours. >> >> So where do these grown up Lamas go. >> Well Georgi Gerganov prefered C++/C >> when he shouted Llama Llama Red Pyjama. >> >> But you also find WebLLM, wrapping the >> underlying C++/C GPU interface via the >> W3C standard WebGPU / WGSL, with JavaScript: >> >> In-Browser LLM Inference Engine >> https://webllm.mlc.ai/ >> >> My experience with WebLLM 6 months >> ago on an iPad Pro 2024, still a little early >> stage performance and robustness. >> >> But hey hardware of AI mobile iGPUs is >> still evolving, and AI laptop, AI smartphones >> and AI tablets, will soon feature Chinese >> >> hardware such some new Kirin AI in 2027. >> >> Bye >> >> Mild Shock schrieb: >>> Hi, >>> >>> Remember when first all local AI was Python >>> and PyTorch APIs. And then suddently people strated >>> using bare metal C/C++ Code. Here is the story: >>> >>> How it started: >>> >>> GPT-J or GPT-J-6B is an open-source large >>> language model (LLM) developed by EleutherAI >>> in 2021. As the name suggests, it is a >>> generative pre-trained transformer model >>> designed to produce human-like text that >>> continues from a prompt. >>> https://www.eleuther.ai/ >>> >>> How it was going [Georgi Gerganov]: >>> >>> So a few days later comes out the LLaMA, I do >>> some calculations and I figure out “Okay, 65 >>> billion parameters. You probably need about >>> 40 gigs of RAM, with 4-bit quantization. So >>> this can run on a MacBook. Why not do it?” >>> >>> Why I was able to do it so quickly - basically, >>> for all that I saw it’s pretty much GPT-J architecture >>> with some modifications, like some extra memorization >>> layers. It’s minor changes. Basically, again, the >>> existing code for the GPT-J, I just simply >>> modified it there, it happened pretty quickly. >>> https://changelog.com/podcast/532 >>> >>> Georgi Gerganov, Bulgarian, now with Hugging >>> Face, ggml-cann also running on Chinese AI chips. >>> ggml Manifesto https://github.com/ggml-org/ggml >>> >>> Bye >>> >> >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-08-01 02:29 +0200 |
| Subject | Can library(ironpaw) repurpose FFT hardware [Glimps into Ryzen AI 7 350] (Was: Actor/Erlang is dead, no Thread and Mailbox conflation) |
| Message-ID | <114jem4$o8qj$1@solani.org> |
| In reply to | #647041 |
Hi, On could believe the AI boom is a kind of Charles Darvin Galapagos Island Evolution Trick of repurposing FFT hardware. But this is of course not true, HPC, high performance computing, has already defined level 3 ops years ago. But look at this rabit hole of Ryzen AI 7 350 NPU design, which is a stripped down Xilinx, stripped of exotic FFT features: Getting peak TOPS on a Ryzen AI 7 350 NPU https://destevez.net/2026/05/getting-peak-tops-on-a-ryzen-ai-7-350-npu/ But the core feature, very long instruction word (VLIW) engines, with hardware accelerated GEMMs, scattered in grids of ASIC tiles, connected by DMA and NoC, is even not very specific to AMD, you find it also in Snapdragon / Qualcomm SoCs for AI Laptops. Bye P.S.: My brain playing tricks, why should I name a library(ironpaw) ? From the same article above. Maybe WebNN is easier to use? "mlir-aie contains a Python framework called IRON that generates LLVM MLIR code representing a workload that runs on the NPU, including the code that runs on each compute tile processor and the configuration of DMAs and other hardware. Kernels for the compute tile processor can be written in C++ and compiled either with the open-source llvm-aie Peano compiler, which is a fork of LLVM that adds support for the Xilinx AI engine processors, or with the closed-source Xilinx CHESS compiler, which is included in Vitis. In simple cases the kernels can also be directly written in Python with IRON." Getting peak TOPS on a Ryzen AI 7 350 NPU https://destevez.net/2026/05/getting-peak-tops-on-a-ryzen-ai-7-350-npu/ Mild Shock schrieb: > Hi, > > Mostlikely for high performance computing à la, > the Actor/Erlang model is dead, they might rely > on MPMC (Multiple Producer, Multiple Consumer) > > queue entities separate from the threads. The > ISO Prolog multi-threading support had also such > threads. But besides that was also Actor/Erlang > > leaning in practice, like SWI, where threads > have some default queues. So an actor is basically > a Thread and Mailbox conflation. While a MPMC queue > > is a kind of separate Mailbox, where multiple > "actors" can read from and write from. A kind of > localized Linda Tuple store. > > Which Programming language did adopted the > non-Actor pi-calculus model? Right golang > with its channels. > > Bye > > Mild Shock schrieb: >> Hi, >> >> Usual question: >> >> > Why implement both pre-emptive threading >> AND cooperative tasks/engines? >> >> I had implemented the ISO proposal in formerly Jekejeke >> Prolog, you find the ISO proposal here: >> >> ISO/IEC DTR 13211–5:2007 >> Prolog multi-threading support >> https://logtalk.org/plstd/threads.pdf >> >> But the ISO proposal doesn't match modern WebGPU APIs, >> where your logical threads can live remotely in a dedicated GPU >> in the VRAM there, and where you would have launch >> >> parameters that say: Hey please run 4096 compute >> shaders for me, that have independet thread state. Using >> cooperative multi-tasking as the orchestrator works well. >> >> Bye >> >> Mild Shock schrieb: >>> Hi, >>> >>> Why does this Lama have a red pyjama. >>> Oh, its a baby Lama. Its still in the cradle >>> and needs some training: >>> >>> RedPajama-Data-v2 >>> https://github.com/togethercomputer/RedPajama-Data >>> >>> But then Andrej Karpathy recently showed >>> GPT-2 training on rented GPUs for less >>> than 100 USD in less then 2 hours. >>> >>> So where do these grown up Lamas go. >>> Well Georgi Gerganov prefered C++/C >>> when he shouted Llama Llama Red Pyjama. >>> >>> But you also find WebLLM, wrapping the >>> underlying C++/C GPU interface via the >>> W3C standard WebGPU / WGSL, with JavaScript: >>> >>> In-Browser LLM Inference Engine >>> https://webllm.mlc.ai/ >>> >>> My experience with WebLLM 6 months >>> ago on an iPad Pro 2024, still a little early >>> stage performance and robustness. >>> >>> But hey hardware of AI mobile iGPUs is >>> still evolving, and AI laptop, AI smartphones >>> and AI tablets, will soon feature Chinese >>> >>> hardware such some new Kirin AI in 2027. >>> >>> Bye >>> >>> Mild Shock schrieb: >>>> Hi, >>>> >>>> Remember when first all local AI was Python >>>> and PyTorch APIs. And then suddently people strated >>>> using bare metal C/C++ Code. Here is the story: >>>> >>>> How it started: >>>> >>>> GPT-J or GPT-J-6B is an open-source large >>>> language model (LLM) developed by EleutherAI >>>> in 2021. As the name suggests, it is a >>>> generative pre-trained transformer model >>>> designed to produce human-like text that >>>> continues from a prompt. >>>> https://www.eleuther.ai/ >>>> >>>> How it was going [Georgi Gerganov]: >>>> >>>> So a few days later comes out the LLaMA, I do >>>> some calculations and I figure out “Okay, 65 >>>> billion parameters. You probably need about >>>> 40 gigs of RAM, with 4-bit quantization. So >>>> this can run on a MacBook. Why not do it?” >>>> >>>> Why I was able to do it so quickly - basically, >>>> for all that I saw it’s pretty much GPT-J architecture >>>> with some modifications, like some extra memorization >>>> layers. It’s minor changes. Basically, again, the >>>> existing code for the GPT-J, I just simply >>>> modified it there, it happened pretty quickly. >>>> https://changelog.com/podcast/532 >>>> >>>> Georgi Gerganov, Bulgarian, now with Hugging >>>> Face, ggml-cann also running on Chinese AI chips. >>>> ggml Manifesto https://github.com/ggml-org/ggml >>>> >>>> Bye >>>> >>> >> >
[toc] | [prev] | [next] | [standalone]
| From | Thomas 'PointedEars' Lahn <PointedEars@web.de> |
|---|---|
| Date | 2026-08-01 04:10 +0200 |
| Subject | Re: Can library(ironpaw) repurpose FFT hardware [Glimps into Ryzen AI 7 350] |
| Message-ID | <114jkjd$9c3u$1@gwaiyur.mb-net.net> |
| In reply to | #647070 |
Mild Shock wrote: ^^^^^^^^^^ Your real name belongs there. > On could believe the AI boom is a kind of > Charles Darvin Galapagos Island Evolution > Trick of repurposing FFT hardware. [...] What is the relation of this to the theories of relativity? If none, then stop crossposting to sci.physics.relativity (before someone makes you to). Also, the "(was: ...)" in the Subject must be written _lowercase_ if it is to be automatically removed by NetNews user agents like Thunderbird/Betterbird on Follow-up. F'up2 sci.physics.relativity -- PointedEars Twitter: @PointedEars2 Please do not cc me. / Bitte keine Kopien per E-Mail.
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-08-01 12:17 +0200 |
| Subject | Tablet and phone UBS-C remote debugging (Re: Lamas in a cradle and Lamas on the edge [Red Pyjama]) |
| Message-ID | <114kh3m$ovdp$3@solani.org> |
| In reply to | #647036 |
Hi, Tablets and phone are more annoying to use with WebGPU. The usual browsers don't have a Chrome DevTools panel integrated, so that one could do JavaScript Debugging directly on the device. Instead one has to use a desktop machine, and connect the device via UBS-C , and start a Chrome Browser there . And then start a Chrome DevTools panel alone, that is pair with the device, via UBS-C cable. So this way I already see where it crashes on the tablets and phone: await output.mapAsync(GPUMapMode.READ) Unhandled Promise Rejection: OperationError The above is the error that one can re-produce already here with this test: 11.4 Giga Lips with a Budget Laptop https://github.com/Jean-Luc-Picard-2021/gigabudget Not sure what exactly happens. Maybe a form of timeout or device lost, that the primitive HTML / JavaScript doesn't handle gracefully yet. Maybe redimensioning the test, so that it consumes less time would help. Who knows? Will see. For production use of a GPU integration I have to anyway provide work slicing it seems. Bye Mild Shock schrieb: > Hi, > > Why does this Lama have a red pyjama. > Oh, its a baby Lama. Its still in the cradle > and needs some training: > > RedPajama-Data-v2 > https://github.com/togethercomputer/RedPajama-Data > > But then Andrej Karpathy recently showed > GPT-2 training on rented GPUs for less > than 100 USD in less then 2 hours. > > So where do these grown up Lamas go. > Well Georgi Gerganov prefered C++/C > when he shouted Llama Llama Red Pyjama. > > But you also find WebLLM, wrapping the > underlying C++/C GPU interface via the > W3C standard WebGPU / WGSL, with JavaScript: > > In-Browser LLM Inference Engine > https://webllm.mlc.ai/ > > My experience with WebLLM 6 months > ago on an iPad Pro 2024, still a little early > stage performance and robustness. > > But hey hardware of AI mobile iGPUs is > still evolving, and AI laptop, AI smartphones > and AI tablets, will soon feature Chinese > > hardware such some new Kirin AI in 2027. > > Bye > > Mild Shock schrieb: >> Hi, >> >> Remember when first all local AI was Python >> and PyTorch APIs. And then suddently people strated >> using bare metal C/C++ Code. Here is the story: >> >> How it started: >> >> GPT-J or GPT-J-6B is an open-source large >> language model (LLM) developed by EleutherAI >> in 2021. As the name suggests, it is a >> generative pre-trained transformer model >> designed to produce human-like text that >> continues from a prompt. >> https://www.eleuther.ai/ >> >> How it was going [Georgi Gerganov]: >> >> So a few days later comes out the LLaMA, I do >> some calculations and I figure out “Okay, 65 >> billion parameters. You probably need about >> 40 gigs of RAM, with 4-bit quantization. So >> this can run on a MacBook. Why not do it?” >> >> Why I was able to do it so quickly - basically, >> for all that I saw it’s pretty much GPT-J architecture >> with some modifications, like some extra memorization >> layers. It’s minor changes. Basically, again, the >> existing code for the GPT-J, I just simply >> modified it there, it happened pretty quickly. >> https://changelog.com/podcast/532 >> >> Georgi Gerganov, Bulgarian, now with Hugging >> Face, ggml-cann also running on Chinese AI chips. >> ggml Manifesto https://github.com/ggml-org/ggml >> >> Bye >> >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-08-01 14:09 +0200 |
| Subject | NPUs doing 2d chess comms (Manhattan Distance or L1 Norm) (Was: Tablet and phone UBS-C remote debugging) |
| Message-ID | <114knku$pjus$1@solani.org> |
| In reply to | #647076 |
Hi,
Looking at the floor plan of a NPU:
Getting peak TOPS on a Ryzen AI 7 350 NPU
https://destevez.net/2026/05/getting-peak-tops-on-a-ryzen-ai-7-350-npu/
It seems to me comms between tiles takes
at least Manhattan Distance or L1 Norm time,
if there is no comms congestion
But how does a packet travel? This way:
+----E
|
|
S
Or this way, from start S to end E:
+-E
+
+
S
And what does the chip do if there is
traffic congestion? Some papers are
here, possibly an old problem giving
that processor "cubes" are nothing new.
But a "cube" would be 3D and not 2D.
This paper is old from 2007 or so:
Routing Algorithms for 2D NoC Architectures
http://cva.stanford.edu/classes/ee382c/research/2DRouting.pdf
Bye
Mild Shock schrieb:
> Hi,
>
> Tablets and phone are more annoying to
> use with WebGPU. The usual browsers don't
> have a Chrome DevTools panel integrated,
>
> so that one could do JavaScript Debugging
> directly on the device. Instead one has to
> use a desktop machine, and connect the
>
> device via UBS-C , and start a Chrome
> Browser there . And then start a Chrome
> DevTools panel alone, that is pair with
>
> the device, via UBS-C cable. So this way
> I already see where it crashes on the
> tablets and phone:
>
> await output.mapAsync(GPUMapMode.READ)
> Unhandled Promise Rejection: OperationError
>
> The above is the error that one can re-produce
> already here with this test:
>
> 11.4 Giga Lips with a Budget Laptop
> https://github.com/Jean-Luc-Picard-2021/gigabudget
>
> Not sure what exactly happens. Maybe
> a form of timeout or device lost, that the
> primitive HTML / JavaScript doesn't handle
>
> gracefully yet. Maybe redimensioning the
> test, so that it consumes less time would
> help. Who knows? Will see. For production
>
> use of a GPU integration I have to anyway
> provide work slicing it seems.
>
> Bye
>
> Mild Shock schrieb:
>> Hi,
>>
>> Why does this Lama have a red pyjama.
>> Oh, its a baby Lama. Its still in the cradle
>> and needs some training:
>>
>> RedPajama-Data-v2
>> https://github.com/togethercomputer/RedPajama-Data
>>
>> But then Andrej Karpathy recently showed
>> GPT-2 training on rented GPUs for less
>> than 100 USD in less then 2 hours.
>>
>> So where do these grown up Lamas go.
>> Well Georgi Gerganov prefered C++/C
>> when he shouted Llama Llama Red Pyjama.
>>
>> But you also find WebLLM, wrapping the
>> underlying C++/C GPU interface via the
>> W3C standard WebGPU / WGSL, with JavaScript:
>>
>> In-Browser LLM Inference Engine
>> https://webllm.mlc.ai/
>>
>> My experience with WebLLM 6 months
>> ago on an iPad Pro 2024, still a little early
>> stage performance and robustness.
>>
>> But hey hardware of AI mobile iGPUs is
>> still evolving, and AI laptop, AI smartphones
>> and AI tablets, will soon feature Chinese
>>
>> hardware such some new Kirin AI in 2027.
>>
>> Bye
>>
>> Mild Shock schrieb:
>>> Hi,
>>>
>>> Remember when first all local AI was Python
>>> and PyTorch APIs. And then suddently people strated
>>> using bare metal C/C++ Code. Here is the story:
>>>
>>> How it started:
>>>
>>> GPT-J or GPT-J-6B is an open-source large
>>> language model (LLM) developed by EleutherAI
>>> in 2021. As the name suggests, it is a
>>> generative pre-trained transformer model
>>> designed to produce human-like text that
>>> continues from a prompt.
>>> https://www.eleuther.ai/
>>>
>>> How it was going [Georgi Gerganov]:
>>>
>>> So a few days later comes out the LLaMA, I do
>>> some calculations and I figure out “Okay, 65
>>> billion parameters. You probably need about
>>> 40 gigs of RAM, with 4-bit quantization. So
>>> this can run on a MacBook. Why not do it?”
>>>
>>> Why I was able to do it so quickly - basically,
>>> for all that I saw it’s pretty much GPT-J architecture
>>> with some modifications, like some extra memorization
>>> layers. It’s minor changes. Basically, again, the
>>> existing code for the GPT-J, I just simply
>>> modified it there, it happened pretty quickly.
>>> https://changelog.com/podcast/532
>>>
>>> Georgi Gerganov, Bulgarian, now with Hugging
>>> Face, ggml-cann also running on Chinese AI chips.
>>> ggml Manifesto https://github.com/ggml-org/ggml
>>>
>>> Bye
>>>
>>
>
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-08-02 02:47 +0200 |
| Subject | Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) |
| Message-ID | <114m42n$qhhl$2@solani.org> |
| In reply to | #647076 |
Hi, Chris M. Thomasson can ask 100 more questions. I will happily answer them. But maybe I should make a Wiki to explain the ever same things: > But, I still don't know what you main goal is? The goal is "Prolog inferencing" > It has textures to work with in the pipeline. I don't need textures for "Prolog inferencing" 98 more questions to go, don't give up! Bye Mild Shock schrieb: > Hi, > > Tablets and phone are more annoying to > use with WebGPU. The usual browsers don't > have a Chrome DevTools panel integrated, > > so that one could do JavaScript Debugging > directly on the device. Instead one has to > use a desktop machine, and connect the > > device via UBS-C , and start a Chrome > Browser there . And then start a Chrome > DevTools panel alone, that is pair with > > the device, via UBS-C cable. So this way > I already see where it crashes on the > tablets and phone: > > await output.mapAsync(GPUMapMode.READ) > Unhandled Promise Rejection: OperationError > > The above is the error that one can re-produce > already here with this test: > > 11.4 Giga Lips with a Budget Laptop > https://github.com/Jean-Luc-Picard-2021/gigabudget > > Not sure what exactly happens. Maybe > a form of timeout or device lost, that the > primitive HTML / JavaScript doesn't handle > > gracefully yet. Maybe redimensioning the > test, so that it consumes less time would > help. Who knows? Will see. For production > > use of a GPU integration I have to anyway > provide work slicing it seems. > > Bye > > Mild Shock schrieb: >> Hi, >> >> Why does this Lama have a red pyjama. >> Oh, its a baby Lama. Its still in the cradle >> and needs some training: >> >> RedPajama-Data-v2 >> https://github.com/togethercomputer/RedPajama-Data >> >> But then Andrej Karpathy recently showed >> GPT-2 training on rented GPUs for less >> than 100 USD in less then 2 hours. >> >> So where do these grown up Lamas go. >> Well Georgi Gerganov prefered C++/C >> when he shouted Llama Llama Red Pyjama. >> >> But you also find WebLLM, wrapping the >> underlying C++/C GPU interface via the >> W3C standard WebGPU / WGSL, with JavaScript: >> >> In-Browser LLM Inference Engine >> https://webllm.mlc.ai/ >> >> My experience with WebLLM 6 months >> ago on an iPad Pro 2024, still a little early >> stage performance and robustness. >> >> But hey hardware of AI mobile iGPUs is >> still evolving, and AI laptop, AI smartphones >> and AI tablets, will soon feature Chinese >> >> hardware such some new Kirin AI in 2027. >> >> Bye >> >> Mild Shock schrieb: >>> Hi, >>> >>> Remember when first all local AI was Python >>> and PyTorch APIs. And then suddently people strated >>> using bare metal C/C++ Code. Here is the story: >>> >>> How it started: >>> >>> GPT-J or GPT-J-6B is an open-source large >>> language model (LLM) developed by EleutherAI >>> in 2021. As the name suggests, it is a >>> generative pre-trained transformer model >>> designed to produce human-like text that >>> continues from a prompt. >>> https://www.eleuther.ai/ >>> >>> How it was going [Georgi Gerganov]: >>> >>> So a few days later comes out the LLaMA, I do >>> some calculations and I figure out “Okay, 65 >>> billion parameters. You probably need about >>> 40 gigs of RAM, with 4-bit quantization. So >>> this can run on a MacBook. Why not do it?” >>> >>> Why I was able to do it so quickly - basically, >>> for all that I saw it’s pretty much GPT-J architecture >>> with some modifications, like some extra memorization >>> layers. It’s minor changes. Basically, again, the >>> existing code for the GPT-J, I just simply >>> modified it there, it happened pretty quickly. >>> https://changelog.com/podcast/532 >>> >>> Georgi Gerganov, Bulgarian, now with Hugging >>> Face, ggml-cann also running on Chinese AI chips. >>> ggml Manifesto https://github.com/ggml-org/ggml >>> >>> Bye >>> >> >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-08-02 03:01 +0200 |
| Subject | npm install webgpu [Google Dawn] (Was: Chris M. Thomasson can ask 100 more questions) |
| Message-ID | <114m4sp$qi3v$1@solani.org> |
| In reply to | #647100 |
Hi, Ok, following the instructions here: npm install webgpu https://github.com/dawn-gpu/node-webgpu I can now run webgpu also from CLI: >node.exe dogelog.mjs Dogelog Spieler 2.2.5, Node, JavaScript 26.4.0 (c) 1985-2026, XLOG Technologies AG, Schweiz ?- ensure_loaded(library(edge/furryhaze)). true. ?- between(1,3,_), time(expedite((between(1,100,_), between(1,100,_), between(1,100,_)), [size(4096)])), fail. % Zeit 1037.994 ms, GC 0.000 ms, Lips 111 k % Zeit 1091.131 ms, GC 0.000 ms, Lips 106 k % Zeit 1045.274 ms, GC 0.000 ms, Lips 110 k fail. Same benchmark result as in the browser. Now I can rent a bigger GPU by the hour and do some easy CLI testing. LoL Bye Mild Shock schrieb: > Hi, > > Chris M. Thomasson can ask 100 more questions. > I will happily answer them. But maybe I should > make a Wiki to explain the ever same things: > > > But, I still don't know what you main goal is? > The goal is "Prolog inferencing" > > > It has textures to work with in the pipeline. > I don't need textures for "Prolog inferencing" > > 98 more questions to go, don't give up! > > Bye > > Mild Shock schrieb: >> Hi, >> >> Tablets and phone are more annoying to >> use with WebGPU. The usual browsers don't >> have a Chrome DevTools panel integrated, >> >> so that one could do JavaScript Debugging >> directly on the device. Instead one has to >> use a desktop machine, and connect the >> >> device via UBS-C , and start a Chrome >> Browser there . And then start a Chrome >> DevTools panel alone, that is pair with >> >> the device, via UBS-C cable. So this way >> I already see where it crashes on the >> tablets and phone: >> >> await output.mapAsync(GPUMapMode.READ) >> Unhandled Promise Rejection: OperationError >> >> The above is the error that one can re-produce >> already here with this test: >> >> 11.4 Giga Lips with a Budget Laptop >> https://github.com/Jean-Luc-Picard-2021/gigabudget >> >> Not sure what exactly happens. Maybe >> a form of timeout or device lost, that the >> primitive HTML / JavaScript doesn't handle >> >> gracefully yet. Maybe redimensioning the >> test, so that it consumes less time would >> help. Who knows? Will see. For production >> >> use of a GPU integration I have to anyway >> provide work slicing it seems. >> >> Bye >> >> Mild Shock schrieb: >>> Hi, >>> >>> Why does this Lama have a red pyjama. >>> Oh, its a baby Lama. Its still in the cradle >>> and needs some training: >>> >>> RedPajama-Data-v2 >>> https://github.com/togethercomputer/RedPajama-Data >>> >>> But then Andrej Karpathy recently showed >>> GPT-2 training on rented GPUs for less >>> than 100 USD in less then 2 hours. >>> >>> So where do these grown up Lamas go. >>> Well Georgi Gerganov prefered C++/C >>> when he shouted Llama Llama Red Pyjama. >>> >>> But you also find WebLLM, wrapping the >>> underlying C++/C GPU interface via the >>> W3C standard WebGPU / WGSL, with JavaScript: >>> >>> In-Browser LLM Inference Engine >>> https://webllm.mlc.ai/ >>> >>> My experience with WebLLM 6 months >>> ago on an iPad Pro 2024, still a little early >>> stage performance and robustness. >>> >>> But hey hardware of AI mobile iGPUs is >>> still evolving, and AI laptop, AI smartphones >>> and AI tablets, will soon feature Chinese >>> >>> hardware such some new Kirin AI in 2027. >>> >>> Bye >>> >>> Mild Shock schrieb: >>>> Hi, >>>> >>>> Remember when first all local AI was Python >>>> and PyTorch APIs. And then suddently people strated >>>> using bare metal C/C++ Code. Here is the story: >>>> >>>> How it started: >>>> >>>> GPT-J or GPT-J-6B is an open-source large >>>> language model (LLM) developed by EleutherAI >>>> in 2021. As the name suggests, it is a >>>> generative pre-trained transformer model >>>> designed to produce human-like text that >>>> continues from a prompt. >>>> https://www.eleuther.ai/ >>>> >>>> How it was going [Georgi Gerganov]: >>>> >>>> So a few days later comes out the LLaMA, I do >>>> some calculations and I figure out “Okay, 65 >>>> billion parameters. You probably need about >>>> 40 gigs of RAM, with 4-bit quantization. So >>>> this can run on a MacBook. Why not do it?” >>>> >>>> Why I was able to do it so quickly - basically, >>>> for all that I saw it’s pretty much GPT-J architecture >>>> with some modifications, like some extra memorization >>>> layers. It’s minor changes. Basically, again, the >>>> existing code for the GPT-J, I just simply >>>> modified it there, it happened pretty quickly. >>>> https://changelog.com/podcast/532 >>>> >>>> Georgi Gerganov, Bulgarian, now with Hugging >>>> Face, ggml-cann also running on Chinese AI chips. >>>> ggml Manifesto https://github.com/ggml-org/ggml >>>> >>>> Bye >>>> >>> >> >
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-08-02 23:59 +0200 |
| Subject | GPU elasticity was already invented in 2008 with CUDA (Re: npm install webgpu [Google Dawn]) |
| Message-ID | <114oej8$rlg9$3@solani.org> |
| In reply to | #647101 |
Hi, Chris M. Thomasson schrieb: > Strive to never make a compute shader wait > on something, like an empty condition of a queue, stack. You are such a moron. GPU elasticity was already invented in 2008 with CUDA. I posted this quote already: "CUDA™ TEChNOLOGY UNLOCkS ThE POWER OF TESLA MANY-CORE PROCESSORS The CUDA C compiler simplifies many-core programming by enabling code development in a high-level language and optimizing code to run on systems without knowledge of how many cores are in the hardware. CUDA applications automatically take advantage of more cores or fewer cores in a system, so they can scale from entry-level notebook GPUs to high end GPUs in technical workstations and further into racks of GPUs in data centers. This allows developers to “code once” and deploy on a range of systems, as well as scale forward in time as future GPUs deliver more performance per watt and more cores per processor. The benefit for software users is the opportunity to boost computing performance simply by adding GPUs or using their existing GPUs in new ways" https://www.nvidia.com/docs/IO/43395/NV_DS_Tesla_S1070_US_Jun08_NV_LR_Final.pdf Today elasticity is on logical thread aka task level, not only on "core" level or something. Don't know exactly what CUDA did back them, maybe only a submit elasticity, like a time sharing system. Today you have quite some run elasticity on modern machines, for your logical threads. Even in budget laptops like a Ryzen AI 7 350 /w Radeon 850M. Bye P.S.: I can demostrate the elasticity, but I didn't write the medium.com article yet. Mild Shock schrieb: > Hi, > > Ok, following the instructions here: > > npm install webgpu > https://github.com/dawn-gpu/node-webgpu > > I can now run webgpu also from CLI: > > >node.exe dogelog.mjs > Dogelog Spieler 2.2.5, Node, JavaScript 26.4.0 > (c) 1985-2026, XLOG Technologies AG, Schweiz > > ?- ensure_loaded(library(edge/furryhaze)). > true. > > ?- between(1,3,_), time(expedite((between(1,100,_), > between(1,100,_), between(1,100,_)), [size(4096)])), fail. > % Zeit 1037.994 ms, GC 0.000 ms, Lips 111 k > % Zeit 1091.131 ms, GC 0.000 ms, Lips 106 k > % Zeit 1045.274 ms, GC 0.000 ms, Lips 110 k > fail. > > Same benchmark result as in the browser. > Now I can rent a bigger GPU by the hour > and do some easy CLI testing. > > LoL > > Bye > > Mild Shock schrieb: >> Hi, >> >> Chris M. Thomasson can ask 100 more questions. >> I will happily answer them. But maybe I should >> make a Wiki to explain the ever same things: >> >> > But, I still don't know what you main goal is? >> The goal is "Prolog inferencing" >> >> > It has textures to work with in the pipeline. >> I don't need textures for "Prolog inferencing" >> >> 98 more questions to go, don't give up! >> >> Bye >> >> Mild Shock schrieb: >>> Hi, >>> >>> Tablets and phone are more annoying to >>> use with WebGPU. The usual browsers don't >>> have a Chrome DevTools panel integrated, >>> >>> so that one could do JavaScript Debugging >>> directly on the device. Instead one has to >>> use a desktop machine, and connect the >>> >>> device via UBS-C , and start a Chrome >>> Browser there . And then start a Chrome >>> DevTools panel alone, that is pair with >>> >>> the device, via UBS-C cable. So this way >>> I already see where it crashes on the >>> tablets and phone: >>> >>> await output.mapAsync(GPUMapMode.READ) >>> Unhandled Promise Rejection: OperationError >>> >>> The above is the error that one can re-produce >>> already here with this test: >>> >>> 11.4 Giga Lips with a Budget Laptop >>> https://github.com/Jean-Luc-Picard-2021/gigabudget >>> >>> Not sure what exactly happens. Maybe >>> a form of timeout or device lost, that the >>> primitive HTML / JavaScript doesn't handle >>> >>> gracefully yet. Maybe redimensioning the >>> test, so that it consumes less time would >>> help. Who knows? Will see. For production >>> >>> use of a GPU integration I have to anyway >>> provide work slicing it seems. >>> >>> Bye >>> >>> Mild Shock schrieb: >>>> Hi, >>>> >>>> Why does this Lama have a red pyjama. >>>> Oh, its a baby Lama. Its still in the cradle >>>> and needs some training: >>>> >>>> RedPajama-Data-v2 >>>> https://github.com/togethercomputer/RedPajama-Data >>>> >>>> But then Andrej Karpathy recently showed >>>> GPT-2 training on rented GPUs for less >>>> than 100 USD in less then 2 hours. >>>> >>>> So where do these grown up Lamas go. >>>> Well Georgi Gerganov prefered C++/C >>>> when he shouted Llama Llama Red Pyjama. >>>> >>>> But you also find WebLLM, wrapping the >>>> underlying C++/C GPU interface via the >>>> W3C standard WebGPU / WGSL, with JavaScript: >>>> >>>> In-Browser LLM Inference Engine >>>> https://webllm.mlc.ai/ >>>> >>>> My experience with WebLLM 6 months >>>> ago on an iPad Pro 2024, still a little early >>>> stage performance and robustness. >>>> >>>> But hey hardware of AI mobile iGPUs is >>>> still evolving, and AI laptop, AI smartphones >>>> and AI tablets, will soon feature Chinese >>>> >>>> hardware such some new Kirin AI in 2027. >>>> >>>> Bye >>>> >>>> Mild Shock schrieb: >>>>> Hi, >>>>> >>>>> Remember when first all local AI was Python >>>>> and PyTorch APIs. And then suddently people strated >>>>> using bare metal C/C++ Code. Here is the story: >>>>> >>>>> How it started: >>>>> >>>>> GPT-J or GPT-J-6B is an open-source large >>>>> language model (LLM) developed by EleutherAI >>>>> in 2021. As the name suggests, it is a >>>>> generative pre-trained transformer model >>>>> designed to produce human-like text that >>>>> continues from a prompt. >>>>> https://www.eleuther.ai/ >>>>> >>>>> How it was going [Georgi Gerganov]: >>>>> >>>>> So a few days later comes out the LLaMA, I do >>>>> some calculations and I figure out “Okay, 65 >>>>> billion parameters. You probably need about >>>>> 40 gigs of RAM, with 4-bit quantization. So >>>>> this can run on a MacBook. Why not do it?” >>>>> >>>>> Why I was able to do it so quickly - basically, >>>>> for all that I saw it’s pretty much GPT-J architecture >>>>> with some modifications, like some extra memorization >>>>> layers. It’s minor changes. Basically, again, the >>>>> existing code for the GPT-J, I just simply >>>>> modified it there, it happened pretty quickly. >>>>> https://changelog.com/podcast/532 >>>>> >>>>> Georgi Gerganov, Bulgarian, now with Hugging >>>>> Face, ggml-cann also running on Chinese AI chips. >>>>> ggml Manifesto https://github.com/ggml-org/ggml >>>>> >>>>> Bye >>>>> >>>> >>> >> >
[toc] | [prev] | [next] | [standalone]
| From | Johann 'Myrkraverk' Oskarsson <johann@myrkraverk.invalid> |
|---|---|
| Date | 2026-08-03 00:34 +0800 |
| Subject | Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) |
| Message-ID | <ekKbS.129884$9jNc.102967@fx16.ams4> |
| In reply to | #647100 |
On 02/08/2026 8:47 AM, Mild Shock wrote:
> Hi,
>
> Chris M. Thomasson can ask 100 more questions.
> I will happily answer them. But maybe I should
> make a Wiki to explain the ever same things:
>
> > But, I still don't know what you main goal is?
> The goal is "Prolog inferencing"
Don't worry about it. There are several regulars here
who
don't
understand
that programming can be done for fun.
>
> > It has textures to work with in the pipeline.
> I don't need textures for "Prolog inferencing"
>
> 98 more questions to go, don't give up!
Here in sci.math, as everyone knows, I'm gearing up for
/linear algebra/ for fun. Still waiting for DVDs because
I'm not in a hurry. The book /Linear Algebra Done Right/
is interesting, and I've yet to go through the other rec-
commendations.[1]
I'm curious if you've ever thought of doing OpenGL with Prolog?
Does that even work?
[1] I have no idea how this word is supposed to be hyphenated,
I just do it anyway, because I'm not an LLM.
>
> Bye
Take care!
--
Johann | email: invalid -> com | http://www.myrkraverk.com/blog/
I'm not from the Internet, I just work there. | via Easynews.com
[toc] | [prev] | [next] | [standalone]
| From | Ross Finlayson <ross.a.finlayson@gmail.com> |
|---|---|
| Date | 2026-08-02 10:41 -0700 |
| Subject | Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) |
| Message-ID | <H-mcnW2o5ZLcHPL3nZ2dnZfqn_idnZ2d@giganews.com> |
| In reply to | #647105 |
On 08/02/2026 09:34 AM, Johann 'Myrkraverk' Oskarsson wrote: > On 02/08/2026 8:47 AM, Mild Shock wrote: >> Hi, >> >> Chris M. Thomasson can ask 100 more questions. >> I will happily answer them. But maybe I should >> make a Wiki to explain the ever same things: >> >> > But, I still don't know what you main goal is? >> The goal is "Prolog inferencing" > > Don't worry about it. There are several regulars here > who > don't > understand > > that programming can be done for fun. > >> >> > It has textures to work with in the pipeline. >> I don't need textures for "Prolog inferencing" >> >> 98 more questions to go, don't give up! > > Here in sci.math, as everyone knows, I'm gearing up for > /linear algebra/ for fun. Still waiting for DVDs because > I'm not in a hurry. The book /Linear Algebra Done Right/ > is interesting, and I've yet to go through the other rec- > commendations.[1] > > I'm curious if you've ever thought of doing OpenGL with Prolog? > > Does that even work? > > > [1] I have no idea how this word is supposed to be hyphenated, > I just do it anyway, because I'm not an LLM. >> >> Bye > > Take care! > You might have good luck looking up reputable university programs and seeing what textbooks they require, these days. Or, you know, just buy old ones when the library retires the old good ones. How about Householder's "The Theory of Matrices in Numerical Analysis". Linear independence and linear spaces inevitably get associated with vector spaces. There are much simpler accounts though of reflections and rotations about the determinantal and the singular and the decompositions and the forms and the echelon forms and reduction with regards to things like cumulants and orthogonants and the matroids, vis-a-vis usual closed categories and so on. The cumulants and orthogonants and so on are lesser-served accounts of the earlier 20'th century, and determinantal analysis, while the matroids are the a bit more obscure accounts of geometrizations with regards to matrices. What "linear" even is is usually enough "linear is linear". Generally considered "ordinary" if through substitution. I'm an anti-reductionist, yet though reduction is one of the most usual results in closed categories, the methods and techniques, point being closed categories aren't allowed to close themselves, only being found so.
[toc] | [prev] | [next] | [standalone]
| From | Ross Finlayson <ross.a.finlayson@gmail.com> |
|---|---|
| Date | 2026-08-02 11:13 -0700 |
| Subject | Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) |
| Message-ID | <VySdnV33ZYBHFfL3nZ2dnZfqn_SdnZ2d@giganews.com> |
| In reply to | #647106 |
On 08/02/2026 10:41 AM, Ross Finlayson wrote: > On 08/02/2026 09:34 AM, Johann 'Myrkraverk' Oskarsson wrote: >> On 02/08/2026 8:47 AM, Mild Shock wrote: >>> Hi, >>> >>> Chris M. Thomasson can ask 100 more questions. >>> I will happily answer them. But maybe I should >>> make a Wiki to explain the ever same things: >>> >>> > But, I still don't know what you main goal is? >>> The goal is "Prolog inferencing" >> >> Don't worry about it. There are several regulars here >> who >> don't >> understand >> >> that programming can be done for fun. >> >>> >>> > It has textures to work with in the pipeline. >>> I don't need textures for "Prolog inferencing" >>> >>> 98 more questions to go, don't give up! >> >> Here in sci.math, as everyone knows, I'm gearing up for >> /linear algebra/ for fun. Still waiting for DVDs because >> I'm not in a hurry. The book /Linear Algebra Done Right/ >> is interesting, and I've yet to go through the other rec- >> commendations.[1] >> >> I'm curious if you've ever thought of doing OpenGL with Prolog? >> >> Does that even work? >> >> >> [1] I have no idea how this word is supposed to be hyphenated, >> I just do it anyway, because I'm not an LLM. >>> >>> Bye >> >> Take care! >> > > You might have good luck looking up reputable university programs > and seeing what textbooks they require, these days. > > Or, you know, just buy old ones when the library retires > the old good ones. > > How about Householder's "The Theory of Matrices in Numerical Analysis". > > Linear independence and linear spaces inevitably > get associated with vector spaces. There are much > simpler accounts though of reflections and rotations > about the determinantal and the singular and the decompositions > and the forms and the echelon forms and reduction with regards > to things like cumulants and orthogonants and the matroids, > vis-a-vis usual closed categories and so on. > > The cumulants and orthogonants and so on are lesser-served > accounts of the earlier 20'th century, and determinantal analysis, while > the matroids are the a bit more obscure accounts of geometrizations with > regards to matrices. > > What "linear" even is is usually enough "linear is linear". > Generally considered "ordinary" if through substitution. > > > I'm an anti-reductionist, yet though reduction is one > of the most usual results in closed categories, the > methods and techniques, point being closed categories > aren't allowed to close themselves, only being found so. > > > > > Sometimes "linear independence" is better read as "linear dependence", this is because words like "abstract" and "general" and "closed" and "regular" and "ordinary" have inverses, matters of perspective and projection, then about the difference from the "non", the "super", for example the "classical". "Truth is regular. Geometry is motion."
[toc] | [prev] | [next] | [standalone]
| From | Johann 'Myrkraverk' Oskarsson <johann@myrkraverk.invalid> |
|---|---|
| Date | 2026-08-03 02:27 +0800 |
| Subject | Pro Gauss-Jordan Reduction, Phigs and Phigs+ (was: Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging)) |
| Message-ID | <K_LbS.25330$BWdc.19278@fx11.ams4> |
| In reply to | #647106 |
On 03/08/2026 1:41 AM, Ross Finlayson wrote: > On 08/02/2026 09:34 AM, Johann 'Myrkraverk' Oskarsson wrote: >> On 02/08/2026 8:47 AM, Mild Shock wrote: >>> Hi, >>> >>> Chris M. Thomasson can ask 100 more questions. >>> I will happily answer them. But maybe I should >>> make a Wiki to explain the ever same things: >>> >>> > But, I still don't know what you main goal is? >>> The goal is "Prolog inferencing" >> >> Don't worry about it. There are several regulars here >> who >> don't >> understand >> >> that programming can be done for fun. >> >>> >>> > It has textures to work with in the pipeline. >>> I don't need textures for "Prolog inferencing" >>> >>> 98 more questions to go, don't give up! >> >> Here in sci.math, as everyone knows, I'm gearing up for >> /linear algebra/ for fun. Still waiting for DVDs because >> I'm not in a hurry. The book /Linear Algebra Done Right/ >> is interesting, and I've yet to go through the other rec- >> commendations.[1] >> >> I'm curious if you've ever thought of doing OpenGL with Prolog? >> >> Does that even work? >> >> >> [1] I have no idea how this word is supposed to be hyphenated, >> I just do it anyway, because I'm not an LLM. >>> >>> Bye >> >> Take care! >> > > You might have good luck looking up reputable university programs > and seeing what textbooks they require, these days. > > Or, you know, just buy old ones when the library retires > the old good ones. > > How about Householder's "The Theory of Matrices in Numerical Analysis". I have several books that have been rescued from libraries. Sometimes even corporate libraries, but now forgot which specimen that was. One of my priced collection is /Phigs and Phigs+, An Introduction to 3D Computer Graphics/, by John W. Blake (1993) and I haven't read a lick of it. Possibly never will. Appending A has a Fortran 77 example, and Appending B has one in C. I have therefore added comp.lang.fortran and comp.lang.c to the discussion. Mostly because the regulars there annoy me. They know who they are. I took sci.physics.relativity out of the discussion, as I don't know anything about relativity at all. A future followup can re-add it if relativity affects this conversation. > Linear independence and linear spaces inevitably > get associated with vector spaces. There are much > simpler accounts though of reflections and rotations > about the determinantal and the singular and the decompositions > and the forms and the echelon forms and reduction with regards > to things like cumulants and orthogonants and the matroids, > vis-a-vis usual closed categories and so on. > > The cumulants and orthogonants and so on are lesser-served > accounts of the earlier 20'th century, and determinantal analysis, while > the matroids are the a bit more obscure accounts of geometrizations with > regards to matrices. > > What "linear" even is is usually enough "linear is linear". > Generally considered "ordinary" if through substitution. I have to admit, I understood some of those words. > I'm an anti-reductionist, yet though reduction is one > of the most usual results in closed categories, the > methods and techniques, point being closed categories > aren't allowed to close themselves, only being found so. I'm on the other hand, pro-Gauss-Jordan reduction. I may even try to code it in C on my own, instead of doing it the coward's way and use Sage like a "normal" mathematician. -- Johann | email: invalid -> com | http://www.myrkraverk.com/blog/ I'm not from the Internet, I just work there. | via Easynews.com
[toc] | [prev] | [next] | [standalone]
| From | Johann 'Myrkraverk' Oskarsson <johann@myrkraverk.invalid> |
|---|---|
| Date | 2026-08-03 02:35 +0800 |
| Subject | Re: Pro Gauss-Jordan Reduction, Phigs and Phigs+ |
| Message-ID | <c6MbS.40961$4Fu9.14809@fx05.ams4> |
| In reply to | #647108 |
On 03/08/2026 2:27 AM, Johann 'Myrkraverk' Oskarsson wrote: > I'm on the other hand, pro-Gauss-Jordan reduction. I may even > try to code it in C on my own, instead of doing it the coward's > way and use Sage like a "normal" mathematician. > And in sci.math, don't think I have anything against "normal" mathe- maticians. It's just a question of how we define "normal." Are we talking about normalized mathematicians like we do to SQL databases, or are we talking about Smith and Hermite normal forms [which I did not know were terms until I looked at the Wolfram website], or just a regular normal form mathematician like we do in /lineal algebra/? Have a nice and normal mathematical day! -- Johann | email: invalid -> com | http://www.myrkraverk.com/blog/ I'm not from the Internet, I just work there. | via Easynews.com
[toc] | [prev] | [next] | [standalone]
| From | "Chris M. Thomasson" <chris.m.thomasson.1@gmail.com> |
|---|---|
| Date | 2026-08-02 12:42 -0700 |
| Subject | Re: Chris M. Thomasson can ask 100 more questions (Re: Tablet and phone UBS-C remote debugging) |
| Message-ID | <114o6ii$ohrk$1@dont-email.me> |
| In reply to | #647100 |
On 8/1/2026 5:47 PM, Mild Shock wrote:
> Hi,
>
> Chris M. Thomasson can ask 100 more questions.
> I will happily answer them. But maybe I should
> make a Wiki to explain the ever same things:
>
> > But, I still don't know what you main goal is?
> The goal is "Prolog inferencing"
>
> > It has textures to work with in the pipeline.
> I don't need textures for "Prolog inferencing"
>
> 98 more questions to go, don't give up!
[...]
Fwiw, I have several compute shaders that do what I want. Mainly
building vector fields, etc.... And yes I use textures for some input
and output, uniforms mainly for the settings, etc. Just, make sure to
code things up to a point where your compute shader never needs to wait
for something... Think of striving for wait-free algorithms.
For instance, this is 100% wait free.
void add_hit(ct_plane2d plane, vec2 p, vec3 weight)
{
vec2 uv = ct_plane2d_unproject(plane, p);
ivec2 px = ivec2(uv * u_resolution);
if (px.x >= 0 && px.x < int(u_resolution.x) &&
px.y >= 0 && px.y < int(u_resolution.y))
{
imageAtomicAdd(accum_r, px, weight.r);
imageAtomicAdd(accum_g, px, weight.g);
imageAtomicAdd(accum_b, px, weight.b);
imageAtomicAdd(accum_hits, px, 1.0f);
}
}
Notice how I separated my accumulation buffer into different textures?
layout(binding = 0, r32f) uniform coherent image2D accum_r;
layout(binding = 1, r32f) uniform coherent image2D accum_g;
layout(binding = 2, r32f) uniform coherent image2D accum_b;
layout(binding = 3, r32f) uniform coherent image2D accum_hits; // alpha
/ hit counter
Works great and runs really fast.
[toc] | [prev] | [next] | [standalone]
| From | Mild Shock <janburse@fastmail.fm> |
|---|---|
| Date | 2026-08-02 23:09 +0200 |
| Subject | You posted that already, but you didn't listen [I NEED BOUNDED QUEUES] (Was: Chris M. Thomasson can ask 100 more questions) |
| Message-ID | <114obm9$rk0b$1@solani.org> |
| In reply to | #647111 |
Hi,
I assure you I have like 3-4 times already
communicated to you that my requirements are
bounded queues. And not the ideally unbounded queues
that you are using, i.e. imageAtomicAdd.
Just check the postings in this forum. I have
like 3-4 times already specified that I need
bounded queues.
> Works great and runs really fast.
You repeating yourself. Whats the motivation
of this spamming. I mean I can officially acknowledge
here that I have seen your imageAtomicAdd code
already. I also responded back then that I
have a Queue prototype that exactly uses that.
But it doesn't work for my purpose because I need:
- bounded queues that can block
- sizes are typically like 4-32 elements
- blocking is not done in GPU
- blocking is done in Hack
- Hack can do work stealing etc..
Because Hack can do a lot of tricks, you shouldn't
worry at all. Also spinning with backoff etc..
could be part of the picture, just check out:
Parallel Programming, Spring 2019, Lecture 16+1:
Spinlocks, Deadlocks, Semaphores
https://spcl.inf.ethz.ch/Teaching/2020-pp/lectures/PP-l17-BeyondLocks.pdf
So just let me do my research, and refrain from
spamming me with always the same nonsense. Better
listen. I assure you I have like 3-4 times already
communicated to you that my requirements are
bounded queues. And not the ideally unbounded queues
that you are using., i.e. imageAtomicAdd.
Just check the postings in this forum. I have
like 3-4 times already specified that I need
bounded queues.
Bye
Chris M. Thomasson schrieb:
> On 8/1/2026 5:47 PM, Mild Shock wrote:
>> Hi,
>>
>> Chris M. Thomasson can ask 100 more questions.
>> I will happily answer them. But maybe I should
>> make a Wiki to explain the ever same things:
>>
>> > But, I still don't know what you main goal is?
>> The goal is "Prolog inferencing"
>>
>> > It has textures to work with in the pipeline.
>> I don't need textures for "Prolog inferencing"
>>
>> 98 more questions to go, don't give up!
> [...]
>
> Fwiw, I have several compute shaders that do what I want. Mainly
> building vector fields, etc.... And yes I use textures for some input
> and output, uniforms mainly for the settings, etc. Just, make sure to
> code things up to a point where your compute shader never needs to wait
> for something... Think of striving for wait-free algorithms.
>
> For instance, this is 100% wait free.
>
> void add_hit(ct_plane2d plane, vec2 p, vec3 weight)
> {
> vec2 uv = ct_plane2d_unproject(plane, p);
> ivec2 px = ivec2(uv * u_resolution);
>
> if (px.x >= 0 && px.x < int(u_resolution.x) &&
> px.y >= 0 && px.y < int(u_resolution.y))
> {
> imageAtomicAdd(accum_r, px, weight.r);
> imageAtomicAdd(accum_g, px, weight.g);
> imageAtomicAdd(accum_b, px, weight.b);
> imageAtomicAdd(accum_hits, px, 1.0f);
> }
> }
>
>
> Notice how I separated my accumulation buffer into different textures?
>
> layout(binding = 0, r32f) uniform coherent image2D accum_r;
> layout(binding = 1, r32f) uniform coherent image2D accum_g;
> layout(binding = 2, r32f) uniform coherent image2D accum_b;
> layout(binding = 3, r32f) uniform coherent image2D accum_hits; // alpha
> / hit counter
>
> Works great and runs really fast.
[toc] | [prev] | [next] | [standalone]
Page 6 of 7 — ← Prev page 1 2 3 4 5 [6] 7 Next page →
Back to top | Article view | sci.math
csiph-web