Download Latest Version node-llama-cpp-electron-example.Windows.3.20.0.arm64.exe (132.0 MB)
Email in envelope

Get an email when there's a new version of node-llama-cpp

Home / v3.20.0
Name Modified Size InfoDownloads / Week
Parent folder
node-llama-cpp-electron-example.Linux.3.20.0.x64.tar.gz 2026-08-11 493.5 MB
node-llama-cpp-electron-example.Linux.3.20.0.arm64.tar.gz 2026-08-11 160.1 MB
node-llama-cpp-electron-example.Linux.3.20.0.arm64.deb 2026-08-11 130.3 MB
node-llama-cpp-electron-example.Linux.3.20.0.amd64.deb 2026-08-11 366.8 MB
node-llama-cpp-electron-example.Linux.3.20.0.amd64.snap 2026-08-11 433.7 MB
node-llama-cpp-electron-example.Linux.3.20.0.x86_64.AppImage 2026-08-11 499.1 MB
node-llama-cpp-electron-example.Linux.3.20.0.arm64.AppImage 2026-08-11 168.3 MB
node-llama-cpp-electron-example.macOS.3.20.0.x64.zip 2026-08-11 174.5 MB
node-llama-cpp-electron-example.Windows.3.20.0.x64.exe 2026-08-11 376.8 MB
node-llama-cpp-electron-example.Windows.3.20.0.arm64.exe 2026-08-11 132.0 MB
node-llama-cpp-electron-example.Windows.3.20.0.exe 2026-08-11 508.1 MB
node-llama-cpp-electron-example.macOS.3.20.0.arm64.zip 2026-08-11 162.0 MB
node-llama-cpp-electron-example.macOS.3.20.0.x64.dmg 2026-08-11 177.2 MB
node-llama-cpp-electron-example.macOS.3.20.0.arm64.dmg 2026-08-11 164.8 MB
README.md 2026-08-11 3.3 kB
v3.20.0 source code.tar.gz 2026-08-11 22.0 MB
v3.20.0 source code.zip 2026-08-11 22.3 MB
Totals: 17 Items   4.0 GB 4

3.20.0 (2026-08-11)

Features

  • Muse Glimmer support (#639) (adb92f2)
  • improve thought segments syntax extraction (#636) (3f686d7)
  • expose download speed and ETA for a model downloader (#636) (3f686d7)
  • expose the model's architecture directly on the model instance (#636) (3f686d7)
  • inspect gpu command: print model parameters count (#639) (adb92f2)

Bug Fixes

  • adapt to breaking llama.cpp changes (#636) (3f686d7)
  • LlamaContextSequence: make .dispose() return a promise (#636) (3f686d7)
  • Qwen chat wrapper auto thought segment opening (#636) (3f686d7)
  • optimize checkpoints with auto opening thought segments (#636) (3f686d7)
  • simulator model dispose while in use race conditions (#636) (3f686d7)
  • Vulkan device memory readings (#636) (3f686d7)
  • topP config when using temperature (#639) (adb92f2)
  • support more quant labels (#639) (adb92f2)

Shipped with llama.cpp release b10361

To use the latest llama.cpp release available, run npx -n node-llama-cpp source download --release latest. (learn more)

Source: README.md, updated 2026-08-11