pi-llama-cpp-stats
Shows prompt processing progress from llama.cpp's SSE stream
Package details
Install pi-llama-cpp-stats from npm and Pi will load the resources declared by the package manifest.
$ pi install npm:pi-llama-cpp-stats- Package
pi-llama-cpp-stats- Version
0.1.6- Published
- Jun 26, 2026
- Downloads
- 148/mo · 38/wk
- Author
- cr4xy
- License
- unknown
- Types
- extension
- Size
- 10.7 KB
- Dependencies
- 0 dependencies · 0 peers
Pi manifest JSON
{
"extensions": [
"./index.ts"
],
"image": "https://cr4xy.dev/pi-llama-cpp-stats/preview.png"
}Security note
Pi packages can execute code and influence agent behavior. Review the source before installing third-party packages.
README
pi-llama-cpp-stats
Pi extension that shows real-time prompt processing statistics from llama.cpp's SSE stream.
Features
- Replaces the "Working..." text with a live progress bar during prompt processing
- Shows progress percentage, estimated time remaining, and tokens/second
Example display:
Prefilling... ████████░░░░░░░░░░░░ 40% · 30s · 1234 tok/s
Installation
pi install npm:pi-llama-cpp-stats
Or install locally for the current project:
pi install npm:pi-llama-cpp-stats -l
Usage
- Restart pi (or
/reload) - Start a chat with a llama.cpp model
- The "Working..." text will be replaced with a live progress bar during prefilling
Requirements
- llama.cpp with OpenAI-compatible API
- llama.cpp must be built with
prompt_progresssupport (enabled by default in recent versions)
