FacelessJim
8 hours ago
Love pi. I tried to run some local models and pi was the only one that actually worked decently because it didn’t have a gargantuan system prompt that would take minutes to prefill on my scrawny ass laptop.
Been running it almost barebones vanilla for a couple of months. Just a bunch of basic extensions and some skills.
Now, if only they could fix the very annoying bug of the history jumping back at the beginning if I am not a the end while the model is reasoning that would great.
RickS
7 hours ago
Couldn't agree more. The vanilla openclaw install was this byzantine mess of MD files talking about souls and identities and such, it really put me off. Stripping back to a bare install of the underlying pi, it was delightfully minimal and easy to reason about. Excellent starting point for building an assistant agent without having to read or fight with a bunch of cruft on top.
Couple skills to integrate with an obsidian MD task tracker, small chat interface on the phone made public via tailscale, and bam, a reminder bot you can text from the grocery store.
AgentMasterRace
3 hours ago
you're comparing apples to oranges here...
simpaticoder
5 hours ago
You inspired me to try Pi out - so far it's worked flawlessly. Plugged it into OpenRouter and ~$.50 of Deepseek later I've installed llama.cpp and Llama 3.1. The local model doesn't work with Pi yet (and I know it will be bad and slow even if it does) but I'm curious to see what you can do on an 8GB consumer GPU these days...
rablackburn
4 hours ago
> I'm curious to see what you can do on an 8GB consumer GPU these days
Running smaller 4B-7B models entirely on the GPU VRAM will get you fast inference, but you will need to scope and define the tasks well. eg, using it the model as a classifier and just feeding it from a queue.
The best performing "agent"-like model to plug into a harness that I have found so far has been Qwen3.6-35B-A3B (mixture of experts) model as I can park most of it in system RAM and CPU, while the VRAM holds the attention/shared weights.
It's definitely workable as a local AI homelab. But expect homelab levels of tuning/fiddling with it.
With the improved support for AMD GPUs I'm finally considering getting a modern 16GB card (and maybe a second one in a few years assuming prices come down)
whatshisface
3 hours ago
If you paid DeepSeek directly, that would have been 1 to 10 cents. OpenRouter has a huge overhead due to their cache logic, I'm surprised they keep business coming in the door for tasks other than system prompt - output pairs.
simpaticoder
2 hours ago
I think it was actually less than that. I was doing something else too in another agent.
sejje
4 hours ago
> I'm curious to see what you can do on an 8GB consumer GPU these days.
Nothing, really. Might be coming soon, but no.
You probably want to try bonsai, I guess, but don't expect good results.
what
3 hours ago
> ~$.50 of Deepseek later I've installed llama.cpp and Llama 3.1
You could install this yourself for free? I get $0.50 isn’t all that much, but still?
simpaticoder
3 hours ago
Sure, but I'm not interested in learning about running cpp, installing CUDA, finding the right URLs for downloading llama weights. It's the best 50 cents I've spent in 20 years.
8n4vidtmkvmk
2 hours ago
Even so, I'm surprised it cost that much. I thought deepseek was cheaper.
But AI for installing tricky opensource software is indeed a good use case. I do that too.
AgentMasterRace
3 hours ago
you're living in the 2020s bro
rsync
8 hours ago
"Been running it almost barebones vanilla for a couple of months. Just a bunch of basic extensions and some skills."
I am also using pi exclusively after having had decent success with openhands but begrudging all of the docker infrastructure ... and all of the emojis.
My only pain point is that in my extremely common and boring workflow, which is pi inside of gnu screen inside of OSX terminal.app ... all reasoning/thinking text is blinking ... like old fashioned ANSI blink on a BBS.
I cannot figure out how to disable the blinking thought/reasoning text ...
rmunn
2 hours ago
Have you tried a different terminal app, such as Ghostty? https://ghostty.org/ has a Mac build, and handles italics properly. That might solve your issue without having to edit any configuration files. Plus, as a side benefit, Ghostty ignores the ANSI color codes for blinking text, so you won't ever see blinking text again.
cbsks
7 hours ago
I just fixed something very similar in my setup. Except in my case the reasoning text was shown in dark grey on a light gray background. Very ugly and hard to read.
If I remember correctly, the reasoning text was being output using the italic ANSI code, which was being formatted funny on my terminal. I fixed it by adding a font that supports italics. I recommend taking a look at the ansi codes.
nine_k
an hour ago
If you love your terminal app and won't switch to Ghostty or WezTerm, at least try tmux instead of the venerable but ancient GNU screen.
huijzer
a few seconds ago
Alacritty plus Zellij works great for me. Much easier to use than tmux
sothatsit
3 hours ago
I was running into similar issues where italics text was blinking. I traced it back to a bug in screen, which I patched in my own screen fork. Not sure if exactly the same bug but could be? https://github.com/Sothatsit/screen
rpdillon
2 hours ago
See if you can replicate this inside of tmux. It might be screen's escape code handling.
gchamonlive
3 hours ago
I use oh-my-pi, not sure how it compares, but I say someone praising antigravity for being a good harness(1), and for the love of good people settle for such low standards of user experience it's almost pitiful.
ziphyrien
3 hours ago
This bug was fixed several months ago; you just need to switch to full-screen mode, though you hadn’t done so previously.
However, they have now set full-screen mode as the default.
wilt_
7 hours ago
Are you using Windows Terminal by any chance? I'm building a personal fork [0] with a patch for this exact bug (plus a few other open PRs from the upstream repo that seemed cool). Haven't tried contributing it upstream, since the patch is fully vibe-coded and I've spent almost no time trying to understand how it works, but the bug hasn't recurred since I've been using it.
berofeev
7 hours ago
You're looking for fullscreen TUI mode: https://www.reddit.com/r/PiCodingAgent/comments/1vh5pys/than...