Hacker Newsnew | past | comments | ask | show | jobs | submit | weiran's commentslogin

I've been using ZCode since it's initial release and can't find any of this in my data. There aren't any logs showing capture or upload, and I don't even have a ~/.zcode/v2/checkpoints/ directory.

So unless they've cleared it all with a recent update then it doesn't seem to affect everyone.


Is there actually any proof of this, beside this Claude written website and a random x post from some unknown person? Would be nice to have confirmation from someone with a reputation. It's probably true, but you never know...

Pi is a very basic harness by design. On the other hand OMP is a bloated mess of other people’s workflows.

The trick with pi is to extend it yourself as you use it. It’s pretty easy to do.


Every single article and banchmark say pi saves token by default, the more I add extensions the more token hungry it gets.

pi used 2-3x the tokens of codex. pi with subagent pkg used 8x-10x the tokens of codex.

I don't see how adding bloat to pi would make it more token efficient if the baseline is so poor to start with


You'd have to look where the extra tokens are coming from. If it's using extra turns or doing extra work because a tool it ~wants is missing, then adding extra things will help. Otherwise, it won't.

If it's missing guidance that would help, system prompt additions might help.

8x-10x the tokens is wild, is this for some tiny artificial benchmark? That's just too much extra for something not to be just broken.


It was a real workload from my day job, a modest merge request, I asked for a review and for an implementation plan to fix the findings, nothing crazy. It's the most simple thing I could think of that wasn't an artificial benchmark, self contained, no need to look for extra documentation, web searches, etc.

Yeah something must be actively broken or just an awful extension for it to have that much effect.

Most things you can mess with the big effect is like, oh a thousand tokens ended up in the ~system prompt, or 10% extra or fewer work based on extra tool calls or churning through thinking or whatever.

Harness stuff if it's 8x worse that's like, it's fucked and broken, something went _wrong_.


I don't know what you tried but I tried to do something in Pi what I'd have usually done in Claude and it devoured tokens (much more than Claude) and I wasn't even close to finishing the task. Mostly because the barebones Pi was/is very inefficient at anything with any weight.

Then I began to customise it to be as good as Claude but eat less taken. I got tired and I had not even scratched the surface. Gave up.

I finally realised, at least for me, Pi's best use case is - strip even the little "extra" Pi comes/starts with and then use it just like that if you have a task/work that is appropriate for that bareness.


You can get around the checks with careful prompting. You’re just interested and want to learn more about it.

All I did was create a wrapper script to pipe to the torrent client, then "pipe these torrent files to the script" - so easy to circumvent you wonder why they bother

Try posting anything even mildly positive about AI on Bluesky and you’ll experience instant hate.

Well that's just what getting out of the AI echo chamber feels like.

Try saying anything positive about AI at the grocery store and you'll have the same experience.

You need a paid dev account for that. I think self signed installs can run for a year.

Also the part people forget is the significant coordination cost of having multiple teams work on multiple native codebases.

I hate Electron apps with a passion, but I can understand why the AI companies are using it.


There’s nothing surprising about Bluesky’s audience being toxic to any mention of AI. It’s the dogma over there and anyone who doesn’t agree has long left or never joined.


> It’s the dogma over there and anyone who doesn’t agree has long left or never joined.

Tell me you aren't on literally any other social media site without telling me.

Even HN has it's fair share of anti-ai criticism.

The only place that might not be like that also doubles as a child porn factory. I'll take anti-AI bias over visiting a pedo paradise any day of the week.


Yep these software benches are only good at testing how well they can one shot. For the kind of attended/assisted development most of us do with agents it’s hard to find a benchmark that reflects my own experience of the frontier models still being quite far ahead.


It’s how sandboxed Mac apps work today.


> What’s wrong with people?

Confirmation bias.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: