Hacker Newsnew | past | comments | ask | show | jobs | submit | faxmeyourcode's commentslogin

Cloudflare mostly because workers and durable objects are so cheap.

This is so cool! I'm having trouble thinking of all the different use cases and potential issues with this, but I don't think anyone has done something like this before. Great work


Thanks for taking a look and for kind words!

We've also had a prototype running multiplayer DOOM between browsers via WebRTC. I wonder what kinds of network applications might be interesting here.



I like rovo -.-


Have you tried others? You're leaving a lot on the table using Rovo, it's just not nearly as capable


Nice! The idea of having a friday leetcode problem for my team has been kicked around to try and keep us on our toes when it comes to solving these style problems and preventing atrophy of this part of our brain. Might try wrapping this cli in some kind of web ui or adding it to our slackbot


It's not clear to me from the article, are they asking sol to output bounding box coordinates with some kind of structured outputs?

Anecdotal but I've seen it use python to crop, zoom, and "enhance" (fiddle with sharpness and brightness) images to read sections of handwritten census data from the 1800s. Feels like that there might just be a mismatch of capabilities when it comes to straight outputting coordinates but I bet the model is better at actually finding the answer given any tools available. Which I get is a bit of an apples and oranges situation.

I've also tried to use it to identify an old pair of glasses and it didn't stand a chance, so I do think it's not quite there yet when it comes to some vision tasks.


I built macrosforhumans.com because I had been using cronometer since 2020 on and off and was familiar with tracking macros but always hated the entry. Measurement is something that you have to do regardless but the annoyance of having to suffer through their search/recipe/diary CRUD UI was unbearable. With MFH there is no mobile or web app. it's just a phone number that you text. This reduced the friction for me significantly and I've already lost a few pounds and, after recently turning 30, hopefully a few more with more diet and exercise.

I've been building this for a while but I'm working on getting my first 100 users now which is my first foray into the marketing/sales side of things and any tips would be helpful. Also I didn't realize until later that the name, while I think is pretty cool, might be extremely poor SEO? :-)

I also think that these kinds of post-app apps are a curious idea and will become more popular, although likely (hopefully not) through sms as the communication medium. I am off this week and plan to write some stuff about it but it feels really cool to remove the terrible CRUD UIs entirely from this part of my life. I tell my friends that even if it doesn't take off I'm never shutting it down because I enjoy using it over myfitnesspal/cronometer/macrofactor etc so much.


Compaction, busting the cache, and other issues like that will lose significance when you're running at 15k tokens per second like chatjimmy. Very interesting to think about what will change in the future.


Apparently there was a gathering of harp guitarists here in North Little Rock at a guitar shop that recently opened just down the street from where I live last year: https://www.harpguitars.net/2025/11/29/hgg23-in-north-little...

I was totally unfamiliar with harp guitar but I am sad that I missed such a special gathering!


I've definitely noticed 5.6 sol being extremely trigger happy in ways other models, even 5.5, we're not. I would definitely categorize a few small incidents at work where it performed "actions a reasonable user would likely not anticipate and strongly object to." Just my anecdotal experience.

For example discussing driver upgrade and subsequent password rotation and it didn't stop and ask me if I wanted to restart the service or install the driver or anything, it immediately took action. It feels like a side effect of pushing more "agency."


I like 5.5 a lot, despite how I feel about OpenAI as a company. In OpenCode it feels about as smart as Opus 4.8, but it's less aggressive about following up on minutiae and getting lost in side quests. Might be a matter of prompt design moreso than model capability. I was looking forward to 5.6 but now this thread is making me quickly lose interest.


I’m still using 5.5 and had it do almost exactly that same example on a task yesterday so doesn’t seem like a clear cut 5.5 vs 5.6 thing. It’s pretty trigger happy already once it gets any kind of “go” without specific restrictions.


I feel like another comparison worth looking at is purely cost.

Capability per dollar is something I care about:

    Opus API    $5/$25
    Sonnet API  $5/$15
    Haiku API   $1/$5

    GLM 5.2 API $1.4/$4.4
So you're really getting near opus level capability for the price of haiku.


Not really, GLM uses more tokens to get work done.


In the article, they claim GLM used almost half the tokens 131,000, and the cost is about a quarter. For the cost to be the same GLM would have to use 4-5 times more tokens.


By how much? At least TFA provided numbers for one example, and they disagree with you (by a lot).


I ran a fairly large experiment last week, and the token usage wasn't bad at all. What softs of use cases are you seeing large token usage by GLM 5.2?


> are you seeing large token usage by GLM 5.2

the statement isn't "GLM 5.2 has large token usage", it's "GLM 5.2 has large token usage vs modern Opus".

I haven't used it, but this wouldn't surprise me. I see ~30% lower token usage for better results with Opus 4.8 vs 4.6 (and i had great results with 4.6)


I'm comparing with GPT5.5 on Codex and it's not even a competition. GLM takes way longer and eats a lot of tokens getting work done, it's easy to rack up a big bill on openrouter. I tried the $20 plan from ollama, too, and ate through half a month of budget in a few hours and blew my daily limit twice and still had to get codex to complete it -- which it did with only 10% of my monthly limit remaining.

GLM is promising but it's pretty costly, all things considered.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: