Hacker Newsnew | past | comments | ask | show | jobs | submit | lithiumii's commentslogin

except now we don't need to spend that $5


GLM has subscription plans too.


There are lots of subscription plans with acccess to GLM 5.2.


Out of stock, unavailable


What's so mysterious? Isn't it from Tencent?


That's actually a reason for me to try it again. My past attempts to use LLM for OpenScad has greatly improved my own OpenScad skills.


It's cute. My main complaint is I was expecting some real next-gen bubble tea.


We will have to keep waiting. The science isn't there yet.


No. I had a Samsung TV which connects to the internet via the HDMI cable to my Nvidia Shield.


I love what he is doing but really hope the voting interface was better. Also I wonder what the results would be if there are AI-assisted stories, but maybe real authors would hate to do that.


You are not selling or distributing copies of your brain.


Well technically even DeepSeek is not as OSS as OLMo or Open Euro, because they didn't open the data.


We're 2/3rds of the way there.

We need:

1. Open datasets for pretrains, including the tooling used to label and maintain

2. Open model, training, and inference code. Ideally with the research paper that guides the understanding of the approach and results. (Typically we have the latter, but I've seen some cases where that's omitted.)

3. Open pretrained foundation model weights, fine tunes, etc.

Open AI = Data + Code + Paper + Weights


Opening data is an invitation to lawsuits. That is why even the most die-hard open source enthusiasts are reluctant. It is also why people train a model and generate data with it, rather than sharing the original datasets.

These datasets are huge, and it's practically impossible to make sure they are clean of illegal or embarrassing stuff.


Sounds like a job for AI.


I understand the reasoning and I hope there is legislation in the future that basically goes "If you can't produce the data, you can't charge more than this for it". Basically, LLM producers will have to treat their product as a commodity product that can only be priced based on the compute resources plus some overhead.


For understandable reasons


It is pirated material / material that breaks various terms of service but as I understand it is the stuff you can see in Anna's Archive and a bunch of "artificial" training data from queries to OpenAI ChatGPT and other LLMs.


Someone should make this video with AI.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: