Hacker News .hnnew | past | comments | ask | show | jobs | submit | InsideOutSanta's commentslogin

The fact that you thought to consider the next 3, 5, or 10 years already makes you a better CEO than most CEOs that I personally know.

Next quarter earnings call is the only thing that's important. Hollow out everything for that goal. My bonus depends on it.

Monkey Island taught me English. I can't tell you how confusing insult sword fighting was initially. I had to create long tables with the correct answers because I didn't get most of the puns, and then I had to start from scratch when I had to fight Carla.

Anyway, thanks, Ron Gilbert.


A pirate I was meant to be, trim the sail and roam the sea

Monkey Island 3 taught me a good deal of english too. I was lucky to get a text-translated version with english voiceover.

We all would avoid scurvy if we eat an orange...


Gandi has started increasing prices like crazy in the last few years.

Yeah, it's crazy that there is no trustworthy source for model reviews. I'd love to know how well the new Deepseek 4 actually performs, for example, but I don't want to spend the next week testing it out. Reddit used to be a somewhat useful gauge, but now there are posts on how 4 is useless right next to posts on how amazing it is. And I have no idea if this is astroturfing, or somebody using a quantized version, or different workloads, or what.

I also find it increasingly difficult to evaluate the models I actually do use. Sometimes each new release seems identical or only marginally better than the previous version, but when I then go back two or three version, I suddenly find that oder model to be dramatically worse. But was that older model always that quality, or am I now being served a different model under the same version name?

It's all just so opaque.


One challenge is that model evaluation is typically domain/application specific. Model performance can also depend on the system prompt and the input/context.

Regarding evaluation, I've found using tools like promptfoo (and in some cases custom tools built on top of that) are useful. These help when evaluating new models/versions and when modifying the system prompt to guide the model. Especially if you can define visualizations and assertions to accurately test what you are trying to achieve.

This can be difficult for tasks like summarization, code generation, or creative writing that don't have clear answers. Though having some basic evaluation metrics and test cases can still be useful, and being able to easily do side-by-side comparisons by hand.


Getting people to have an undifferentiated distrust of news organizations in general is an important aspect of technofeudalism.

If "today" were random, our universe would be pretty fricken weird.

What is today right now in Australia? How about where you live? You have not thought enough about what you’re saying and are probably not aware of all the weird time issues we have in our world.

That's not what "random" means.

Totally sounds like an LLM wrote it. Should have been two paragraphs instead of this verbose drivel.


Also, isn't this just a huge fire hazard of they actually do what they claim? Or will they remove the batteries from these old, continually plugged in, poorly cooled laptops?


This is addressed on their site.

> We might modify your laptop to remove or power down the battery, wireless radios, etc. to ensure it can be used safely in the data center.


I'm curious about what actually happened, but I can'make it past the second section, the writing is so terrible.


My theory is that YouTube blocks some accounts for publishing LLM-generated music, and people who wanted to earn ad money from it get burned and publish LLM-generated posts about it.

I would be on YouTube's side here, except it's possible that their motivation is simply to avoid poisoning their dataset while they train their models off creators videos. Also, the question is how they tell apart what's LLM-generated without false positives.

Maybe there were also artificial listens fraud (it's a problem with their competitor Spotify), but we'll never know because no one who was blocked would publish that honestly.


"No wireless. Less space than a Nomad. Lame."

No, wait, that was that other site.


Consider applying for YC's Summer 2026 batch! Applications are open till May 4

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: