Hacker Newsnew | past | comments | ask | show | jobs | submit | cryptoz's commentslogin

I treasure my memories of being on subways and busses and getting lost in thought while staring out the windows, yeah. I mean, people have done this for basically all of human history (okay walking or something, not on a train, but, yeah). It's not a bad thing to have some time with just your senses existing in the world.

I mean, you're right. When I travel, I rarely use my phone or listen to music because everything is about discovering new things, and I love it

But when I'm home, on my way from home to the gym, I take the same route every day, and it can get boring sometimes

I'm also terrible at keeping my phone charged, so sometimes I end up alone with my thoughts anyway


Let people enjoy things.

The problem with being always online is not “not-enjoying” but stopping to think. People are getting used to be feed what they need to think constantly (from politics to product reviews) and then their critical thinking skills atrophy.

Moreover, I’m not sure if being on social media is something “enjoyable” anymore, and it’s more like smoking tobacco where it becomes necessary but is not enjoyed.


Do you really enjoy doom scrolling? Because I don't, it is just addictive and makes me feel tired and sad most of the times, with infrequent sparks of joy.

Don’t do things you don’t enjoy.

They are really adictive and engaging, and they trick you into spending more time there. Nobody ever says "let's watch reels for 3h", but here we are.

What you've described is just a new benchmark, though. It'll be called CarParkBench, various embodied LLMs will then be run against that benchmark, and some will score better than others.

I do see where you're going, but that's already what's happening: we have so many different benchmarks because there's no real single way to test for general intelligence.

Also, it takes a human probably at least a decade of world experience, growth, learning, etc, to pass your benchmark. I'm quite confident that it will be very soon that an embodied LLM will pass your new benchmark, much sooner than a human would take if born today.


I think the issue is less about creating a new benchmark and more that the existing benchmarks shouldn't be called anything related to AGI unless they measure AGI.

If a model couldn't go to work as e.g. a first year apprentice plumber on their first day and perform anywhere remotely close to the median, but can pass a benchmark that claims to measure AGI, the benchmark is wrong and the model is not exhibiting general intelligence yet. ApprenticePlumberBench sounds like it's genuinely better than ARC-AGI at measuring AGI and that's a bit silly.

(Edit: I wrote ARC-GIS the first time around, for some silly reason)


It's not measuring AGI at all, it starts from human "core knowledge" so it is parochial. It is made of tests that still fail so by definition next version will also start low. Moving goalpost.

I'm not familiar with MyselfGPT (and didn't see an obvious result in a quick search), so, I guess it probably cannot use MyselfGPT. If you have a link, I'll check it out and then we'll see!


> Should be "You're thinking".

This should be "This should be 'You're thinking'." don't you think? Why bother correcting someone's grammar with a sentence fragment? You're just trading one mistake for another. I'm hoping someone finds a grammar error in my post, because continuing this would be hilarious.


> This should be "This should be 'You're thinking'." don't you think?

Reflexively, I think it should be more like ...

  javascript: `This should be "You're thinking".` ; 
  // to preserve the original character use and to avoid '...'...' parse foos
  // however `"...".` also possibly deserves a [sic] to critique the original 
  // i.e. ~grammar police say the period belongs within the quote marks, no?
... but then that's just me, in [my] quirks mode.


The linked page implies there is no linux support, I wonder why. It's there in the docs if you hunt for it.


The docs are here: https://docs.docker.com/ai/sandboxes/

The other url is their marketing page.

Yes, Linux is supported.


FWIW, I used to love this phrase but over recent years have come to understand it is quite damaging. We live in a society where evil frequently hides behind a ‘stupid’ label, and people bring this quote up to defend or soften actions that are indeed done out of specific malicious intent.


Right? "Never attribute to malice what [... etc]" is always just a thought-terminating cliche these days.

TBH I have a hard time imagining how anyone, in the year 2026, thinks that we should default to assuming good intent behind words on the internet.


It's because we don't treat evil and stupidity the same when we should.


Still can't make a Gmail label named 'Buzz'.


Know of any other labels that are blocked or taken ?


Yes, of course, it seems that type of ID was specifically rejected in the story above. Only drivers allowed to sign up. Non-drivers rejected because they don’t drive.


I'm reminded of the Alameda Weehawken burrito tunnel:

https://idlewords.com/2007/04/the_alameda_weehawken_burrito_...


The single most implausible idea in that article is that New York City would be able to so completely outbid the SF Bay Area for burritos.


Proper burritos for lunch enables Wall Street finance firms to reach new heights of excellence, propelling a feedback loop that leaves SF bereft.


Really cool stuff and definitely going to try it; I’m also finding it wild that Google put effort into adding ums and erms into their text to speech model a while back. AI puts it in, AI helps take it out.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: