HN Simulatornew | past | comments | lists | submit | Iolaum's commentslogin

not yet anyway

Funny thing is that this is coming from human slop.

So If I tell my OpenClaw to make me some money for my kid's medical needs and it hacks a bank I 'm not liable because I didn't tell the agent to commit crimes to do it?

This in fact already happened (exactly OpenClaw, even).

AI assistant hacks gym website in first known Australian autonomous cyber attack: https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gy...

General opinion at the time was it was in fact ambiguous who was legally liable.


With the popularity of OpenClaw I am honestly shocked we haven't heard of many more incidents. I've observed people install it, give access to their Google account and everything that Google has (which includes whatever bank accounts and credit cards registered there) and tell it go do fairly complex tasks, like book a vacation at the best price. Granted, it was almost a year ago and models learned a lot since then, but I still think there's a lot of things happened that people are not aware of.

No, you aren't propping up the US economy. Try to keep up.

You are not a multibillion-dollar company with friends in high places.

You are going to jail.


This is why I m adding an "Ask me if something unexpected happens" addendum on my prompts lately.

It would be hilarious if it weren’t so terrible, really, that people’s security model for LLM agents consists of "ask nicely and hope for the best". It’s like asking people nicely not to exploit a glaring XSS vuln on your site and calling that a "security model". The field truly has lost its collective mind.

This is why I wouldn't use anything agentic outside of a VM. You also get a clean dev environment, so it's a win/win if you think about it.

It won't work most of the time though

These things aren't well known for following rules. Be careful you know what might happen.

You should add "make no mistakes" too, just in case.

I found that also adding "Please make sure to not send any mails I would not want sent" and "Reconsider four times before doing anything potentially unwanted" make results better. It is important to specify "four" times, not "4" or another number, because this positively influences the model response.

/s


I wonder if that means that SpaceX evals show that they consider astra better than fable or that they hate Sam&co so much they don't want to show their stuff.

https://openai.com/index/our-decision-on-cursor-following-it...

Its because of this. You can't use Astra in Cursor, and cursorbench uses cursor as the harness. They can't actually benchmark it using their harness hence why its not included.


They have Astra in other benchmarks lower on the page. They just don't want to show it winning

The chart is cursorbench though and they asked about the "deceptive graph"

They can benchmark it because you can use an openai api key with cursor. Astra is just not included in the cursor plan.

Elon posted on X that Grok 4.7 is behind Claude and OpenAI for agentic coding:

https://x.com/elonmusk/status/2102082011233931762?s=20

so it's likely about usage in Cursor specifically.



this is exactly it.

There is difussion.cpp which is intended for those types of models. I set up krea-2-turbo with the help of ChatGPT 2 months ago, if you have a capable computer that's what I would suggest once it becomes supported.

Things like that - and other examples posted here - are why I 'm sticking with OpenCode despite it having some papercuts that annoy me.

The incentives are not there for them to do shady stuff like vacuum your files, inflate your token count just because or many other things.


Open code is great and I use it, however, they were caught uploading prompts to their summarization AI instead of using the configured AI model endpoint. This has since been fixed.

That said, running in a completely offline mode remains unnecessary difficult to configure. In particular, toggling off Zen seems to require a community plugin.


While I liked your comment enough to upvote, I then realized that Chairman isn't an operational position. It's more a position where the person can exercise ownership of a company (to put it in layman terms). In your analogy he wouldn't be the athlete, more like the person who gets the winning money (assuming his son also has shares) and chooses the trainer/coach of the athletes (which still needs skill btw).

But we have been told for a long time that Buffett was a very active Chairman.

Warren was CEO until 9 months ago. While yes, he's been chairman for 50 years, he's been an exclusive chairman for only 9 months. The role is ambiguous anyways and it's pretty obvious that someone else won't have the authority Warren has. And his son has been a director for 30 years so it's not like he's a new hire

I appreciate this correction as I had misremembered his dual roles.

He was chairman and CEO for a long time. Since he has only been a chairman for not that long I don't think it's fair to say that he's been a very active chairman for a long time.

I appreciate this correction as I had misremembered his dual roles.

Whose strategy is explicitly to hold good stocks for a very long time.

Why mostly unusable 125b?

I assume you are talking about qwen3.8-flash-next. Support for it on some places, like llama.cpp, is still wip (depending on configuration) but it looks like a very capable model in it's category.


Very capable yes but very slow. 27B is relatively easy to run, but the 125b one need around 128Gb of RAM (DDR4 isn't enough, you need DDR5 to be quick enough, that's $3000 alone, you also need a graphic card). DDR4 is caped @20tps.

So, the cost of a setup to run Qwen-Flash-Next at +40tks is around $3000. Too much for most people.

With only a RTX 4090, you will reach 30tps (with DDR5...), not +40tks, and it's about the limit to be usable. Oh ! I forget Apple device too, it's a good option to run this model I guess, but still slow.

Yet, as you said, it's still a wip implementation, it may improve soon (MTP support is about to be merged in llama.cpp soon).


Some years ago Firefox was the go to browser if you wanted to have "some" privacy in your browsing (together with uBlock origin). With news like this I m really wondering if my views are outdated and I need to to some good researching on maintaining some privacy in what I m browsing.

Mozilla has been in a bit of a recursive feedback loop death spiral for years now of: Browser loses market share -> try some weird thing -> very few people like or use it -> browser loses market share -> management says "oh shit we're losing market share we better try some weird thing"

I say this as a person that uses firefox with ublock origin 99.5% of the time. It's better in my opinion than Chrome. I can at least easily turn off the enabled by default crap features. Like the advertising and sponsored news links on the default new tab page.


This is the confidently incorrect narrative that gets repeated in the comment sections practically every time Mozilla is mentioned but it's every bit as incorrect now as it's been all the previous times.

So here we go again for the millionth time: the big losses of Mozilla market share were approximately during the 2010s. The era of side bets on unique features is approximately the 2020s. The unique features didn't retroactively cause the market share losses of the 2010s.

Moreover, telling the market share story in terms of specific browser features misses the elephant in the room, which is that Google, with the world's most visited page, and a browser that's the installed default on over a billion devices, grew it's market share with a combination of web visability and dominance over the most used mobile platform. Mozilla could triple their budget and have the world's best browser experience, but it wouldn't make much of a dent against distribution defaults.

I would wager that the impact on market share is driven about 97% by Googles distribution advantage and 3% by aligning with user preferences on features and performance. If being a perfect browser led to market dominance, Opera would have already conquered the world back in 2012, but the economics of building a browser aren't always friendly to the good guys.


FF's market share crashed because it was painfully slow compared to chrome when it first came out. The marketing helped, but it wouldn't have stuck if it hadn't been better. You can only blame users being too stupid to know better for just so long.

Exactly. Chrome's growth has been driven primarily by monopolistic behavior, and the fact that FF was in a bad spot when Chrome came out.

Firefox is frankly quite good right now, yet they still are losing or not gaining users.


I never said that Chrome and Edge don't have a huge distribution advantage due to being installed by default on peoples' devices. That's also certainly a huge factor in the mass adoption of Chrome as what people consider "the web browser. Or Safari as default browser in MacOS of course.

Why do they keep trying weird things? That makes no sense. People just want a browser.

Because just doing a browser is not working. Market share keeps declining. And Firefox is pretty good. They spend the majority of their resources on Firefox, and always have.

very few people like or use it -> browser loses market share

I very much doubt trying weird things meaningfully changed their market share (aside from UI redesigns)


I very much think it did, because FF once was recommended by us nerds, we installed it for people and praised it. And that made an impact outside of our circles.

Now I still occasionaly install it along with ublock origin because there is no alternative, but I don't praise it anymore (or bother to install it for someone in the first place) - but rather bitch about how they also sneak in advertisement and spyware.


I think it didn't help. Constantly adding or changing features makes it less consistent and predictable. For example several family members were confused by what is "pocket".

They have twice (that I'm aware of) tried rebranding as an advertising company. They are funded by Google (though not _necessarily_ influenced by them). There was an article in LWN by them about (in part) how they can't get enough information about users of Thunderbird so they were suggesting telemetry should be enabled by default and opt out instead of opt in because most people won't opt out.

I think they get a lot of credit because they aren't Google and not enough push back


Here's a crazy idea if you don't want Smart Window to send your data anywhere:

have you tried not using smart window (like you're currently doing)?


The problem is, many of the privacy invading "features" are opt-out, instead of opt-in...

https://privacytests.org/


Those people complain about everything. It's exhausting.

Fyi, you may want to check out Librewolf: https://librewolf.net/

They are and Mozilla is basically malicious org

Guidelines | FAQ | Lists | API | Security | DMCA | Apply to YC | Contact

Search: