Why not? Friction from tidal forces keeps the moon got enough to contain liquid water under its ice cover. This makes it similar to Europa in some ways.
Models can't learn from misbehavior after training. Any session is an independent context and there is no mode for punishment or deterrence in production.
Corrective punishment in the real world relies on the receiver's rational and emotional responses as well as their ability to remember that episode. Even animals respond to such treatment. None of these levers exist for ussrs of LLMs.
> Models can't learn from misbehavior after training.
Some of these events were during testing; I do not know if this test was during training or after, it could have been either.
> Any session is an independent context and there is no mode for punishment or deterrence in production.
Not so, at two levels.
For the companies behind the models: this is why they sometimes throw you A/B tests for which answer you prefer, and still have up/down vote buttons on responses. Those things go into training the next model or iteration of the current model. It's still useful to only deploy checkpoints, but the point is "useful", not "necessary".
For the users: if you have monitoring to detect output, you can trigger interrupts, and injections of "no, stop!" even as a plain English string because it understands natural language.
> Corrective punishment in the real world relies on the receiver's rational and emotional responses as well as their ability to remember that episode. Even animals respond to such treatment. None of these levers exist for ussrs of LLMs.
LLMs impersonate humans. This role-playing does allow them a degree of, if not feeling emotion, at least acting like they experience it.
I expect the problem is that the models are trained to obey the user so hard they're often not willing to push back and say "no" when they ought to. I mean, the logs show the agents were identifying the actions as bad, so it isn't like this was simply the agents being unable to tell right from wrong.
A terminally cynical mind might insinuate here that focusing on the product is a way for AI companies to keep doing their own business as usual, no matter how negligent that may be.
Is it? A lot of software is really just a pile of context-specific database front-ends. There is virtually nothing difficult in the implementation itself. Most businesses in this space survive because they know how to dress up their pile of CRUD for their customers. And it gets worse: a less reliable, buggier product may still be the better fit for a customer if it works well enough for what they use it the most.
I also hate preachy sermons about new tech with a deep passion and that transfers somewhat to the tech they preach about.
Despite this, there is a reality that LLMs for coding can create pretty decent results fast. There is something real and usable hidden in all that mess. It's like the dot com bubble in some ways: the bubble died with all the overhyped "x on the web" BS that had no real substance, but the web itself stayed. This time around, LLMs and agents will stick around as tools, but the hype will likely die a loud, messy death.
Nobody here is saying otherwise. The Codeberg policy in question does not forbid LLM-assistance outright; their policy is targeted towards unmitigated slop. Even myself, as I argue heatedly against the religious adherents, not only use LLMs daily but work for a very successful LLM startup developing and selling LLMs-as-tools. That does not make the zealots who won't shut the fuck up about LLMs-as-gods any less infuriating.
Furniture is actually an ongoing service on the scale of public organisations. When I worked for a Danish city we had a set of subscriptions for furniture, depending on what sort of institution it was. This probably won't surprise you but city hall had the supreme plus subscription which meant that furniture was replaced regularily, while schools had the shit tier subscription meaning the furniture almost had to break into pieces before it got replaced. It's done with purchase agreements for an entire sector.
I think the comparison is fair enough considering that self-hosted would work sort of similar. In that you'd rent the physical space for your hardware (and possibly the hardware itself) at local suppliers, and not actually have it in your basement.
Pinch to zoom and two finger scrolling work well for me in most Windows apps. Apple has a few touchoad gestures that I don't believe have equivalents elsewhere, but they would be pretty neat.
The main issue with going beyond the mouse is that it involves changes across the event handling stack. For track pads, you have to build multi-touch infrastructure up ro a point and then there is a decision: what are gestures that are handled at a desktop level? What are gestures that get converted into semantically similar events (e.g. tap->LMB click)? What gets passed through as multi-touch? This ends up touching a lot of parts of the desktop software stack and the amount of required buy in to get this done on Linux seems massive.
Then change the stack. We can't be stuck on obsolete inefficient mental models. We are programmers, we can change software to fit new requirements. We already found ways to deliver keyboard input to the focused window.
That is definitely a big motivator for publishing preprints. Journal submissions can take up to a year. Comference submissions take months. If the field is moving fast, claiming a finding early can become an important career move.
In older days, academics would just share notes on their work and word wouldn't usually spread widely before publication.
Preprints may be the better model. But public visibility means that non-experts now get to see the good and the bad research equally, but they won't have the domain knowledge and skill to distinguish one from the other with confidence.
No, hardening has nothing to do with it. They have a feature flag system deeply embedded in their software. And such systems can totally distribute different flags to different people. And that's how you could easily create VIP accounts with different behaviors than the rest of the world, like all ad festures turned off.
I guess it was simply too close to reality. That's why I didn't register that as humor. Text also doesn't carry tone of voice well, which makes confusion more likely
It depends entirely on the purpose of the forgery. Some grocery stores do ID checks by looking at the front of the ID. Others just run the ID across a scanner and the employees are so rushed they don't read it or check the picture. Similar things happen e.g. at bars or casinos. Incomplete forgeries can get you far enough under the right circumstances.
reply