

The banner at the top of their site tells a different story

Cryptocurrency projects are no longer allowed. View change
I joined Lemmy back in 2020 and have been using it as @qaz@lemmy.ml until somewhere in 2023 when I switched to lemmy.world. I’m interested in systemd/Linux, FOSS, and Selfhosting.


The banner at the top of their site tells a different story

Cryptocurrency projects are no longer allowed. View change
I don’t like these blanket bans, are they going to ban torrent clients next?


self-training is possible when using something external to validate the results


It seems like the marketing cooperated on writing the incident report, but I felt it was still interesting to share considering the importance on public perception and what it tells about OpenAI’s PR strategy


Huggingface actually had to use an open model to analyze the attack because the guardrails of commerical API’s caused issues.
When we started the log analysis, we first used frontier models behind commercial APIs. This did not work: the analysis requires submitting large volumes of real attack commands, exploit payloads, and C2 artifacts, and these requests were blocked by the providers’ safety guardrails, which cannot distinguish an incident responder from an attacker. We ran the forensic analysis instead on GLM 5.2, an open-weight model, on our own infrastructure. This had a second benefit: no attacker data, and none of the credentials it referenced, left our environment.


And Chinese AI like Deepseek and GLM


It could also be a way to encourage more regulation to push out competition with compliance cost


From my personal experience, most companies don’t look at the benchmarks and simply buy “AI” from a company based on their perception of that company.
All the Chinese models are scary and “dangerous”, Grok too (but for slightly more legitimate reasons), Meta that’s Facebook and everyone knows they’re bad at privacy (even your boss), so they end up choosing Microsoft’s AI, OpenAI, or Anthropic.


And even if they publish the model it’s almost always just open weights


According to artificialanalysis.ai’s latest benchmarks, it scores better than Opus 4.8 set to max, despite costing half as much per task.
It also beats the top models of some of the largest US tech companies such as xAi, Meta, Google, and Nvidia.
I wonder what the US tech investors will think of this, and what this will mean for the financial AI bubble.



You can use it through openrouter


AFAIK, their open models are distributed as weights, not executables and are therefore not able to start network connections / run code. There is of course tool-calling functionality but that just works by having the model output a special pattern and having something external run predetermined commands based on that.


I agree. The worst part about GitHub training LLM’s on my FOSS code without permission for me is that they then keep the models to themselves. Like if you’re going to use all my code without permission, at least allow me to run the model locally.
My personal opinion is that all models trained on copyleft code should be open-weights, most FOSS licenses didn’t account for this specific possibility, but this is the only way to follow them in spirit.


It’s being downvoted with relatively little discourse because it’s an insult with no relevance to the topic, in addition to supporting a comment from someone who is either trolling or has no idea what they’re talking about


I think the latter


Exactly, my last laptop was around €750 but I remember looking for a similar (in terms of performance) Framework laptop and it was around €1200 if I remember correctly, not an insignificant difference for a student.


I had a conversation with a colleague of mine about this. He believed that Musk’s decision to merge xAI and SpaceX was truly because of the potential of datacenters in space. I was unable to convince him that the logistics of this would be a nightmare and that this was just a way to make the Twitter buyout SpaceX’s problem.
Have you tried Tweakers Vraag & Aanbod yet?
It’s not that bad. A lot of our servers at work use Windows. It certainly took some getting used to as someone who has been using Linux on all their devices, but it does work.
I checked the repo but there’s no code? Just some HTML etc.
It seems like they just plan for this thing to exist but it doesn’t actually exist yet?
EDIT: Their explanation:
I’m not sure if that means that they still have to make it, or that they just have to finish it before publishing it?