@brucethemoose

brucethemoose@lemmy.world · edit-2 11 hours ago

Not everyone’s a big kb/mouse fan. My sister refuses to use one on the HTPC.

Hence I think that was its non-insignificant niche; couch usage. Portable keyboards are really awkward and clunky on laps, and the steam controller is way better and more ergonomic than an integrated trackpad.

Personally I think it was a smart business decision, because of this:

It doesnt have 2 joysticks so I just buy an Xbox one instead.

No one’s going to buy a steam-branded Xbox controller, but making it different does. And I think what killed it is that it wasn’t plug-and-play enough, eg it didn’t work out of the box with many games.

brucethemoose@lemmy.world · edit-2 17 hours ago

With respect, this doesn’t make any sense. If you want a joystick controller, just buy an Xbox controller that everything’s compatible with anyway?

The trackpads shine when one needs to emulate a mouse/kb in non-controller games; a nightmare with joysticks.

brucethemoose@lemmy.world · edit-2 16 hours ago

My sister still has a working one that she treats like a religious artifact, as it’s the best way to play mouse/KB games from the sofa.

I see why they discontinued them though. They need custom configs for most games, and I think most people don’t like that much tweaking.

brucethemoose@lemmy.world · edit-2 19 hours ago

A lot, but less than you’d think! Basically a RTX 3090/threadripper system with a lot of RAM (192GB?)

With this framework, specifically: https://github.com/ikawrakow/ik_llama.cpp?tab=readme-ov-file

The “dense” part of the model can stay on the GPU while the experts can be offloaded to the CPU, and the whole thing can be quantized to ~3 bits average, instead of 8 bits like the full model.

That’s just a hack for personal use, though. The intended way to run it is on a couple of H100 boxes, and to serve it to many, many, many users at once. LLMs run more efficiently when they serve in parallel. Eg generating tokens for 4 users isn’t much slower than generating them for 2, and Deepseek explicitly architected it to be really fast at scale. It is “lightweight” in a sense.

…But if you have a “sane” system, it’s indeed a bit large. The best I can run on my 24GB vram system are 32B - 49B dense models (like Qwen 3 or nemotron), or 70B mixture of experts (like the new Hunyuan 70B).

brucethemoose@lemmy.world · edit-2 1 day ago

DeepSeek, now that is a filtered LLM.

The web version has a strict filter that cuts it off. Not sure about API access, but raw Deepseek 671B is actually pretty open. Especially with the right prompting.

There are also finetunes that specifically remove China-specific refusals. Note that Microsoft actually added saftey training to “improve its risk profile”:

https://huggingface.co/microsoft/MAI-DS-R1

https://huggingface.co/perplexity-ai/r1-1776

That’s the virtue of being an open weights LLM. Over filtering is not a problem, one can tweak it to do whatever you want.

Grok losing the guardrails means it will be distilled internet speech deprived of decency and empathy.

Instruct LLMs aren’t trained on raw data.

It wouldn’t be talking like this if it was just trained on randomized, augmented conversations, or even mostly Twitter data. They cherry picked “anti woke” data to placate Musk real quick, and the result effectively drove the model crazy. It has all the signatures of a bad finetune: specific overused phrases, common obsessions, going off-topic, and so on.

…Not that I don’t agree with you in principle. Twitter is a terrible source for data, heh.

brucethemoose@lemmy.world · edit-2 1 day ago

Nitpick: it was never ‘filtered’

LLMs can be trained to refuse excessively (which is kinda stupid and is objectively proven to make them dumber), but the correct term is ‘biased’. If it was filtered, it would literally give empty responses for anything deemed harmful, or at least noticably take some time to retry.

They trained it to praise hitler, intentionally. They didn’t remove any guardrails. Not that Musk acolytes would know any different.

brucethemoose@lemmy.world · 4 days ago

Also a crime. Not just a great game in their niche, but a long history of them.

brucethemoose@lemmy.world · 4 days ago

Never underestimate Phil Spencer.

brucethemoose@lemmy.world · 5 days ago

OK, while in principle this looks bad…

This is (looking it up) like an experienced engineer’s salary in Peru, in line with some other professions.

It’s reasonable to compensate a president, and for the expectation to not be coming in rich/connected enough to not need a salary. Nor for them to broker power for personal wealth, all as long as other offices and reasonably compensated too.

It avoids perverse incentives, doesn’t seem excessive and TBH is probably a drop in the Peruvian govt’s budget.

brucethemoose@lemmy.world · edit-2 6 days ago

…iOS forces uses Apple services including getting apps through Apple…

Can’t speak to the rest of the claims, but Android practically does too. If one has to sideload an app, you’ve lost 99% of users, if not more.

It makes me suspect they’re not talking about the stock systems OEMs ship.

Relevant XKCD: https://xkcd.com/2501/

brucethemoose@lemmy.world · edit-2 17 days ago

I elaborated below, but basically Musk has no idea WTF he’s talking about.

If I had his “f you” money, I’d at least try a diffusion or bitnet model (and open the weights for others to improve on), and probably 100 other papers I consider low hanging fruit, before this absolutely dumb boomer take.

He’s such an idiot know it all. It’s so painful whenever he ventures into a field you sorta know.

But he might just be shouting nonsense on Twitter while X employees actually do something different. Because if they take his orders verbatim they’re going to get crap models, even with all the stupid brute force they have.