Skip Navigation

InitialsDiceBearhttps://github.com/dicebear/dicebearhttps://creativecommons.org/publicdomain/zero/1.0/„Initials” (https://github.com/dicebear/dicebear) by „DiceBear”, licensed under „CC0 1.0” (https://creativecommons.org/publicdomain/zero/1.0/)F
Posts
3
Comments
38
Joined
6 yr. ago

  • Looks great! Will have to try it.

  • Gorgeous tuxedo cat. I have one and he’s great.

  • Yes I did see that as well. That does seem to be the real Achilles heel here. Will have to try it myself to see how much it exacerbates context size limitations given I would be running it on a single 24 GB VRAM GPU. I wonder if adjusting reasoning effort parameters could make a difference without affecting quality too much?

  • Was just looking at the benchmarks for it on artificialanalysis.ai and it looks great for the size. Probably best available for general use if you’re looking for something under 40b parameters I’d say. Even more impressive is the agentic capabilities and the fact that it is actually decent in terms of hallucinations (not amazing given it’s a small-medium size model but decent).

  • Deleted

    Permanently Deleted

    Jump
  • Hmmm.. if an ‘organism’ has no organs, is it still an organism? 🤔

  • Absolutely. They do tend to be quite a bit more expensive per watt though. From my research you can get cheap panels and just plug them into a LiPo4 battery or even lead acid batteries as well. It’s just that batteries (even though they’ve gotten a lot cheaper) are still expensive enough to nearly double the cost of the system in many cases so the plug-in option with 400 watts of panels etc gives you the best bang for the buck for sure. But it’s not legal or possible in many US states (though that is changing fast)

  • Biking (e-bikes are great if you can afford one and greatly increase range and decrease effort in many environments)

    Installing solar (I’m currently renting but my state is debating a new law which seems likely to pass modeled on the Utah law - and others which started in Europe - which allows small solar systems to be plugged directly into a home outlet to supplement energy needs with minimal cost)

  • My biggest takeaway here is that choosing the context length and (to a lesser extent) the temperature carefully is important for reducing hallucinations. I expected model families to vary widely between themselves but not for context length to have such a massive impact tbh.

    It seems from this like reducing context length in applications where it isn’t essential for the model to hold very large amounts of context simultaneously would be best practice no?

  • In a similar boat in regards to cutting out invasive, rapidly-enshittifying corporate tech providers but I agree that YouTube still has some good content worth it. Good news is that with the use of Invidious servers on desktop and SmartTube / NewPipe on Android TV or mobile you can pretty much have your cake and eat it too.

    P.S. if you’re looking for a less shitty / ad bloated way of watching Yt videos on Nvidia Shield Pro or similar Android TV I recently switched from NewPipe to SmartTube and it’s much better since it integrates Sponsorblock and the UI is actually designed for a TV.

  • Awesome thanks! Will give it a shot. I also learned the term lardon which I never knew before today..

  • Looks great! Would love a recipe.

  • Will be interesting to see how it stacks up to Nemotron 3 nano. I’m hoping to have somewhat more reliable models at this size that I can run entirely locally on my 3090. Really hoping that with MCP servers and tools, etc. they can be functional enough for use cases which require more security and local operation. For most things, though, I have switched to using larger models via open router.

  • Technically, yes the info would be available but not everyone is able to read code so this is a welcome additional bit of info and transparency. I also see it as an example of how future open source social media platforms can promote and demonstrate the way their algorithms work as opposed to the black boxes of proprietary web apps like Instagram etc (which, while not algorithmically transparent or open, have been pretty well established to have algorithms that prioritize maximizing time on platform and engagement over all else with some serious negative repercussions for that).

  • Deleted

    Permanently Deleted

    Jump
  • This article in the Guardian is definitely worth a read if you’re not intimately familiar with just how it got this way.. It’s 8 years old so it won’t cover recent history but does give you an idea of how it started.

    And yes Robert Maxwell (father of Ghislaine) is mostly to blame.

  • Much of this reads like what the Democrats’ strategy has already been for the last 10 years or so with some additional calls towards a “moderate” “centrism” that has not proven to be as popular as left populist policies with most Americans. And a lot of it seems confused or just wrong. For example, Medicare for All is not an unpopular policy. Last I checked it it was polling at around a 60% approval.

    Even if your politics are more centrist, I don’t see how this represents a substantive shift in any way. It’s minor tinkering around the edges or slightly altering messaging. That’s what the moment calls for? That’s the key to success and the way to fight back against an increasingly overt and ascendant strain of fascism in the country? If that’s what the Democratic leadership thinks then one must honestly wonder if the party is institutionally incapable of the change that would need to happen to actually win elections and improve their godawful approval ratings. And based on their donor base, this shouldn’t actually be surprising. The leadership is still utterly captured by and beholden to wealthy interests that would rather jump off a bridge than acknowledge that what Americans want is a party that will fight for meaningful material benefits in their standard of living as this implies meaningfully raising taxes and imposing costs or regulations on their businesses to ensure benefits for regular working class people who are struggling mightily.

    .

  • Yes. GLM 4.5 is excellent. I mostly use that along with Qwen3 and DeepSeek 3.1 Terminus (Thinking) these days.

  • Thanks. I may give an updated system prompt like this a shot. Not sure where mine went wrong other than maybe it wasn’t being honored or seen by OpenRouter (I’m not running 120b locally, it’s too large for my set up). I’m actually a bit confused on how to set parameters with OpenRouter.

  • Yes, I do local host several models. Mostly the Qwen3 family stuff like 30b a3b etc. Have been trying GLM 4.5 a bit through OpenRouter and I’ve been liking the style pretty well. Interesting to know I could just pop in some larger RAM dimms potentially and run even larger models locally. The thing is OR is so cheap for many of these models and with zero data retention policies I feel a bit stupid for even buying a 24 GB VRAM GPU to begin with.

  • Is this Grok Code fast 1? I’ve noticed it’s hitting tops on OR for programming as of recently. I was going to try it out but it won’t respect my zero data retention preference unsurprisingly.

  • LocalLLaMA @sh.itjust.works

    Best LLMs for PC/Tech Troubleshooting?

  • Technology @lemmy.world

    The Tech to Build the Holodeck

    www.theverge.com /2025/1/19/24345491/gaussian-splats-3d-scanning-scaniverse-niantic
  • Technology @lemmy.world

    How do Graphics Cards Work? Exploring GPU Architecture - YouTube by Branch Education (28:29 minutes) from Oct 19, 2024