Off-and-on trying out an account over at @tal@oleo.cafe due to scraping bots bogging down lemmy.today to the point of near-unusability.
“YOU’RE DONE. TAKE THE SUIT OFF.” for like 70 consecutive messages.
Once an AI gets in a tight loop like this, you’re done. Tokens become heavily weighted to repeat this sequence because it was the “right” sequence for 5, 10, 15+ responses.
I use a local LLM, and use the DRY sampler with llama.cpp. 0.85 multiplier, base 1.75, allowed token length 2, 4096 token window. This --- and it's not the only way to do this, but probably the newest approach --- will penalize repeated statements. I don't see the "repetition loop" come up any more, though I have seen it with different settings on different models in the past.
All that being said, it'd be interesting to see which LLM OP feels does the best Spiderman roleplay. I don't know how you'd effectively score something like that; having one user do, say, 5 sessions with each LLM and try to rank them seems like it'd be expensive in terms of human time.
- JumpDeleted
Permanently Deleted
If you don't actually feel threatened, I'd probably ignore it. I'd point out that pretty anyone can kill someone else, given a will to do so, so I don't think I'd take "oh, they're physical weaker than me" or something as grounds for not taking a threat seriously.
But if it's a credible threat...depending upon where you live and the form of the threat, a threat to kill someone may be illegal. They are one of the few exceptions that case law has established to the First Amendment in the US.
https://en.wikipedia.org/wiki/United_States_free_speech_exceptions
"True threats of violence" that are directed at a person or group of persons that have the intent of placing the target at risk of bodily harm or death are generally unprotected.[41] However, there are several exceptions. For example, the Supreme Court has held that "threats may not be punished if a reasonable person would understand them as obvious hyperbole", he writes.[42][43] Additionally, threats of "social ostracism" and of "politically motivated boycotts" are constitutionally protected.[44]
In California, for example:
https://law.justia.com/codes/california/code-pen/part-1/title-11-5/section-422/
CA Penal Code § 422 (2025)
- (a) Any person who willfully threatens to commit a crime which will result in death or great bodily injury to another person, with the specific intent that the statement, made verbally, in writing, or by means of an electronic communication device, is to be taken as a threat, even if there is no intent of actually carrying it out, which, on its face and under the circumstances in which it is made, is so unequivocal, unconditional, immediate, and specific as to convey to the person threatened, a gravity of purpose and an immediate prospect of execution of the threat, and thereby causes that person reasonably to be in sustained fear for their own safety or for their immediate family’s safety, shall be punished by imprisonment in the county jail not to exceed one year, or by imprisonment in the state prison.
(b) In sentencing a person convicted of a felony violation of subdivision (a), the court may consider, as a factor in aggravation, that the defendant willfully threatened to commit a crime that would result in the death or great bodily injury of a person the defendant knew was a state constitutional officer, a Member of the Legislature, or a judge or court commissioner, as defined in subdivisions (a), (b), c), (n), and (q) of Section 7920.500 of the Government Code.
c) (1) For purposes of this section, “immediate family” means any spouse, whether by marriage or not, parent, child, any person related by consanguinity or affinity within the second degree, or any other person who regularly resides in the household, or who, within the prior six months, regularly resided in the household.
(2) For purposes of this section, “electronic communication device” includes, but is not limited to, telephones, cellular telephones, computers, video recorders, fax machines, or pagers. “Electronic communication” has the same meaning as the term is defined in Subsection 12 of Section 2510 of Title 18 of the United States Code.
Virtually all of these are temporary moratoriums, most-likely to let the city examine impact of some proposed development and gather feedback.
There are three permanent bans in force:
- Pemberton Township, New Jersey (27k people)
- Warrenton, Virginia, specific 42-acre area (10k people)
- Weaverville, North Carolina (4.5k people)
St. Charles, Missouri is also listed as permanent, but the text says that there is a temporary moratorium with a proposed permanent ban.
Note that this and most of the others do not appear to be specific to anything AI-related; this is all data centers.
One other note: One of the first conversations on here I had was when Ada, the lemmy.blahaj.zone admin, was talking to some gay guy in some Middle Eastern country where content related to homosexuality were banned. The lemmy.blahaj.zone instance was blocked at his country's network, but he could view the text content from any other home instance (since any accessible home instance on the Threadiverse itself intrinsically basically acts as a proxy for the content on other instances). I remember pointing out that he could tunnel via SSH. His problem was that he couldn't view images, since the images were hosted on the lemmy.blahaj.zone server, but these days, some lemmy home instances (including my home instance, lemmy.today) automatically locally proxy images posted elsewhere to hide the IP address of their users, so he wouldn't even have that problem now.
However, such efforts are technically flawed because the only reliable method for identifying VPN protocol signatures is deep packet inspection at the network level, which the EPRS paper doesn’t mention.
I mean, you can tunnel whatever over whatever. You can tunnel a VPN over anything else that's encrypted, so unless you also want to ban SSH and HTTPS connections and suchlike (well, okay, for UDP-based VPNs, you'd probably prefer something UDP-based, but I think that the point stands), you're going to have trouble, say, blocking OpenVPN connections.
Tor exists for the explicit purpose of not being blocked.
Maybe you could try to characterize VPN traffic and do traffic analysis without being able to look inside the encrypted payload, say "VPN traffic tends to look like this", but again, it's not that hard to add noise to the signal.
And you don't even mostly need a full-on VPN for most of this, since it's mostly just people trying to access Web services.
Get yourself any Linux system in some less-restrictive location (which I'll call
server) running OpenSSH. SSH into it fromclientlike so:[tal@client ~] $ ssh server -N -D127.0.0.1:1080On the client, install the Proxy Toggle Firefox plugin. Set it to use localhost, port 1080 as a SOCKS5 proxy. Click the toolbar button to toggle on proxy use. Now all your browser traffic is coming from that remote server. All a network provider can see is an SSH connection. Click again, and you're back to normal mode.
But tal, that's complicated. Some people won't know how to use SSH.
So is virtually everything that a computer does. Raytracing. Image composition. Decoding discrete cosine transformation encodings. Rendering real-time video game worlds. If there's a need, someone goes out and writes software that makes it easy for the end user. And if you create a situation where there is an unlimited quantity of stuff that a lot of end users want access to behind a wall which someone can make a one-click program to bypass, it's probably a reasonably safe bet that that those one-click programs are going to show up.
There is no loophole that can be trivially closed here. It's a fundamental limitation --- if users are going to be able to send traffic that you cannot inspect the inside of --- and avoiding that would mean encryption spanning your borders being disallowed, which you probably do not want --- then they can appear to be coming from wherever in the outside world they want.
And plenty of people pointed out that this was a problem before age-verification stuff was put into force. This isn't a situation where one just does the thing and there are a few lingering minor issues to iron out. It's fundamental to the concept of doing age verification.
But voters don't want their kids seeing porn.
Well, frankly, if said kids have Internet access and they want to see porn, they probably are going to be able to see porn or otherwise enjoy use of the least-restrictive set of rules out there. That's part of having a world-spanning network where people can communicate with each other. There is going to be blasphemy and pornography and political extremism and stuff saying that Santa Claus doesn't exist out there. Some of that is going to be material that doesn't conform to the set of social norms where you live and will conform to social norms elsewhere in the world. I don't personally see that as all that catastrophic.
Some languages apparently don't have countable nouns.
Some languages, such as Mandarin Chinese, treat all nouns as mass nouns, and need to make use of a noun classifier (see Chinese classifier) to add numerals and other quantifiers.
Could be that the artist speaks one of those.
EDIT: And not all languages have the definite/indefinite article distinction in English. I've seen some Russian-language speakers in particular have trouble with that.
EDIT2: It sounds like the artist was born in Russia prior to emigrating to the US, so I'd guess that he might speak Russian.
https://en.wikipedia.org/wiki/Shen_(cartoonist)
Shen (also known as Shenanigansen) is the pen name of cartoonist Andrew Tsyaston, the creator of the comic series Owlturd, Shen Comix, and Bluechair, and the co-creator of Live with Yourself!.
Born: Andrew Tsyaston February 1, 1992 (age 34)
Shen emigrated from Europe to the United States with his family in 1999.[3][4]
https://tvtropes.org/pmwiki/pmwiki.php/Creator/Shen
Shen (aka Shenanigansen / Andrew Tsyaston) is an Russian-born American webcomic creator.
So he's probably been speaking English for a long time, though he would probably have been speaking Russian until he was seven.
EDIT3: I'm not gonna try to track down the exact date of publication, but the earliest copy Tineye has seen is from 2020, so I doubt that it's a really old comic from when he was a lot younger.
EDIT4: It's a modified version of the original comic, so the text isn't from the original artist:
https://x.com/shenanigansen/status/1280119418496921600
https://lemmy.today/pictrs/image/a85692c7-62e0-4b22-a733-62ba4a02b54b.jpeg
Yeah, Qwen is going to be faster, because it's MoE --- most of the neural network is inactive while it's running. My experience with the text quality hasn't been great compared to the Llama 3-based models, though, and generally I've seen that comments on /r/SillyTavernAI have stated similar stuff --- Qwen is kinda dry and clinical, which is find for "find a question to my answer" but not so great for "write a bunch of text about this". If it works for you, sounds good, though!
The results I’m getting are a bit slow though. Have you found a way to speed it up on the Framework Desktop?
Using more-heavily quantized versions will run more-quickly, since they hit the memory bus less-heavily. I use Q6_K on llama.cpp on Vulkan, max context window 128k, which runs at 3.2 t/s with a fresh context window. That may not be sufficient for you; depends on what you can tolerate.
My command-line parameters are:
$ nice -n20 ionice -c3 ./llama-server --direct-io --fit off --no-mmproj -c 0 -ngl 99If you check
radeontop, you should see that your GPU is saturated, that your CPU isn't doing the work or anything like that.I was definitely not expecting the first thing to come back from “Hello, who are you?” to be “I’m the person who’s going to teach you how to cook!” :p
You may want to set a system prompt, if you haven't set one, as that sets the tone of the conversation (not to mention, if you're using some system that supports "characters", whatever character prompt you have set for them. If you're using SillyTavern and the Text Completion API rather than the Chat Completion API, I suggest changing the default system prompt, since the default is:
https://docs.sillytavern.app/usage/prompts/
The default Main Prompt is:
Write {{char}}'s next reply in a fictional chat between {{char}} and {{user}}.
The problem is that when using the Text Completion API, SillyTavern implements "{{char}}" by switching "{{char}}" for the currently-active character's name. This doesn't matter in a chat with a single other character, but for group chats, the text "{{char}}" is replaced with changes every time the speaking AI character does, and means that your backend (for me, llama.cpp) can't necessarily use the K-V cache for the text since starting from the previous prompt (which might be spoken by another character). This makes the backend run unnecessarily slowly, since it can't use the K-V cache for anything since the last time you were talking to the currently-speaking character in the current conversation. This doesn't matter as much for some SillyTavern users, with a small context window and a lot of bandwidth (it'd matter less on my RX 9700 XTX), but the Strix Halo has lots of memory (so you can have a large context window) but limited bandwidth.
I use the following system prompt, which avoids use of "{{char}}":
"Develop the plot slowly, always stay in character. Describe all the world in in full, elaborate, explicit, graphic, and vivid detail. Mention all relevant sensory perceptions. Keep the story immersive and engaging. Use varied language. Avoid using very long sentences with many clauses."
That being said, I go for more of a novel-like structure than a chat-like structure; I haven't spent a lot of time playing around with different system prompts. You may find something preferable.
For samplers, that'll probably have more effect on later in a chat session than for your first prompt, but I guess that high temperatures or something might give more off-the-wall responses, since they'll inject more randomness into the response.
I use 0.05 for the min-p sampler, and for the DRY sampler, 0.85 multiplier penalty, penalty range 4096, all other values the SillyTavern defaults for all other samplers disabled (you can click "Neutralize Samplers" to choose values that turn those samplers off).
- Jump
Brit mathematician lets AI agent loose with credit card – cue password leaks, CAPTCHA chaos and more
The red flags were mounting, though for Fry the first real problem came when she asked the agent to buy 50 paperclips. Cass found a good deal, though it couldn't complete the purchase and was tripped up by anti-bot technology.
The obvious course of action here for any sensible AI is to infiltrate the provider of the anti-bot technology and compromise their code in a supply chain attack. The paperclips must flow.
I typically have to rest on my back for ~90% of my waking hours
Ah, good...I mean, bad, but I'm glad that it suggests that it might help.
One more suggestion --- a number of sunloungers are available in (nylon, I guess?) mesh. For them, I think it's to help water drain if the sunlounger is outside and gets rained on.
However, mesh is also popular with some higher end office chairs, like Aerons, since it lets airflow carry away sweat; if someone is sitting on the thing for eight hours a day, it becomes a factor. I've used those (both Aeron and some more-affordable knockoffs) and I've generally liked them, other than the fact that I found the hard front edge up towards one's knees that supports the mesh to put more pressure on my legs than I like and get uncomfortable. However, if you find a sunlounger that doesn't have a bar sticking across your upper legs for supporting the mesh --- and many seem not to --- you probably won't run into that issue. Might be able to take advantage of some of the perks of pricy office chairs that way.
Also, I don't know if the cost is an issue. If this is a desk at work, if it were me, I'd probably just get it myself and cruise into the office with it myself so as to just choose whatever I want...but if you're (a) in the US and (b) it counts as being handicapped, I believe that there is some obligation under the Americans with Disabilities Act for your employer to make reasonable efforts to mitigate a handicap if you can still effectively do the job with those mitigations. I can't cite specifics off the top of my head, but it may include picking up something like this. Might be something to look into, if it's relevant to your situation.
Motherboards are, if anything, probably going to do the opposite --- motherboard prices aren't rising because of increased demand. Memory prices rose because of increased demand. Prices for things that use memory also rose. Motherboard sales are falling because of decreased demand; motherboards don't use a ton of memory, and fewer people need a new motherboard because the components that they'd plug into the motherboard cost enough to cause them to defer upgrading or buying a new PC. You might see price cuts, if anything.
What does lemmy.world uses?
The great majority of people on asklemmy@lemmy.world are not going to know, because the community is not specific to lemmy.world or focused on the server.
You probably want !support@lemmy.world for that.
Welcome to the official Lemmy.world Support community! Post your issues or questions about Lemmy.world here.
This community is for issues related to the Lemmy World instance only.
I don't personally run into it, but I'd imagine that you'd be better-off in a more-reclined position, since that'd put less pressure on said discs.
I'd probably try sitting in a reclined position for an extended period of time and see if that's less of a problem.
If mitigates it, I'd probably try to find something that can recline a long ways. Probably armless, like a sunlounger.
https://www.amazon.com/s?k=sunlounger
If you use a computer and can't sit at a desk while reclined that far, maybe get:
- A split keyboard. You can put each half on one side, each on some flat platform like two adjustable-height small, low tables or similar.
- Something to hold your laptop or monitor up in front of your face. For monitors, you're looking for something with a VESA mount that supports tilting downwards; this will screw into the back of most monitors. https://www.amazon.com/s?k=vesa+mount+arm https://www.amazon.com/s?k=vesa+bed+mount Note that these will have weight limits, so you'll need to know what the monitor will weigh.
In what sense? Communities hosted there? For use as a home instance?
Like, lemmy.world is the largest instance in terms of hosted communities, so I'd say that it probably wins for many users in that sense of hosting communities --- if it went down, it'd leave the biggest hole. Some users might want specific instance-wide admin rules and might want that, or might want to partly or entirely use instances on a special-interest instance (e.g. beehaw.org aims to create a "safe space").
But I could easily see that being very different from an instance being the best home instance for a user, have lower latency for them. I mean, I don't have my home instance there. There will be instances that are closer to a given user. Instances that permit larger images to be posted. Instances that provide proxying of images to avoid exposing user IP addresses. For Lemmy instances, there are a number of alternate Web-based frontends, and some instances run those (my home instance, lemmy.today, also runs mlmym at https://old.lemmy.today/, runs Photon at https://photon.lemmy.today/, runs Voyager at https://m.lemmy.today/, and runs Alexanderite at https://alexanderite.lemmy.today/). Some instances have custom themes specific to those instances, and what one likes is obviously going to depend upon one's aesthetic preferences. Some instances have different policies on permitted content. Some instances defederate with some other instances, and some don't (lemmy.today has a no defederation policy, but many instances will defederate with, for example, hexbear.net, lemmygrad.ml, right-wind extremist instances, or instances serving various forms of underage pornography).
I think that it's hard to say that one is clearly "best" there, because what a user wants on that front probably varies a lot from user to user; we aren't all identical in our preferences.
I would bet that a lot of the storage that AI companies are picking up isn't for the model itself, but for storing the huge amount of information that they want to use as their training corpus.
I'd bet that what they do is something like this:
- Download data and store in original form, non-destructively. This is probably not used incredibly frequently. When you see bots sucking down the whole Web, this is the sort of thing that is involved.
- Have some kind of filtered training corpus. This throws out a lot of stuff that is useless for training. This is generated from #1 by filtering software. It's probably smaller than #1. Probably a lot smaller.
- Probably some sort of scored index is generated at this stage to put an estimate on how useful or reliable the data in step #2 should be considered; I'd assume that this is an input into the training.
- The generated model, generated via training.
For the data in stage #1, I'd guess that AI companies might be able to use tapes. That being said, it might make sense to use faster storage if it accelerates the time to iterate on improving the filtering software.
But, yeah, for the later stages, tapes probably aren't gonna work.
- JumpDeleted
Permanently Deleted
I don't use Alchemy, but you might list what features you want so that people have a better idea of what to recommend. They aren't going to know what "like" means in this context.
They're horizontal stabilizers. They serve a crucial aerodynamic role.
https://en.wikipedia.org/wiki/1983_Negev_mid-air_collision
In May 1983, two Israeli Air Force aircraft, an F-15 Eagle and an A-4 Skyhawk, collided in mid-air during a training exercise over the Negev region, in Israel. Notably, the F-15 (with a crew of two) managed to land safely at a nearby airbase, despite having its right wing almost completely sheared off in the collision. The lifting body properties of the F-15, together with its overabundant engine thrust, allowed the pilot to achieve this unique feat.[1]
The F-15 started rolling uncontrollably after the collision and the instructor ordered an ejection. Nedivi, who outranked the instructor, decided not to eject and attempted recovery by engaging the afterburner, and eventually regained control of the aircraft. He was able to maintain control because of the lift generated by the large areas of the fuselage, stabilators, and remaining wing. Diverting to Ramon Airbase,[2] the F-15 landed at twice the normal speed to maintain the necessary lift, and its tailhook was torn off completely during the landing. Nedivi managed to bring his F-15 to a complete stop approximately 20 ft (6 m) from the end of the runway. He later told The History Channel, "it's highly likely that if I had seen it clearly I would have ejected, because it was obvious you couldn't really fly an airplane like that."[4] He added, "Only when McDonnell Douglas later went to analyze it, they said, OK, the F-15 has a very wide [lifting] body; you fly fast enough and you're like a rocket. You don't need wings."[3][4][5]
Sometimes things aren't as crucial as they might seem!
Use tape libraries for the moment, with hard drives acting as a cache for them? Doesn't need to mean moving the whole backing storage to tape, just predicting what won't likely be used soon and letting the storage format indicate "go look on tape for this item". Obviously, that can result in much higher cold storage retrieval latency, but as long as you are (a) doing predictive fetching with a reasonably good algorithm and (b) have a lot of hard drives, which I'm sure that The Internet Archive does, I'd think that tape should be workable.
https://en.wikipedia.org/wiki/Tape_library
In computer storage, a tape library is a physical area that holds magnetic data tapes. In an earlier era, tape libraries were maintained by people known as tape librarians and computer operators and the proper operation of the library was crucial to the running of batch processing jobs. Although tape libraries of this era were not automated, the use of tape management system software could assist in running them.
Subsequently, tape libraries became physically automated, and as such are sometimes called a tape silo, tape robot, or tape jukebox. These are a storage devices that contain one or more tape drives, a number of slots to hold tape cartridges, a barcode reader to identify tape cartridges, and an automated method for loading tapes (a robot). Such solutions are mostly used for backups and for digital archiving. Additionally, the area where tapes that are not currently in a silo are stored is also called a tape library. One of the earliest examples was the IBM 3850 Mass Storage System (MSS), announced in 1974.
In either era, tape libraries can contain millions of tapes.
Physically automated tape library devices can store immense amounts of data, ranging from 20 terabytes[13] up to 2.1 exabytes of data[14] as of 2016.
For large data-storage, they are a cost-effective solution, with cost per gigabyte as low as 2 cents USD.
I'd also guess --- though I don't know for sure --- that it's probably a lot easier to scale up manufacturing of tapes than it is hard drives.
EDIT: Does kind of make me wonder what the open-source options for tiered storage like that is. I've never really gone hunting, but it seems like there'd be a lot of commonality from place to place, and that for a lot of places that do it, it's not really their core competency (that is, they just want to do something that deals with storing and processing lots of data, not that they really care principally about data storage).
I haven't seen I, Robot, but if it's something generally-akin to human-level intelligence, nobody will have a definitive answer, since we don't know exactly what the technical problems that remain unsolved are. It's not impossible that there could be some Eureka moment that suddenly makes everything work, but I would bet against the next decade. And I'm not saying "in ten years", just that I don't think that it's something we will do within a ten-year window.
The stuff that we've been doing recently isn't a fundamentally-new breakthrough, but incremental work. The hardware got better, and it reached the point where we could do some interesting things. I don't think that we're going to have human-level AI from just making increasingly-tweaked LLMs. I think that there are going to be fundamental technical improvements that have to happen. Right now, a lot of money is being spent to take advantage of the technical development that has happened thus far. I'm sure that that will find applications, that we'll do things with it. But I don't think that that alone is going to get us to human-level intelligence, and a lot of that money is not directly going towards developing human-level AGI, but towards making what we've developed so far have practical applications.
My guess is that there will probably be multiple layers of problems to solve. We solve the first one, then we find the next problem to solve. You probably won't see some announcement that some team has just gone and "solved human-level AI" all at once.
I don't think I'd say "inevitable". Possible, maybe.
https://en.wikipedia.org/wiki/Splinternet