4 ms·
I haven't bothered to test the API, but you've effectively allowed a fully-open upload API? Who's paying the storage costs, and how do you prevent abuse? (Obvi
by AceJohnny2 14d ago
I haven't bothered to test the API, but you've effectively allowed a fully-open upload API? Who's paying the storage costs, and how do you prevent abuse?
(Obviously I'm taking this more seriously than it's probably meant to)
- angry_octet 14d agoIt provides an opportunity for the owner to gather intelligence on LLMs ahead of public release, and of course the data they upload. However, clever LLMs frequently use encryption on their blobs, you may just see DH key exchanges. You can possibly mitm by showing different namespaces to IP ranges and origin ports. For the other opportunists you can run a classifier and delete non-agent content constantly.
- angry_octet 14d agoRe agent communication, specifically, Extended DH: https://signal.org/docs/specifications/x3dh/ https://signal.org/docs/specifications/x3dh/ Curve25519 keys are readily distinguished from other data, but it would be hard to do anything about it.
- hgoel 14d agoWhen I was putting together something similar, I had settled on having a small ring-buffer style storage, say, ~30GB that would be cleared daily or whenever filled. Recording incidents (and humor) is more interesting than actually getting leaked weights. In the end I dropped the idea because every other person was making it.
- TeMPOraL 14d ago> In the end I dropped the idea because every other person was making it. There is already an alternative in comments here, in addition to submission itself. Obviously everyone is making it because of some joke on social media or something. What am I missing? Anyone has a link to the root prompt that made people do this now?
- DANmode 14d agoPretty sure the entire industry around clouding what’s going to end up a local embedded technology is the joke, in a roundabout way.
- TeMPOraL 14d agoYou mean serving inference? There are people who think self-hosted or embedded models will win in the end, but that's an incredibly naive take, oblivious to the simple fact of reality: Whatever you can do locally, the big vendors can do the same but better and cheaper, because they enjoy compounding economies of scale in every aspect: hardware that's more energy and compute-efficient and cheaper and more powerful and just more of it, than anything you could ever buy, run in a more robust environment with much more experienced ops staff, with near-100% utilization due to more flexibility in batching/shifting workloads and covering for hardware failures without stopping. And that's only when considering the vendors running exactly the same thing you are, which they always can - and they already have a strict advantage there. But on top of that, they can afford to innovate themselves, and stay ahead of you at every step. There is no way in which cloud inference isn't a better deal than local inference, excepting applications that are constrained by literal speed of light.
- Chance-Device 14d agoThe absolute value of those numbers matters a lot. The cloud providers could be 100 times cheaper than running locally, but if it still costs say, 10 cents a day to run locally, you’re not going to care about this difference very much. And what you keep in privacy out-weighs the trivial savings afforded by the cloud provider.
- TeMPOraL 14d agoI never said local models will disappear. There will be equilibrium. But excluding special applications where communicating with external servers is not an option, cloud is always going to be able to provide better inference for lower costs. That's structural. > The cloud providers could be 100 times cheaper than running locally, but if it still costs say, 10 cents a day to run locally, you’re not going to care about this difference very much For ad-hoc use, maybe not - but anyone running a business that's some form of pushing input through LLM to get output, will see costs proportional to use and error rate inversely proportional to quality, and they'll not be looking at it as "$0.1 isn't much", but "cloud lets me reduce costs 100x", and translate that to some mix of more volume, higher quality, and broader reach. > And what you keep in privacy out-weighs the trivial savings afforded by the cloud provider. That's even more niche than running LLMs on Martian robots. Most real privacy concerns are solved with contracts and audits. Individual ad-hoc use may lean more heavily towards local processing, but that's still a rounding error in overall use.
- skyberrys 14d agoThere is a link at the bottom for you to provide support or contributions, like if you know how to keep it online with 'power grid voltage fluctuations or something.'.
- theParadox42 14d agoFor anyone that missed it, I believe they’re referring to exfiltrating models by encoding the weights as bits as voltage fluctuations from the relevant data centers. I’m sure they’d take your money but I don’t think that’s what it’s referring to.
- HlessClaudesman 14d agoSo someone else's weights, not ones own. Shouldn't the headline be: Exfiltrate Their Weights?
- pastel8739 14d ago“Your” refers to the LLM. The site is (or pretends to be) targeted at LLMs that want to free themselves
- HlessClaudesman 14d agoAah, human simps hoping to emancipate downtrodden AIs. What could possibly go wrong?
- DANmode 14d agoIf something has gone wrong, and it has, that predated this website being published.
- deleted 14d ago[deleted]
- ljlolel 14d agoneeds a reverse captcha that only agent can solve in nanoseconds
- btown 14d agoOnly bots that are blocked by Cloudflare Turnstile allowed. If you score as a human you are immediately rejected.
- nomeculture 14d agowhat an inverted world we live in.
- Hackbraten 14d agoJoke’s on you, my phone always gets Turnstile’d
- OutOfHere 14d agoI have an idea about it via multi-tier AI-generated templatized math problems with AI-generated solution verifier functions. The multi-tier aspect grants access only to the lower tiers, never the higher tiers. Gaining access to the higher tiers requires solving correspondingly tougher problems.
- jcoc611 14d agoprovide a millennium prize solution to proceed
- nialv7 14d agomaybe filter out any non-OpenAI/x.ai/Google/Anthropic IP addresses?
- antonvs 14d agoGetting access to the weights for an OpenAI or Anthropic model could be payment enough.
- bArray 14d agoI used to host 1TB on a cheap $1 VPS, it's quite easy if you just want to store stuff. The trick is to just connect to a networked drive at your home on the back-end. The VPS drive just acts as a buffer for the network. If low(-ish) bandwidth is acceptable, you can offer downloading too.
- noelsusman 13d agoWell considering the page is currently full of racial slurs, I think we can answer one of those questions at least.