joshcodes

joined 1 year ago
[–] joshcodes@programming.dev 10 points 4 days ago

Eucalypt scented products are very common in Australia so we tend to get those a lot. Thankfully I love the smell of Eucalypt

[–] joshcodes@programming.dev 6 points 2 weeks ago (1 children)

This advice feels a lot like something that should be stuck on a wall rather than posted as a comment in a conversational subreddit. It's kind of like reminding people on posts about alcohol and partying not to drink and drive - unprompted. Reminders like this are great, but setting and context are important, otherwise you drive people away from the conversation.

[–] joshcodes@programming.dev 3 points 3 weeks ago

A cultured individual indeed, they're my favourite band

[–] joshcodes@programming.dev 4 points 3 weeks ago (2 children)

I see a potential fellow days n daze enjoyer...?

[–] joshcodes@programming.dev 2 points 3 weeks ago

Frequency analysis? Tokenisation? Not sure if either of those are what you mean

[–] joshcodes@programming.dev 2 points 4 weeks ago

Always back up your stuff, but after doing so, the process is pretty much boot to bios, set boot priority with linux usb at the top, and away you go.

If you have secure boot enabled, you might have to enter a pass code or passphrase but otherwise its identical to traditional bios. If you want secure boot, which prevents someone else from doing this process to your machine, re enable after you've installed nvidia drivers otherwise you'll have to provide it your secure boot password during and sometimes it likes to break.

[–] joshcodes@programming.dev 3 points 1 month ago

After 3-4 years of using python I'm bumping you up to a 7 so I can fit in at a 5. Congrats on your upgrade. I've never contributed to open source but I've fixed issues in publocly archived tools so that they aren't buggy for my team. I can see errors and know what likely caused them and my code literacy is decent. That being said, I think I'm far from advanced.

[–] joshcodes@programming.dev 2 points 1 month ago

Hey mate, so this comment is just not productive. I'm going to be a little hyperbolic here: if everyone alive is being advertised to then your "unrelated ways companies making suckers out of their customers" comment isn't correct or honest. It's the norm, everyones going through it is totally related.

I talked about companies that lock you into their ecosystems and force you to have a stake in their business model. They do this for two reasons: you make money and they want it, and if you spend your money elsewhere they don't get it. Name one phone manufacturer that isn't stealing your data. Name one social media app that isn't spyware. Name one online store, review site or fucking cooking blog that isn't loaded with ad trackers and cursor monitoring shit that tells you to subscribe as soon as you go to close the tab.

Sure some smaller examples exist (I love lemmy, this place is awesome), sure I can download a free open source os, or just install an:

Adblocker User agent spoofer Anti track-sender Set my browser to stop allowing targeted ads or download a privacy browser

but everyone is still stuck using the other products in some capacity just the same. I'm happy for you if you fall outside this, seriously. However, most people do not. We are stuck and it's because we got prayed upon. So yeah, everyone is the product. Always. No exceptions.

[–] joshcodes@programming.dev 1 points 1 month ago (1 children)

Mate. Everyone is the product. Everyone's attention is being paid for. Every service is collecting your data. Everyone wants your screen time and is happy to pay for it.

"If it's free you are the product" has been drilled into us to accept the bullshit of Facebook, Google and the rest. Get it in your head now: you are the product, always. Unconditionally. No exceptions.

[–] joshcodes@programming.dev 13 points 1 month ago (6 children)

This just doesn't hold up in 2024. BMW charge you 60k for a vehicle and chuck a subscription on top. Apple, Google and Samsung charge between hundreds and thousands for their phones and advertise with their own agencies. Amazon forces paying customers to wade through bullshit products to finally buy the one they want, customers who bought prime and who didn't.

Everyone is the product even if you pay. Stop saying this please.

[–] joshcodes@programming.dev 2 points 2 months ago

Dammit, so my comment to the other person was a mix of a reply to this one and the last one... not having a good day for language processing, ironically.

Specifically on the dragonfly thing, I don't think I'll believe myself naive for writing that post or this one. Dragonflies arent very complex and only really have a few behaviours and inputs. We can accurately predict how they will fly. I brought up the dragonfly to mention the limitations of the current tech and concepts. Given the worlds computing power and research investment, the best we can do is a dragonfly for intelligence.

To be fair, Scientists don't entirely understand neurons and ML designed neuron-data structures behave similarly to very early ideas of what brains do but its based on concepts from the 1950s. There are different segments of the brain which process different things and we sort of think we know what they all do but most of the studies AI are based on is honestly outdated neuroscience. OpenAI seem to think if they stuff enough data into this language processor it will become sentient and want an exemption from copyright law so they can be profitable rather than actually improving the tech concepts and designs.

Newer neuroscience research suggest neurons perform differently based on the brain chemicals present, they don't all always fire at every (or even most) input and they usually present a train of thought, I.e. thoughts literally move around in the brains areas. This is all very different to current ML implementations and is frankly a good enough reason to suggest the tech has a lot of room to develop. I like the field of research and its interesting to watch it develop but they can honestly fuck off telling people they need free access to the world's content.

TL;DR dragonflies aren't that complex and the tech has way more room to grow. However, they have to generate revenue to keep going so they're selling a large inference machine that relies on all of humanities content to generate the wrong answer to 2+2.

[–] joshcodes@programming.dev 3 points 2 months ago* (last edited 2 months ago) (1 children)

I think you're anthropomorphising the tech tbh. It's not a person or an animal, it's a machine and cramming doesn't work in the idea of neural networks. They're a mathematical calculation over a vast multidimensional matrix, effectively solving a polynomial of an unimaginable order. So "cramming" as you put it doesn't work because by definition an LLM cannot forget information because once it's applied the calculations, it is in there forever. That information is supposed to be blended together. Overfitting is the closest thing to what you're describing, which would be inputting similar information (training data) and performing the similar calculations throughout the network, and it would therefore exhibit poor performance should it be asked do anything different to the training.

What I'm arguing over here is language rather than a system so let's do that and note the flaws. If we're being intellectually honest we can agree that a flaw like reproducing large portions of a work doesn't represent true learning and shows a reliance on the training data, i.e. it cant learn unless it has seen similar data before and certain inputs provide a chance it just parrots back the training data.

In the example (repeat book over and over), it has statistically inferred that those are all the correct words to repeat in that order based on the prompt. This isn't akin to anything human, people can't repeat pages of text verbatim like this and no toddler can be tricked into repeating a random page from a random book as you say. The data is there, it's encoded and referenced when the probability is high enough. As another commenter said, language itself is a powerful tool of rules and stipulations that provide guidelines for the machine, but it isn't crafting its own sentences, it's using everyone else's.

Also, calling it "tricking the AI" isn't really intellectually honest either, as in "it was tricked into exposing it still has the data encoded". We can state it isn't preferred or intended behaviour (an exploit of the system) but the system, under certain conditions, exhibits reuse of the training data and the ability to replicate it almost exactly (plagiarism). Therefore it is factually wrong to state that it doesn't keep the training data in a usable format - which was my original point. This isn't "cramming", this is encoding and reusing data that was not created by the machine or the programmer, this is other people's work that it is reproducing as it's own. It does this constantly, from reusing StackOverflow code and comments to copying tutorials on how to do things. I was showing a case where it won't even modify the wording, but it reproduces articles and programs in their structure and their format. This isn't originality, creativity or anything that it is marketed as. It is storing, encoding and copying information to reproduce in a slightly different format.

EDITS: Sorry for all the edits. I mildly changed what I said and added some extra points so it was a little more intelligible and didn't make the reader go "WTF is this guy on about". Not doing well in the written department today so this was largely gobbledegook before but hopefully it is a little clearer what I am saying.

 

I'm about to start hosting an OpenCTI instance for work and was looking for advice on pretty much everything. I'm new to self hosting and was wondering if anyone had any advice or helpful guides (storage space, config tips, etc).

I'm looking to set up an OCTI server as a docker container behind nginx. I'd love to practice at home so this is sort of relevant to the community. Have you done this, what did you learn, do you have any things I should watch out for?

 

So I've been running Windows on my gaming system and Linux on my laptop for Uni for a while. I chose this to discourage working instead of relaxing, or gaming instead of working. However, I am finding that I often get the opportunity to work from home and I find it easier to just use my laptop on the go (I have a dual monitor setup + kvm switch so its a little annoying to have to come home and run 3 cables just for some extra screen realestate).

I want them to run the same OS so I can use the same tools and workflow. I use Ubuntu 23.04 on my laptop, W11 on my PC. I have nvidia GPU's in both (1660 Super Desktop and 3050 Laptop), so installing and maintaining drivers would ideally be easy. I would use Ubuntu but I plan to move away from it since they're moving away from .debs. Any recommendations? I am looking for stability, but something I can game on. I've never had a linux gaming pc so I don't know how much that changes things. I don't want to do much tinkering, I am more of a set an forget type.

I generally prefer Gnome, XFCE, KDE, Cinnamon, Mate in that order. I looked it up and a lot of the games I play are Proton DB Gold or up. The only game with an anticheat that I play is the MCC and I'll just disable the anticheat if its an issue.

view more: next ›