@coolin

coolin@beehaw.org · 2 years ago

I think your job in your current form is likely in danger.

SOTA Foundation Models like GPT4 and Gemini Ultra can write code, execute, and debug with special chain of thought prompting techniques, and large acale process verification on synthetic data and RL search for correct outputs will make this 10x better. The silver lining to this is that I expect this to require an absolute shit ton of compute to constantly generate LLM output hundreds of times for each internal prompt over multiple prompts, requiring immense compute and possibly taking longer than an ordinary software engineer to run. I suspect early full stack developer LLMs will mainly be used to do a few very tedious coding tasks and SWEs will be cheaper for a fair length of time.

I expect it will be 2-3 years before this happens, so for that short period I expect workers to be “super-productive” by using LLMs in the coding process, but I expect the crossover point when the LLM becomes better is quite soon, perhaps in the next 5 years as compute requirements go down.

coolin@beehaw.org · 2 years ago

NFTs are stupid AF for most of the tasks people currently use them for and definitely shouldn’t be used as proof of ownership of physical assets.

However, I think NFTs make a lot of sense as proof of ownership of purely digital assets, especially those which are scarce.

For example, there are several projects for domain name resolution based on NFT ownership (e.g you look up crypto.eth, your browser checks that the site is signed by the owner of the crypto.eth NFT, then you are connected to the site), as it could replace our current system, which has literally 7 guys that hold a private key that is the backbone of the DNS system and a bunch of registrars you have to go through to get a domain. This won’t happen anytime soon but it is an interesting concept.

Then I think an NFT would also be good as a decentralized alternative to something like Google sign in, where you sign up for something with the NFT and sign in by proving your ownership of it.

In general though I find NFTs to be a precarious concept. I mean the experience I’ve had with crypto is you literally have a seed phrase for your wallet, and if it gets stolen all your funds are drained. And then for an NFT, if you click on the wrong smart contract, all your monkeys could be gone in an instant. There is in general no legal recourse to reverse crypto transactions, and I think that is frankly the biggest issue with the technology as it stands today.

coolin@beehaw.org · 2 years ago

“I use Signal to hide my data from the US government and big tech”

“Wait, you seriously still use Reddit? Everyone switched to the Fediverse!”

“Wow, can’t believe you use Apple! Android is so much better.”

No one who isn’t terminally online understands what these statements mean. If you want people to use something else, don’t make it about privacy and choose something with fancy buttons and cool features that looks close enough to what they have. They do not care about privacy and are literally of the mindset “if I have nothing to hide I have nothing to fear”. They sleep well at night.

coolin@beehaw.org · 2 years ago

Only thing really missing is Wallet and NFC support. Other than that I think Graphene and Lineage OS cover it all

coolin@beehaw.org · 2 years ago

Hello, kids! Pirates are very bad! Never use qBittorent to download copyrighted material, and certainly do NOT connect it to a VPN to avoid getting caught. Additionally, you should also NEVER download illegal material via an https connection because it is fully encrypted and you won’t get caught!

coolin@beehaw.org · 2 years ago

Reddit since changed the UI again which killed my interest in scrolling r/all. I still have to go there to view r/localllama, r/singularity and r/UFOs, none of which have a sizeable Feddit equivalent. I could do without the speculation of the latter 2 in my life, but I need LocalLlama because it is a great source for news and advice on LLMs.

coolin@beehaw.org · 2 years ago

deleted by creator

coolin@beehaw.org · 2 years ago

This is another reminder that the anomalous magnetic moment of the muon was recalculated by two different groups using higher precision lattice QCD techniques and wasn’t found to be significantly different from the Brookhaven/Fermilab “discrepancy”. More work needs to be done to check for errors in the original and newer calculations, but it seems quite likely to me that this will ultimately confirm the standard model exactly as we know it and not provide any new insight or the existence of another force particle.

My hunch is that unknown particles like dark matter rely on a relatively simple extension of the standard model (e.g. supersymmetry, axioms, etc.) and the new physics out there that combines gravity and QM is something completely different from what we are currently working on and can’t be observed with current colliders or any other experiments on Earth.

So probably we will continue finding nothing interesting for quite some time until we can get a large ML model crunching every single possible model to check for fit on the data, and hopefully derive some better insight from there.

Though I’m not an expert and I’m talking out of my ass so take this all with a grain of salt.

coolin@beehaw.org · 2 years ago

Sam Altman: We are moving our headquarters to Japan

coolin@beehaw.org · 2 years ago

I think this is downplaying what LLMs do. Yeah, they are not the best at doing things in general, but the fact that they were able to learn the structure and semantic context of language is quite impressive, even if it doesn’t know what the words converted into tokens actually mean. I suspect that we will be able to use LLMs as one part of a full digital “brain”, with some model similar to our own prefrontal cortex calling the LLM (and other things like vision model, sound model, etc.) and using its output to reason about a certain task and take an action. That’s where I think the hype will be validated, is when you put all these parts we’ve been working on together and Frankenstein a new and actually intelligent system.

coolin@beehaw.org · 2 years ago

For the love of God please stop posting the same story about AI model collapse. This paper has been out since May, been discussed multiple times, and the scenario it presents is highly unrealistic.

Training on the whole internet is known to produce shit model output, requiring humans to produce their own high quality datasets to feed to these models to yield high quality results. That is why we have techniques like fine-tuning, LoRAs and RLHF as well as countless datasets to feed to models.

Yes, if a model for some reason was trained on the internet for several iterations, it would collapse and produce garbage. But the current frontier approach for datasets is for LLMs (e.g. GPT4) to produce high quality datasets and for new LLMs to train on that. This has been shown to work with Phi-1 (really good at writing Python code, trained on high quality textbook level content and GPT3.5) and Orca/OpenOrca (GPT-3.5 level model trained on millions of examples from GPT4 and GPT-3.5). Additionally, GPT4 has itself likely been trained on synthetic data and future iterations will train on more and more.

Notably, by selecting a narrow range of outputs, instead of the whole range, we are able to avoid model collapse and in fact produce even better outputs.

coolin@beehaw.org · edit-2 2 years ago

We have no moat and neither does OpenAI is the leaked document you’re talking about

It’s a pretty interesting read. Time will tell if it’s right, but given the speed of advancements that can be stacked on top of each other that I’m seeing in the open source community, I think it could be right. If open source figured out scalable distributed training I think it’s Joever for AI companies.

coolin@beehaw.org · edit-2 2 years ago

Based NixOS user

I love NixOS but I really wish it had some form of containerization by default for all packages like flatpak and I didn’t have to monkey with the config to install a package/change a setting. Other than that it is literally the perfect distro, every bit of my os config can be duplicated from a single git repo.

coolin@beehaw.org · 2 years ago

I don’t know what type of chatbots these companies are using, but I’ve literally never had a good experience with them and it doesn’t make sense considering how advanced even something like OpenOrca 13B is (GPT-3.5 level) which can run on a single graphics card in some company server room. Most of the ones I’ve talked to are from some random AI startup that have cookie cutter preprogrammed text responses that feel less like LLMs and more like a flow chart and a rudimentary classifier to select an appropriate response. We have LLMs that can do the more complex human tasks of figuring out problems and suggesting solutions and that can query a company database to respond correctly, but we don’t use them.

coolin@beehaw.org · 2 years ago

Blocking out the sun with aerosols is a good idea if you know with high confidence how it will impact the climate system and environment. That’s why they’re trying to simulate it with the supercomputer, so they know if it fucks stuff up or not.

coolin@beehaw.org · 2 years ago

Cool meme but Reuters doesn’t own AP and Rothschild doesn’t own Reuters. It is quite ironic to be pushing against the very real problem of media disinfo/government propaganda trickled down through AP/Reuters, while at the same time spreading misinformation.

coolin@beehaw.org · 2 years ago

This makes sense for any other company but OpenAI is still technically a non profit in control of the OpenAI corporation, the part that is actually a business and can raise capital. Considering Altman claims literal trillions in wealth would be generated by future GPT versions, I don’t think OpenAI the non profit would ever sell the company part for a measly few billions.

coolin@beehaw.org · 2 years ago

Lmao Twitter is not that hard to create. Literally look at the Mastodon code base and “transform” it and you’re already most of the way there.

coolin@beehaw.org · 2 years ago

My bad you are correct I’m just talking out of my ass

coolin@beehaw.org · 2 years ago

The natural next place for people to go to once they can’t block ads on YouTube’s website is to go to services that exploit the API to serve free content (NewPipe, Invidious, youtube-dl, etc.). If that happens at a large scale, YouTube might shut off its API just like Reddit did and we’ll end up in scenario where creators are forced to move to Peertube, and, given how costly hosting is for video streaming, it could be much worse than Reddit->Lemmy+KBin or Twitter->Mastodon. Then again, YouTube has survived enshittiffication for a long time, so we’ll have to wait and see.