@AIhasUse

AIhasUse@lemmy.world · 1 year ago

I was mainly doing python with gpt4, but now im working on an android project, so kotlin. Gpt4 wasn’t much use for kotlin, especially for questions involving more than a couple files. Sonnet is crushing it though, even when I give it 2k+ LoC. I’d say I’ve done about 2 months of pre-llm work in the last week, granted I am no professional, just a hobbyist.

AIhasUse@lemmy.world · 1 year ago

For programming it is Sonnet 3.5, there is no remotely close 2nd place that I have tried or heard of, and I am always looking. I personally don’t really have any interest in measuring them in other ways. But for coding, Sonnet 3.5 is in a distant lead. Abacus.ai is a nice way to try various models for cheap. Really, some sort of agent setup like mixture of agents that uses Claude and got and maybe some others may do better than Claude alone. Matthew Berman shows Mixture of Agents with local models beating gpt4o, so doing it with sonnet3.5 and others of the best closed models would probably be pretty great.

AIhasUse@lemmy.world · 1 year ago

AIhasUse@lemmy.world · 1 year ago

Your mistake is thinking that you are talking to people who think more than 2 feet in front of their noses.

AIhasUse@lemmy.world · 1 year ago

Yeah, then we would all have so much more money all of a sudden, that would mean we could all buy so much more stuff. That’s definitely how money works.

AIhasUse@lemmy.world · 1 year ago

My dude, no, I’m not the creator, settle down. Mixture of agents is free and open to anyone to use. Here is a demo of it by Matthew Berman. It isnt hard to set up.

https://youtu.be/aoikSxHXBYw

Believe it or not, openai is no longer making the best models. Claude Sonnet 3.5 is much better than openai’s best models by a considerable amount.

AIhasUse@lemmy.world · 1 year ago

It takes a lot of energy to train the models in the first place, but very little once you have them. I run mixture of agents on my laptop, and it outperforms anything openai has released on pretty much every benchmark, maybe even every benchmark. I run it quite a bit and have noticed no change in my electricity bill. I imagine inference on gpt4 must almost be very efficient, if not, they should just switch to piping people open sourced llms run through MoA.

AIhasUse@lemmy.world · 1 year ago

Yeah. It’s really interesting because juniors and hobbyist are the ones getting used to how to interact with it. Since it is rapidly improving, it won’t be long until it will outpace the grunt work ability of seniors and the new seniors will be the ones willing and able to use it. Programming is switching away from being able to write tedious code and into being able to come up with ideas and convey them clearly to an llm. There’s going to be a real leveling of the playing field when even the best seniors won’t have any use for most of their grunt work coding skills. The jump up from Opus 3 to Sonnet 3.5 is absolutely insane, and Opus 3.5 should be here before too long.

AIhasUse@lemmy.world · 1 year ago

That’s really interesting. For android studio it’s been absolutely crushing it for me. It’s taken some getting used to, but I’ve had it build an app with about 60 files. I’m no master programmer, but I’ve been a hobbyist for a couple decades. What it’s done in the last 5 days for me would have taken me 2 months easy, and there’s lots of extra touches that I probably wouldn’t have taken time to do if it wasn’t as simple as loading in a few files and telling it what I want.

Usually when I work on something like this, my todo list grows much faster than my ability to actually put it together, but with this project I’m quickly running out of even any features that I can imagine. I’ve not had any of the issues of it running in circles like I would often get it gpt4.

AIhasUse@lemmy.world · edit-2 1 year ago

Have you coded with Claude Sonnet 3.5 yet? It is mind-blowingly better than Opus 3, which was already noticeably better than anything openAI has put out yet. Gpt 4 was nice to code with, but this is on a whole other level. I can’t imagine what Opus 3.5 will be able to do.

AIhasUse@lemmy.world · 1 year ago

Good answer, no way AI will possibly ever catch up to such brilliant responses as this. Certainly, there is no reason to want to have our views represented in the next generation of technology.

AIhasUse@lemmy.world · 1 year ago

You could have a much more complex understanding of what they are. It isn’t nearly as simple as you are imagining. If you genuinely are curious about what you’re overlooking, then here is a link.

https://situational-awareness.ai/

AIhasUse@lemmy.world · 1 year ago

If you are genuinely open to understanding the path we are on, the new situational awareness paper would be very eye-opening. It is 160 pages, so it’s probably a bit too much to get through, but there are really good videos that explain it. Matthew Berman has a great video about it. I’m not interested in swaying you and not going to debate, I’m 100s of hours deep into this and have been absolutely obsessed with it. Nobody doubted its impact as much as me. Education on the matter will undeniably change your mind tremendously. The information is there if you want a peak at the future.

https://situational-awareness.ai/

AIhasUse@lemmy.world · 1 year ago

Thanks so much for taking the time to explain this. I was just going to give them a link.

AIhasUse@lemmy.world · 1 year ago

It’s a much much bigger issue than this. Would you rather live in a world where other countries have good AI and you do not? Would you like it if only China has powerful AI? I get the copyright issue, but some things are more important than other things. This is an arms race, and everyone slowing down isn’t exactly an option.

AIhasUse@lemmy.world · 1 year ago

Is it definitely a W that EU perspectives won’t be as represented in the AI programs that we are all using?

AIhasUse@lemmy.world · 1 year ago

This is related, but I just realized that I’ve been conditioned to expect a Rick Roll every time I click a link in a comment, even though it hardly ever happens. My brain just always says “Here comes Rick Astel or whatever his name is”.

AIhasUse@lemmy.world · edit-2 1 year ago

What about Good Friday? 🔨😵

Edit: My bad, I just assumed murder was one.

AIhasUse@lemmy.world · 1 year ago

This is especially interesting, considering he left Google 3 years ago, according to his website. It’s a bit misleading to put this old tweet up alongside a recent Google screenshot.

AIhasUse@lemmy.world · 1 year ago

This website is amazing! How have I never heard of this before?! Did you know that using glue will make your cheese extra stretchy? Who would have guessed it? This is my new favorite site.