Interesting! To me 80% hitrate sounds actually pretty good and awesome if it actually improves productivity, though understandably not something that could be left on it's own devices.
I had no idea about Make.com or n8n, they seem interesting. Thanks for the tip! Will check them out.
My take is that if the LLM outputs text for humans to read, that's not an agent. If it's making API calls and doing things with the results, that's an agent. But given the way "AI" has stretched to become the new "radium" [1], I'm sure "agent" will shortly become almost meaningless.
The definition of agent is blurry. I prefer to avoid that term because it does not mean anything in particular. These are implemented as chat completion API calls + parsing + interpretation.
Seems like a great service!
We could definitely consider this as an viable option for our startup's application.
I am not an expert regarding the OSM data and maps in general, but how customizable is your library, can I somehow relatively easy add housenumbers of the addresses from OSM data on the buildings in the map?
This is totally customizable and the house numbers are already present in the original data, it's just not displayed in the default styles. You can customize the style by using the Maputnik editor, house numbers are visible in the "Inspector" view.
https://maputnik.github.io/editor?style=https://tiles.openfr...
Sorry for the slightly off topic question, but can someone enlighten me which Claude model is more capable, Opus or Sonnet 3.5? I am confused because I see people fuzzing about Sonnet 3.5 being the best and yet somehow I seem to read again and again in factual texts and some benchmarks that Claude Opus is the most capable. Is there a simple answer to the question, what do I not understand? Please, thank you.
I.e. Opus is the largest and best model of each family but Sonnet is the first model of the 3.5 family and can beat 3's Opus in most tasks. When 3.5 Opus is released it will again outpace the 3.5 Sonnet model of the same family universally (in terms of capability) but until then it's a comparison of two different families without a universal guarantee, just a strong lean towards the newer model.
Opus is the largest model, but of the Claude 3 family. Claude 3.5 is the newest family of models, with Sonnet being the middle sized 3.5 model - and also the only available one. Regardless, it's better than Opus (the largest Claude 3 one).
Presumably, a Claude 3.5 Opus will come out at some point, and should be even better - but maybe they've found that increasing the size for this model family just isn't cost effective. Or doesn't improve things that much. I'm unsure if they've said anything about it recently.
These two are great success stories, but they’re also the only Finnish unicorn exits in the post-Nokia era.
The exits were somewhat less exciting to founders than these numbers suggest. Supercell sold 51% to SoftBank already in 2013 for 1.1B EUR. And Wolt’s purchase price was paid entirely in DoorDash stock which was down 75% by the time the lockups expired.
Startups generating low-hundreds of millions in annual revenue just aren’t unicorns anymore, unless they happen to be AI.
Both Supercell and Wolt have their headquarters steadily in Finland. The founders and Finnish early investors have gained hundreds of millions or billions of euros wealth for themselves which they have further spended and invested in Finland. They have paid huge amounts of taxes and keep on doing all of these since they are still located in Finland. It's hard to downplay the value of those IMO. Overall Rovio wasn't a complete disaster either. First made billions of euros for many years and was later sold to Sega for >$700 million. Still has HQ in Finland.
There's plenty of interesting and fast growing startups still left here. For example Supermetrics, Varjo, Smartly, Iceye, Aiven to name a few. IMO you are being pessimistic.
In any case, I agree in that the acquisition is great news and the economy is in a depression. :) Huge part of it is because Finnish mortgages are mostly straight tied to Euribor unlike in other Euro countries and since post covid the interest rates went up, Finns got f*cked. Hopefully the Euribor interest rate will be going down and the mortgages will start to become smaller, at least when they are being paid off.
Now if Meta would only get that website layout of theirs working on this Chrome browser that runs on my two year old Samsung Android phone, they would be a whole lot credible to me.
API access is not coupled with being a ChatGPT+ subscriber, and I believe GPT-4 API access is also not included in that. That means you can use GPT-3.5-turbo via API right now, and GPT-4 through API is available via waitlist.
I'm not yet, because I haven't got confidence in ChatGPT's answers or abilities so far (e.g. the fact that it 'lies' with a straight face when it doesn't know the answer is especially troubling for me). I hope ChatGPT 5 or 6 are much better in this regard, and I will very gladly pay for it then.
At least as a non-native english speaker I was a bit confused as to what an agent would do compared to using plain ChatGPT. I tried then the examples "Plan a detailed trip to Hawaii" and "Write some code to make a platformer game". I tried the same Hawaii sentence with plain ChatGPT and told it that it can browse the web. Now I am thinking that AgentGPT does seem a good tool on top of ChatGPT. As they say on the Github page it "It will attempt to reach the goal by thinking of tasks to do, executing them, and learning from the results". This was a whole lot more thorough service than what plain ChatGPT did with the prompt. I'm just thinking that maybe they should emphasize more those points already on the app page, about the core value it adds and how it does it. Perhaps even explain it some more than what that sentence from their Github page does.
I had no idea about Make.com or n8n, they seem interesting. Thanks for the tip! Will check them out.