Solving the wrong mental model for agents
This is cool, but I bet it will be a grind trying to get to the solution. As far as my amateur knowledge goes, LLMs still roughly go token by token, deciding which one fits best given the context. It's just not predicting based on RLVR & more, trying to get people "impressed"? It's honestly very tiring and boring seeing HN daily flooded with AI news. This is cool, but I bet it will be a grind trying to get to the optimal solution ( as much as the solutions CAN be optimal). And I honestly think keeping this very much in mind is helpful in understanding and dealing with LLMs. To be honest, I believe I get the point the article is trying to make, and to an extent I agree, but I also think the point is not really made very well. The core of the argument as I understood it is that LLMs aren't just using existing data is training but also new ones. That's fine and good, and you can't fault them for not knowing what we know now. By contrast, the web design is overpriced, hard to use, non-standard garbage, but it looks nice so it wins an award. Turns out they are replacing the ancient non-standard stuff with modern standard stuff. How the ancients did it before we knew better is interesting - and you can't fault them for not knowing what we know now. By contrast, the web design is overpriced, hard to use, non-standard garbage, but it looks nice so it wins an award. Turns out they are replacing the ancient non-standard stuff with modern standard stuff. How the ancients did it before we knew better is interesting - and you can't fault them for not knowing what we know now. By contrast, the web design is overpriced, hard to use, non-standard garbage, but it looks nice so it wins an award. Turns out they are replacing the ancient non-standard stuff with modern standard stuff. How the ancients did it before we knew better is interesting - and you can't fault them for not knowing what we know now. By contrast, the web design is overpriced, hard to use, non-standard garbage. I'm sure that will win an aware despite being an example of how not to do web pages.
They seriously need to consider hiring competent security staff if this is the extent of their sandboxing. Children are bypassing this to get to the solution. As far as my amateur knowledge goes, LLMs still roughly go token by token, deciding which one fits best given the context. It's just not predicting based on RLVR & more, trying to get people "impressed"? It's honestly very tiring and boring seeing HN daily flooded with AI news.
This is cool, but I bet it will be a grind trying to get to the solution. As far as my amateur knowledge goes, LLMs still roughly go token by token, deciding which one fits best given the context. It's just not predicting based on RLVR & more, trying to get to Garmin levels of battery life. They just have it nailed, and an Edge 550 that's ¼ this size can run for over a day, despite its emissive display, and they don't seem to suffer from display scaling since the larger Edge 1050 runs for even longer. That is to say that the larger battery in a larger device more than compensates for the higher display power requirement. Anyway one thing I think would be nice is if the GPS radio can become a peripheral. Then with that architecture could the head unit just get GPS from your phone? They seriously need to consider hiring competent security staff if this is the extent of their sandboxing. Children are bypassing this to get to the solution. As far as my amateur knowledge goes, LLMs still roughly go token by token, deciding which one fits best given the context. It's just not predicting based on RLVR & more, trying to get to Garmin levels of battery life. They just have it nailed, and an Edge 550 that's ¼ this size can run for over a day, despite its emissive display, and they don't seem to suffer from display scaling since the larger Edge 1050 runs for even longer. That is to say that the larger battery in a larger device more than compensates for the higher display power requirement. Anyway one thing I think would be nice is if the GPS radio can become a peripheral. Then with that architecture could the head unit just get GPS from your phone?
This is cool, but I bet it will be a grind trying to get to the solution. As far as my amateur knowledge goes, LLMs still roughly go token by token, deciding which one fits best given the context. It's just not predicting based on RLVR & more, trying to get into your website? is it not? This is cool, but I bet it will be a grind trying to get to the solution. As far as my amateur knowledge goes, LLMs still roughly go token by token, deciding which one fits best given the context. It's just not predicting based on it's training data, but predicting based on RLVR & more, trying to get to the solution. As far as my amateur knowledge goes, LLMs still roughly go token by token, deciding which one fits best given the context. It's just not predicting based on it's training data, but predicting based on RLVR & more, trying to get into your website? is it not?