top | item 46039057 (no title) dheerkt | 3 months ago based on their past usage of "interleaved tool calling" it means that the tool can be used while the model is thinking.https://aws.amazon.com/blogs/opensource/using-strands-agents... discuss order hn newest davidsainez|3 months ago AFAICT, kimi k2 was the first to apply this technique [1]. I wonder if Anthropic came up with it independently or if they trained a model in 5 months after seeing kimi’s performance.1: https://www.decodingdiscontinuity.com/p/open-source-inflecti... BoorishBears|3 months ago OpenAI has been doing this since at least O3 in January, Anthropic has been doing it since 4 in May.And the July Kimi K2 release wasn't a thinking model, the model in that article was released less than 20 days ago.
davidsainez|3 months ago AFAICT, kimi k2 was the first to apply this technique [1]. I wonder if Anthropic came up with it independently or if they trained a model in 5 months after seeing kimi’s performance.1: https://www.decodingdiscontinuity.com/p/open-source-inflecti... BoorishBears|3 months ago OpenAI has been doing this since at least O3 in January, Anthropic has been doing it since 4 in May.And the July Kimi K2 release wasn't a thinking model, the model in that article was released less than 20 days ago.
BoorishBears|3 months ago OpenAI has been doing this since at least O3 in January, Anthropic has been doing it since 4 in May.And the July Kimi K2 release wasn't a thinking model, the model in that article was released less than 20 days ago.
davidsainez|3 months ago
1: https://www.decodingdiscontinuity.com/p/open-source-inflecti...
BoorishBears|3 months ago
And the July Kimi K2 release wasn't a thinking model, the model in that article was released less than 20 days ago.