We cannot find the code for Evaluation metrics: "Task Offloading", "Exploration Efficiency", and the "Extraneous Effort" of the paper to evaluate the LLM ReAct experiments with them. Particularly in the decentralized setting, using ReAct for both agents. The current code evaluates these metrics for the hitl experiments, as far as we understand, but not for the LLM-LLM experiments. We saw in the paper, in Table 12, that you used these metrics with the LLM-LLM setting as well. Could you please help us? Thanks!
We cannot find the code for Evaluation metrics: "Task Offloading", "Exploration Efficiency", and the "Extraneous Effort" of the paper to evaluate the LLM ReAct experiments with them. Particularly in the decentralized setting, using ReAct for both agents. The current code evaluates these metrics for the hitl experiments, as far as we understand, but not for the LLM-LLM experiments. We saw in the paper, in Table 12, that you used these metrics with the LLM-LLM setting as well. Could you please help us? Thanks!