The current oracle is designed by @EvanTian233 through an extensive discussion,
SREGym#651
Personally, I like the current oracle for its simplicity. It's certainly not perfect. I wish we could have a grounded model of failures, but I myself fail to do so.
Unfortunately, I don't have an idea how good or bad it is -- do you encounter a problem that the oracle misgrades the agent's results? If so, please report so we can take a look at it and improve the oracle.
There is at least one outstanding issue to close
SREGym#745
There was also a historical issue here,
SREGym#747
@SuMayaBee Could you take charge of this task and hopefully to close SREGym#745?
The current oracle is designed by @EvanTian233 through an extensive discussion,
SREGym#651
Personally, I like the current oracle for its simplicity. It's certainly not perfect. I wish we could have a grounded model of failures, but I myself fail to do so.
Unfortunately, I don't have an idea how good or bad it is -- do you encounter a problem that the oracle misgrades the agent's results? If so, please report so we can take a look at it and improve the oracle.
There is at least one outstanding issue to close
SREGym#745
There was also a historical issue here,
SREGym#747
@SuMayaBee Could you take charge of this task and hopefully to close SREGym#745?