According to the settings in DCI-Agent-Lite/scripts/bcplus_eval/run_bcplus_eval_openai.sh , when thinking-level is set to high and using the gpt-5.4-nano model, severe refusal-to-answer behavior occurs. Specifically, the model eventually refuses to answer due to insufficient data, resulting in metrics significantly lower than the results in the repository.
According to the settings in
DCI-Agent-Lite/scripts/bcplus_eval/run_bcplus_eval_openai.sh, whenthinking-levelis set tohighand using the gpt-5.4-nano model, severe refusal-to-answer behavior occurs. Specifically, the model eventually refuses to answer due to insufficient data, resulting in metrics significantly lower than the results in the repository.