Repository navigation
fix(gooddata-eval): read set_skills' result rather than its arguments to determine active skills - #1865
Open
cobanfurkanx wants to merge 1 commit into
Open
fix(gooddata-eval): read set_skills' result rather than its arguments to determine active skills#1865cobanfurkanx wants to merge 1 commit into
cobanfurkanx wants to merge 1 commit into
Conversation
…active skills Prefer the authoritative post-replacement active skills echoed back in the set_skills tool result over its request arguments in both conversation evaluation and visualization evaluation. When the tool result is missing, unparseable, or indicates an error, fall back to the requested skill_names arguments to preserve compatibility with legacy traces. Closes gooddata#1779
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Closes #1779
Description
_set_skills_declarations()/_final_skill_declaration()inpackages/gooddata-eval/src/gooddata_eval/core/agentic/conversation.pyand_check_visualization_skill_activatedincore/evaluators/visualization.pypreviously derived the active skill set solely from theset_skillstool call's arguments (what the agent requested).However, what the agent requests is not always what actually becomes active:
The tool call's own result echoes back the authoritative post-replacement active set (e.g.
skills_to_activate). Reading the result is strictly truer to what actually became active than re-deriving it from the request arguments.Changes
_extract_active_skills(tc)to prefer the authoritative post-replacement set intc.parsed_result()(skills_to_activate,skill_names,skills), falling back totc.parsed_arguments().get("skill_names")when the result is absent, unparseable, or errored._set_skills_declarations()inconversation.pyto use_extract_active_skills()._check_visualization_skill_activated()invisualization.pyto use_extract_active_skills().test_agentic_conversation.pyandtest_visualization_evaluator.pycovering:Summary by CodeRabbit