style(evals): improve CI summary layout and harden inline Python#2550
Merged
Conversation
Mason Daugherty (mdrxy)
marked this pull request as ready for review
April 8, 2026 02:54
Mason Daugherty (mdrxy)
requested a review
from Eugene Yurtsev (eyurtsev)
as a code owner
April 8, 2026 02:54
james8814
pushed a commit
to james8814/deepagents
that referenced
this pull request
May 1, 2026
…gchain-ai#2550) Clean up the eval CI summary rendering: per-model results were full-width vertical tables that dominated the page. Now they're compact horizontal tables inside collapsed `<details>` toggles, radar charts are constrained to 500px, and the aggregate section has a frontmatter callout clarifying these are the final results. Also hardens the inline Python against unclosed `<details>` tags and malformed `experiment_links` entries.
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Clean up the eval CI summary rendering: per-model results were full-width vertical tables that dominated the page. Now they're compact horizontal tables inside collapsed
<details>toggles, radar charts are constrained to 500px, and the aggregate section has a frontmatter callout clarifying these are the final results. Also hardens the inline Python against unclosed<details>tags and malformedexperiment_linksentries.