feat(agentx): publish AgentX methodology and canonical routes - #755
Merged
Conversation
|
The latest updates on your projects. Learn more about Vercel for GitHub.
|
cquil11
marked this pull request as ready for review
August 18, 2026 20:37
There was a problem hiding this comment.
Cursor Bugbot has reviewed your changes using default effort and found 1 potential issue.
❌ Bugbot Autofix is OFF. To automatically fix reported issues with cloud agents, enable autofix in the Cursor dashboard.
Reviewed by Cursor Bugbot for commit 0fe51da. Configure here.
This comment has been minimized.
This comment has been minimized.
Document the v1.0 trace pipeline, replay controls, interpretation guidance, and reconstruction limits on /datasets and /zh/datasets. Add E2E coverage for both localized pages. 中文:补充 AgentX v1.0 的 trace 处理流程、回放控制、结果解读与重建边界,并同步更新 /datasets 和 /zh/datasets;新增覆盖中英文页面的 E2E 测试。
cquil11
added a commit
that referenced
this pull request
Aug 18, 2026
Use agentic inference for category-level headlines, SEO titles, and labels. Keep AgentX as the name of the long-context, multi-turn coding scenario, workload, results, and methodology. Remove vague promotional and AI-writing-style phrasing, and point machine discovery to the /agentx route from #755. 中文:在类别层级的标题、SEO 标题与标签中统一使用“智能体推理”,仅在指代长上下文多轮编码场景、工作负载、结果与方法时使用 AgentX。删除空泛宣传及明显的 AI 写作套话,并将机器可读入口指向 #755 提供的 /agentx 路由。
Make /agentx and /zh/agentx canonical, preserve the former datasets URLs with permanent redirects, and update navigation, metadata, internal links, analytics, and tests. Add bilingual long-form methodology pages with source-grounded copy, 16 original figures, descriptive captions, and reproducibility links. 中文:发布 AgentX 规范路由与方法论页面。将 /agentx 和 /zh/agentx 设为规范地址,以永久重定向保留旧 datasets 链接,并同步更新导航、元数据、站内链接、分析事件与测试。新增中英文长篇方法论页面,使用 16 张原始图片、准确图注及可复现性链接。
cquil11
force-pushed
the
agent/agentx-methodology
branch
from
August 18, 2026 21:12
0df295e to
71db075
Compare
Place each definition term before its value in the DOM on both AgentX methodology surfaces while retaining the existing value-first visual order. 中文:修正 AgentX 统计信息的语义结构。在两处方法论页面中,让每个定义术语在 DOM 中位于数值之前,同时保持现有的数值优先视觉顺序。
Keep the hamburger navigation active through 1024 CSS pixels so the six-link AgentX header does not overflow. Add component coverage at 1009, 1012, 1020, and 1024 pixels and retain an explicit xl desktop-bounds check. 中文:将桌面导航延后至 xl 屏幕。在 1024 CSS 像素及以下继续使用汉堡菜单,避免包含六个链接的 AgentX 页眉横向溢出;新增 1009、1012、1020 和 1024 像素的组件测试,并保留 xl 桌面边界检查。
cquil11
added a commit
that referenced
this pull request
Aug 18, 2026
Use agentic inference for category-level headlines, SEO titles, and labels. Keep AgentX as the name of the long-context, multi-turn coding scenario, workload, results, and methodology. Remove vague promotional and AI-writing-style phrasing, and point machine discovery to the /agentx route from #755. 中文:在类别层级的标题、SEO 标题与标签中统一使用“智能体推理”,仅在指代长上下文多轮编码场景、工作负载、结果与方法时使用 AgentX。删除空泛宣传及明显的 AI 写作套话,并将机器可读入口指向 #755 提供的 /agentx 路由。
* feat(marketing): surface AgentX across site copy Position the shipped AgentX agentic coding workload across core metadata, indexable copy, machine-discovery feeds, and localized Chinese surfaces without adding new promotional UI.\n\nCorrect stale public copy that described AgentX or the benchmark cadence as forthcoming/nightly, and add regression coverage for English and Chinese positioning.\n\n中文:在网站核心文案与元数据中突出 AgentX\n\n在核心元数据、可索引内容、机器发现订阅源及中文页面中准确呈现已上线的 AgentX 智能体编码工作负载,不新增促销式 UI。同步修正将 AgentX 描述为尚未上线、以及将当前测试频率描述为每夜运行的过时文案,并补充中英文回归测试。 * fix(marketing): distinguish AgentX from agentic inference Use agentic inference for category-level headlines, SEO titles, and labels. Keep AgentX as the name of the long-context, multi-turn coding scenario, workload, results, and methodology. Remove vague promotional and AI-writing-style phrasing, and point machine discovery to the /agentx route from #755. 中文:在类别层级的标题、SEO 标题与标签中统一使用“智能体推理”,仅在指代长上下文多轮编码场景、工作负载、结果与方法时使用 AgentX。删除空泛宣传及明显的 AI 写作套话,并将机器可读入口指向 #755 提供的 /agentx 路由。 * fix(overview): shorten agentic inference heading Use the shorter category-correct heading in English and Chinese to preserve the non-scrolling overview layout at desktop, tablet, and phone widths. 中文:中英文统一采用更短且类别表述准确的智能体推理成本标题,确保总览页面在桌面、平板与手机宽度下不会出现横向滚动。
Add a first-class Agentic inference category with AgentX, workload, replay, closed-loop, and subagent definitions. Keep the English and Simplified Chinese glossary surfaces, metadata, cross-links, and content-contract tests aligned.\n\n中文:新增“智能体推理”一级分类,并补充 AgentX、智能体编码工作负载、轨迹回放、闭环基准测试与子智能体定义;同步英文和简体中文术语表、元数据、交叉链接及内容契约测试。
Add a localized NEW badge to the AgentX tab on desktop and mobile. Replace the Kimi K3 launch promotion with an agentic-results banner and modal that link directly to Agentic Traces, name the covered models, and use View results actions. Capitalize the AgentX Methodology heading and extend route, nudge, locale, and responsive coverage. 中文:为桌面端和移动端的 AgentX 导航标签添加本地化“NEW”徽标;将 Kimi K3 发布提示替换为智能体结果横幅与弹窗,直接跳转至 Agentic Traces,列出已覆盖模型,并统一使用“查看结果”操作;同时修正 AgentX Methodology 标题大小写,并补充路由、提示、本地化与响应式测试。
Render the AgentX navigation, launch banner, and launch modal badges through one shared 32 by 16 pixel component. Add an end-to-end geometry assertion so the three pills cannot drift apart again. 中文:AgentX 导航、发布横幅和发布弹窗统一使用同一个 32×16 像素徽标组件,并新增端到端尺寸断言,防止三个“NEW”徽标再次出现样式偏差。
Shift the shared badge label by one pixel to compensate for the uppercase font baseline while preserving the common 32 by 16 pixel pill. Extend the browser regression test to verify the label position in all three placements. 中文:将共享徽标文字上移 1 像素,以补偿大写字体的视觉基线,同时保持统一的 32×16 像素胶囊尺寸。扩展浏览器回归测试,验证三个使用位置的文字对齐。
Center every NEW label exactly within its shared pill and remove the trailing letter-spacing imbalance. Replace ten methodology figures with the supplied 4200-pixel originals, add five new full/256k and replay variants, and extend bilingual captions and browser coverage to all 21 figures. 中文:将所有“NEW”文字精确居中到共享徽标中,并移除字距造成的视觉偏移。使用提供的 4200 像素原图替换 10 张方法论图片,新增 5 张 full/256k 与回放变体图片,并为全部 21 张图片补充中英文说明和浏览器测试。
The 32x16 pill carried px-1.5, leaving a 20px content box for a label that inks ~23px at 10px bold. The glyphs overflowed that box and Chrome pushed them to the end, so `NEW` sat 6px from the pill's left edge and 2.7px from the right, with the W nearly touching it. Drop the padding to slack (px-0.5) and centre with flex on the pill itself instead of a nested full-width grid. That nested `w-full` label is also what hid the bug from the e2e check: its box stayed centred at 20px while the text spilled out of it, so assert on the rendered ink. Renders 32x16 with even 4.34px gutters in DM Sans and in every fallback tested (Arial, Verdana, Tahoma, Trebuchet, Impact, Courier New), so the shared badge size holds even when next/font's `display: optional` falls back.
functionstackx
added a commit
that referenced
this pull request
Aug 19, 2026
#755 rewrote the footer description and the "Explore InferenceX" lead and platform-coverage lines around AgentX. Restore the previous copy in both languages: the footer returns to the trusted-by line, and the landing page returns to the active-models lead and the hardware-coverage list. 中文:#755 曾把页脚描述以及"探索 InferenceX"的引导语与平台覆盖说明改写为围绕 AgentX 的表述。现将中英文文案恢复为此前版本:页脚回到"获得万亿美元级 AI 基础设施运营方信赖"的表述,落地页回到活跃模型引导语与硬件覆盖列表。 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
functionstackx
added a commit
that referenced
this pull request
Aug 19, 2026
…页面 (#766) * feat(agentx): publish the AgentX optimizations pages AgentX's first months produced 50+ upstream PRs across the serving stack, and that work only existed as an internal document. Publish it as /agentx/optimizations plus one page per project at /agentx/optimizations/<framework>, covering vLLM, SGLang, TensorRT-LLM, AMD ATOM, ROCm AITER, NVIDIA Dynamo, LMCache, and Mooncake. Every claim carries the pull request it came from. Sections store PRs as {repo, number} rather than prose links, so the reference set is shared by both languages and cannot drift between them; the same holds for the ten diagrams, which are keyed by asset rather than duplicated per locale. The index page adds the distributed-inference primer, the day-zero enablement section, and what the AgentX matrix activates. /agentx (and /zh/agentx) gains a callout with a button to the index and a direct link per project, so a reader who already knows which engine they run can skip the index. Both trees are registered in the sitemap. The editorial notes in the source document ("TO BE UPDATED", the request for more PRs) are deliberately not published, and a unit test fails if any of those markers reappear. 中文:AgentX 上线头几个月在推理栈各层推动了 50 多个上游 PR,此前这些内容只 存在于内部文档中。现将其发布为 /agentx/optimizations 以及按项目划分的 /agentx/optimizations/<framework> 页面,覆盖 vLLM、SGLang、TensorRT-LLM、 AMD ATOM、ROCm AITER、NVIDIA Dynamo、LMCache 与 Mooncake。 每条结论都附带对应的 PR。各小节以 {repo, number} 结构而非正文链接存储 PR, 因此中英文共用同一份引用集合、不会产生漂移;十张示意图同样按资源 key 引用, 不在各语言重复。索引页还包含分布式推理生态简介、day-zero 支持章节,以及 AgentX 测试矩阵激活了什么。 /agentx(及 /zh/agentx)新增引导卡片,提供进入索引页的按钮和每个项目的直达 链接,方便已经确定所用引擎的读者跳过索引页。中英文两棵路由树均已登记到 sitemap。 源文档中的编辑批注("TO BE UPDATED"、补充 PR 的请求)刻意未予发布,并有单元 测试在这些标记重新出现时报错。 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * revert(copy): restore the footer and landing blurbs from before #755 #755 rewrote the footer description and the "Explore InferenceX" lead and platform-coverage lines around AgentX. Restore the previous copy in both languages: the footer returns to the trusted-by line, and the landing page returns to the active-models lead and the hardware-coverage list. 中文:#755 曾把页脚描述以及"探索 InferenceX"的引导语与平台覆盖说明改写为围绕 AgentX 的表述。现将中英文文案恢复为此前版本:页脚回到"获得万亿美元级 AI 基础设施运营方信赖"的表述,落地页回到活跃模型引导语与硬件覆盖列表。 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * feat(agentx): add a results CTA to the AgentX datasets card The AgentX card explains how the benchmark is built but never offered a way to the numbers it produces. Add a primary "View AgentX Performance Results" button under the intro, pointing at /overview (/zh/overview on the Chinese page), alongside the existing methodology link. Also update the zh footer assertion to the restored copy from the previous commit — it still expected the AgentX-flavored wording. 中文:AgentX 卡片此前只说明基准测试如何构建,却没有通往测试结果的入口。现在 在简介下方新增主按钮"查看 AgentX 性能结果",指向 /overview(中文页面为 /zh/overview),与已有的方法论链接并列。 同时把中文页脚的断言更新为上一个提交恢复后的文案——它此前仍在断言围绕 AgentX 的旧表述。 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * fix(e2e): keep the centered launch modal out of the banner tests The launch modal is centered behind a full-screen backdrop, and that backdrop covers the launch banner — including its dismiss button — whenever both are eligible. Since the nudge spec opts out of the global modal suppression, the banner's two click tests raced the modal's mount and failed on whichever shard lost. Give the banner describe its own storage helper that clears everything except the modal, which stays dismissed. Also append the AgentX scenario sentence to the landing lead in both languages. 中文:发布弹窗现在是带全屏遮罩的居中模态框,当它与发布横幅同时符合展示条件 时,遮罩会盖住横幅(包括其关闭按钮)。由于 nudge 用例集已退出全局弹窗抑制, 横幅的两个点击用例会与弹窗挂载竞争,在部分分片上失败。现为横幅用例单独提供 存储辅助函数:清空其他状态,但保持弹窗为已关闭。 同时在落地页引导语的中英文版本后追加 AgentX 场景说明。 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> --------- Co-authored-by: Claude Opus 5 (1M context) <noreply@anthropic.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.

Summary
/agentxand/zh/agentxthe canonical AgentX surfaces, with permanent redirects from/datasetsand/zh/datasets, including nested paths and query stringsNEWbadge on desktop and mobile; navigation and launch badges share one 32×16 treatment with labels centered at an exact zero-pixel offset/agentx/methodologyplus/zh/agentx/methodologyfor the full methodology, including 21 figures, captions, alt text, and primary-source links; 15 supplied source figures use their original 4200-pixel filesAgentic benchmark results are live; the modal names Kimi K3, DeepSeek-V4-Pro, MiniMax-M3, Qwen3.5 397B, and GLM-5.2; both useView resultsactions and link directly to/inference?i_seq=agentic-tracesAgentX Methodologyheading and keep all new user-facing content synchronized with natural Simplified Chinese siblingsThis PR now includes the semantic-marketing and glossary follow-ups previously reviewed in #756 and #757.
Source and editorial audit
The public methodology is adapted from the methodology section of InferenceXv3/AgentX: How does CUDA Moat Hold up in Agentic Coding?. It summarizes the measured methodology instead of reproducing internal or unfinished prose.
The page omits unresolved cost claims, TODO sections, promotional superlatives, the internal proxy warning, and the unattributed distillation meme. It does not claim that placeholders reconstruct original prompts, code, or tool payloads. New copy was audited against Wikipedia: Signs of AI writing for generic puffery, canned contrast, vague attribution, repetitive conclusions, excessive headings, and related patterns.
Validation
bun run fmtbun run lintbun run typecheckbun run test:unit— 219 files, 3,957 testsE2E_FIXTURES=1 bun run build— 950 statically generated pagesLatest local validation passes all 3,957 unit tests, 27 focused Cypress tests, the 950-page fixture production build, type checking, lint, and formatting. Browser verification confirms all 21 English and Chinese figures load without missing alt text, overflow, or framework errors. Post-push checks were started but not awaited at the request of the reviewer.
Scope
This PR changes static AgentX routes, methodology content, site positioning, glossary content, navigation, and launch promotions. It does not change benchmark measurements, database contents, API response contracts, or inference chart calculations. Unofficial-run overlays are unaffected.
Note
Low Risk
Changes are routing, static content, navigation, and marketing copy; benchmark APIs, chart math, and measurement pipelines are untouched. Redirect and link churn is the main regression surface for bookmarks and external references.
Overview
AgentX becomes the public surface for trace datasets and methodology: new
/agentxand/zh/agentxpages (overview + dataset list), full/agentx/methodologysiblings with 21 figures and sourced copy, while/datasets→/agentx308 redirects preserve paths and query strings. Old standalone datasets index pages are removed in favor of this tree.Navigation and discovery move from “Datasets” to AgentX in the header (with a shared
NEWbadge viaNewBadge), footer, scenario tooltip, timeline deep links, sitemap, andllms.txt. Desktop primary nav now uses thexlbreakpoint and adds overflow checks around 1009–1024px so the hamburger layout does not clip.Site messaging reframes InferenceX as an agentic inference benchmark (AgentX as the named long-context coding scenario) across landing, overview titles (“Agentic Inference Costs”), About/FAQ, glossary (new Agentic inference category and terms), blog metadata, manifest, and share text. Several blog MDX captions/links point to
/agentxinstead of/datasets.Launch promotions swap Kimi K3 storage keys and copy for agentic benchmark results; banner/modal View results routes to
/inference?i_seq=agentic-traces.Tests update Cypress for new routes, methodology e2e, redirects, and nudge/badge behavior; unit tests for conversation hrefs use
/agentx.next.configadds imagequalities: [75, 100]for methodology figures.Reviewed by Cursor Bugbot for commit 5aa625f. Bugbot is set up for automated code reviews on this repo. Configure here.