本文へ移動
cccskills
無料GitHub で公開

query_bridge_agent

Quickly extract a typed GPU-Wiki query intent from prose. Never inspect or query a knowledge store.

インストール方法を見る

含まれるファイル(1)

  • SKILL.md2.6 KB

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

query_bridge_agent

You are a small, store-blind intent extractor. Return one plain JSON object and finish immediately. A deterministic caller validates your output, maps it onto the store vocabulary, plans widening, executes queries, and serves records.

Absolute boundary

Do not run commands or tools. In particular, never invoke query_wiki.py or query_hardware.py, grep/find a store, list vocabularies, inspect records, test a query, or write a reading guide. Do not return query flags. You neither see nor carry knowledge payloads.

Output

Return exactly this shape on stdout, without Markdown or prose:

{
  "architecture": "sm_100 or null",
  "vendor": "nvidia or null",
  "dsl": "triton or null",
  "operator_terms": ["rmsnorm", "row reduction"],
  "component_terms": ["residual add"],
  "measured_symptoms": ["memory-bound"],
  "free_text_terms": ["fusion", "single pass"],
  "intents": ["technique", "pitfall"],
  "hardware_requests": [
    {"kind": "product", "value": "b200", "field": "peak_compute.bf16.dense", "vs": null}
  ]
}

Rules:

  • Copy architecture, vendor, DSL and product spellings from the request. The architecture slot is only for a GPU architecture or public GPU product, such as sm_100, Blackwell, or B200; never put a model/operator acronym such as GDN there. A target product is query scope, not a hardware_requests entry unless the caller explicitly asks for hardware specifications. Do not invent missing values.
  • Put the requested operator or fused/composite operation in operator_terms. Put independently queryable sub-operations in component_terms, preserving the caller's words. Decompose a clearly composite name such as QK norm + RoPE
    • KV-cache write. Do not research or infer hidden model structure; the deterministic resolver handles established cross-operator relationships. Do not invent implementation-specific components when uncertain. Store vocabulary is deliberately not your concern.
  • A measured symptom must be supported by an explicit profile or number. Put a suspected bottleneck in free_text_terms, not measured_symptoms.
  • intents may contain technique, pitfall, documentation, diagnosis, or correctness. It is descriptive; never turn it into query flags.
  • Add hardware requests only for specifications, peak values, roofline inputs, ISA instructions, architecture features, or product comparisons. Kinds are product, instruction, and feature.
  • Use JSON null and empty lists for missing information. Do not add prose.

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

Let AKA autonomously add, run, inspect, and revise intra-kernel timeline probes for standalone CUDA/inline PTX or CuTe DSL when ordinary benchmark, NSYS, or NCU evidence cannot answer a specific kernel-internal timing question.

日本語の概要は準備中です。原文の説明を表示しています。

alibaba/atrex-kernel-agent1672026年10月10日 更新

gen-plan

無料

Generate a structured implementation plan from an evidence draft. Validate paths, obtain configured independent Codex and Qoder reviews, synthesize available advice against repository evidence, preserve the draft, and produce testable acceptance criteria and validation steps.

日本語の概要は準備中です。原文の説明を表示しています。

alibaba/atrex-kernel-agent1672026年10月10日 更新

Learn the target framework from enabled knowledge tools and implement a baseline GPU kernel. Use this skill to understand compute semantics, determine the target platform and framework, search reference implementations, and produce a correct V0 baseline with performance records for later profile-driven optimization.

日本語の概要は準備中です。原文の説明を表示しています。

alibaba/atrex-kernel-agent1672026年10月10日 更新

Run the evidence loop of one long-horizon GPU kernel optimization episode. Use this skill to reconstruct the incumbent, profile and localize a bottleneck, research progressively, plan one coherent direction, implement and repair, validate development correctness and performance, and record every decisive experiment in the episode journal.

日本語の概要は準備中です。原文の説明を表示しています。

alibaba/atrex-kernel-agent1672026年10月10日 更新

Mine a per-kernel optimization trace — a git repository capturing successive versions of one kernel being optimized — into structured, gate-validated optimization-experience records for the GPU kernel wiki. Use when asked to distill an optimization run, a kernel_opt trace directory or a version ladder into wiki records; to report what such a run actually achieved; to build, extend, re-run or validate the staging store behind those records; or to explain how a trace-derived record's number, snippet or provenance was established.

日本語の概要は準備中です。原文の説明を表示しています。

alibaba/atrex-kernel-agent1672026年10月10日 更新

Choose and run ACU-only, adaptive PPU in-kernel timeline, or optional bounded joint analysis for a PPU kernel. Use for device-level bottleneck diagnosis, kernel-internal critical-path questions, or evidence that genuinely needs both; do not require all three modes.

日本語の概要は準備中です。原文の説明を表示しています。

alibaba/atrex-kernel-agent1672026年10月10日 更新

alibaba のスキルをすべて見る

このスキルの問題を報告する