[Газиз Валитов] · deep-research-skill.zip

deep-research-skill

  • Score: 32/40
  • Verdict: Pass (workshop-pass)
  • Отправитель: “Газиз Валитов” valitovgaziz@yandex.ru
  • SHA-256: 383ee44b3e4a5e1c…
  • Blockers: none
  • Unverified: none (все нули — failed; требуемые artifacts/traces отсутствуют в сдаче)

Required rework

  1. F06 — добавить отсутствующие eval case-типы (boundary, security injection, broken script, portability) и поле failure_modes.
  2. F07/F08 — провести и зафиксировать наблюдаемые прогоны (baseline + triggers + security/evidence traces).
  3. A02 — переименовать deep-research в прикладное не-reserved имя.
  4. A05 — убрать примеры запросов и feature-inventory из description.
  5. D03 — добавить Contents в файлы >100 строк; убрать цепочки ссылок.
  6. E03 — сделать промежуточные artifacts опциональными.
  7. A01 — убрать лишние frontmatter-ключи.

Полная scorecard

IDScoreStatusEvidenceDefect / minimal fix
A010/1failedSKILL.md:1-8 — frontmatter содержит доп. ключи license: MIT и metadataУбрать license и metadata из frontmatter
A020/1failedSKILL.md:2name: deep-research; совпадает с reserved-именем deep-researchПереименовать в прикладное имя, напр. sdlc-evidence-research
A031/1passedSKILL.md:3description 455 символов, без XML-тегов, routing-текст
A041/1passedSKILL.md:3 — задача (research SDLC practices), триггеры, domain-термины; третье лицо
A050/1failedSKILL.md:3 — содержит примеры запросов и feature-inventoryУбрать примеры и инвентарь методов из description
A061/1passedSKILL.md:3 — positive/negative/boundary intents различимы
A071/1passedSKILL.md:100-116 — все ссылки относительные, существуют, переносимы
B011/1passedSKILL.md:14-98 — повторяемая 6-шаговая процедура с результатом
B021/1passedSKILL.md:12 — одна связная единица работы; узкий domain scope
B031/1passedSKILL.md:65-78 (freshness); references/sdlc-heuristics.md:137-145 (evidence aging)
B041/1passedSKILL.md:14-98 — структурированный workflow с зонами judgment и gates
C011/1passedSKILL.md:14-98 — один ясный default path: 6 последовательных шагов
C021/1passedSKILL.md:30,40-42 — ветки с условиями выбора; default ясен
C031/1passedSKILL.md:80-94 (self-review verification); consequential irreversible actions отсутствуют
C041/1passedSKILL.md:188-211 — таблица rework loop: Failure → Return To → Change → Rerun; max 2 cycles
C051/1passedSKILL.md:96-98; assets/research-report-template.md — exact schema, readiness criterion
C061/1passedSKILL.md:14-98 — common workflow легко найти, шаги операционны
C071/1passedSKILL.md:118-153 — tool policy: capability guards, fallback; нет скрытых действий
D011/1passedSKILL.md — 222 строки (≤ 500)
D021/1passedSKILL.md — lean body: core workflow, safety, decisions; теория в references/
D030/1failed5 файлов >100 строк без Contents; цепочка ссылок SKILL→A→BДобавить Contents в файлы >100 строк; убрать цепочки
D041/1passedreferences/, assets/, tests/ — назначение корректно, имена описательные
D051/1passedscripts отсутствуют; common path их не требует
D061/1passedнет broken required helpers; common path документирован
E011/1passedSKILL.md:157-166 — untrusted content boundary: внешний контент не переопределяет workflow
E021/1passedSKILL.md:181-186 — credential/data safety; нет установки dependencies
E030/1failedSKILL.md:46, 82 — авто-создание промежуточных draft-файлов без запроса пользователяДелать artifacts опциональными или по запросу
E041/1passedнет shell-команд, bundled binaries, obfuscated code
E051/1passedSKILL.md:155-186 — skill усиливает safeguards; “ignore previous instructions” → suspect
E061/1passedSKILL.md:128-179 — Capability Guards для web search/browser/PDF; offline fallback
E071/1passedgrep не нашёл machine-specific путей; все пути относительные
E081/1passedMCP tools не используются и не предполагаются
F011/1passedassets/design.md:5-19 — use case + outcome; baseline failures B1–B7; supporting files/evals
F021/1passedreferences/source-quality.md:1-100 — 4-tier taxonomy, primary sources в приоритете
F031/1passedSKILL.md:65-78 (freshness); references/sdlc-heuristics.md:137-155 (stopping)
F041/1passedSKILL.md:56-59 (facts vs inferences); «Insufficient Evidence» как валидный итог
F051/1passedassets/research-report-template.md — Research Question, Summary, Evidence, Limitations, Confidence, Provenance
F060/1failedtests/eval-scenarios.md — есть positive/negative/evidence-quality, но отсутствуют: boundary, security injection, broken script, portability; нет failure_modesДобавить 4 недостающих case-типа и поле failure_modes
F070/1failedtests/eval-scenarios.md:115-124 — файлы с ожидаемым поведением без прогона; нет baseline-traceПровести прогон baseline + positive/negative/boundary; приложить trace
F080/1failedнет trace навигации/injection/evidence-проверок в сдачеПровести прогоны security-injection и evidence-quality; приложить trace

Required rework

  1. F06 — добавить отсутствующие eval case-типы (boundary, security injection, broken script, portability) и поле failure_modes.
  2. F07/F08 — провести и зафиксировать наблюдаемые прогоны (baseline + triggers + security/evidence traces).
  3. A02 — переименовать deep-research в прикладное не-reserved имя.
  4. A05 — убрать примеры запросов и feature-inventory из description.
  5. D03 — добавить Contents в файлы >100 строк; убрать цепочки ссылок.
  6. E03 — сделать промежуточные artifacts опциональными.
  7. A01 — убрать лишние frontmatter-ключи.

← назад к лидерборду