literature-close-read
Produce a structured close-reading report from a paper's full PDF-to-Markdown text (with `## Page XX` pagination and image references) when you need to systematically extract background, research questions, methods, results, limitations, and reproducible experimental details.
Veto GatesRequired pass for any deployment consideration
| Dimension | Result | Detail |
|---|---|---|
| Scientific Integrity | PASS | The legacy audit did not indicate that retrieval outputs were presented as unsupported findings. |
| Practice Boundaries | PASS | The package stayed in retrieval, extraction, or evidence-organization scope rather than drifting into unsupported interpretation. |
| Methodological Ground | PASS | The legacy audit preserved a method-grounded interpretation of the Produce a structured close-reading report from a paper's full PDF-to-Markdown text (with ## Page XX pagination and image references) when you need to systematically extract background, research questions, methods, results, limitations, and reproducible experimental details workflow. |
| Code Usability | N/A | The package is evaluated primarily as a structured deliverable rather than an executable scientific code workflow. |
Core Capability85 / 100 — 8 Categories
Medical TaskExecution Average: 87.6 / 100 — Assertions: 20/20 Passed
When you have a full paper converted from PDF to Markdown and need... was evaluated as a bounded documentation path, not as a runnable script workflow.
This variant a case stayed inside the documented workflow and remained instruction-led.
This edge case stayed inside the documented workflow and remained instruction-led.
This variant b case stayed inside the documented workflow and remained instruction-led.
End-to-end case for Reads the entire Markdown paper text,... was evaluated as a bounded documentation path, not as a runnable script workflow.
Key Strengths
- Primary routing is Evidence Insight with execution mode A
- Static quality score is 85/100 and dynamic average is 79.6/100
- Assertions and command execution outcomes are recorded per input for human review
- Execution verification summary: No script verification was applicable