dogfood
nousresearch/hermes-agent
Systematic exploratory QA testing for web apps: navigate, interact, capture bugs, and generate structured reports.
What is dogfood?
Dogfood guides you through systematic exploratory testing of web applications using browser tools. You navigate pages, interact with elements, check for console errors and visual issues, collect evidence via screenshots, and produce a structured bug report organized by severity and category.
- Navigate web pages and capture DOM snapshots to understand structure
- Take annotated screenshots with element labels for interactive testing
- Check JavaScript console for errors after navigation and interactions
- Test interactive elements (buttons, forms, links) with valid and invalid inputs
- Collect and de-duplicate issues with severity and category classification
- Generate structured bug reports with evidence screenshots and reproduction steps
How to install dogfood
npx skills add https://github.com/nousresearch/hermes-agent --skill dogfood- Browser toolset must be available (browser_navigate, browser_snapshot, browser_click, browser_type, browser_vision, browser_console, browser_scroll, browser_back, browser_press)
- Target URL and testing scope from the user
How to use dogfood
- 1.Plan your testing scope by identifying pages, features, and user flows to test
- 2.Navigate to each page and take snapshots to understand the DOM structure
- 3.Use browser_vision with annotate=true to identify interactive elements and visual issues
- 4.Test interactive elements systematically (clicks, form inputs, keyboard navigation)
- 5.Check browser_console() after each navigation and interaction for JavaScript errors
- 6.Capture screenshots of any issues found with browser_vision
- 7.Collect issue details including URL, steps to reproduce, expected vs actual behavior, and console errors
- 8.Review and de-duplicate all collected issues, assigning severity and category
Use cases
- Test a new web application feature before release to find bugs and regressions
- Perform comprehensive QA on a specific section or user flow (sign-up, checkout, search)
- Validate form handling, validation rules, and error states across a web app
- Check for console errors, accessibility issues, and visual rendering problems
- Document bugs with screenshots and reproduction steps for developer handoff
- QA engineers and testers performing exploratory testing
- Product managers validating feature quality before launch
- Developers dogfooding their own applications before release
- Technical leads reviewing application stability and user experience
dogfood FAQ
Record the error message, the URL where it occurred, and the steps that triggered it. Include it in the issue report under the Console category. Silent JS errors are high-value findings.
Fill form fields with both valid and invalid inputs, then submit. Check for appropriate error messages, validation feedback, and console errors. Test empty submissions and edge cases like special characters.
It overlays numbered labels [N] on interactive elements in the screenshot, mapping each to a ref like @eN that you can use in subsequent browser commands like browser_click or browser_type.
Create an output directory with a screenshots/ subdirectory for evidence images and a report.md file for the final report. The default location is ./dogfood-output.
Critical issues block core functionality or prevent users from completing key flows. High severity issues significantly degrade experience but don't completely block functionality. Refer to the issue taxonomy in the skill documentation for detailed classification guidelines.
Full instructions (SKILL.md)
Source of truth, from nousresearch/hermes-agent.
name: dogfood description: "Exploratory QA of web apps: find bugs, evidence, reports." version: 1.0.0 platforms: [linux, macos, windows] metadata: hermes: tags: [qa, testing, browser, web, dogfood] related_skills: []
Dogfood: Systematic Web Application QA Testing
Overview
This skill guides you through systematic exploratory QA testing of web applications using the browser toolset. You will navigate the application, interact with elements, capture evidence of issues, and produce a structured bug report.
Prerequisites
- Browser toolset must be available (
browser_navigate,browser_snapshot,browser_click,browser_type,browser_vision,browser_console,browser_scroll,browser_back,browser_press) - A target URL and testing scope from the user
Inputs
The user provides:
- Target URL — the entry point for testing
- Scope — what areas/features to focus on (or "full site" for comprehensive testing)
- Output directory (optional) — where to save screenshots and the report (default:
./dogfood-output)
Workflow
Follow this 5-phase systematic workflow:
Phase 1: Plan
- Create the output directory structure:
{output_dir}/ ├── screenshots/ # Evidence screenshots └── report.md # Final report (generated in Phase 5) - Identify the testing scope based on user input.
- Build a rough sitemap by planning which pages and features to test:
- Landing/home page
- Navigation links (header, footer, sidebar)
- Key user flows (sign up, login, search, checkout, etc.)
- Forms and interactive elements
- Edge cases (empty states, error pages, 404s)
Phase 2: Explore
For each page or feature in your plan:
-
Navigate to the page:
browser_navigate(url="https://example.com/page") -
Take a snapshot to understand the DOM structure:
browser_snapshot() -
Check the console for JavaScript errors:
browser_console(clear=true)Do this after every navigation and after every significant interaction. Silent JS errors are high-value findings.
-
Take an annotated screenshot to visually assess the page and identify interactive elements:
browser_vision(question="Describe the page layout, identify any visual issues, broken elements, or accessibility concerns", annotate=true)The
annotate=trueflag overlays numbered[N]labels on interactive elements. Each[N]maps to ref@eNfor subsequent browser commands. -
Test interactive elements systematically:
- Click buttons and links:
browser_click(ref="@eN") - Fill forms:
browser_type(ref="@eN", text="test input") - Test keyboard navigation:
browser_press(key="Tab"),browser_press(key="Enter") - Scroll through content:
browser_scroll(direction="down") - Test form validation with invalid inputs
- Test empty submissions
- Click buttons and links:
-
After each interaction, check for:
- Console errors:
browser_console() - Visual changes:
browser_vision(question="What changed after the interaction?") - Expected vs actual behavior
- Console errors:
Phase 3: Collect Evidence
For every issue found:
-
Take a screenshot showing the issue:
browser_vision(question="Capture and describe the issue visible on this page", annotate=false)Save the
screenshot_pathfrom the response — you will reference it in the report. -
Record the details:
- URL where the issue occurs
- Steps to reproduce
- Expected behavior
- Actual behavior
- Console errors (if any)
- Screenshot path
-
Classify the issue using the issue taxonomy (see
references/issue-taxonomy.md):- Severity: Critical / High / Medium / Low
- Category: Functional / Visual / Accessibility / Console / UX / Content
Phase 4: Categorize
- Review all collected issues.
- De-duplicate — merge issues that are the same bug manifesting in different places.
- Assign final severity and category to each issue.
- Sort by severity (Critical first, then High, Medium, Low).
- Count issues by severity and category for the executive summary.
Phase 5: Report
Generate the final report using the template at templates/dogfood-report-template.md.
The report must include:
- Executive summary with total issue count, breakdown by severity, and testing scope
- Per-issue sections with:
- Issue number and title
- Severity and category badges
- URL where observed
- Description of the issue
- Steps to reproduce
- Expected vs actual behavior
- Screenshot references (use
MEDIA:<screenshot_path>for inline images) - Console errors if relevant
- Summary table of all issues
- Testing notes — what was tested, what was not, any blockers
Save the report to {output_dir}/report.md.
Tools Reference
| Tool | Purpose |
|---|---|
browser_navigate | Go to a URL |
browser_snapshot | Get DOM text snapshot (accessibility tree) |
browser_click | Click an element by ref (@eN) or text |
browser_type | Type into an input field |
browser_scroll | Scroll up/down on the page |
browser_back | Go back in browser history |
browser_press | Press a keyboard key |
browser_vision | Screenshot + AI analysis; use annotate=true for element labels |
browser_console | Get JS console output and errors |
Tips
- Always check
browser_console()after navigating and after significant interactions. Silent JS errors are among the most valuable findings. - Use
annotate=truewithbrowser_visionwhen you need to reason about interactive element positions or when the snapshot refs are unclear. - Test with both valid and invalid inputs — form validation bugs are common.
- Scroll through long pages — content below the fold may have rendering issues.
- Test navigation flows — click through multi-step processes end-to-end.
- Check responsive behavior by noting any layout issues visible in screenshots.
- Don't forget edge cases: empty states, very long text, special characters, rapid clicking.
- When reporting screenshots to the user, include
MEDIA:<screenshot_path>so they can see the evidence inline.
Related skills
More from nousresearch/hermes-agent and the wider catalog.
find-skills
Discover and install agent skills to extend your coding agent's capabilities on demand
frontend-design
Build visually distinctive UI with opinionated aesthetic direction, typography, and layout choices that avoid templated defaults.
vercel-react-best-practices
70 React/Next.js performance rules from Vercel Engineering, prioritized by impact for writing, reviewing, and refactoring code.
agent-browser
Fast browser automation CLI for AI agents — navigate, click, scrape, screenshot, and test via Chrome CDP
web-design-guidelines
Review UI code against Web Interface Guidelines for accessibility, UX, and design best practices
finetuning
Fine-tune models on Azure AI Foundry with SFT, DPO, or RFT training methods.