← Back to the report
Agent Readiness

Youku

youku.com · Streaming, Video & Music

25
Readiness /100
6
Checks passed
7
Checks failed
62
Checks run

Discoverability

Can agents find your site and its capabilities?

failWebMCP / agent-protocol tools discoveredAgent manifests

The page registers no tools via the WebMCP browser API (navigator.modelContext).

Fix: Expose your actions to AI agents with the WebMCP browser API: register tools via navigator.modelContext.registerTool. Optionally also publish a /.well-known/webmcp.json tool catalog as a discovery signpost.

failllms.txt does not follow recommendationsLighthouse Agentic

If your llms.txt file does not follow recommendations, large language models may not be able to understand how you want your website to be crawled or used for training. The [llms.txt](https://llmstxt.org/) file should be a Markdown file containing at least one H1 header.

warnAGENTS.md operating guide for agentsAgent manifests

Found /AGENTS.md but it is served as "text/html;charset=utf-8", not Markdown.

Fix: Serve AGENTS.md as Markdown (text/markdown). See agents.md.

passKey pages discoveredKey pages crawl

Crawled 12 of 12 discovered pages beyond the homepage.

infoLighthouse Agentic Browsing runLighthouse Agentic

Lighthouse Agentic Browsing audits ran.

infoWebsite platformplatform

Could not identify a known website platform (custom stack or unrecognised).

Understanding

Can agents understand your content?

fail<html lang> attributeStructured data & HTML

No lang attribute on <html>.

Fix: Add lang to the <html> root (e.g. lang="en") so screen readers and AI translation engines pick the right voice / model.

failImages have alt textStructured data & HTML

Only 4 of 159 images have alt attributes (3%).

Fix: Add alt text to every <img>. Screen readers and AI vision models both rely on it to caption images.

failSemantic landmark elementsStructured data & HTML

No semantic landmark elements (main, header, footer, nav) detected on the homepage.

Fix: Wrap your page structure in semantic landmarks. AI structural extractors rely on them to separate navigation from content from boilerplate.

failSingle H1 on homepageStructured data & HTML

No <h1> tag found on the homepage.

Fix: Add exactly one <h1> that names the page. AI summarisers use it as the document title.

warnMarkdown content negotiationOpen agent protocols

Homepage ignored Accept: text/markdown and returned HTML.

Fix: Serve a markdown version of high-value pages when the client asks for text/markdown. AI summarisers, chatbots and IDE agents prefer markdown — fewer tokens, no DOM noise.

warnDOM size is reasonableStructured data & HTML

Homepage contains ~1525 elements — heavy.

Fix: Aim for under 1500 elements on the homepage. Heavier DOMs slow first paint and overflow some AI crawlers' parse buffers.

warnImages have explicit dimensionsStructured data & HTML

Only 0/159 images have width+height — large CLS risk.

Fix: Set width and height on every <img> so the browser reserves space before the image loads. A layout that shifts while loading makes an AI agent mis-click the element it targeted.

warnInternal links on homepageStructured data & HTML

Only 0 internal links on the homepage.

Fix: Add at least 5-10 internal links pointing to your top product, pricing, docs, blog and about pages so an AI agent can navigate to them from the homepage.

warnKey pages have <title> and <h1>Key pages crawl

12 of 12 pages are missing <title> and/or <h1>: /contact, /contact-us, /pricing, /plans, /about, /about-us, /services, /products, /blog, /news, /faq, /help.

Fix: Make sure every page sets a unique <title> and exactly one <h1>. The title is how an AI agent confirms it landed on the right page after navigating.

warnKey pages are server-renderedKey pages crawl

12 of 12 pages return <80 words of visible text — agents without JavaScript see an empty page: /contact (0w), /contact-us (0w), /pricing (0w), /plans (0w), /about (0w), /about-us (0w), /services (0w), /products (0w), /blog (0w), /news (0w), /faq (0w), /help (0w).

Fix: Pre-render or server-render these pages so AI crawlers (which usually don't execute JS) can read them. Frameworks: Next.js Server Components, Nuxt SSR, Astro, or build-time prerendering.

passHomepage content is server-renderedCrawlability (Google)

Homepage server response contains 377 words of visible text — content is reachable without executing JavaScript.

passHeading hierarchy is intactStructured data & HTML

6 headings, no level skips.

Operability

Can agents actually operate the page?

failLinks do not have a discernible nameAccessibility

Lighthouse flagged this audit — agents may be unable to perceive or operate the affected elements.

Fix: Give every link discernible text (avoid bare icons / "click here") so agents know where it goes.

failForm fields have labelsStructured data & HTML

Only 0 of 1 visible interactive form fields have labels.

Fix: Associate every input with a label. Form-filling agents (and screen readers) need it to know what to type into each field.

failWebMCP tools registered with the browserWebMCP tools

The page registers no tools via the navigator.modelContext browser API, so an AI agent has no structured way to operate it.

Fix: Expose your site’s actions to AI agents with the WebMCP browser API: call navigator.modelContext.registerTool({ name, description, inputSchema, execute }) from your page so an agent can invoke them. Implement it directly, or with a library like the @mcp-b polyfill (https://mcp-b.ai). Spec: https://github.com/webmachinelearning/webmcp.

warnAccessibility tree is agent-navigableAccessibility

1 of 6 agent-critical accessibility audits failed (Links do not have a discernible name) — an agent may be unable to identify or operate the affected elements.

Fix: Fix the failing agent-accessibility audits listed below (accessible names on controls, valid ARIA roles/relationships, nothing interactive hidden from the tree).

warnOverall accessibility scoreAccessibility

Lighthouse accessibility score: 78/100. Agents read the page through its accessibility tree, so this is a proxy for how navigable your site is to an AI agent.

Fix: Resolve the failing accessibility audits below — each one removes an element or relationship an agent would otherwise be blind to.

pass`[aria-*]` attributes match their rolesAccessibility

`[aria-*]` attributes match their roles passed.

pass`[aria-hidden="true"]` is not present on the document `<body>`Accessibility

`[aria-hidden="true"]` is not present on the document `<body>` passed.

pass`[aria-*]` attributes are valid and not misspelledAccessibility

`[aria-*]` attributes are valid and not misspelled passed.

pass`[aria-*]` attributes have valid valuesAccessibility

`[aria-*]` attributes have valid values passed.

passDocument has a `<title>` elementAccessibility

Document has a `<title>` element passed.

passAccessibility tree is well-formedLighthouse Agentic

All audits passed

passCumulative Layout ShiftLighthouse Agentic

0

passCumulative Layout ShiftPerformance

0.090 (field data from CrUX) — good.

passTime to First BytePerformance

786ms (field data from CrUX) — good.

infoWebMCP form coverageLighthouse Agentic

1 form missing annotations

infoWebMCP tools registeredLighthouse Agentic

Lists the [WebMCP tools](http://goo.gle/webmcp-docs) registered at the time of analysis.

infoTool catalog signpost publishedWebMCP tools

No tool catalog was found at /.well-known/webmcp.json (or /.well-known/webmcp). The catalog is a community convention, not part of the WebMCP standard, so this does not affect WebMCP presence.

Fix: Consider also publishing a tool catalog at /.well-known/webmcp.json: a JSON document with a "spec" of "webmcp/0.1" and a "tools" array, where each tool declares a name and a clear description. It is a community convention (optional, not part of the WebMCP standard) that lets crawlers and agents discover your tools without executing JavaScript.

Trust & Security

Can agents safely transact?

failNo mixed-content referencesSecurity & trust

Found 5 http:// references in the homepage HTML — browsers will block these resources.

Fix: Update every http:// resource URL to https:// (or use protocol-relative `//`). Common offenders: image CDNs and legacy embed snippets.

warnOAuth authorization server metadataOpen agent protocols

/.well-known/oauth-authorization-server returned fetch failed.

Fix: Publish /.well-known/oauth-authorization-server so AI agents discovering your OAuth setup can negotiate flows automatically. Required if your site offers an authenticated API.

warnOAuth protected-resource metadataOpen agent protocols

/.well-known/oauth-protected-resource returned fetch failed.

Fix: Publish /.well-known/oauth-protected-resource so AI agents discovering your OAuth setup can negotiate flows automatically. Required if your site offers an authenticated API.

warnWeb Bot Auth signature headerOpen agent protocols

No Web Bot Auth signature headers — sites can't verify agent identity.

Fix: Web Bot Auth (IETF HTTP Message Signatures over Signature / Signature-Input) lets you cryptographically verify which agent is hitting you. Several CDNs offer turn-key support; otherwise skip until vendor support matures.

warnContent-Security-Policy headerSecurity & trust

No Content-Security-Policy header on the homepage.

Fix: Add a Content-Security-Policy. Even a strict default-src directive cuts XSS blast radius dramatically. Start in report-only mode to find violations.

warnHTTP Strict-Transport-Security headerSecurity & trust

HSTS configured (max-age=31536000) but missing includeSubDomains.

Fix: Add `includeSubDomains` so subdomains inherit the policy.

warnhttp:// redirects to https://Security & trust

http://youku.com did not respond: fetch failed.

Fix: Listen on port 80 and 301-redirect every request to the https:// origin.

warnReferrer-Policy headerSecurity & trust

No Referrer-Policy header on the homepage.

Fix: Add `Referrer-Policy: strict-origin-when-cross-origin` so outbound links don't leak full URLs (including query strings) to third parties.

warnTLS certificate is not near expirySecurity & trust

TLS handshake failed: .

passSite is served over HTTPSSecurity & trust

Homepage scheme is https:.

passHomepage shows no error outputSecurity & trust

No stack-trace markers in the first 5KB.

passNo tech-stack disclosure in headersSecurity & trust

Server / X-Powered-By headers don't leak product version.

passMIME-sniff protectionSecurity & trust

X-Content-Type-Options: nosniff.

passClick-jacking protectionSecurity & trust

X-Frame-Options: SAMEORIGIN.