Skip to content

nytimes.com

Agents are likely to struggle

What should the prompt cover?

16 findings · 5 selected

Failures (8)

Warnings (8)

Task

What does nytimes.com do and who is it for? Explain it back to me.

Waiting for the agent’s first step…

Critical access blockers remain

These checks describe whether an ordinary agent can enter, read, and operate the public site.

  • Agents can reach the site

    Crawler access and bot defenses.

    0 / 2 passed
  • Core content is available

    Useful content remains accessible without a fragile browser-only path.

    0 / 1 passed
  • Navigation fails safely

    Redirects and missing pages give agents a recoverable path.

    1 / 2 passed
  • Controls are understandable

    Forms and interactive controls expose usable names and structure.

    4 / 4 passed

Evaluated surfaces have material gaps

The public website is always evaluated. Optional surfaces appear when the scan finds positive evidence that they apply.

Public website

Needs work

59%

5 of 15 mature checks passed

API

Blocked

35%

2 of 7 mature checks passed

Authentication

Needs work

67%

2 of 3 mature checks passed

Fix these gaps first

Critical access gaps come first, followed by other applicable readiness gaps.

  1. 01

    Agent crawler reachability

    Verify that major agent User-Agents can reach the homepage. If your WAF or bot rules block them, remove or narrow the blocking rule. Add an allow rule only when your security setup denies them by default.

    Critical access
  2. 02

    Not blocked by bot detection

    Allowlist known AI agent User-Agents (ChatGPT-User, ClaudeBot, Google-Extended, DeepSeekBot) in your WAF or bot-detection rules.

    Critical access
  3. 03

    Agent-friendly 404s

    Return a real HTTP 404 (or 410) status for nonexistent paths - never a 200 with your app shell, which makes agents believe every path exists. For full credit, give the 404 response a short markdown body pointing agents at your sitemap, llms.txt, or docs index. Verify with curl -s -o /dev/null -w "%{http_code}" https://yourdomain.com/some-path-that-does-not-exist - it must print 404.

    Critical access
  4. 04

    Content without JavaScript

    Server-side render your homepage so AI crawlers see meaningful content without JavaScript. Ensure an H1 and 500+ chars of text in raw HTML.

    Critical access
  5. 05

    OpenAPI spec published

    Publish an OpenAPI (Swagger) specification at /openapi.json or /api/openapi.yaml. This is how agents understand your API surface automatically.

    Other readiness checks

Audit the checks behind the score

Applicable evidence is grouped by how it contributes to this preview model. Bonus checks appear only when they add points.

Essential2 of 10 passed · 29.3 / 80 points
  • Content without JavaScriptPartial (67%)

    5896 chars with H1 but flat heading structure

    Recommendation

    Server-side render your homepage so AI crawlers see meaningful content without JavaScript. Ensure an H1 and 500+ chars of text in raw HTML.

  • Not blocked by bot detectionPartial (50%)

    Some agents blocked: GPTBot, ClaudeBot, ChatGPT-User, PerplexityBot

    Recommendation

    Allowlist known AI agent User-Agents (ChatGPT-User, ClaudeBot, Google-Extended, DeepSeekBot) in your WAF or bot-detection rules.

  • Redirect hygienePassed

    No meta-refresh stubs, JavaScript-redirect stubs, or cross-domain hops across 1 checked page

    Recommendation

    Replace meta-refresh and JavaScript-only redirects with real HTTP 301/302 redirects. Non-JS agents never execute location.href or wait for a meta refresh - they see only the stub page. Verify with curl -sI <url> - you should see a Location header, not a 200 with a near-empty body.

  • OpenAPI spec publishedFailed

    No OpenAPI/Swagger specification found

    Recommendation

    Publish an OpenAPI (Swagger) specification at /openapi.json or /api/openapi.yaml. This is how agents understand your API surface automatically.

  • Markdown content negotiation (acceptmarkdown.com)Failed

    Not acceptmarkdown.com compliant: Accept: text/markdown returned text/html; charset=utf-8; Vary header missing Accept (got "accept-encoding, fastly-ssl")

    Recommendation

    On the responses that serve text/markdown via Accept negotiation, add Accept to the Vary header (Vary: Accept, Accept-Encoding). Without it, CDNs can serve the cached HTML variant to an agent asking for markdown (or vice versa), depending on which variant landed in cache first.

  • Agent crawler reachabilityFailed

    Some AI crawlers are blocked - ChatGPT-User: blocked, ClaudeBot: blocked, Google-Extended: reachable, ora-agent: reachable, DeepSeekBot: reachable

    Recommendation

    Verify that major agent User-Agents can reach the homepage. If your WAF or bot rules block them, remove or narrow the blocking rule. Add an allow rule only when your security setup denies them by default.

  • OAuth 2.0 supportPassed

    OAuth endpoint found at https://help.nytimes.com/oauth/authorize

    Recommendation

    Implement OAuth 2.0 for API authentication. Publish your authorization server metadata at /.well-known/oauth-authorization-server.

  • Scoped permissionsFailed

    No declared OAuth scopes, security schemes, or scoped-permission documentation found

    Recommendation

    Declare scoped API permissions where machines can read them: named OAuth scopes in your OpenAPI security schemes, or scopes_supported in RFC 9728 protected-resource metadata. Prose descriptions of roles help humans, but agents need the machine-readable declaration to request least-privilege access.

  • JSON error responsesFailed

    API does not return JSON error responses (or no API detected)

    Recommendation

    Return structured JSON error responses with error codes, messages, and resolution hints. Agents can't parse HTML error pages.

  • Agent-friendly 404sPartial (50%)

    Nonexistent paths return a real HTTP 404. For full credit, include a short markdown body (site map links, where to look next) so agents can recover.

    Recommendation

    Return a real HTTP 404 (or 410) status for nonexistent paths - never a 200 with your app shell, which makes agents believe every path exists. For full credit, give the 404 response a short markdown body pointing agents at your sitemap, llms.txt, or docs index. Verify with curl -s -o /dev/null -w "%{http_code}" https://yourdomain.com/some-path-that-does-not-exist - it must print 404.

Recommended7 of 15 passed · 12.8 / 20 points
  • Developer resource discoverabilityPartial (67%)

    Agent found developer resources by name including developer portal (3 relevant pages). Not searchable: MCP server. Not found via search: API docs, OpenAPI spec, auth docs

    Recommendation

    Make your developer resources (API docs, OpenAPI spec, auth docs, webhooks, MCP server) discoverable by name. Publish them at predictable URLs, list them in llms.txt, and include your product name in page titles and headings so search engines surface them for name-based queries.

  • Brand name discoverabilityPassed

    nytimes.com appears at position #1 in a clean brand-name search for "The New York Times news intelligence" (7 total matches)

    Recommendation

    Make sure a clean search for your brand name returns your own domain in the top results. If it does not, your brand may be too generic, conflict with a more established term, or not yet indexed. Strengthen brand-name search by claiming consistent NAP across listings, earning press mentions that link to the canonical domain, and avoiding redirect chains that mask the apex domain in search results.

  • Sitemap existsPassed

    Valid sitemap found at https://www.nytimes.com/sitemaps/new/news.xml.gz with 383 entries

    Recommendation

    Add a valid XML sitemap at /sitemap.xml listing all indexable URLs. Include lastmod dates and keep it under 50MB.

  • JSON-LD structured dataPartial (50%)

    JSON-LD has Organization type but missing key fields (name, description)

    Recommendation

    Add JSON-LD structured data to your homepage using the identity type that matches your site - SoftwareApplication for products, Organization or LocalBusiness for companies, Person for personal sites, Article for blogs - with name, description, url, and type-appropriate fields (offers, sameAs, author) so AI can parse your identity programmatically.

  • Public API/docs linked from homepagePassed

    Documentation site found at https://developer.nytimes.com

    Recommendation

    Publish API documentation at a discoverable URL (/docs, /api, /developers). Include authentication, endpoints, and example requests.

  • Agent instruction / when-to-useFailed

    No agent instruction file with when-to-use guidance found

    Recommendation

    Tell agents when to reach for you: add a 'when to use this' section to your llms.txt (or a dedicated agent-instructions file) that names your best-fit use cases and how an agent should call you. Be specific about the jobs you are right for - generic marketing copy does not read as guidance.

  • Metadata completenessPassed

    All metadata signals present: canonical URL, lang="en", og:image, og:type

    Recommendation

    Add all four signals to your homepage: , , , and . Agents use these for entity resolution and attribution.

  • Organization schema completenessPartial (50%)

    Organization schema found but missing: contactPoint, address

    Recommendation

    Add Organization JSON-LD that includes both contactPoint (with email/phone and contactType) and address (PostalAddress). This lets AI verify your business legitimacy and answer contact queries.

  • Trust anchor pagesPartial (50%)

    Contact, Privacy pages verified - missing: About

    Recommendation

    Publish real /about, /contact, and /privacy pages with at least 500 characters of content each. These are the pages AI agents check to verify your business is legitimate before recommending you.

  • Page token budgetPassed

    All 1 measured page fit an agent context budget (largest ~1K tokens)

    Recommendation

    Keep each page's extracted text under ~100K characters (~25K tokens) so it fits an agent's context window without truncation. Split oversized reference pages into focused per-topic documents and link them from an index. Check a page with curl -s <url> | wc -c and remember agents read the extracted text, not the raw HTML.

  • Developer portalPassed

    Developer portal found at https://developer.nytimes.com

    Recommendation

    Create a developer portal at /developers with API keys, documentation, quickstart guides, and a sandbox environment.

  • Public API with reachable endpointsPartial (43%)

    API described in documentation at https://developer.nytimes.com but no machine-verifiable API surface confirmed. Best-of-protocols score: 3/7.

    Recommendation

    Expose a public REST or GraphQL API. AI agents need programmatic access - not just a web UI - to integrate with your product.

  • Agent onboarding frictionPassed

    Low friction onboarding verified live: sandbox/test environment, zero-friction self-serve registration at /signup

    Recommendation

    Offer a free tier or trial, self-serve API key generation, and a sandbox environment. Agents can't fill out 'contact sales' forms.

  • API schema complexity analysisFailed

    No API schema detected

    Recommendation

    Make your API spec self-describing: a unique operationId and a description on every operation, typed parameters, and response schemas. For GraphQL, a fully typed schema with a documented cost or rate limit reads best.

  • Function calling compatibilityFailed

    No API spec found - function calling requires discoverable endpoints

    Recommendation

    Ensure API endpoints have unique operation IDs, typed schemas, and descriptions compatible with LLM function-calling formats.

Bonus signals9 positive · +2.1 points
  • NPM/PyPI SDK packagePassed

    NPM package found: @nytimes/react-prosemirror - "<p align="center"> <img src="https://github.com/nytimes/react-prosemirror/raw/main/react-prosemirror-logo.png" alt="React ProseMirror Logo" width="120px" height="120px"/> <br> <em>A fully featured library for safely integrating ProseMirror and React"

    Recommendation

    Publish a JavaScript/TypeScript SDK package on npm so developers can integrate your API programmatically. In package.json set repository to your source repo and homepage to your product domain - these links are how agents confirm the package is your official SDK rather than a third-party tool with a similar name.

  • MCP well-known discoveryPartial (50%)

    MCP server at https://help.nytimes.com/mcp - consider adding /.well-known/mcp for standard discovery

    Recommendation

    Serve your MCP server at /.well-known/mcp, publish a server-card.json at /.well-known/mcp/server-card.json, or reference it in llms.txt so agents can discover it automatically without manual URL input.

  • Sitemap freshness (lastmod)Passed

    100% of 383 sampled sitemap entries carry lastmod; newest is 0 day(s) old

    Recommendation

    Add dates (W3C datetime, e.g. 2026-08-01) to your sitemap entries and update them when content actually changes. Aim for lastmod on at least half your entries with the newest within the last year. Verify with curl https://yourdomain.com/sitemap.xml | grep lastmod.

  • JSON-LD entity linking (sameAs)Passed

    Strong entity linking via sameAs: linkedin.com, wikidata.org, wikipedia.org

    Recommendation

    Add sameAs links in your JSON-LD structured data pointing to your Wikipedia page, Wikidata entry, GitHub org, and social profiles. This helps AI disambiguate your brand from similarly named entities.

  • Accessible document structurePassed

    Server HTML is a well-structured document (main=true, landmarks=4/4, h1=1, maxHeadingSkip=1).

  • Native interactive controlsPassed

    131 native controls, 0 non-native div-soup affordances (100% native).

  • Accessible names on controlsPassed

    118/131 interactive elements have a computable accessible name (90%).

  • Form control labelingPassed

    1/1 form controls have an associated label (100%).

  • Accessibility-tree injection safety (bonus)Passed

    No hidden instruction text detected in accessibility-tree attributes or off-screen content.

Inspect the underlying audit

The complete Ora audit uses evidence from the scan on . After applying changes, run another scan from the homepage to refresh these recommendations.

Sign me up for Vercel product updates and marketing emails.

Unsubscribe anytime. Privacy Notice

Source: Ora API

Snapshot 2026-08-24T04-18-50-102Z