nytimes.com
Agents are likely to struggle
Task
What does nytimes.com do and who is it for? Explain it back to me.
Critical access blockers remain
These checks describe whether an ordinary agent can enter, read, and operate the public site.
- 0 / 2 passed
Agents can reach the site
Crawler access and bot defenses.
- 0 / 1 passed
Core content is available
Useful content remains accessible without a fragile browser-only path.
- 1 / 2 passed
Navigation fails safely
Redirects and missing pages give agents a recoverable path.
- 4 / 4 passed
Controls are understandable
Forms and interactive controls expose usable names and structure.
Evaluated surfaces have material gaps
The public website is always evaluated. Optional surfaces appear when the scan finds positive evidence that they apply.
Public website
Needs work
5 of 15 mature checks passed
API
Blocked
2 of 7 mature checks passed
Authentication
Needs work
2 of 3 mature checks passed
Fix these gaps first
Critical access gaps come first, followed by other applicable readiness gaps.
- 01Critical access
Agent crawler reachability
Verify that major agent User-Agents can reach the homepage. If your WAF or bot rules block them, remove or narrow the blocking rule. Add an allow rule only when your security setup denies them by default.
- 02Critical access
Not blocked by bot detection
Allowlist known AI agent User-Agents (ChatGPT-User, ClaudeBot, Google-Extended, DeepSeekBot) in your WAF or bot-detection rules.
- 03Critical access
Agent-friendly 404s
Return a real HTTP 404 (or 410) status for nonexistent paths - never a 200 with your app shell, which makes agents believe every path exists. For full credit, give the 404 response a short markdown body pointing agents at your sitemap, llms.txt, or docs index. Verify with
curl -s -o /dev/null -w "%{http_code}" https://yourdomain.com/some-path-that-does-not-exist- it must print 404. - 04Critical access
Content without JavaScript
Server-side render your homepage so AI crawlers see meaningful content without JavaScript. Ensure an H1 and 500+ chars of text in raw HTML.
- 05Other readiness checks
OpenAPI spec published
Publish an OpenAPI (Swagger) specification at /openapi.json or /api/openapi.yaml. This is how agents understand your API surface automatically.
Audit the checks behind the score
Applicable evidence is grouped by how it contributes to this preview model. Bonus checks appear only when they add points.
Essential2 of 10 passed · 29.3 / 80 points
- Content without JavaScriptPartial (67%)
5896 chars with H1 but flat heading structure
Recommendation
Server-side render your homepage so AI crawlers see meaningful content without JavaScript. Ensure an H1 and 500+ chars of text in raw HTML.
- Not blocked by bot detectionPartial (50%)
Some agents blocked: GPTBot, ClaudeBot, ChatGPT-User, PerplexityBot
Recommendation
Allowlist known AI agent User-Agents (ChatGPT-User, ClaudeBot, Google-Extended, DeepSeekBot) in your WAF or bot-detection rules.
- Redirect hygienePassed
No meta-refresh stubs, JavaScript-redirect stubs, or cross-domain hops across 1 checked page
Recommendation
Replace meta-refresh and JavaScript-only redirects with real HTTP 301/302 redirects. Non-JS agents never execute
location.hrefor wait for a meta refresh - they see only the stub page. Verify withcurl -sI <url>- you should see a Location header, not a 200 with a near-empty body. - OpenAPI spec publishedFailed
No OpenAPI/Swagger specification found
Recommendation
Publish an OpenAPI (Swagger) specification at /openapi.json or /api/openapi.yaml. This is how agents understand your API surface automatically.
- Markdown content negotiation (acceptmarkdown.com)Failed
Not acceptmarkdown.com compliant: Accept: text/markdown returned text/html; charset=utf-8; Vary header missing Accept (got "accept-encoding, fastly-ssl")
Recommendation
On the responses that serve text/markdown via Accept negotiation, add Accept to the Vary header (Vary: Accept, Accept-Encoding). Without it, CDNs can serve the cached HTML variant to an agent asking for markdown (or vice versa), depending on which variant landed in cache first.
- Agent crawler reachabilityFailed
Some AI crawlers are blocked - ChatGPT-User: blocked, ClaudeBot: blocked, Google-Extended: reachable, ora-agent: reachable, DeepSeekBot: reachable
Recommendation
Verify that major agent User-Agents can reach the homepage. If your WAF or bot rules block them, remove or narrow the blocking rule. Add an allow rule only when your security setup denies them by default.
- OAuth 2.0 supportPassed
OAuth endpoint found at https://help.nytimes.com/oauth/authorize
Recommendation
Implement OAuth 2.0 for API authentication. Publish your authorization server metadata at /.well-known/oauth-authorization-server.
- Scoped permissionsFailed
No declared OAuth scopes, security schemes, or scoped-permission documentation found
Recommendation
Declare scoped API permissions where machines can read them: named OAuth scopes in your OpenAPI security schemes, or scopes_supported in RFC 9728 protected-resource metadata. Prose descriptions of roles help humans, but agents need the machine-readable declaration to request least-privilege access.
- JSON error responsesFailed
API does not return JSON error responses (or no API detected)
Recommendation
Return structured JSON error responses with error codes, messages, and resolution hints. Agents can't parse HTML error pages.
- Agent-friendly 404sPartial (50%)
Nonexistent paths return a real HTTP 404. For full credit, include a short markdown body (site map links, where to look next) so agents can recover.
Recommendation
Return a real HTTP 404 (or 410) status for nonexistent paths - never a 200 with your app shell, which makes agents believe every path exists. For full credit, give the 404 response a short markdown body pointing agents at your sitemap, llms.txt, or docs index. Verify with
curl -s -o /dev/null -w "%{http_code}" https://yourdomain.com/some-path-that-does-not-exist- it must print 404.
Recommended7 of 15 passed · 12.8 / 20 points
- Developer resource discoverabilityPartial (67%)
Agent found developer resources by name including developer portal (3 relevant pages). Not searchable: MCP server. Not found via search: API docs, OpenAPI spec, auth docs
Recommendation
Make your developer resources (API docs, OpenAPI spec, auth docs, webhooks, MCP server) discoverable by name. Publish them at predictable URLs, list them in llms.txt, and include your product name in page titles and headings so search engines surface them for name-based queries.
- Brand name discoverabilityPassed
nytimes.com appears at position #1 in a clean brand-name search for "The New York Times news intelligence" (7 total matches)
Recommendation
Make sure a clean search for your brand name returns your own domain in the top results. If it does not, your brand may be too generic, conflict with a more established term, or not yet indexed. Strengthen brand-name search by claiming consistent NAP across listings, earning press mentions that link to the canonical domain, and avoiding redirect chains that mask the apex domain in search results.
- Sitemap existsPassed
Valid sitemap found at https://www.nytimes.com/sitemaps/new/news.xml.gz with 383 entries
Recommendation
Add a valid XML sitemap at /sitemap.xml listing all indexable URLs. Include lastmod dates and keep it under 50MB.
- JSON-LD structured dataPartial (50%)
JSON-LD has Organization type but missing key fields (name, description)
Recommendation
Add JSON-LD structured data to your homepage using the identity type that matches your site - SoftwareApplication for products, Organization or LocalBusiness for companies, Person for personal sites, Article for blogs - with name, description, url, and type-appropriate fields (offers, sameAs, author) so AI can parse your identity programmatically.
- Public API/docs linked from homepagePassed
Documentation site found at https://developer.nytimes.com
Recommendation
Publish API documentation at a discoverable URL (/docs, /api, /developers). Include authentication, endpoints, and example requests.
- Agent instruction / when-to-useFailed
No agent instruction file with when-to-use guidance found
Recommendation
Tell agents when to reach for you: add a 'when to use this' section to your llms.txt (or a dedicated agent-instructions file) that names your best-fit use cases and how an agent should call you. Be specific about the jobs you are right for - generic marketing copy does not read as guidance.
- Metadata completenessPassed
All metadata signals present: canonical URL, lang="en", og:image, og:type
Recommendation
Add all four signals to your homepage: , , , and . Agents use these for entity resolution and attribution.
- Organization schema completenessPartial (50%)
Organization schema found but missing: contactPoint, address
Recommendation
Add Organization JSON-LD that includes both contactPoint (with email/phone and contactType) and address (PostalAddress). This lets AI verify your business legitimacy and answer contact queries.
- Trust anchor pagesPartial (50%)
Contact, Privacy pages verified - missing: About
Recommendation
Publish real /about, /contact, and /privacy pages with at least 500 characters of content each. These are the pages AI agents check to verify your business is legitimate before recommending you.
- Page token budgetPassed
All 1 measured page fit an agent context budget (largest ~1K tokens)
Recommendation
Keep each page's extracted text under ~100K characters (~25K tokens) so it fits an agent's context window without truncation. Split oversized reference pages into focused per-topic documents and link them from an index. Check a page with
curl -s <url> | wc -cand remember agents read the extracted text, not the raw HTML. - Developer portalPassed
Developer portal found at https://developer.nytimes.com
Recommendation
Create a developer portal at /developers with API keys, documentation, quickstart guides, and a sandbox environment.
- Public API with reachable endpointsPartial (43%)
API described in documentation at https://developer.nytimes.com but no machine-verifiable API surface confirmed. Best-of-protocols score: 3/7.
Recommendation
Expose a public REST or GraphQL API. AI agents need programmatic access - not just a web UI - to integrate with your product.
- Agent onboarding frictionPassed
Low friction onboarding verified live: sandbox/test environment, zero-friction self-serve registration at /signup
Recommendation
Offer a free tier or trial, self-serve API key generation, and a sandbox environment. Agents can't fill out 'contact sales' forms.
- API schema complexity analysisFailed
No API schema detected
Recommendation
Make your API spec self-describing: a unique operationId and a description on every operation, typed parameters, and response schemas. For GraphQL, a fully typed schema with a documented cost or rate limit reads best.
- Function calling compatibilityFailed
No API spec found - function calling requires discoverable endpoints
Recommendation
Ensure API endpoints have unique operation IDs, typed schemas, and descriptions compatible with LLM function-calling formats.
Bonus signals9 positive · +2.1 points
- NPM/PyPI SDK packagePassed
NPM package found: @nytimes/react-prosemirror - "<p align="center"> <img src="https://github.com/nytimes/react-prosemirror/raw/main/react-prosemirror-logo.png" alt="React ProseMirror Logo" width="120px" height="120px"/> <br> <em>A fully featured library for safely integrating ProseMirror and React"
Recommendation
Publish a JavaScript/TypeScript SDK package on npm so developers can integrate your API programmatically. In package.json set
repositoryto your source repo andhomepageto your product domain - these links are how agents confirm the package is your official SDK rather than a third-party tool with a similar name. - MCP well-known discoveryPartial (50%)
MCP server at https://help.nytimes.com/mcp - consider adding /.well-known/mcp for standard discovery
Recommendation
Serve your MCP server at /.well-known/mcp, publish a server-card.json at /.well-known/mcp/server-card.json, or reference it in llms.txt so agents can discover it automatically without manual URL input.
- Sitemap freshness (lastmod)Passed
100% of 383 sampled sitemap entries carry lastmod; newest is 0 day(s) old
Recommendation
Add dates (W3C datetime, e.g. 2026-08-01) to your sitemap entries and update them when content actually changes. Aim for lastmod on at least half your entries with the newest within the last year. Verify with
curl https://yourdomain.com/sitemap.xml | grep lastmod. - JSON-LD entity linking (sameAs)Passed
Strong entity linking via sameAs: linkedin.com, wikidata.org, wikipedia.org
Recommendation
Add sameAs links in your JSON-LD structured data pointing to your Wikipedia page, Wikidata entry, GitHub org, and social profiles. This helps AI disambiguate your brand from similarly named entities.
- Accessible document structurePassed
Server HTML is a well-structured document (main=true, landmarks=4/4, h1=1, maxHeadingSkip=1).
- Native interactive controlsPassed
131 native controls, 0 non-native div-soup affordances (100% native).
- Accessible names on controlsPassed
118/131 interactive elements have a computable accessible name (90%).
- Form control labelingPassed
1/1 form controls have an associated label (100%).
- Accessibility-tree injection safety (bonus)Passed
No hidden instruction text detected in accessibility-tree attributes or off-screen content.
Inspect the underlying audit
The complete Ora audit uses evidence from the scan on . After applying changes, run another scan from the homepage to refresh these recommendations.
Sign me up for Vercel product updates and marketing emails.
Unsubscribe anytime. Privacy Notice
Source: Ora API
Snapshot 2026-08-24T04-18-50-102Z