The registry
Sources
Every external fetch this system makes is described by a row here: what the source is, how we reach it, what its word is worth on the ten-tier priority scale, and whether it is switched on. A source we do not use is listed with the reason, not left out.
- Registered
- 18
- Enabled
- 17
- Disabled
- 1
Tier is the priority scale described in how we research — 1 is legislation and regulators, 10 is unattributed web content — and it travels with every citation. No source is crawled in bulk. Where a directory is the best route to a fact, it identifies the issue and the fact is then checked against the primary source.
01Queried
These are reached when a research run needs them. Which of them a given run actually queried is printed under that answer, with the status of each call.
One named page is retrieved and stored with its hash — the legislation page, the vendor’s own DPA.
Primary source for every compliance FACT (§25 tier 1).
- Kind
- Regulator
- Access
- LIGHT_PAGE_FETCH
- Used for
- Compliance · Verification
- Terms
- Permitted
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domains
- pcpd.org.hk · elegislation.gov.hk · hkma.gov.hk · sfc.hk · digitalpolicy.gov.hk · cac.gov.cn · npc.gov.cn · gov.cn · ppc.go.jp · meti.go.jp · soumu.go.jp · cao.go.jp · digital.go.jp · pipc.go.kr · law.go.kr · msit.go.kr · law.moj.gov.tw · moda.gov.tw · ndc.gov.tw · eur-lex.europa.eu · edpb.europa.eu · digital-strategy.ec.europa.eu · europa.eu · legislation.gov.uk · ico.org.uk · gov.uk
- 02
Vendor legal and technical documentation
vendor_docsTier 2official vendor legal documentationOne named page is retrieved and stored with its hash — the legislation page, the vendor’s own DPA.
A vendor's own published pages are tier-2 evidence about that vendor (§25).
- Kind
- Vendor
- Access
- LIGHT_PAGE_FETCH
- Used for
- Verification · Compliance
- Terms
- Permitted
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
Queried through the source’s own public API.
- Kind
- Repository
- Access
- OFFICIAL_API
- Used for
- Discovery · Verification · Enrichment
- Terms
- Permitted
- Bulk ingestion
- yes
- Crawled
- no
- State
- enabled
- Domains
- github.com · api.github.com
Queried through the source’s own public API.
- Kind
- Model registry
- Access
- OFFICIAL_API
- Used for
- Discovery · Verification · Enrichment
- Terms
- Permitted
- Bulk ingestion
- yes
- Crawled
- no
- State
- enabled
- Domain
- huggingface.co
Queried through the source’s own public API.
- Kind
- Model registry
- Access
- OFFICIAL_API
- Used for
- Discovery · Enrichment
- Terms
- Permitted
- Bulk ingestion
- yes
- Crawled
- no
- State
- enabled
- Domain
- openrouter.ai
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Priority discovery source for governance tooling, but secondary for conclusions (§12).
- Kind
- Compliance directory
- Access
- SEARCH_INDEX
- Used for
- Discovery · Compliance
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- aiactdirectory.com
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Discovers self-hostable candidates; authoritative facts come from GitHub (§14).
- Kind
- Directory
- Access
- SEARCH_INDEX
- Used for
- Discovery
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- enterprisedna.co
Queried through the source’s own public API.
READMEs are read through the GitHub API and diffed; links are candidates, not facts (§17).
- Kind
- Repository
- Access
- OFFICIAL_API
- Used for
- Discovery
- Terms
- Permitted
- Bulk ingestion
- yes
- Crawled
- no
- State
- enabled
- Domain
- raw.githubusercontent.com
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Self-hosting and hardware sizing intelligence (§15).
- Kind
- Directory
- Access
- SEARCH_INDEX
- Used for
- Discovery · Enrichment
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- ownmetal.com
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Use its dimensions as research prompts; never reproduce its scores as ours (§13).
- Kind
- Compliance directory
- Access
- SEARCH_INDEX
- Used for
- Discovery · Compliance
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- trustkit.co
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Editorial discovery signal; bulk copying is prohibited (§11).
- Kind
- Directory
- Access
- SEARCH_INDEX
- Used for
- Discovery
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- futurepedia.io
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Candidate discovery only — descriptions are not authoritative evidence (§10A).
- Kind
- Directory
- Access
- SEARCH_INDEX
- Used for
- Discovery
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- theresanaiforthat.com
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Market intelligence, not canonical truth (§10B).
- Kind
- Directory
- Access
- SEARCH_INDEX
- Used for
- Discovery
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- toolify.ai
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Skill demand and role taxonomy signal (§23).
- Kind
- Job market
- Access
- SEARCH_INDEX
- Used for
- Talent
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domains
- fdejobs.com · fde.jobs · fwddeploy.com
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Understand the FDE profession; do not copy community content wholesale (§21).
- Kind
- Community
- Access
- SEARCH_INDEX
- Used for
- Community
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- rocketlane.com
Reached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
Marketplace competitive intelligence only (§22).
- Kind
- Job market
- Access
- SEARCH_INDEX
- Used for
- Talent
- Terms
- Search only
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
- Domain
- workgenius.com
- 17
Web search
webTier 10unattributed web contentReached through a web-search provider scoped to the source’s domain, not by crawling it. Skipped entirely when no search key is configured.
- Kind
- Web search
- Access
- SEARCH_INDEX
- Used for
- Discovery · Verification
- Terms
- Permitted
- Bulk ingestion
- no
- Crawled
- no
- State
- enabled
02Registered but not queried
Registered and visible, with the reason. A source is not allowed to become live by nobody having looked at it.
- 01
Product Hunt
product_huntTier 7general AI directoryNot queried.
awaiting_commercial_permission
- Kind
- Directory
- Access
- DISABLED
- Used for
- Discovery
- Terms
- Awaiting permission
- Bulk ingestion
- no
- Crawled
- no
- State
- disabled