For AI teams
License internal documentation for AI training
Short answer
SourceX sources internal documentation from U.S. businesses that own them. Each dataset is rights-reviewed, prepared with employee names, internal hostnames and credentials and customer references removed, and approved by the supplying company before delivery under a use-limited license. Teams use it for knowledge assistants, onboarding tools and retrieval systems.
What's included#
Typically wikis, how-to pages, runbooks and process docs.
Why it matters for models#
These records are valuable because they record how a company actually does its work.
Common uses#
- Knowledge assistants
- Onboarding tools
- Retrieval systems
Removed before delivery#
- Employee names
- Internal hostnames and credentials
- Customer references
- Confidential strategy
How a request works#
- Describe the records, volume and uses you need.
- SourceX matches suitable U.S. suppliers and runs a rights review.
- A prepared sample is shared under NDA after supplier approval.
- The license sets allowed uses, term and deletion terms; resale and re-identification are barred.
Frequently asked questions
Who supplies the data?
U.S. businesses that own the records. Every release is approved by the supplying company.
Can we see a sample?
Yes, once an NDA is in place and the supplier approves a prepared sample.
Is pricing published?
No. Terms depend on volume, history, exclusivity and allowed uses, and are agreed per deal.